49 cloud reliability engineer jobs at 33 companies in Sanborn, NY
1mo
Save
Mark Applied
Hide
1mo
Cloud Performance Engineering - Site Reliability Engineer
Toronto, Ontario, Canada
$110k-$125k/yrRemoteFull Time
Smile Digital Health: Provides software for healthcare data management and interoperability.
Expertise with cloud providers (Azure), performance testing, observability, autoscaling, Kafka tuning, IaC (Terraform/Ansible/Chef), and production Linux operations; strong troubleshooting and security/compliance experience.
iManage: Intelligent document and email management software for professionals.
Experience in reliability engineering with automation, cloud platforms, observability, and on-call responsibility; strong collaboration and architectural skills.
7+ YOEExpert in reliability engineering, SLO/SLI frameworks, incident and problem management, observability, automation, cloud platforms, and production operations; 7+ years systems analysis/application development or equivalent.
SimCorp: Provides integrated software solutions for investment and asset managers.
3+ YOE3+ years in Site Reliability, DevOps, or Cloud Engineering; Azure expertise; IaC with Bicep/ARM/Terraform; monitoring/logging tools; IdP onboarding and security; Kubernetes/Docker; ITIL familiarity.
Domino Data Lab: Enterprise MLOps platform for developing and managing AI models.
Deep SRE/platform engineering experience with Kubernetes, Linux, cloud platforms, observability, Python or Go, incident response, SLO/SLI definition, and mentoring/technical leadership.
BeyondTrust: Provides identity security and privileged access management software solutions.
7+ YOERequires 7+ years in SRE, DevOps, or platform engineering, including 2+ years at senior or staff level; expertise in cloud and on-prem infrastructure, Docker, Kubernetes, CI/CD, observability, GitOps, and a systems language.
AWS, Azure, api gateways, service meshes, Terraform, OpenTofu, Ansible, GitOps, Grafana Cloud, Datadog, OpenTelemetry, Docker, Kubernetes, Go, Java, C#, Linux, Windows
Site Reliability Engineer - Cloud & Platform Engineering, Manulife Bank Technology
Waterloo or Toronto
$86k-$136k/yrHybridFull Time
ManulifeTSX: MFC: Provides insurance, wealth management, and investment services globally.
2+ YOERequires 2–5+ years in SRE, DevOps, platform engineering, cloud operations, or related technology; cloud, automation, observability, troubleshooting, communication, and collaboration experience.
Microsoft Azure, AWS, Google Cloud Platform, Python, Bash, PowerShell, New Relic, Grafana, Azure Data Explorer (ADX), Docker, Kubernetes, ITIL, Agile, Microsoft Power BI
Electronic ArtsNASDAQ: EA: Develops and publishes video games and interactive entertainment software.
7+ YOE7+ years experience with cloud, containers, virtualization, Linux, automation and distributed systems; strong scripting/programming in Python, Golang or Java; experience with Terraform, Helm, Chef, Puppet, Packer and Kubernetes.
Staff Site Reliability Engineer - Confluent Incident Management & Reliability
Markham or Toronto
$134k-$248k/yrRemoteFull Time
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
10+ YOE10+ years SRE/incident management experience, cloud experience (AWS, GCP, or Azure), deep incident tooling knowledge (Rootly, PagerDuty), Kubernetes and observability expertise, strong communication and coaching skills.
Royal Bank of CanadaTSX: RY: Provides personal, commercial, and investment banking services worldwide.
5+ YOERequires 5–7 years in site reliability engineering or cloud development, cross-functional leadership, Kubernetes and cloud experience, CI/CD and DevOps knowledge, production support, and proficiency with listed SRE technologies.
Kubernetes, CI/CD, DevOps, Agile Methodology, Python, YAML, Shell, OpenShift, Linux, MongoDB, Dynatrace, Prometheus, PagerDuty, Moog, Splunk, Elastic, Ansible, Grafana, Chaos Engineering, MQ, Kafka, GitHub, Elastic Stack (ELK), Red Hat Ansible, Red Hat OpenShift
Senior Site Reliability Engineer (Cloud Networking & Infrastructure as Code)
Waterloo or Toronto or Ottawa
$120k-$170k/yrHybridFull Time
Magnet Forensics: Provides software for digital forensics and evidence recovery.
Requires networking or computer science education or equivalent experience, strong AWS networking expertise, multi-account and multi-region architecture experience, IaC, CI/CD, scripting, troubleshooting, and communication skills.
Site Reliability Engineer (SRE) - Senior Vice President
Mississauga, Ontario, Canada
$145k-$218k/yrHybridFull Time
CitiNYSE: C: Providing global banking, investment, and wealth management services.
10+ YOE10+ years in system development or platform engineering, strong SRE/DevOps experience (CI/CD, IaC, observability), leadership experience, AI/ML applied to SRE, cloud, containers, microservices; Bachelor's or equivalent.
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development and distributed systems experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud platform expertise (AWS, GCP, Azure); Linux and networking knowledge; participate in 24/7 on-call.
DoceboNASDAQ: DCBO: Cloud platform for enterprise learning management and training delivery.
4+ YOERequires 4–8 years in SRE, DevOps, or production engineering in SaaS, with Linux, containers, cloud infrastructure and networking, infrastructure-as-code, CI/CD, observability, version control, and incident response experience.
Site Reliability Engineer (SRE) - Senior Vice President
Mississauga, Ontario, Canada
$145k-$218k/yrHybridFull Time
CitiNYSE: C: A global financial services providing banking and credit services.
10+ YOE10+ years in system development, platform engineering, or related technology; DevOps/SRE, cloud, containers, microservices, IaC, CI/CD, AI/ML, system design, leadership, and communication skills; bachelor's degree required.
Electronic ArtsNASDAQ: EA: Develops and publishes interactive entertainment software and video games.
7+ YOERequires 7+ years in production infrastructure, cloud computing, Kubernetes, Docker, Linux, networking, automation, and distributed systems, plus Python, Golang, or Java coding experience.
BMOTSX: BMO: Provides personal and commercial banking, investment, and wealth services.
Requires cloud and on-premises deployment and support, deep observability, incident/problem/change management, automation, AIOps, and leadership experience; Agile, financial services, SRE, DevOps, and ITIL knowledge preferred.