49 cloud reliability engineer jobs at 33 companies in Sanborn, NY

1mo
Save
Mark Applied
Hide
Cloud Performance Engineering - Site Reliability Engineer
Toronto, Ontario, Canada
$110k-$125k/yr RemoteFull Time
Smile Digital Health
Smile Digital Health: Provides software for healthcare data management and interoperability.
Expertise with cloud providers (Azure), performance testing, observability, autoscaling, Kafka tuning, IaC (Terraform/Ansible/Chef), and production Linux operations; strong troubleshooting and security/compliance experience.
FHIR, Otel, Grafana, Prometheus, JMeter, Gatling, Azure Load Testing, Terraform, Ansible, Chef, Kubernetes, OpenShift, Azure Monitor, Application Insights, Log Analytics, Kafka, Java
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Toronto, Ontario, Canada
HybridFull Time
iManage
iManage: Intelligent document and email management software for professionals.
Experience in reliability engineering with automation, cloud platforms, observability, and on-call responsibility; strong collaboration and architectural skills.
Kubernetes, Docker, Terraform, Prometheus, Grafana, ELK, EFK, CI/CD, Bash, Python, Java
2w
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Buffalo, New York, United States
$140k-$233k/yr OnsiteFull Time
M&T Bank
M&T BankNYSE: MTB: Provides retail, commercial, and wealth management banking services.
7+ YOEExpert in reliability engineering, SLO/SLI frameworks, incident and problem management, observability, automation, cloud platforms, and production operations; 7+ years systems analysis/application development or equivalent.
AWS, Azure, CI/CD, SDLC, SLO/SLI
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Warsaw or Toronto
HybridFull Time
SimCorp
SimCorp: Provides integrated software solutions for investment and asset managers.
3+ YOE3+ years in Site Reliability, DevOps, or Cloud Engineering; Azure expertise; IaC with Bicep/ARM/Terraform; monitoring/logging tools; IdP onboarding and security; Kubernetes/Docker; ITIL familiarity.
Microsoft Azure, Terraform, Bicep, ARM, Kubernetes, Docker, DataDog, Log Analytics, Application Insights, OpenTelemetry, Playwright, SAML, OAuth, OIDC, KQL, Defender for Cloud
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer
New York City or Toronto
$184k-$240k/yr OnsiteFull Time
Movable Ink
Movable Ink: Provides AI-powered content personalization for digital marketing campaigns.
6+ YOE6+ years SRE/Software Engineering experience designing and operating scalable, multi-cloud distributed systems; expertise with observability, IaC, Kubernetes, and multiple programming languages.
Apache Pulsar, Apache Kafka, Grafana Loki, ScyllaDB, Cassandra, Prometheus, Thanos, Grafana Alloy, Tempo, Terraform, Chef, EKS, GKE, NodeJS, Golang, Ruby, Python, shell
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Argentina or Toronto or United States
$200k-$230k/yr RemoteFull Time
Domino Data Lab
Domino Data Lab: Enterprise MLOps platform for developing and managing AI models.
Deep SRE/platform engineering experience with Kubernetes, Linux, cloud platforms, observability, Python or Go, incident response, SLO/SLI definition, and mentoring/technical leadership.
Kubernetes, Linux, Python, Go, LLM
2w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Canada or United States or Toronto
RemoteFull Time
BeyondTrust
BeyondTrust: Provides identity security and privileged access management software solutions.
7+ YOERequires 7+ years in SRE, DevOps, or platform engineering, including 2+ years at senior or staff level; expertise in cloud and on-prem infrastructure, Docker, Kubernetes, CI/CD, observability, GitOps, and a systems language.
AWS, Azure, api gateways, service meshes, Terraform, OpenTofu, Ansible, GitOps, Grafana Cloud, Datadog, OpenTelemetry, Docker, Kubernetes, Go, Java, C#, Linux, Windows
4d
Save
Mark Applied
Hide
Site Reliability Engineer - Cloud & Platform Engineering, Manulife Bank Technology
Waterloo or Toronto
$86k-$136k/yr HybridFull Time
Manulife
ManulifeTSX: MFC: Provides insurance, wealth management, and investment services globally.
2+ YOERequires 2–5+ years in SRE, DevOps, platform engineering, cloud operations, or related technology; cloud, automation, observability, troubleshooting, communication, and collaboration experience.
Microsoft Azure, AWS, Google Cloud Platform, Python, Bash, PowerShell, New Relic, Grafana, Azure Data Explorer (ADX), Docker, Kubernetes, ITIL, Agile, Microsoft Power BI
3w
Save
Mark Applied
Hide
Site Reliability Engineer III
Vancouver or Toronto or Edmonton or Victoria
$122k-$171k/yr HybridFull Time
Electronic Arts
Electronic ArtsNASDAQ: EA: Develops and publishes video games and interactive entertainment software.
7+ YOE7+ years experience with cloud, containers, virtualization, Linux, automation and distributed systems; strong scripting/programming in Python, Golang or Java; experience with Terraform, Helm, Chef, Puppet, Packer and Kubernetes.
AWS, Kubernetes, Docker, Terraform, Helm, Chef, Puppet, Packer, Python, Golang, Java, Linux
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer - Confluent Incident Management & Reliability
Markham or Toronto
$134k-$248k/yr RemoteFull Time
IBM
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
10+ YOE10+ years SRE/incident management experience, cloud experience (AWS, GCP, or Azure), deep incident tooling knowledge (Rootly, PagerDuty), Kubernetes and observability expertise, strong communication and coaching skills.
Rootly, PagerDuty, Jira, Confluence, Slack, Kubernetes, AWS, GCP, Azure
2w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Toronto, Ontario, Canada
OnsiteFull Time
Royal Bank of Canada
Royal Bank of CanadaTSX: RY: Provides personal, commercial, and investment banking services worldwide.
5+ YOERequires 5–7 years in site reliability engineering or cloud development, cross-functional leadership, Kubernetes and cloud experience, CI/CD and DevOps knowledge, production support, and proficiency with listed SRE technologies.
Kubernetes, CI/CD, DevOps, Agile Methodology, Python, YAML, Shell, OpenShift, Linux, MongoDB, Dynatrace, Prometheus, PagerDuty, Moog, Splunk, Elastic, Ansible, Grafana, Chaos Engineering, MQ, Kafka, GitHub, Elastic Stack (ELK), Red Hat Ansible, Red Hat OpenShift
6d
Save
Mark Applied
Hide
Senior Site Reliability Engineer (Cloud Networking & Infrastructure as Code)
Waterloo or Toronto or Ottawa
$120k-$170k/yr HybridFull Time
Magnet Forensics
Magnet Forensics: Provides software for digital forensics and evidence recovery.
Requires networking or computer science education or equivalent experience, strong AWS networking expertise, multi-account and multi-region architecture experience, IaC, CI/CD, scripting, troubleshooting, and communication skills.
AWS, VPC, Transit Gateway, Route 53, VPN, Terraform, AWS CDK, CloudFormation, CI/CD, Python, Bash, PowerShell, TCP/IP, DNS, ISO 27001, SOC 2, NIST
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE) - Senior Vice President
Mississauga, Ontario, Canada
$145k-$218k/yr HybridFull Time
Citi
CitiNYSE: C: Providing global banking, investment, and wealth management services.
10+ YOE10+ years in system development or platform engineering, strong SRE/DevOps experience (CI/CD, IaC, observability), leadership experience, AI/ML applied to SRE, cloud, containers, microservices; Bachelor's or equivalent.
CI/CD, infrastructure-as-code (IaC), observability (logging, metrics, tracing), containerization, container orchestration, microservices, AI, machine learning
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, AI Infrastructure
Mississauga or Salt Lake City
$139k-$155k/yr HybridFull Time
PointClickCare
PointClickCare: Develops cloud-based healthcare software for senior care providers.
5+ YOE5+ years SRE/platform experience; strong observability and incident response; Terraform and GitOps; cloud platform experience (Databricks, Azure ML, Kubernetes); platform security and automation skills; strong communication.
Databricks, Azure AI, Azure ML, Kubernetes, Terraform, GitOps, CI/CD
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Edmonton or Vancouver or Kitchener-Waterloo or Toronto
$146k-$197k/yr HybridFull Time
Jobber
Jobber: Software for scheduling, invoicing, and managing home service businesses.
Senior cloud infrastructure engineer with AWS, Terraform, continuous deployment, programming, incident management, automation, and collaboration experience; Azure, Kubernetes, security, and observability are advantageous.
AWS, Infrastructure-as-Code, Terraform, CircleCI, Azure, Ruby, Python, Bash, Ruby on Rails, GQL, React, Kubernetes, Cloudflare, DNS, AI
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Toronto or New York City or North America
$144k-$200k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development and distributed systems experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud platform expertise (AWS, GCP, Azure); Linux and networking knowledge; participate in 24/7 on-call.
Python, Go, Argo Workflows, ArgoCD, Kubernetes, AWS, Google Cloud Platform (GCP), Azure, Linux, CI/CD
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Toronto, Ontario, Canada
$103k-$137k/yr HybridFull Time
Docebo
DoceboNASDAQ: DCBO: Cloud platform for enterprise learning management and training delivery.
4+ YOERequires 4–8 years in SRE, DevOps, or production engineering in SaaS, with Linux, containers, cloud infrastructure and networking, infrastructure-as-code, CI/CD, observability, version control, and incident response experience.
Linux, CI/CD, version control
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE) - Senior Vice President
Mississauga, Ontario, Canada
$145k-$218k/yr HybridFull Time
Citi
CitiNYSE: C: A global financial services providing banking and credit services.
10+ YOE10+ years in system development, platform engineering, or related technology; DevOps/SRE, cloud, containers, microservices, IaC, CI/CD, AI/ML, system design, leadership, and communication skills; bachelor's degree required.
DevOps, Site Reliability Engineering (SRE), Infrastructure as Code (IaC), CI/CD, Artificial Intelligence (AI), Machine Learning (ML), cloud platforms, microservices, containerization, container orchestration, logging, metrics, tracing
3w
Save
Mark Applied
Hide
Site Reliability Engineer III
Vancouver or Toronto or Edmonton or Victoria
$122k-$171k/yr HybridFull Time
Electronic Arts
Electronic ArtsNASDAQ: EA: Develops and publishes interactive entertainment software and video games.
7+ YOERequires 7+ years in production infrastructure, cloud computing, Kubernetes, Docker, Linux, networking, automation, and distributed systems, plus Python, Golang, or Java coding experience.
AWS, Kubernetes, Docker, Linux, Terraform, Helm, Chef, Puppet, Packer, Python, Golang, Java
1w
Save
Mark Applied
Hide
Lead, Reliability Engineering
Toronto, Ontario, Canada
$122k-$212k/yr OnsiteFull Time
BMO
BMOTSX: BMO: Provides personal and commercial banking, investment, and wealth services.
Requires cloud and on-premises deployment and support, deep observability, incident/problem/change management, automation, AIOps, and leadership experience; Agile, financial services, SRE, DevOps, and ITIL knowledge preferred.
AWS, Azure, Dynatrace, Ansible, AIOps, NoOps, Site Reliability Engineering (SRE), DevOps, ITIL, Agile