48 platform reliability engineer jobs at 36 companies in Sanborn, NY

2w
Save
Mark Applied
Hide
Lead Platform Reliability Engineer, Global AI Platform & Solutions
Toronto, Ontario, Canada
$113k-$210k/yr HybridFull Time
Manulife
ManulifeTSX: MFC: Provides insurance, wealth management, and investment services globally.
5+ YOE5+ years platform/DevOps experience operating cloud-native distributed systems; experience with Azure, Kubernetes, Terraform/Ansible; on-call and incident response; knowledge of LLM/AI infrastructure and backend services.
Azure, Kubernetes, Terraform, Ansible, Python, Java, Scala, TypeScript, CI/CD, GitOps
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Toronto, Ontario, Canada
HybridFull Time
iManage
iManage: Intelligent document and email management software for professionals.
Experience in reliability engineering with automation, cloud platforms, observability, and on-call responsibility; strong collaboration and architectural skills.
Kubernetes, Docker, Terraform, Prometheus, Grafana, ELK, EFK, CI/CD, Bash, Python, Java
2w
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Buffalo, New York, United States
$140k-$233k/yr OnsiteFull Time
M&T Bank
M&T BankNYSE: MTB: Provides retail, commercial, and wealth management banking services.
7+ YOEExpert in reliability engineering, SLO/SLI frameworks, incident and problem management, observability, automation, cloud platforms, and production operations; 7+ years systems analysis/application development or equivalent.
AWS, Azure, CI/CD, SDLC, SLO/SLI
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Argentina or Toronto or United States
$200k-$230k/yr RemoteFull Time
Domino Data Lab
Domino Data Lab: Enterprise MLOps platform for developing and managing AI models.
Deep SRE/platform engineering experience with Kubernetes, Linux, cloud platforms, observability, Python or Go, incident response, SLO/SLI definition, and mentoring/technical leadership.
Kubernetes, Linux, Python, Go, LLM
3w
Save
Mark Applied
Hide
Staff Software Reliability Engineer - Data Platform
Toronto, Ontario, Canada
$160k-$220k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
5+ YOE5+ years experience building and operating scalable distributed data-platform services; strong OO skills (Java); experience with messaging, data processing, storage systems, reliability, observability, and incident management.
Kinesis, Flink, ElasticSearch, Snowflake, Kafka, Spark, Beam, Databricks, Hadoop, Kubernetes, Mesos, Java, AWS
1mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
San Francisco or Toronto
OnsiteFull Time
Cerebras Systems
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
15+ YOE15+ years in SRE/infrastructure/platform engineering with large-scale fleets; experience in capacity management, orchestration, observability, SLOs/SLIs, incident response, and cross-team architecture.
Wafer-Scale Engine (WSE), Bazel
2w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Canada or United States or Toronto
RemoteFull Time
BeyondTrust
BeyondTrust: Provides identity security and privileged access management software solutions.
7+ YOERequires 7+ years in SRE, DevOps, or platform engineering, including 2+ years at senior or staff level; expertise in cloud and on-prem infrastructure, Docker, Kubernetes, CI/CD, observability, GitOps, and a systems language.
AWS, Azure, api gateways, service meshes, Terraform, OpenTofu, Ansible, GitOps, Grafana Cloud, Datadog, OpenTelemetry, Docker, Kubernetes, Go, Java, C#, Linux, Windows
2mo
Save
Mark Applied
Hide
Head of Platform Engineering, Reliability & Control
Toronto or Vancouver
$170k-$185k/yr HybridFull Time
Connor, Clark & Lunn Financial Group
Connor, Clark & Lunn Financial Group: Independent asset management firm providing traditional and alternative investment strategies.
Senior leadership in platform engineering/SRE with experience scaling shared engineering capabilities and establishing reliability standards.
Datadog, Grafana, Prometheus, Elasticsearch, OpenSearch, CI/CD, Infrastructure as Code, Docker, Kubernetes
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (265177)
Toronto, Ontario, Canada
OnsiteFull Time
Scotiabank
ScotiabankToronto Stock Exchange: BNS: Provides global personal, commercial, and investment banking services.
3+ YOE3+ years experience in ETL platforms and application support, Unix shell scripting, Java, and SQL. Experience with observability tools, incident management, automation, IaC, and on-call rotations. Undergraduate degree in CS or equivalent required.
ETL, iWay, Informatica, Talend, DataStage, Unix Shell Scripting, Java, SQL, Dynatrace, Grafana, Splunk, Infrastructure as Code (IaC)
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, AI Infrastructure
Mississauga or Salt Lake City
$139k-$155k/yr HybridFull Time
PointClickCare
PointClickCare: Develops cloud-based healthcare software for senior care providers.
5+ YOE5+ years SRE/platform experience; strong observability and incident response; Terraform and GitOps; cloud platform experience (Databricks, Azure ML, Kubernetes); platform security and automation skills; strong communication.
Databricks, Azure AI, Azure ML, Kubernetes, Terraform, GitOps, CI/CD
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Observability
United States or Vancouver or Toronto or Buenos Aires or Brazil or Colombia or Mexico
$129k-$304k/yr RemoteFull Time
Chainlink Labs
Chainlink Labs: Building decentralized oracle networks for blockchain smart contracts.
7+ YOE7+ years in DevOps/SRE/platform roles; experience with observability (metrics, logs, traces), Kubernetes, monitoring stacks, real-time systems, and proficiency in one or more languages (C, C++, Java, Python, Go, Perl, Ruby).
OTEL, Prometheus, Grafana, ELK Stack, Splunk, Grafana Stack, AWS, Terraform, Terragrunt, Kubernetes, Calico, ArgoCD, GitHub Actions, Packer, GitOps, C, C++, Java, Python, Go, Perl, Ruby
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE) - Senior Vice President
Mississauga, Ontario, Canada
$145k-$218k/yr HybridFull Time
Citi
CitiNYSE: C: Providing global banking, investment, and wealth management services.
10+ YOE10+ years in system development or platform engineering, strong SRE/DevOps experience (CI/CD, IaC, observability), leadership experience, AI/ML applied to SRE, cloud, containers, microservices; Bachelor's or equivalent.
CI/CD, infrastructure-as-code (IaC), observability (logging, metrics, tracing), containerization, container orchestration, microservices, AI, machine learning
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Toronto, Ontario, Canada
$140k-$182k/yr RemoteFull Time
Movable Ink
Movable Ink: Provides AI-powered content personalization for digital marketing campaigns.
4+ YOE4+ years SRE/Software Engineering experience on cloud platforms, observability and SLO design, IaC and Kubernetes, on-call readiness, strong Linux and programming skills.
AWS, GCP, Prometheus, Thanos, Grafana Alloy, Loki, Tempo, Terraform, Kubernetes, EKS, GKE, NodeJS, Go, Ruby, Python, Shell, Linux
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Toronto or New York City or North America
$144k-$200k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development and distributed systems experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud platform expertise (AWS, GCP, Azure); Linux and networking knowledge; participate in 24/7 on-call.
Python, Go, Argo Workflows, ArgoCD, Kubernetes, AWS, Google Cloud Platform (GCP), Azure, Linux, CI/CD
3d
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Toronto, Ontario, Canada
$140k-$155k/yr RemoteFull Time
Caseware
Caseware: AI-powered audit and financial reporting software platform.
8+ YOERequires 8+ years in SRE, platform engineering, DevOps, or related roles; advanced AWS and Kubernetes expertise; Istio, IaC, CI/CD, observability, TypeScript, Node.js, and incident management experience.
AWS, Amazon EKS, AWS IAM, Amazon VPC, AWS Lambda, Amazon CloudFront, Amazon S3, Kubernetes, Istio, AWS CDK, GitHub Actions, AWS CloudWatch, OpenTelemetry, AWS X-Ray, TypeScript, Node.js, Gateway API, mTLS, Certn.co
20h
Save
Mark Applied
Hide
Data Platform Engineer
United States or Canada or San Francisco or Toronto
$190k-$220k/yr RemoteFull Time
Owner
Owner: AI-powered business management platform for independent restaurants.
5+ YOERequires 5+ years in data or analytics engineering, strong SQL and Python, production Snowflake, dbt and modern orchestration experience, plus reliable systems design and cross-functional communication.
Fivetran, Portable, dbt, Hex, Snowflake, Dagster, Metaplane, Select.dev, PostHog, GA4, Hightouch, Census, Sigma, SQL, Python, Airflow, Databricks
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE) - Senior Vice President
Mississauga, Ontario, Canada
$145k-$218k/yr HybridFull Time
Citi
CitiNYSE: C: A global financial services providing banking and credit services.
10+ YOE10+ years in system development, platform engineering, or related technology; DevOps/SRE, cloud, containers, microservices, IaC, CI/CD, AI/ML, system design, leadership, and communication skills; bachelor's degree required.
DevOps, Site Reliability Engineering (SRE), Infrastructure as Code (IaC), CI/CD, Artificial Intelligence (AI), Machine Learning (ML), cloud platforms, microservices, containerization, container orchestration, logging, metrics, tracing
1mo
Save
Mark Applied
Hide
Sr. Software Engineer, Platform
Ontario or British Columbia or Toronto or Victoria
$180k-$233k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
Strong backend engineering experience building production services and APIs; technical judgment for reliable, secure, auditable integrations; ability to turn ambiguous workflows into incremental designs and collaborate across teams.
Salesforce
1w
Save
Mark Applied
Hide
Senior AI Platform Engineer
Toronto, Ontario, Canada
$126k-$175k/yr HybridFull Time
Epiq
Epiq: Provides technology-enabled legal services and eDiscovery solutions.
7+ YOE7+ years building production software; strong Python and TypeScript experience; production AI/LLM systems, retrieval-augmented generation, distributed systems, reliability, and incident response experience required.
Python, TypeScript, LangGraph, LangChain, MCP, Langfuse, OpenTelemetry, PostgreSQL, React
4w
Save
Mark Applied
Hide
Quality Platform Lead
Welland, Ontario, Canada
$105k-$155k/yr OnsiteFull Time
INNIO
INNIO: Manufactures gas engines for decentralized power and gas compression.
3+ YOEBachelor's in engineering or related field preferred, 3+ years relevant experience (5+ preferred), Lean Six Sigma, RCA and reliability methods, strong data analysis and stakeholder skills.
Microsoft Office, Microsoft Excel, Microsoft PowerPoint, Creo