38 platform reliability engineer jobs at 28 companies in Newfane, NY

1w
Save
Mark Applied
Hide
Lead Platform Reliability Engineer, Global AI Platform & Solutions
Toronto, Ontario, Canada
$113k-$210k/yr HybridFull Time
Manulife
ManulifeTSX: MFC: Provides insurance, wealth management, and investment services globally.
5+ YOE5+ years platform/DevOps experience operating cloud-native distributed systems; experience with Azure, Kubernetes, Terraform/Ansible; on-call and incident response; knowledge of LLM/AI infrastructure and backend services.
Azure, Kubernetes, Terraform, Ansible, Python, Java, Scala, TypeScript, CI/CD, GitOps
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Node Platform
United States or Vancouver or Toronto or Argentina or Mexico or Colombia or Brazil
$129k-$304k/yr RemoteFull Time
Chainlink Labs
Chainlink Labs: Building decentralized oracle networks for blockchain smart contracts.
6+ YOE6+ years SRE/platform experience, production Kubernetes scaling, Terraform/Crossplane, GitOps (ArgoCD/Flux), CI/CD reliability, AWS, automation-first mindset; Go preferred; CKA is desirable.
Chainlink Runtime Environment (CRE), Kubernetes, Kubernetes Operators, StatefulSets, HPA, VPA, KEDA, Terraform, Crossplane, ArgoCD, Flux, GitOps, CI/CD, AWS, Go
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Toronto, Ontario, Canada
HybridFull Time
iManage
iManage: Intelligent document and email management software for professionals.
Experience in reliability engineering with automation, cloud platforms, observability, and on-call responsibility; strong collaboration and architectural skills.
Kubernetes, Docker, Terraform, Prometheus, Grafana, ELK, EFK, CI/CD, Bash, Python, Java
1w
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Buffalo, New York, United States
$140k-$233k/yr OnsiteFull Time
M&T Bank
M&T BankNYSE: MTB: Provides retail, commercial, and wealth management banking services.
7+ YOEExpert in reliability engineering, SLO/SLI frameworks, incident and problem management, observability, automation, cloud platforms, and production operations; 7+ years systems analysis/application development or equivalent.
AWS, Azure, CI/CD, SDLC, SLO/SLI
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Argentina or Toronto or United States
$200k-$230k/yr RemoteFull Time
Domino Data Lab
Domino Data Lab: Enterprise MLOps platform for developing and managing AI models.
Deep SRE/platform engineering experience with Kubernetes, Linux, cloud platforms, observability, Python or Go, incident response, SLO/SLI definition, and mentoring/technical leadership.
Kubernetes, Linux, Python, Go, LLM
2w
Save
Mark Applied
Hide
Staff Software Reliability Engineer - Data Platform
Toronto, Ontario, Canada
$160k-$220k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
5+ YOE5+ years experience building and operating scalable distributed data-platform services; strong OO skills (Java); experience with messaging, data processing, storage systems, reliability, observability, and incident management.
Kinesis, Flink, ElasticSearch, Snowflake, Kafka, Spark, Beam, Databricks, Hadoop, Kubernetes, Mesos, Java, AWS
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Toronto, Ontario, Canada
OnsiteFull Time
Royal Bank of Canada
Royal Bank of CanadaTSX: RY: Provides personal, commercial, and investment banking services worldwide.
5+ YOE5+ years SRE/Production/Platform engineering experience, strong automation and scripting (Bash, Python, PowerShell), Ansible expertise, observability tooling experience, and SLO/SLI practice.
Elasticsearch, Ansible, GitHub Actions, Dynatrace, PagerDuty, Moogsoft, Kubernetes, OpenShift, Kafka, Bash, Python, PowerShell, OpenTelemetry, Prometheus, Grafana, Splunk, Catchpoint, Azure Automation, Jenkins, Artifactory, Vault, Docker
1mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
San Francisco or Toronto
OnsiteFull Time
Cerebras Systems
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
15+ YOE15+ years in SRE/infrastructure/platform engineering with large-scale fleets; experience in capacity management, orchestration, observability, SLOs/SLIs, incident response, and cross-team architecture.
Wafer-Scale Engine (WSE), Bazel
2mo
Save
Mark Applied
Hide
Head of Platform Engineering, Reliability & Control
Toronto or Vancouver
$170k-$185k/yr HybridFull Time
Connor, Clark & Lunn Financial Group
Connor, Clark & Lunn Financial Group: Independent asset management firm providing traditional and alternative investment strategies.
Senior leadership in platform engineering/SRE with experience scaling shared engineering capabilities and establishing reliability standards.
Datadog, Grafana, Prometheus, Elasticsearch, OpenSearch, CI/CD, Infrastructure as Code, Docker, Kubernetes
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (265177)
Toronto, Ontario, Canada
OnsiteFull Time
Scotiabank
ScotiabankToronto Stock Exchange: BNS: Provides global personal, commercial, and investment banking services.
3+ YOE3+ years experience in ETL platforms and application support, Unix shell scripting, Java, and SQL. Experience with observability tools, incident management, automation, IaC, and on-call rotations. Undergraduate degree in CS or equivalent required.
ETL, iWay, Informatica, Talend, DataStage, Unix Shell Scripting, Java, SQL, Dynatrace, Grafana, Splunk, Infrastructure as Code (IaC)
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Toronto, Ontario, Canada
$140k-$182k/yr RemoteFull Time
Movable Ink
Movable Ink: Provides AI-powered content personalization for digital marketing campaigns.
4+ YOE4+ years SRE/Software Engineering experience on cloud platforms, observability and SLO design, IaC and Kubernetes, on-call readiness, strong Linux and programming skills.
AWS, GCP, Prometheus, Thanos, Grafana Alloy, Loki, Tempo, Terraform, Kubernetes, EKS, GKE, NodeJS, Go, Ruby, Python, Shell, Linux
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Toronto or New York City or North America
$144k-$200k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development and distributed systems experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud platform expertise (AWS, GCP, Azure); Linux and networking knowledge; participate in 24/7 on-call.
Python, Go, Argo Workflows, ArgoCD, Kubernetes, AWS, Google Cloud Platform (GCP), Azure, Linux, CI/CD
1mo
Save
Mark Applied
Hide
Sr. Software Engineer, Platform
Ontario or British Columbia or Toronto or Victoria
$180k-$233k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
Strong backend engineering experience building production services and APIs; technical judgment for reliable, secure, auditable integrations; ability to turn ambiguous workflows into incremental designs and collaborate across teams.
Salesforce
1w
Save
Mark Applied
Hide
API Platform Operations Engineer
Toronto or Waterloo
$65k-$105k/yr HybridFull Time
Sun Life
Sun LifeToronto Stock Exchange: SLF: Offers global insurance, wealth management, and asset management services.
3+ YOEBachelor’s degree in Computer Science, Engineering, or related area; 3–6 years in related IT, including 2+ years in application support and operations and 1–2 years of Java development. Requires Reliability Security Clearance.
Java, Docker, Kubernetes, REST API, Open API Spec, RAML, AWS, Amazon EKS, Amazon EC2, Kafka, Confluent, OAuth, OpenID, AWS Lambda, AWS API Gateway, Kustomize, Helm Charts, Jenkins, Ansible, Continuous Delivery Director (CDD), Spring Boot
1d
Save
Mark Applied
Hide
Senior AI Platform Engineer
Toronto, Ontario, Canada
$126k-$175k/yr HybridFull Time
Epiq
Epiq: Provides technology-enabled legal services and eDiscovery solutions.
7+ YOE7+ years building production software; strong Python and TypeScript experience; production AI/LLM systems, retrieval-augmented generation, distributed systems, reliability, and incident response experience required.
Python, TypeScript, LangGraph, LangChain, MCP, Langfuse, OpenTelemetry, PostgreSQL, React
1w
Save
Mark Applied
Hide
API Platform Operations Engineer
Toronto or Waterloo or Canada
$65k-$105k/yr HybridFull Time
Sun Life
Sun LifeToronto Stock Exchange: SLF: Provides insurance, retirement, and asset management services globally.
3+ YOEBachelor’s degree and 3–6 years in IT, including API operations and support, Java development, CI/CD, DevOps, Docker, Kubernetes, and REST API standards; Reliability Security Clearance required.
Java, Kafka, Docker, Kubernetes, Open API Spec, RAML, AWS, EKS, EC2, Confluent, OAuth, OpenID, AWS Lambda, AWS API Gateway, Kustomize, Helm Charts, Jenkins, Ansible, Continuous Delivery Director (CDD), Spring Boot
2w
Save
Mark Applied
Hide
Quality Platform Lead
Welland, Ontario, Canada
$105k-$155k/yr OnsiteFull Time
INNIO
INNIO: Manufactures gas engines for decentralized power and gas compression.
3+ YOEBachelor's in engineering or related field preferred, 3+ years relevant experience (5+ preferred), Lean Six Sigma, RCA and reliability methods, strong data analysis and stakeholder skills.
Microsoft Office, Microsoft Excel, Microsoft PowerPoint, Creo
2w
Save
Mark Applied
Hide
Lead, Site Reliability Engineering (Application Support)
Toronto, Ontario, Canada
$86k-$130k/yr HybridFull Time
Oxford Properties
Oxford Properties: Global real estate investment, development, and management.
5+ YOE5+ years SRE/Platform/DevOps experience with strong Azure, incident response, CI/CD (GitHub Actions), observability (Datadog/Azure Monitor/Log Analytics), container and networking knowledge, and scripting (PowerShell/Bash/Python).
Azure Container Apps, Azure Active Directory (Entra ID), Key Vault, Storage Accounts, Azure SQL, API Management (APIM), Azure Functions, GitHub Actions, Datadog, Azure Monitor, Log Analytics, PowerShell, Bash, Azure CLI, Python, Kubernetes
1mo
Save
Mark Applied
Hide
Staff Software Engineer, Storage Platform
Bellevue or Menlo Park or Toronto
$230k-$270k/yr HybridFull Time
Robinhood
RobinhoodNASDAQ: HOOD: Provides a commission-free platform for investing and financial services.
5+ YOEDeep expertise in PostgreSQL/Aurora, distributed systems (sharding, replication, transactions), proficiency in Go or Rust, experience with Kubernetes and AWS services, and strong reliability/performance engineering skills.
PostgreSQL, Aurora PostgreSQL, Go, Rust, Kubernetes, RDS, DynamoDB
1mo
Save
Mark Applied
Hide
Copy of Senior Software Engineer, Billing Platform
Toronto, Ontario, Canada
$200k-$295k/yr HybridFull Time
Sentry
Sentry: Developer platform for error tracking and performance monitoring.
Senior-level engineer with experience building and operating high-accuracy distributed systems, mentoring engineers, and ensuring reliability and financial correctness at scale.