54 site reliablity engineer jobs at 38 companies in Lockport, NY
3mo
Save
Mark Applied
Hide
3mo
Senior Site Reliability Engineer
Warsaw or Toronto
HybridFull Time
SimCorp: Provides integrated software solutions for investment and asset managers.
3+ YOE3+ years in Site Reliability, DevOps, or Cloud Engineering; Azure expertise; IaC with Bicep/ARM/Terraform; monitoring/logging tools; IdP onboarding and security; Kubernetes/Docker; ITIL familiarity.
iManage: Intelligent document and email management software for professionals.
Experience in reliability engineering with automation, cloud platforms, observability, and on-call responsibility; strong collaboration and architectural skills.
7+ YOEExpert in reliability engineering, SLO/SLI frameworks, incident and problem management, observability, automation, cloud platforms, and production operations; 7+ years systems analysis/application development or equivalent.
Domino Data Lab: Enterprise MLOps platform for developing and managing AI models.
Deep SRE/platform engineering experience, hands-on operational ownership, fluency with Kubernetes and Linux, strong software engineering skills in Python or Go, observability and SLO/SLI expertise, mentoring and cross-team influence.
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
15+ YOE15+ years in SRE/infrastructure/platform engineering with large-scale fleets; experience in capacity management, orchestration, observability, SLOs/SLIs, incident response, and cross-team architecture.
BeyondTrust: Provides identity security and privileged access management software solutions.
7+ YOERequires 7+ years in SRE, DevOps, or platform engineering, including 2+ years at senior or staff level; expertise in cloud and on-prem infrastructure, Docker, Kubernetes, CI/CD, observability, GitOps, and a systems language.
AWS, Azure, api gateways, service meshes, Terraform, OpenTofu, Ansible, GitOps, Grafana Cloud, Datadog, OpenTelemetry, Docker, Kubernetes, Go, Java, C#, Linux, Windows
Royal Bank of CanadaTSX: RY: Provides personal, commercial, and investment banking services worldwide.
5+ YOERequires 5–7 years in site reliability engineering or cloud development, cross-functional leadership, Kubernetes and cloud experience, CI/CD and DevOps knowledge, production support, and proficiency with listed SRE technologies.
Kubernetes, CI/CD, DevOps, Agile Methodology, Python, YAML, Shell, OpenShift, Linux, MongoDB, Dynatrace, Prometheus, PagerDuty, Moog, Splunk, Elastic, Ansible, Grafana, Chaos Engineering, MQ, Kafka, GitHub, Elastic Stack (ELK), Red Hat Ansible, Red Hat OpenShift
5+ YOE5+ years SRE/DevOps experience, Bachelor’s in CS/Engineering or equivalent, strong AWS and IaC (Terraform/CDK/CloudFormation) skills, CI/CD and container expertise, Python/Bash scripting, monitoring and SRE practices, on-call experience.
Boson AI: Develops generative AI and large-scale audio foundation models.
4+ YOE4+ years in SRE or related production-operations; deep expertise in networking, cluster scheduling, storage, GPU systems or AI infra; strong Linux and scripting; production availability, performance, security, and automation focus.
Dominion Dynamics: Developing autonomous defense platforms and Arctic sensing technology.
Senior-level experience running Linux on embedded/edge devices, DDIL design, ARM/Jetson/embedded toolchains, networking fundamentals, secure boot/attestation and provisioning pipelines; ability to define reliability and observability for constrained field devices.
Linux, SELinux, AuraNet SDK, ROS, TPM, Jetson, ARM
Electronic ArtsNASDAQ: EA: Develops and publishes video games and interactive entertainment software.
7+ YOE7+ years experience with cloud, containers, virtualization, Linux, automation and distributed systems; strong scripting/programming in Python, Golang or Java; experience with Terraform, Helm, Chef, Puppet, Packer and Kubernetes.
Toronto or Pune or Kraków or Stockholm or Gothenburg
$140k-$180k/yrOnsiteFull Time
Tripstack: B2B travel technology provider for virtual interlining and booking.
8+ YOE2+ MgmtRequires 8+ years in production infrastructure, 2+ years leading teams, deep Kubernetes and GCP experience, Terraform, Helm, GitOps, observability, security operations, migration leadership, and English proficiency.
ScotiabankToronto Stock Exchange: BNS: Provides global personal, commercial, and investment banking services.
3+ YOE3+ years experience in ETL platforms and application support, Unix shell scripting, Java, and SQL. Experience with observability tools, incident management, automation, IaC, and on-call rotations. Undergraduate degree in CS or equivalent required.
United States or Vancouver or Toronto or Buenos Aires or Brazil or Colombia or Mexico
$129k-$304k/yrRemoteFull Time
Chainlink Labs: Building decentralized oracle networks for blockchain smart contracts.
7+ YOE7+ years in DevOps/SRE/platform roles; experience with observability (metrics, logs, traces), Kubernetes, monitoring stacks, real-time systems, and proficiency in one or more languages (C, C++, Java, Python, Go, Perl, Ruby).
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development and distributed systems experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud platform expertise (AWS, GCP, Azure); Linux and networking knowledge; participate in 24/7 on-call.
Cloud Performance Engineering - Site Reliability Engineer
Toronto, Ontario, Canada
$110k-$125k/yrRemoteFull Time
Smile Digital Health: Provides software for healthcare data management and interoperability.
Expertise with cloud providers (Azure), performance testing, observability, autoscaling, Kafka tuning, IaC (Terraform/Ansible/Chef), and production Linux operations; strong troubleshooting and security/compliance experience.
Caseware: AI-powered audit and financial reporting software platform.
8+ YOERequires 8+ years in SRE, platform engineering, DevOps, or related roles; advanced AWS and Kubernetes expertise; Istio, IaC, CI/CD, observability, TypeScript, Node.js, and incident management experience.