Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
15+ YOE15+ years in SRE/infrastructure/platform engineering with large-scale fleets; experience in capacity management, orchestration, observability, SLOs/SLIs, incident response, and cross-team architecture.
BeyondTrust: Provides identity security and privileged access management software solutions.
7+ YOERequires 7+ years in SRE, DevOps, or platform engineering, including 2+ years at senior or staff level; expertise in cloud and on-prem infrastructure, Docker, Kubernetes, CI/CD, observability, GitOps, and a systems language.
AWS, Azure, api gateways, service meshes, Terraform, OpenTofu, Ansible, GitOps, Grafana Cloud, Datadog, OpenTelemetry, Docker, Kubernetes, Go, Java, C#, Linux, Windows
Toronto or Pune or Kraków or Stockholm or Gothenburg
$140k-$180k/yrOnsiteFull Time
Tripstack: B2B travel technology provider for virtual interlining and booking.
8+ YOE2+ MgmtRequires 8+ years in production infrastructure, 2+ years leading teams, deep Kubernetes and GCP experience, Terraform, Helm, GitOps, observability, security operations, migration leadership, and English proficiency.
Senior Site Reliability Engineer (Cloud Networking & Infrastructure as Code)
Waterloo or Toronto or Ottawa
$120k-$170k/yrHybridFull Time
Magnet Forensics: Provides software for digital forensics and evidence recovery.
Requires networking or computer science education or equivalent experience, strong AWS networking expertise, multi-account and multi-region architecture experience, IaC, CI/CD, scripting, troubleshooting, and communication skills.
DoceboNASDAQ: DCBO: Cloud platform for enterprise learning management and training delivery.
4+ YOERequires 4–8 years in SRE, DevOps, or production engineering in SaaS, with Linux, containers, cloud infrastructure and networking, infrastructure-as-code, CI/CD, observability, version control, and incident response experience.
Lead Platform Reliability Engineer, Global AI Platform & Solutions
Toronto, Ontario, Canada
$113k-$210k/yrHybridFull Time
ManulifeTSX: MFC: Provides insurance, wealth management, and investment services globally.
5+ YOE5+ years platform/DevOps experience operating cloud-native distributed systems; experience with Azure, Kubernetes, Terraform/Ansible; on-call and incident response; knowledge of LLM/AI infrastructure and backend services.
Electronic ArtsNASDAQ: EA: Develops and publishes interactive entertainment software and video games.
7+ YOERequires 7+ years in production infrastructure, cloud computing, Kubernetes, Docker, Linux, networking, automation, and distributed systems, plus Python, Golang, or Java coding experience.
Senior Software Engineer, Site Reliability Engineering
Toronto or Victoria or Ontario or British Columbia
$180k-$233k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure; extensive AWS and Linux fluency; coding in Python, Go, PHP, JavaScript; expertise in DNS, TLS, HTTP/S, TCP/IP; experience operating and observing distributed microservices; rotating on-call.
Austin or California or New York or Denver or Calgary or Toronto
$221k-$260k/yrRemoteFull Time
IntersectNASDAQ: GOOGL: Develops-located data centers and renewable energy infrastructure.
12+ YOEBachelor's degree in a related engineering field and 12+ years commissioning and reliability experience with high-voltage infrastructure, utility-scale power, or mission-critical industrial facilities.
Epiq: Provides technology-enabled legal services and eDiscovery solutions.
3+ YOERequires 3+ years in infrastructure, platform, or site reliability engineering; production cloud and Kubernetes experience; Terraform, CI/CD, observability, security, incident response, system design, and Python or Go proficiency.
The Ocean Cleanup: Developing advanced technologies to remove plastic from global waters.
5+ YOEBachelor's degree in software engineering, computer science, or similar; 5+ years of software development; production distributed-systems experience; automation, Linux, containers, infrastructure as code, observability, Git, testing, code review, and CI.
Klue: AI platform for competitive and market intelligence.
Senior-level software engineering experience building backend and retrieval systems, working with LLM/agentic systems, Python, cloud infrastructure, search/vector DBs, and production reliability/observability.
CGINYSE: GIB: Provides information technology and business consulting services.
8+ YOERequires 8+ years in SRE, DevOps, cloud, or infrastructure engineering; Azure, Kubernetes, IaC, DevSecOps, CI/CD, observability, automation, troubleshooting, and AI-assisted development experience.
Microsoft Azure, Kubernetes, Microsoft Visual Studio Code, GitLab, JFrog Artifactory, Nexus, Ansible, Infrastructure as Code (IaC), OpenShift Pipelines, Argo CD, SonarQube, Terraform, Bicep, Azure Resource Manager (ARM), Azure Monitor, Prometheus, Grafana, Elastic, GitHub, Azure DevOps, BMAD
Lisbon or Atlanta or London or San Francisco or Santiago or Sydney or Tokyo or Toronto or Australia or Canada or United States
HybridFull Time
PagerDutyNYSE: PD: Platform for real-time incident response and digital operations automation.
5+ YOE5+ years of software engineering experience building production distributed systems and AI systems, with expertise in LLMs, agents, retrieval, cloud infrastructure, containers, Kubernetes, reliability, and evaluation.
Wealthsimple: Provides digital investment, trading, and saving services for Canadians.
8+ YOE8+ years software engineering experience with platform/SRE or infrastructure focus; proven reliability improvements, distributed systems and load testing experience, strong communication and technical influence; familiarity with Kubernetes, Helm, and Argo.
Principal Engineer, Cyber Technology Operations SRE (Global Security)
Toronto or Vancouver
OnsiteFull Time
Royal Bank of CanadaTSX: RY: Provides personal, commercial, and investment banking services worldwide.
10+ YOERequires 10+ years in information security, software engineering, SRE, DevOps, or infrastructure; 5+ years in SRE/reliability work; cloud, Kubernetes, Docker, Helm, CI/CD, security, risk, and compliance expertise.
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
7+ YOE7+ years of software development experience; expertise designing scalable event-driven systems, database fundamentals, Kafka, and reliable platform infrastructure. Golang or Java preferred.