1,050 site reliability engineering jobs at 552 companies in United States
2w
Save
Mark Applied
Hide
2w
Director, Site Reliability Engineering
Denver or United States
$175k-$220k/yrHybridFull Time
VertaforeNYSE: ROP: Provides cloud-based software solutions for the insurance industry.
15+ YOE8+ MgmtBachelor's degree and 15+ years in software engineering/SRE with 8+ years leadership; expertise in CI/CD, observability, incident response, AWS, container orchestration, and operating SaaS products.
Thomson ReutersNASDAQ: TRI: Provides professional software, data, and news services globally.
10+ YOE10+ years in SRE or related tech leadership with experience leading global teams, observability, incident management, automation, and resilience engineering.
LexisNexis Risk SolutionsNYSE: RELX: Provides data and analytics for risk management and compliance.
Lead SRE teams; implement infrastructure as code and DevOps practices; manage production reliability; cloud (AWS/Azure); Kubernetes and Docker; security tooling; incident management; FinOps cost optimization; collaboration with cross-functional teams.
Amazon Web Services, Microsoft Azure, Kubernetes, Docker, GitHub Advanced Security, Qualys, Wiz, Trufflehog
DocuSignNASDAQ: DOCU: Provider of e-signature and intelligent agreement management software.
10+ YOE4+ Mgmt10+ years in Infrastructure/SRE/Software Engineering,4+ years managing engineering teams; experience with automation, incident management, SLOs/Error Budgets, and cloud multi-region architectures.
DocuSignNASDAQ: DOCU: Provides electronic signature and agreement management software solutions.
10+ YOE4+ Mgmt10+ years in Infrastructure/SRE/Software Engineering, 4+ years managing engineering teams, experience with automation, incident management, SLOs/Error Budgets, and modern languages (Go or Python).
Aya Healthcare: Provides healthcare staffing and workforce management software solutions.
10+ YOE4+ Mgmt10+ years in SRE/DevOps/Platform roles, 4+ years people management, deep Azure/AKS and observability experience (Datadog), incident command, AIOps/automation experience, and executive communication skills.
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
3+ Mgmt3+ years technical leadership experience; experience with cloud-native architectures, Kubernetes, Terraform, CI/CD, observability platforms; strong software development and automation background; US Person status required.
Amazon Web Services (AWS), Kubernetes, Terraform, Grafana, Splunk, APM, CI/CD
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
10+ YOE10+ years experience leading SRE teams with capacity planning, incident management, automation, and cross-functional collaboration for reliable scalable infrastructure.
Claritas Rx: Provides data analytics for specialty biopharmaceutical product performance.
7+ YOE3+ Mgmt7+ years SRE/DevOps or infrastructure engineering experience with 3+ years managing teams; deep AWS, IaC, SLOs, incident management, CI/CD, and compliance (HIPAA/SOC2/HITRUST) experience required.
United States or United Kingdom or Hong Kong or New Zealand or North America
$187k-$243k/yrRemoteFull Time
Counterpart HealthNASDAQ: CLOV: AI-powered physician enablement platform for value-based care.
10+ YOE6+ Mgmt6+ years managing SRE teams and 10+ years SRE/infrastructure experience; deep experience with Kubernetes, GCP, Terraform, Helm, ArgoCD, PostgreSQL, Prometheus/Grafana; strong Python/Go skills; CI/CD (GitHub Actions); FinOps and platform engineering experience.
Electrolux GroupNasdaq Stockholm: ELUX-B: Global manufacturer of household appliances and consumer kitchen equipment.
6+ YOE6+ years in infrastructure/site reliability/cloud engineering; experience with cloud platforms, IaC, CI/CD, observability, troubleshooting, and strong collaboration skills.
Microsoft Azure, AWS, Google Cloud Platform, Akamai CDN, Terraform, CloudFormation, Ansible, Puppet, Chef, Microsoft Azure DevOps, GitHub, Argo CD
O.C. Tanner: Provides employee recognition software and corporate award manufacturing services.
5+ YOE2+ Mgmt5+ years in SRE/DevOps or platform engineering with 2+ years in technical leadership; hands-on AWS and Kubernetes experience; observability (OpenTelemetry, Datadog, Coralogix); incident management and SLO/SLI expertise.
FlywireNASDAQ: FLYW: Platform for processing complex global payments across specialized industries.
5+ YOE2+ Mgmt5+ years SRE experience, 2+ years managing SRE teams; programming experience; familiarity with containers, cloud, CI/CD, and testing methodologies; strong communication and incident response skills.
Horizon3.ai: Autonomous penetration testing platform for continuous security assessment.
Proven experience building and leading SRE teams, defining incident management and on-call programs, SLO/SLA and runbooks, strong observability and cloud (AWS/GCP/Azure) knowledge, hiring and people management skills.
United States or United Kingdom or Hong Kong or New Zealand
$187k-$243k/yrRemoteFull Time
Clover HealthNasdaq: CLOV: Provide Medicare Advantage plans and AI-powered clinical decision tools.
10+ YOE6+ Mgmt6+ years managing SRE teams and 10+ years hands-on SRE/infrastructure experience; strong with Kubernetes, GCP (GKE, Cloud SQL, Pub/Sub, GCS), Terraform, Helm, ArgoCD, PostgreSQL, Prometheus/Grafana; Python or Go; GitHub Actions; leadership across time zones.
London Stock Exchange GroupLondon Stock Exchange: LSEG: Provides financial market infrastructure and global data analytics services.
10+ YOESenior SRE/Platform Engineer with 10+ years of hands-on experience in Azure, Kubernetes, observability, and security; strong leadership and collaboration.