1,001 site reliablity engineer jobs at 554 companies in United States

1mo
Save
Mark Applied
Hide
Member of Technical Staff, Site Reliablity Engineer
San Francisco, California, United States
$200k-$270k/yr HybridFull Time
Vapi
Vapi: Build and deploy AI-powered voice agents via flexible APIs.
Experience running incident command and postmortems, operating SLOs/error budgets, capacity planning and load testing, Kubernetes production ops, KEDA autoscaling, and shipping services in Go or TypeScript.
Go, TypeScript, Bash, Chronosphere, Prometheus, Grafana, Datadog, OpenTelemetry, Kubernetes, EKS, KEDA
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
San Francisco or South San Francisco
$150k/yr OnsiteFull Time
VantageScore
VantageScore: Provides credit scoring and data analytics solutions.
5+ YOEExperienced Site Reliability Engineer with a DevSecOps focus; patch management, vulnerability remediation; AWS and CI/CD, security tooling.
AWS, EC2, ECS, Lambda, EKS, S3, RDS, IAM, VPC, CloudTrail, Config, GuardDuty, GitHub Actions, CodePipeline, Terraform, CloudFormation, AWS CDK, Kubernetes, Snyk, Wiz, Prisma Cloud, Kong, HashiCorp Vault, Secrets Manager, CloudWatch, Datadog, Grafana
2d
Save
Mark Applied
Hide
Site Reliability Engineer
Charlotte, North Carolina, United States
HybridFull Time
Electrolux Group
Electrolux GroupNasdaq Stockholm: ELUX-B: Global manufacturer of household appliances and consumer kitchen equipment.
6+ YOE6+ years in infrastructure/site reliability/cloud engineering; experience with cloud platforms, IaC, CI/CD, observability, troubleshooting, and strong collaboration skills.
Microsoft Azure, AWS, Google Cloud Platform, Akamai CDN, Terraform, CloudFormation, Ansible, Puppet, Chef, Microsoft Azure DevOps, GitHub, Argo CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Atlanta, Georgia, United States
$110k-$209k/yr RemoteFull Time
AbbVie
AbbVieNYSE: ABBV: Develops and sells innovative pharmaceutical and biopharmaceutical medicines.
7+ YOE7+ years in site reliability engineering / information security, cloud platforms (AWS/GCP/Azure), CI/CD, Kubernetes, GitOps, Linux/Windows administration; strong security practices.
AWS, GCP, Azure, Kubernetes, Docker, GitOps, Terraform, Crossplane, ArgoCD, Helm, Prometheus, Grafana, OpenTelemetry, Jenkins, GitHub Actions, Azure DevOps, Python, Go, NoSQL, Relational Databases
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
San Francisco, California, United States
$167k-$226k/yr HybridFull Time
Drata
Drata: Automated security and compliance platform for businesses.
6+ YOE6+ years of Site Reliability Engineering or related cloud engineering experience; strong cloud, Terraform, Docker, Linux; Datadog monitoring; CI/CD with GitHub Actions; incident management; AI experience preferred.
Terraform, Docker, Git, Linux, Datadog, GitHub Actions, Python, Bash, AWS, ECS, Kubernetes, MySQL
2w
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States
$175k-$200k/yr RemoteFull Time
Tern
Tern: Software platform for travel advisors to manage business operations.
Proven production reliability ownership, end-to-end cloud migration experience (GCP preferred), strong observability and monitoring skills, incident leadership, infrastructure-as-code familiarity, and experience coaching engineers.
Ruby on Rails, Hotwire, Postgres, Heroku, Google Cloud Platform, Fivetran, BigQuery, Hex, AppSignal, Bugsnag, Canny, Claude Code
3mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer Market Risk
Houston, Texas, United States
OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years of software engineering and site reliability experience; strong reliability, scalability, security, and architecture knowledge; proficiency in observability and CI/CD; container orchestration and programming skills
Grafana, Dynatrace, Prometheus, Datadog, Splunk, Jenkins, GitLab, Terraform, ECS, Kubernetes, Docker, Python, Java, .NET
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
United States
$175k-$200k/yr RemoteFull Time
Order.co
Order.co: AI-powered platform automating business procurement and payment workflows.
Senior-level SRE with strong focus on reliability, automation, and platform engineering underpinned by software development skills.
Ruby, Ruby on Rails, Linux, AWS, Kubernetes, Terraform, CloudFormation, Datadog, OpenTelemetry, Python, Go, Bash
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Santa Clara, California, United States
$163k-$214k/yr OnsiteFull Time
IonQ
IonQNYSE: IONQ: Develops and sells trapped-ion quantum computers and cloud services.
7+ YOE7+ years production engineering experience; hands-on AWS/GCP reliability, observability and SLO ownership, incident command, resilience testing, and multi-team technical leadership.
AWS, GCP, Amazon Bedrock Agent Core
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with GCP exposure, Terraform/IaC, CI/CD, observability, DevSecOps practices, incident response, and automation for reliability and resiliency.
GCP, Azure, Terraform, Terraform Enterprise, Log Analytics, Dynatrace, Resource Graph, CI/CD, IAM
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States
RemoteFull Time
Supabase
Supabase: Provides cloud-hosted databases and authentication tools for developers.
7+ YOE7+ years in SRE/production engineering, experience shaping SRE practices, defining and operationalizing SLOs/SLIs, incident response and postmortems, software engineering mindset, cloud infra (AWS) and IaC (Pulumi/Terraform/CDK).
Postgres, AWS, Pulumi, Terraform, CDK, Kubernetes, OpenTelemetry, VictoriaMetrics, Grafana, DORA metrics
2d
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Redmond, Washington, United States
$120k-$235k/yr HybridFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
6+ YOE6+ years software/network/systems engineering experience or equivalent degree+experience; experience with large-scale cloud services, observability, automation, and customer-facing reliability work.
Microsoft Fabric, Azure SQL DB, Azure Cosmos DB, Azure PostgreSQL, Azure Data Factory, Azure Synapse Analytics, Azure Service Bus, Azure Event Grid, Power BI, Logic Apps, Jupyter Notebooks
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Cambridge or United States
$76k-$136k/yr HybridFull Time
Akamai
AkamaiNASDAQ: AKAM: Provides content delivery, cybersecurity, and cloud computing services globally.
Bachelor's in Computer Science/Engineering or equivalent; experience in SRE/Software Engineering for large-scale distributed systems; Terraform and IAC experience; familiarity with SaltStack/Ansible/Chef/Puppet; Linux, CI/CD, observability, and participation in on-call rotation.
Terraform, SaltStack, Ansible, Chef, Puppet, Linux, CI/CD, Infrastructure as Code (IAC)
1w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Los Angeles, California, United States
$140k-$199k/yr HybridFull Time
Green Dot
Green DotNYSE: GDOT: Provides mobile banking and payment solutions to consumers and businesses.
7+ YOE7+ years in release/reliability engineering, cloud platform experience (AWS/Azure/GCP), automated deployment and observability proficiency, scripting with PowerShell/Bash/Python, excellent troubleshooting and communication skills.
AWS, Azure, GCP, PowerShell, Bash, Python
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Boston, Massachusetts, United States
$131k-$174k/yr HybridFull Time
Klaviyo
KlaviyoNYSE: KVYO: Customer data platform for e-commerce marketing and automation.
2+ YOEMaster's degree in a relevant engineering field with 2 years engineering experience. Requires building and operating distributed systems, Python/Bash/Shell, Linux/Ubuntu, AWS (CloudFormation, Terraform, EC2, IAM, CloudWatch, CloudTrail, S3, Lambda), Kubernetes, Docker, Grafana, Prometheus, filebeat, and logstash.
Python, Bash, Shell, Linux, Ubuntu, AWS CloudFormation, Terraform, EC2, IAM, CloudWatch, CloudTrail, S3, Lambda, Grafana, Prometheus, filebeat, logstash, Kubernetes, Docker, Covey Scout for Inbound
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
United States
$81k-$187k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
3+ YOE3+ years SRE or reliability engineering experience; design and operate cloud-native infrastructure on Oracle Cloud; incident response, automation, observability; strong Python skills; security clearance may be required.
OCI - Kubernetes Engine, Oracle Cloud Infrastructure (OCI), Python
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Oakland, California, United States
$175k-$210k/yr HybridFull Time
Fivetran
Fivetran: Automates data movement into cloud data warehouses.
5+ YOE5+ years SaaS experience; managed Kubernetes, cloud platforms (AWS/GCP/Azure), Terraform/Ansible/ArgoCD; Python/Shell scripting, Linux admin, PostgreSQL; incident response and reliability engineering experience.
Kubernetes, EKS, AKS, GKE, PostgreSQL, ArgoCD, Terraform, Ansible, Python, Shell, Go, Java, AWS, GCP, Azure, Grafana, Buildkite, Temporal, Pulumi, Linux, VPN, PrivateLink, Private Service Connect (GCP)
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York, New York, United States
$177k-$209k/yr HybridFull Time
Peloton
PelotonNASDAQ: PTON: Sells connected fitness equipment and streaming exercise classes.
Kubernetes expertise; observability/monitoring mindset; CI/CD experience; IaC (Terraform/Pulumi); cloud (AWS); security and reliability focus; programming (Python/Go/Java/C).
Kubernetes, Observability, Jenkins, ArgoCD, Harness, Tekton, Terraform, Pulumi, AWS, Python, Golang, Java, C, Nginx, Ubuntu, Chef