1,011 site reliable engineer jobs at 534 companies in United States

2mo
Save
Mark Applied
Hide
Site Reliability Engineer
San Francisco or South San Francisco
$150k/yr OnsiteFull Time
VantageScore
VantageScore: Provides credit scoring and data analytics solutions.
5+ YOEExperienced Site Reliability Engineer with a DevSecOps focus; patch management, vulnerability remediation; AWS and CI/CD, security tooling.
AWS, EC2, ECS, Lambda, EKS, S3, RDS, IAM, VPC, CloudTrail, Config, GuardDuty, GitHub Actions, CodePipeline, Terraform, CloudFormation, AWS CDK, Kubernetes, Snyk, Wiz, Prisma Cloud, Kong, HashiCorp Vault, Secrets Manager, CloudWatch, Datadog, Grafana
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Charlotte, North Carolina, United States
HybridFull Time
Electrolux Group
Electrolux GroupNasdaq Stockholm: ELUX-B: Global manufacturer of household appliances and consumer kitchen equipment.
6+ YOE6+ years in infrastructure/site reliability/cloud engineering; experience with cloud platforms, IaC, CI/CD, observability, troubleshooting, and strong collaboration skills.
Microsoft Azure, AWS, Google Cloud Platform, Akamai CDN, Terraform, CloudFormation, Ansible, Puppet, Chef, Microsoft Azure DevOps, GitHub, Argo CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
San Francisco, California, United States
$167k-$226k/yr HybridFull Time
Drata
Drata: Automated security and compliance platform for businesses.
6+ YOE6+ years of Site Reliability Engineering or related cloud engineering experience; strong cloud, Terraform, Docker, Linux; Datadog monitoring; CI/CD with GitHub Actions; incident management; AI experience preferred.
Terraform, Docker, Git, Linux, Datadog, GitHub Actions, Python, Bash, AWS, ECS, Kubernetes, MySQL
3w
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States
$175k-$200k/yr RemoteFull Time
Tern
Tern: Software platform for travel advisors to manage business operations.
Proven production reliability ownership, end-to-end cloud migration experience (GCP preferred), strong observability and monitoring skills, incident leadership, infrastructure-as-code familiarity, and experience coaching engineers.
Ruby on Rails, Hotwire, Postgres, Heroku, Google Cloud Platform, Fivetran, BigQuery, Hex, AppSignal, Bugsnag, Canny, Claude Code
3mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer Market Risk
Houston, Texas, United States
OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years of software engineering and site reliability experience; strong reliability, scalability, security, and architecture knowledge; proficiency in observability and CI/CD; container orchestration and programming skills
Grafana, Dynatrace, Prometheus, Datadog, Splunk, Jenkins, GitLab, Terraform, ECS, Kubernetes, Docker, Python, Java, .NET
1w
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Nashville, Tennessee, United States
$85k-$210k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
3+ YOEExperience in site reliability, Windows/Linux administration, scripting/automation, cloud infrastructure, patching and incident response; 3+ years relevant experience; strong communication and documentation skills.
PowerShell, Bash, Python, Ansible, Chef, Oracle Cloud Infrastructure, Citrix, Security Technical Implementation Guides (STIG)
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
United States
$175k-$200k/yr RemoteFull Time
Order.co
Order.co: AI-powered platform automating business procurement and payment workflows.
Senior-level SRE with strong focus on reliability, automation, and platform engineering underpinned by software development skills.
Ruby, Ruby on Rails, Linux, AWS, Kubernetes, Terraform, CloudFormation, Datadog, OpenTelemetry, Python, Go, Bash
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with GCP exposure, Terraform/IaC, CI/CD, observability, DevSecOps practices, incident response, and automation for reliability and resiliency.
GCP, Azure, Terraform, Terraform Enterprise, Log Analytics, Dynatrace, Resource Graph, CI/CD, IAM
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Santa Clara, California, United States
$163k-$214k/yr OnsiteFull Time
IonQ
IonQNYSE: IONQ: Develops and sells trapped-ion quantum computers and cloud services.
7+ YOE7+ years production engineering experience; hands-on AWS/GCP reliability, observability and SLO ownership, incident command, resilience testing, and multi-team technical leadership.
AWS, GCP, Amazon Bedrock Agent Core
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States
RemoteFull Time
Supabase
Supabase: Provides cloud-hosted databases and authentication tools for developers.
7+ YOE7+ years in SRE/production engineering, experience shaping SRE practices, defining and operationalizing SLOs/SLIs, incident response and postmortems, software engineering mindset, cloud infra (AWS) and IaC (Pulumi/Terraform/CDK).
Postgres, AWS, Pulumi, Terraform, CDK, Kubernetes, OpenTelemetry, VictoriaMetrics, Grafana, DORA metrics
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Redmond, Washington, United States
$120k-$235k/yr HybridFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
6+ YOE6+ years software/network/systems engineering experience or equivalent degree+experience; experience with large-scale cloud services, observability, automation, and customer-facing reliability work.
Microsoft Fabric, Azure SQL DB, Azure Cosmos DB, Azure PostgreSQL, Azure Data Factory, Azure Synapse Analytics, Azure Service Bus, Azure Event Grid, Power BI, Logic Apps, Jupyter Notebooks
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Cambridge or United States
$76k-$136k/yr HybridFull Time
Akamai
AkamaiNASDAQ: AKAM: Provides content delivery, cybersecurity, and cloud computing services globally.
Bachelor's in Computer Science/Engineering or equivalent; experience in SRE/Software Engineering for large-scale distributed systems; Terraform and IAC experience; familiarity with SaltStack/Ansible/Chef/Puppet; Linux, CI/CD, observability, and participation in on-call rotation.
Terraform, SaltStack, Ansible, Chef, Puppet, Linux, CI/CD, Infrastructure as Code (IAC)
2w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Los Angeles, California, United States
$140k-$199k/yr HybridFull Time
Green Dot
Green DotNYSE: GDOT: Provides mobile banking and payment solutions to consumers and businesses.
7+ YOE7+ years in release/reliability engineering, cloud platform experience (AWS/Azure/GCP), automated deployment and observability proficiency, scripting with PowerShell/Bash/Python, excellent troubleshooting and communication skills.
AWS, Azure, GCP, PowerShell, Bash, Python
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Oakland, California, United States
$175k-$210k/yr HybridFull Time
Fivetran
Fivetran: Automates data movement into cloud data warehouses.
5+ YOE5+ years SaaS experience; managed Kubernetes, cloud platforms (AWS/GCP/Azure), Terraform/Ansible/ArgoCD; Python/Shell scripting, Linux admin, PostgreSQL; incident response and reliability engineering experience.
Kubernetes, EKS, AKS, GKE, PostgreSQL, ArgoCD, Terraform, Ansible, Python, Shell, Go, Java, AWS, GCP, Azure, Grafana, Buildkite, Temporal, Pulumi, Linux, VPN, PrivateLink, Private Service Connect (GCP)
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York, New York, United States
$177k-$209k/yr HybridFull Time
Peloton
PelotonNASDAQ: PTON: Sells connected fitness equipment and streaming exercise classes.
Kubernetes expertise; observability/monitoring mindset; CI/CD experience; IaC (Terraform/Pulumi); cloud (AWS); security and reliability focus; programming (Python/Go/Java/C).
Kubernetes, Observability, Jenkins, ArgoCD, Harness, Tekton, Terraform, Pulumi, AWS, Python, Golang, Java, C, Nginx, Ubuntu, Chef
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
$151k-$191k/yr HybridFull Time
Alloy
Alloy: Identity and fraud decisioning platform for financial institutions.
5+ YOE5+ years in infrastructure/SRE or software engineering; experience with Kubernetes, Terraform, Docker, observability tools; coding in Python/Go/JavaScript; on-call experience.
Kubernetes, Docker, Terraform, Datadog, CloudWatch, ELK, EFK, Python, Go, JavaScript
2w
Save
Mark Applied
Hide
Site Reliability Engineer
United States
$150k-$200k/yr RemoteFull Time
Runpod
Runpod: Cloud platform for AI training and model deployment.
5+ YOE5+ years SRE or production engineering experience; strong Linux, networking, container, distributed systems, SLI/SLO, incident response, and scripting skills.
Prometheus, Grafana, Python, Go, Bash, Linux, Slack
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
United States
RemoteFull Time
CertifyOS
CertifyOS: Automated healthcare provider credentialing and network data infrastructure.
5+ YOE5+ years SRE/DevOps experience operating production systems at scale; strong GCP, IaC (Terraform/Pulumi), observability, CI/CD, and scripting (Python/Bash/Go) skills; incident response and reliability engineering experience.
GCP, GKE, Cloud Run, BigQuery, Cloud Monitoring, Terraform, Pulumi, Docker, Kubernetes, GitHub Actions, Cloud Build, Prometheus, Grafana, Datadog, Python, Bash, Go, Sentry, Snyk, SonarQube, Jira, Slack