1,182 site reliability engineer jobs at 578 companies in United States

1w
Save
Mark Applied
Hide
Site Reliability Engineer Intern (Data Infra) - 2027 Fall
San Jose, California, United States
OnsiteInternship
ByteDance
ByteDance: Global technology specializing in AI-powered content platforms.
Currently pursuing a bachelor's degree in computer science or related technical discipline; programming experience in C, C++, Java, Python, Go, or Rust; knowledge of Unix/Linux internals, networking, and distributed systems.
C, C++, Java, Python, Go, Rust, Unix, Linux, Kubernetes, Redis, MySQL, Flink, Nginx, Docker, OpenStack, Hadoop, Spark
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
San Francisco or South San Francisco
$150k/yr OnsiteFull Time
VantageScore Solutions, LLC
VantageScore Solutions, LLC: A Higher Level of Confidence
5+ YOEExperienced Site Reliability Engineer with a DevSecOps focus; patch management, vulnerability remediation; AWS and CI/CD, security tooling.
AWS, EC2, ECS, Lambda, EKS, S3, RDS, IAM, VPC, CloudTrail, Config, GuardDuty, GitHub Actions, CodePipeline, Terraform, CloudFormation, AWS CDK, Kubernetes, Snyk, Wiz, Prisma Cloud, Kong, HashiCorp Vault, Secrets Manager, CloudWatch, Datadog, Grafana
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Charlotte, North Carolina, United States
HybridFull Time
Electrolux Group
Electrolux GroupNasdaq Stockholm: ELUX B: Global home appliance manufacturer reinventing taste, care, and wellbeing.
6+ YOE6+ years in infrastructure/site reliability/cloud engineering; experience with cloud platforms, IaC, CI/CD, observability, troubleshooting, and strong collaboration skills.
Microsoft Azure, AWS, Google Cloud Platform, Akamai CDN, Terraform, CloudFormation, Ansible, Puppet, Chef, Microsoft Azure DevOps, GitHub, Argo CD
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Austin or Reston
HybridFull Time
Seekr Technologies
Seekr Technologies: Private American enterprise AI providing explainable, secure AI software and hardware to government and critical-infrastructure customers.
5+ YOERequires 5+ years in site reliability engineering and Linux systems, monitoring and logging expertise, programming or scripting skills, Docker, Kubernetes, configuration automation, and incident response experience.
Linux, ELK, Prometheus, InfluxDB, Grafana, Python, Ruby, Bash, Java, Docker, Kubernetes, Puppet, Chef, Ansible, Terraform, Elasticsearch, Kafka, Aerospike, Git, GitHub, GitLab, ArgoCD
4d
Save
Mark Applied
Hide
Site Reliability Engineer
Lehi, Utah, United States
OnsiteFull Time
Enzo Health
Enzo Health: AI-native home health EHR and operations platform for U.S. home health agencies.
5+ YOEAt least 5 years in site reliability, platform, infrastructure, or production engineering; hands-on AWS, Kubernetes, Terraform, Postgres, CI/CD, observability, automation, and incident response experience.
AWS, Kubernetes, Terraform, Postgres, GitHub, CI/CD, SOC 2, OASIS
2w
Save
Mark Applied
Hide
Site Reliability Engineer
Birmingham, Alabama, United States
HybridFull Time
PRADCO Outdoor Brands
PRADCO Outdoor Brands: Family-owned U.S. manufacturer of hunting, fishing, and pet products for hunters, anglers, and pet owners.
3+ YOERequires 3–5 years in site reliability or similar engineering, Azure expertise, distributed-systems support, CI/CD, performance diagnostics, automation, and a bachelor's degree or equivalent experience.
Microsoft Azure, App Service Plans, Azure Functions, Event Grid, Event Hub, Service Bus, Azure SQL Service, Azure DevOps, CI/CD, SQL
6d
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Nashville, Tennessee, United States
$81k-$187k/yr OnsiteFull Time
Oracle Corporation
Oracle CorporationNYSE: ORCL: Cloud infrastructure and enterprise software solutions provider.
3+ YOEBachelor’s degree in Computer Science or equivalent experience; 3+ years in Site Reliability Engineering, DevOps, or Systems Engineering; cloud operations, incident management, automation, programming, and infrastructure tooling experience.
AWS, Azure, GCP, OCI, Chef, Ansible, Jenkins, Terraform, Docker, RESTful APIs, CI/CD, Agile
2w
Save
Mark Applied
Hide
Site Reliability Engineer
Bentonville, Arkansas, United States
$70k-$80k/yr OnsiteFull Time
Cognizant
CognizantNASDAQ: CTSH: Global professional services providing technology and consulting services.
3+ YOERequires 3+ years in software engineering focused on reliability, infrastructure, or platforms; Java, microservices, MVC, JDBC, REST, frameworks, Azure, SRE, troubleshooting, and on-call support.
Java, Microservice Architecture, MVC (Model-View-Controller), JDBC (Java Database Connectivity), RESTful web services, Azure
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: Private supply-chain AI software serving corporations, government agencies, and banks with risk and compliance technology.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
2w
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Santa Clara, California, United States
$50-$60/hr RemoteContract
ServiceNow
ServiceNowNYSE: NOW: Enterprise software providing cloud-based workflow automation platforms.
3+ YOEBachelor's degree in computer science or related field; 3+ years in site reliability engineering; 2+ years with AWS and cloud automation; Kubernetes, Linux, Terraform, networking, GitOps, monitoring, and customer support experience.
AWS, Kubernetes, Helm, Linux, Terraform, GitOps, Prometheus, Grafana, Bazel, CueLang, Version Control, Okta, Snowflake, Google
3w
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
San Francisco, California, United States
$350k-$475k/yr OnsiteFull Time
Thinking Machines Lab
Thinking Machines Lab: Private AI research and product building customizable multimodal systems for researchers and the wider public.
Experience in distributed systems/cloud/site reliability, software automation for reliability, incident response and postmortems, strong communication and coordination skills.
Tinker, Kubernetes, LoRA, CI/CD
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Customer engagement platform for cross-channel marketing and analytics.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer (200046807)
United States
OnsiteFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Multinational technology providing software, cloud, and AI solutions.
Senior-level site reliability engineering role at Microsoft (requisition posted). Specific qualifications not provided in posting.
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
Europe or Canada or Bellevue or Los Angeles
RemoteFull Time
Shopify
ShopifyNasdaq: SHOP: Provides internet infrastructure and tools for commerce.
Experienced SRE/engineer with on-call experience, ability to build resilient production tooling, improve observability, respond to alerts, and collaborate across engineering teams.
IDE
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States
RemoteFull Time
Supabase
Supabase: Developer platform providing Postgres databases, authentication, storage, realtime, REST APIs, and edge functions for application developers.
7+ YOE7+ years in SRE/production engineering, experience shaping SRE practices, defining and operationalizing SLOs/SLIs, incident response and postmortems, software engineering mindset, cloud infra (AWS) and IaC (Pulumi/Terraform/CDK).
Postgres, AWS, Pulumi, Terraform, CDK, Kubernetes, OpenTelemetry, VictoriaMetrics, Grafana, DORA metrics
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Cambridge or United States
$76k-$136k/yr HybridFull Time
Akamai Technologies
Akamai TechnologiesNASDAQ: AKAM: Cloud and edge computing platform for secure digital experiences.
Bachelor's in Computer Science/Engineering or equivalent; experience in SRE/Software Engineering for large-scale distributed systems; Terraform and IAC experience; familiarity with SaltStack/Ansible/Chef/Puppet; Linux, CI/CD, observability, and participation in on-call rotation.
Terraform, SaltStack, Ansible, Chef, Puppet, Linux, CI/CD, Infrastructure as Code (IAC)
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Riverwoods, Illinois, United States
$110k-$120k/yr HybridFull Time
QualityAI
QualityAI: AI-first quality engineering and software testing firm.
6+ YOE6+ years SRE experience with AWS, monitoring/observability, Linux/Unix, scripting, CI/CD, and building automated reliability and performance solutions.
AWS, AWS Lambda, Linux/Unix, Datadog, Dynatrace, Grafana, Kibana, ELK, APM tools, Python, Java, Shell Scripting, Go, Ansible, Jenkins, CI/CD, Kubernetes, OpenShift, JIRA, ServiceNow (SNOW), SQL, MySQL
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States
$150k-$200k/yr RemoteFull Time
Runpod
Runpod: AI cloud computing platform providing on-demand GPUs and serverless compute to developers, researchers, and AI companies.
5+ YOE5+ years SRE or production engineering experience; strong Linux, networking, container, distributed systems, SLI/SLO, incident response, and scripting skills.
Prometheus, Grafana, Python, Go, Bash, Linux, Slack
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
United States
$175k-$200k/yr RemoteFull Time
Order.co
Order.co: AI-powered B2B procurement platform helping businesses manage purchasing, approvals, payments, and reporting.
Senior-level SRE with strong focus on reliability, automation, and platform engineering underpinned by software development skills.
Ruby, Ruby on Rails, Linux, AWS, Kubernetes, Terraform, CloudFormation, Datadog, OpenTelemetry, Python, Go, Bash
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
California, United States
OnsiteFull Time
Arena Intelligence
Arena Intelligence: AI model evaluation platform serving enterprises, AI labs, and independent researchers in real-world workflows.
6+ YOE6+ years backend engineering with distributed systems, proficiency in Go or Rust, experience with LLM provider APIs, cloud (AWS/GCP), Kubernetes, Terraform, Postgres, and Redis.
Go, Rust, OpenAI, Anthropic, Google, AWS, GCP, Kubernetes, Terraform, Postgres, Redis, Bifrost, Kong, Envoy, Tyk, Stripe, Metronome, Orb, vLLM, LiteLLM, LangChain