465 aws site reliability engineer jobs at 307 companies in United States

2mo
Save
Mark Applied
Hide
Site Reliability Engineer
San Francisco or South San Francisco
$150k/yr OnsiteFull Time
VantageScore
VantageScore: Provides credit scoring and data analytics solutions.
5+ YOEExperienced Site Reliability Engineer with a DevSecOps focus; patch management, vulnerability remediation; AWS and CI/CD, security tooling.
AWS, EC2, ECS, Lambda, EKS, S3, RDS, IAM, VPC, CloudTrail, Config, GuardDuty, GitHub Actions, CodePipeline, Terraform, CloudFormation, AWS CDK, Kubernetes, Snyk, Wiz, Prisma Cloud, Kong, HashiCorp Vault, Secrets Manager, CloudWatch, Datadog, Grafana
2mo
Save
Mark Applied
Hide
AWS Cloud Site Reliability Engineer
Basking Ridge, New Jersey, United States
$73k-$130k/yr HybridFull Time
UnitedHealth Group
UnitedHealth GroupNYSE: UNH: Provides health insurance and technology-enabled health care services.
3+ YOE3+ years in AWS cloud with Java Spring Boot, Scala, Python; DynamoDB/Athena; AWS services (S3, CloudWatch, ECS, Lambda, RDS, EMR); CI/CD with GitHub Actions.
AWS, DynamoDB, CloudWatch, ECS, Lambda, RDS, EMR, GitHub Actions, Kubernetes, Scala, Java, Spring Boot, Python, Hadoop, HBase, Hive, Elastic APM
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (AWS)
Jacksonville or Atlanta or Milwaukee
HybridFull Time
FIS
FISNYSE: FIS: Provides technology solutions for merchants, banks, and capital markets
7+ YOE7+ years SRE/Cloud Engineering experience with hands-on AWS, Linux, CI/CD (Jenkins/Harness), Docker/Kubernetes (EKS), Terraform, scripting (Python/Bash/Shell), monitoring tools, on-call experience, and required AWS certifications.
AWS, EC2, EKS, RDS, S3, KMS, Secrets Manager, IAM, Route53, Security Groups, Linux, Git, Docker, Kubernetes, OpenShift, Helm, Jenkins, Harness, Terraform, Python, Bash, Shell, Dynatrace, CloudWatch, Splunk, Prometheus, Grafana, CheckMarx, SonarQube, Maven, Node, Artifactory, FlyWay, KeyFactor, HashiCorp Vault, CyberArk, SNOW, Jira, Confluence, Oracle DB, PostgreSQL DB, Postgres DB, Redis, SFTP, Tivoli
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Charlotte, North Carolina, United States
HybridFull Time
Electrolux Group
Electrolux GroupNasdaq Stockholm: ELUX-B: Global manufacturer of household appliances and consumer kitchen equipment.
6+ YOE6+ years in infrastructure/site reliability/cloud engineering; experience with cloud platforms, IaC, CI/CD, observability, troubleshooting, and strong collaboration skills.
Microsoft Azure, AWS, Google Cloud Platform, Akamai CDN, Terraform, CloudFormation, Ansible, Puppet, Chef, Microsoft Azure DevOps, GitHub, Argo CD
18h
Save
Mark Applied
Hide
Senior Site Reliability Engineer – Telephony & Communications Platform (AWS)
United States
$175k-$195k/yr RemoteFull Time
Filevine
Filevine: Legal case management software and AI platform for lawyers.
8+ YOE8+ years software engineering, 4+ years SRE, strong AWS and Terraform experience, CI/CD, observability, scripting (Python/Bash/PowerShell), incident response and reliability engineering.
EC2, ECS, EKS, VPC, Route 53, IAM, Lambda, S3, CloudWatch, Terraform, CI/CD, Python, Bash, PowerShell, Amazon Connect, Twilio, SIP/RTP
3w
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1w
Save
Mark Applied
Hide
Site Reliability Engineer (Cloud-Native / AWS)
Charlotte or Raleigh
OnsiteFull Time
Infosys
InfosysNYSE: INFY: Provides IT consulting, software development, and business outsourcing services.
Bachelor's degree or equivalent, experience with AWS and Python, familiarity with Grafana/Datadog/Splunk and ServiceNow, strong analytical and communication skills, experience defining non-functional requirements and reliability engineering concepts.
AWS, Python, Grafana, Datadog, Splunk, ServiceNow, Ansible, Puppet
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States
RemoteFull Time
Supabase
Supabase: Provides cloud-hosted databases and authentication tools for developers.
7+ YOE7+ years in SRE/production engineering, experience shaping SRE practices, defining and operationalizing SLOs/SLIs, incident response and postmortems, software engineering mindset, cloud infra (AWS) and IaC (Pulumi/Terraform/CDK).
Postgres, AWS, Pulumi, Terraform, CDK, Kubernetes, OpenTelemetry, VictoriaMetrics, Grafana, DORA metrics
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York, New York, United States
$177k-$209k/yr HybridFull Time
Peloton
PelotonNASDAQ: PTON: Sells connected fitness equipment and streaming exercise classes.
Kubernetes expertise; observability/monitoring mindset; CI/CD experience; IaC (Terraform/Pulumi); cloud (AWS); security and reliability focus; programming (Python/Go/Java/C).
Kubernetes, Observability, Jenkins, ArgoCD, Harness, Tekton, Terraform, Pulumi, AWS, Python, Golang, Java, C, Nginx, Ubuntu, Chef
1w
Save
Mark Applied
Hide
Site Reliability Engineer
California, United States
OnsiteFull Time
Arena
Arena: Benchmarks AI models using crowdsourced human preference evaluations.
6+ YOE6+ years backend engineering with distributed systems, proficiency in Go or Rust, experience with LLM provider APIs, cloud (AWS/GCP), Kubernetes, Terraform, Postgres, and Redis.
Go, Rust, OpenAI, Anthropic, Google, AWS, GCP, Kubernetes, Terraform, Postgres, Redis, Bifrost, Kong, Envoy, Tyk, Stripe, Metronome, Orb, vLLM, LiteLLM, LangChain
1w
Save
Mark Applied
Hide
Site Reliability Engineer
New York City or United States
$135k-$160k/yr RemoteFull Time
Fabric
Fabric: Provides clinical automation and care enablement software for healthcare.
5+ YOE5+ years SRE or platform engineering experience with AWS/EKS, production Kubernetes, Terraform, Datadog, Helm, GitHub Actions, and coding in Python/Bash/Go; HIPAA compliance experience preferred.
AWS, EKS, EC2, RDS, S3, Kubernetes (EKS), Terraform, Datadog, Helm, GitHub Actions, Python, Bash, Go
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Santa Clara, California, United States
$163k-$214k/yr OnsiteFull Time
IonQ
IonQNYSE: IONQ: Develops and sells trapped-ion quantum computers and cloud services.
7+ YOE7+ years production engineering experience; hands-on AWS/GCP reliability, observability and SLO ownership, incident command, resilience testing, and multi-team technical leadership.
AWS, GCP, Amazon Bedrock Agent Core
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Vienna, Virginia, United States
RemoteFull Time
Knexus
Knexus: Provider of applied artificial intelligence for government missions.
6+ YOE6+ years in infrastructure engineering; strong cloud expertise (GCP/AWS/Azure); security controls (NIST 800-53/800-171); DoD Cloud SRG; ATO/SSP experience; GCP certification desired; US citizen eligible for security clearance.
Google Cloud Platform, Amazon Web Services, Microsoft Azure, Kubernetes, IAM, SSP, ATO, NIST 800-53/800-171, DoD Cloud SRG, Google Cloud Professional certifications
2w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Los Angeles, California, United States
$140k-$199k/yr HybridFull Time
Green Dot
Green DotNYSE: GDOT: Provides mobile banking and payment solutions to consumers and businesses.
7+ YOE7+ years in release/reliability engineering, cloud platform experience (AWS/Azure/GCP), automated deployment and observability proficiency, scripting with PowerShell/Bash/Python, excellent troubleshooting and communication skills.
AWS, Azure, GCP, PowerShell, Bash, Python
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Alpharetta, Georgia, United States
OnsiteFull Time
Morgan Stanley
Morgan StanleyNYSE: MS: Provides global investment banking, wealth management, and advisory services.
5+ YOE5+ years production experience; strong scripting (Python, Perl, Shell, Ruby, Java, C#); DB2/Sybase/Oracle, Autosys, CI/CD, containers/VMs, Splunk/IP Soft/Sockeye, Jenkins/Train; cloud (Azure/AWS); BS in CS/Engineering required.
Python, Perl, Shell, Ruby, Java, C#, DB2, Sybase, Oracle, Autosys, Splunk, IP Soft, Sockeye, Jenkins, Train, Azure, AWS, MQ, UNIX, Linux, Windows
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Oakland, California, United States
$175k-$210k/yr HybridFull Time
Fivetran
Fivetran: Automates data movement into cloud data warehouses.
5+ YOE5+ years SaaS experience; managed Kubernetes, cloud platforms (AWS/GCP/Azure), Terraform/Ansible/ArgoCD; Python/Shell scripting, Linux admin, PostgreSQL; incident response and reliability engineering experience.
Kubernetes, EKS, AKS, GKE, PostgreSQL, ArgoCD, Terraform, Ansible, Python, Shell, Go, Java, AWS, GCP, Azure, Grafana, Buildkite, Temporal, Pulumi, Linux, VPN, PrivateLink, Private Service Connect (GCP)
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States or Canada
$124k-$177k/yr HybridFull Time
TextNow
TextNow: Provides free, ad-supported mobile calling and texting services.
5+ YOE5+ years in SRE/DevOps or Infrastructure, AWS, Terraform/Ansible, incident management, automation, observability, collaboration.
AWS, GitHub, Terraform, Ansible
2mo
Save
Mark Applied
Hide
Founding Engineer - Site Reliability
San Francisco or United States
$185k-$285k/yr RemoteFull Time
uRun
uRun: Infrastructure cloud for interactive, stateful AI inference.
7+ YOE7+ years in site reliability or infrastructure engineering; strong SLOs, incident response, and observability; Kubernetes and cloud (AWS); software engineering fundamentals; first SRE at a company.
Kubernetes, AWS, Prometheus, Grafana, Datadog, Automation, VPC, GPU compute
2w
Save
Mark Applied
Hide
Site Reliability Engineer
Wilmington, Delaware, United States
$120k-$140k/yr HybridFull Time
Best Egg
Best Egg: Provides personal loans and credit cards for prime borrowers.
Hands-on production support and incident response experience; Datadog, cloud (AWS) and scripting (Python, PowerShell, Bash) familiarity; strong communication and mentoring skills.
Datadog, JAMS, GoAnywhere, xMatters, ServiceNow, Jira, CI/CD, AWS, Python, PowerShell, Bash, Linux, AIOps
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States or Washington
$130k-$160k/yr RemoteFull Time
Berkeley Research Group
Berkeley Research Group: Provides expert testimony and specialized business consulting services.
5+ YOEBachelor's in computer science or similar, 5+ years SRE or similar, programming in Golang/Ruby/Python, Kubernetes, cloud experience (Azure/AWS/GCP), observability tools, and incident management expertise.
Microsoft Azure Cloud Services, GitHub Actions, GitLab CI, Golang, Ruby, Python, Kubernetes, AWS, GCP, Datadog, OpsGenie, PagerDuty