65 cloud reliability engineer jobs at 33 companies in Renton, WA
1mo
Save
Mark Applied
Hide
1mo
Tech Lead Cloud Site Reliability Engineer - DCS Cloud
Seattle, Washington, United States
OnsiteFull Time
ByteDance: Developing AI-driven content platforms and mobile applications.
5+ YOEBachelor's in CS or related,5+ years SRE/Linux/DevOps experience,proficient in Go/Python/C++,familiar with public cloud platforms,monitoring,incident response,and strong troubleshooting and communication skills.
T-MobileNASDAQ: TMUS: Provides wireless network services and mobile communication devices.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, scripting, APIs, cybersecurity, and reliability engineering experience, plus U.S. work authorization.
C, C#, Java, Perl, Python, Go, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Linux, Shell
Site Reliability Engineer (SRE) - Hybrid Cloud Storage
Seattle, Washington, United States
$140k-$210k/yrHybridFull Time
Qumulo: Unified file and object storage for hybrid cloud environments.
3+ YOE3+ years building/operating automated testing for complex software; strong C and Python skills; experience with on‑prem and cloud (AWS/GCP/Azure); Linux/Ubuntu fluency; familiarity with Kubernetes, Ansible, Terraform, and observability tooling.
Senior Site Reliability Engineer - Core Cloud Platform
San Francisco or San Jose or Bellevue
$240k-$356k/yrHybridFull Time
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
7+ YOE7+ years SRE or production infrastructure experience, deep Kubernetes and Terraform knowledge, experience with observability and SLOs, proficiency in Go or Python, on-call and incident leadership experience.
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, Python, APIs, Power Platform, identity governance, and reliability engineering experience.
C, C#, Java, Perl, Python, Go, Shell, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Microsoft Power Platform, Linux, VMs
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOEBachelor's in CS/engineering/math with 5 years software experience or 7+ years SRE/DevOps experience; Linux experience required; Kubernetes, Kafka, cloud-native tooling, and programming in Python/Go/Java/C#/Scala preferred.
Senior Site Reliability Engineer, Production Engineer - ThousandEyes
San Francisco or Seattle or Austin or New York City
$165k-$241k/yrHybridFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE5+ years experience; proficiency in Python or Go; expertise with Kubernetes, cloud (AWS), Unix/Linux; strong SRE principles, incident response, and security-minded engineering.
Python, Go, Kubernetes, Service Mesh, Prometheus, OpenTelemetry, ArgoCD, CNCF, AWS, Unix, Linux
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yrHybridFull Time
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
Senior Engineer 2 - Site Reliability Engineering (Hybrid, Seattle)
Seattle, Washington, United States
$166k-$258k/yrHybridFull Time
Nordstrom: Operates luxury department stores and off-price retail outlets.
10+ YOEBachelor's in CS/Engineering or equivalent,10+ years software engineering experience in SRE/infrastructure,proficiency with Kubernetes,cloud providers,networking,strong problem-solving and communication skills.
Principal Site Reliability Engineer - CTJ - Secret
Redmond, Washington, United States
$143k-$275k/yrOnsiteFull Time
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
2+ YOEDegree in CS/IT (or equivalent experience) with minimum 2+ years technical experience (Doctorate path) and ability to obtain required background investigations (T3/CJIS) for government cloud environments; SRE, incident response, and cloud systems experience.
RobinhoodNASDAQ: HOOD: Provides a commission-free platform for investing and financial services.
1+ YOEMaster's in Computer Science or related field plus one year of experience; expertise in production network infrastructure, BGP/OSPF, network automation (Python or Golang), and monitoring.
Senior Principal Network Reliability Engineer - Network Region Build (NRB)
Seattle or United States
$126k-$264k/yrOnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
6+ YOEExpertise in hyperscale cloud networking, routing and transport protocols, reliability engineering, automation, and cross-organizational technical leadership; 6+ years experience preferred; English required.
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experienced engineering manager required to lead and grow reliability engineers delivering highly available, scalable, resilient, fault-tolerant global network services.
Seattle or San Francisco or Detroit or United States
$180k-$279k/yrHybridFull Time
Rocket CompaniesNYSE: RKT: Provides digital mortgage, real estate, and personal finance services.
7+ YOE7+ years AWS/cloud infra; 5+ years PostgreSQL/AWS services; Linux admin/scripting; mentoring; infrastructure as code and security; AI code generation tools; on-call readiness.
AWS, PostgreSQL, Aurora/RDS, S3, ElastiCache, OpenSearch, DynamoDB, Linux, Python, Infrastructure as Code, Security practices, AI code generation tools
Rippling: Unified platform managing workforce HR, IT, and finance operations
8+ YOE4+ Mgmt8+ years infrastructure or software engineering, 4+ years engineering management, deep AWS and Kubernetes experience, Terraform and Helm proficiency, production incident and reliability expertise, strong cross-functional leadership.
Aurelian: AI-powered voice automation for 9-1-1 non-emergency calls.
4+ YOE4+ years in infrastructure/platform/backend engineering, experience with reliability and scale, comfortable across backend and cloud, experience building analytics/observability/developer tooling.
DocuSignNASDAQ: DOCU: Provider of e-signature and intelligent agreement management software.
8+ YOE8+ years backend engineering experience, B.S. in CS or similar, 3+ years building cloud-native microservices, experience with SQL/NoSQL, service reliability, and strong communication and design skills.
Redis, Cassandra, Elasticsearch, Microsoft Azure, Microsoft Azure App Services, Microsoft Azure Kubernetes Service (AKS), Microsoft Azure Blob Storage, Microsoft Azure SQL Database, C#, .NET, ASP.NET, PostgreSQL, Cosmos DB, GraphQL, GRPC, REST