57 cloud reliability engineer jobs at 36 companies in Bellevue, WA
1w
Save
Mark Applied
Hide
1w
Tech Lead Cloud Site Reliability Engineer - DCS Cloud
Seattle, Washington, United States
OnsiteFull Time
ByteDance: Developing AI-driven content platforms and mobile applications.
5+ YOEBachelor's in CS or related,5+ years SRE/Linux/DevOps experience,proficient in Go/Python/C++,familiar with public cloud platforms,monitoring,incident response,and strong troubleshooting and communication skills.
Site Reliability Engineer (SRE) - Hybrid Cloud Storage
Seattle, Washington, United States
$140k-$210k/yrHybridFull Time
Qumulo: Unified file and object storage for hybrid cloud environments.
3+ YOE3+ years building/operating automated testing for complex software; strong C and Python skills; experience with on‑prem and cloud (AWS/GCP/Azure); Linux/Ubuntu fluency; familiarity with Kubernetes, Ansible, Terraform, and observability tooling.
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
6+ YOE6+ years software/network/systems engineering experience or equivalent degree+experience; experience with large-scale cloud services, observability, automation, and customer-facing reliability work.
Microsoft Fabric, Azure SQL DB, Azure Cosmos DB, Azure PostgreSQL, Azure Data Factory, Azure Synapse Analytics, Azure Service Bus, Azure Event Grid, Power BI, Logic Apps, Jupyter Notebooks
7+ YOE7+ years in DB engineering/DBA/DBRE roles; strong SQL, performance tuning, HA, backup/recovery; bachelor's or equivalent experience; experience with cloud DBs, automation, observability, and on-call rotations.
United States or Kansas or Washington or California or Texas or Illinois or North Carolina or Colorado or Massachusetts or Pennsylvania or Virginia or Oregon or Nevada or Hawaii or New York or Georgia or Ohio or Arizona or Seattle or San Francisco or New York City
$110k-$183k/yrRemoteFull Time
Veeam: Data resilience and security for hybrid cloud environments
3+ YOE3+ years in software engineering with 1+ year in SRE/Platform/DevOps, cloud experience (Azure or comparable), observability (Prometheus, Grafana, OpenTelemetry, ELK), IaC (Terraform/Terragrunt/Pulumi), Kubernetes, CI/CD tooling, programming in TypeScript/JS, Go, Java, or C#, and experience in compliance-oriented environments.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yrHybridFull Time
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
RobinhoodNASDAQ: HOOD: Provides a commission-free platform for investing and financial services.
1+ YOEMaster's in Computer Science or related field plus one year of experience; expertise in production network infrastructure, BGP/OSPF, network automation (Python or Golang), and monitoring.
Senior Principal Network Reliability Engineer - Network Region Build (NRB)
Seattle or United States
$126k-$264k/yrOnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
6+ YOEExpertise in hyperscale cloud networking, routing and transport protocols, reliability engineering, automation, and cross-organizational technical leadership; 6+ years experience preferred; English required.
Senior Site Reliability Engineer - Video Platform - USDS
Seattle, Washington, United States
$178k-$342k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
5+ YOEBachelor's or equivalent,5+ years SRE/software development experience in large-scale online services; programming in C,C++,Java,Python,C# or Go; networking, OS, DB, container and cloud knowledge; incident response and capacity management.
Senior Site Reliability Engineer, Data Infrastructure
New York or Bellevue
$165k-$242k/yrHybridFull Time
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
5+ YOE5+ years in SRE/Platform/Infrastructure roles; Kubernetes; CI/CD with Argo CD and GitHub Actions; high availability (≥99.99%); multi-region, security, observability; IaC (Helm, Terraform, Pulumi); performance tuning; security best practices in cloud environments.
Airwallex: Global financial platform for business payments and money management.
5+ YOE5+ years software engineering; strong PostgreSQL and Redis experience; cloud (GCP) and IaC; building automation and platform tooling; strong system design.
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
3+ YOEMaster's+3 years or Bachelor's+5 years in CS or related; experience with CI/CD, Jenkins, GitLab/GitHub, AppDynamics, Splunk, Windows/Unix/RHEL, and CQL/SQL; strong DevOps and cloud skills.
Seattle or San Francisco or Detroit or United States
$180k-$279k/yrHybridFull Time
Rocket CompaniesNYSE: RKT: Provides digital mortgage, real estate, and personal finance services.
7+ YOE7+ years AWS/cloud infra; 5+ years PostgreSQL/AWS services; Linux admin/scripting; mentoring; infrastructure as code and security; AI code generation tools; on-call readiness.
AWS, PostgreSQL, Aurora/RDS, S3, ElastiCache, OpenSearch, DynamoDB, Linux, Python, Infrastructure as Code, Security practices, AI code generation tools
IT Spec (ENTARCH) "Platform/site Reliability Engineer", GS-2210-14 FPL GS-14 (DH)
San Francisco or Denver or Washington or Atlanta or Chicago or Salt Lake City or Seattle
$128k-$197k/yrHybridFull Time
Federal Student Aid: Administers federal student loans and financial aid programs.
1+ YOEDesign and implement scalable cloud platforms, IaC, CI/CD, containers, observability, SRE practices; lead platform engineering and security; must be U.S. citizen and pass background/fingerprint check.
Rippling: Unified platform managing workforce HR, IT, and finance operations
8+ YOE4+ Mgmt8+ years infrastructure or software engineering, 4+ years engineering management, deep AWS and Kubernetes experience, Terraform and Helm proficiency, production incident and reliability expertise, strong cross-functional leadership.