113 cloud reliability engineer jobs at 41 companies in Washington
1mo
Save
Mark Applied
Hide
1mo
Tech Lead Cloud Site Reliability Engineer - DCS Cloud
Seattle, Washington, United States
OnsiteFull Time
ByteDance: Global technology specializing in AI-powered content platforms.
5+ YOEBachelor's in CS or related,5+ years SRE/Linux/DevOps experience,proficient in Go/Python/C++,familiar with public cloud platforms,monitoring,incident response,and strong troubleshooting and communication skills.
T-MobileNASDAQ: TMUS: The Un-carrier providing wireless and home internet services.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, scripting, APIs, cybersecurity, and reliability engineering experience, plus U.S. work authorization.
C, C#, Java, Perl, Python, Go, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Linux, Shell
Site Reliability Engineer (SRE) - Hybrid Cloud Storage
Seattle, Washington, United States
$140k-$210k/yrHybridFull Time
Qumulo: The world's most advanced file system – any data, any location, total control.
3+ YOE3+ years building/operating automated testing for complex software; strong C and Python skills; experience with on‑prem and cloud (AWS/GCP/Azure); Linux/Ubuntu fluency; familiarity with Kubernetes, Ansible, Terraform, and observability tooling.
OpenEye: Private commercial cloud video surveillance serving businesses with AI-driven analytics, business intelligence, and loss-prevention tools.
1+ YOERequires 1–5 years of related experience, cloud, CI/CD, infrastructure automation, monitoring, scripting or development experience, TCP/IP knowledge, Agile familiarity, and strong problem-solving and communication skills.
Senior Site Reliability Engineer - Core Cloud Platform
San Francisco or San Jose or Bellevue
$240k-$356k/yrHybridFull Time
Lambda: AI infrastructure building GPU cloud services and supercomputers for researchers, enterprises, and hyperscalers.
7+ YOE7+ years SRE or production infrastructure experience, deep Kubernetes and Terraform knowledge, experience with observability and SLOs, proficiency in Go or Python, on-call and incident leadership experience.
Practice by Numbers: Dental software providing an all-in-one operations, analytics, communications, payments, and marketing platform for dental practices.
6+ YOEEngineering degree (BS/MS) required, 6+ years software/SRE experience, production-quality programming in Go/Python/Java/TypeScript, cloud experience (AWS preferred), on-call and incident leadership experience, SLO/SLI/observability skills.
T-MobileNASDAQ: TMUS: The Un-carrier providing wireless and home internet services.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, Python, APIs, Power Platform, identity governance, and reliability engineering experience.
C, C#, Java, Perl, Python, Go, Shell, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Microsoft Power Platform, Linux, VMs
SpaceXNasdaq: SPCX: Designing, manufacturing, and launching advanced rockets and spacecraft.
5+ YOEBachelor's in CS/engineering/math with 5 years software experience or 7+ years SRE/DevOps experience; Linux experience required; Kubernetes, Kafka, cloud-native tooling, and programming in Python/Go/Java/C#/Scala preferred.
Senior Site Reliability Engineer, Production Engineer - ThousandEyes
San Francisco or Seattle or Austin or New York City
$165k-$241k/yrHybridFull Time
ThousandEyes: ThousandEyes is a Cis-owned digital experience assurance platform helping organizations monitor networks, applications, and cloud services.
5+ YOE5+ years experience; proficiency in Python or Go; expertise with Kubernetes, cloud (AWS), Unix/Linux; strong SRE principles, incident response, and security-minded engineering.
Python, Go, Kubernetes, Service Mesh, Prometheus, OpenTelemetry, ArgoCD, CNCF, AWS, Unix, Linux
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yrHybridFull Time
OktaNASDAQ: OKTA: Identity management and access control software provider.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
Alibaba CloudNYSE, HKEX: BABA, 9988: Global cloud computing and data intelligence service provider.
5+ YOEBachelor's degree in computer science or related field, 5+ years developing or operating large-scale distributed systems, expert Linux administration, Kubernetes, Python, Shell, and big data architecture expertise.
Senior Engineer 2 - Site Reliability Engineering (Hybrid, Seattle)
Seattle, Washington, United States
$166k-$258k/yrHybridFull Time
NordstromNew York Stock Exchange: JWN: Fashion specialty retailer offering apparel, footwear, and accessories.
10+ YOEBachelor's in CS/Engineering or equivalent,10+ years software engineering experience in SRE/infrastructure,proficiency with Kubernetes,cloud providers,networking,strong problem-solving and communication skills.
Principal Site Reliability Engineer - CTJ - Secret
Redmond, Washington, United States
$143k-$275k/yrOnsiteFull Time
MicrosoftNASDAQ: MSFT: Multinational technology providing software, cloud, and AI solutions.
2+ YOEDegree in CS/IT (or equivalent experience) with minimum 2+ years technical experience (Doctorate path) and ability to obtain required background investigations (T3/CJIS) for government cloud environments; SRE, incident response, and cloud systems experience.
Senior Principal Network Reliability Engineer - Network Region Build (NRB)
Seattle or United States
$126k-$264k/yrOnsiteFull Time
Oracle CorporationNYSE: ORCL: Cloud infrastructure and enterprise software solutions provider.
6+ YOEExpertise in hyperscale cloud networking, routing and transport protocols, reliability engineering, automation, and cross-organizational technical leadership; 6+ years experience preferred; English required.
HP Inc.NYSE: HPQ: Global leader in personal computing, printing, and technology solutions.
15+ YOE8+ MgmtRequires a bachelor's or master's degree, 15+ years in software, infrastructure, platform, or reliability engineering, and 8+ years leading engineering organizations. VP-level experience, cloud-native expertise, and global-scale operations required.
San Francisco or Denver or Austin or Jacksonville or Bridgeport or Seattle or Boston or New York City
$126k-$205k/yrRemoteFull Time
Palo Alto NetworksNASDAQ: PANW: Global cybersecurity platform providing network, cloud, and AI-driven security solutions.
8+ YOERequires 8+ years of relevant experience, backend programming proficiency, cloud-native and distributed systems expertise, Linux and networking knowledge, debugging skills, and experience with AWS or GCP and Kubernetes.
Alibaba GroupNYSE, HKEX: BABA, 9988: Global technology focused on AI, cloud, and consumption.
3+ YOERequires 3+ years in SRE, DevOps, or backend development; proficiency in Python, Go, Java, or C++; Linux, networking, databases, cloud-native systems, incident response, and Chinese-English fluency.
AppleNASDAQ: AAPL: Designing and manufacturing consumer electronics, software, and digital services.
Experienced engineering manager required to lead and grow reliability engineers delivering highly available, scalable, resilient, fault-tolerant global network services.
Blue Origin: Developing reusable space vehicles and infrastructure.
3+ YOEBS or higher in a technical field or equivalent experience, 3+ years software development/SRE experience, proficiency with CI/CD, IaC (Terraform/Pulumi), cloud (AWS/Azure/GCP), Kubernetes, Linux, and at least one language (Python, Go, Rust, Java); U.S. work authorization required.