51 senior site reliability engineer jobs at 28 companies in Washington

1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Bellevue or San Francisco
$147k-$226k/yr OnsiteFull Time
Okta
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
5+ YOE5+ years SRE/DevOps experience; expert AWS multi-account governance; Terraform and Python automation; Kubernetes and observability experience; strong networking, Linux, security and documentation skills.
AWS, AWS Orgs, IAM, Identity Center, StackSets, Terraform, Python, GitLab, GitHub Actions, Kubernetes, Splunk, CloudWatch, Grafana, BGP, IPsec, VPCs, TGWs, VPC endpoints, Linux
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Seattle or Austin or Reston
$81k-$187k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
5+ YOETS/SCI with Polygraph and U.S. citizenship required. Bachelor's/Master's in CS or related, 5+ years SRE/Systems experience, expertise with Oracle Linux, Ansible, Terraform, Python, Bash, observability and distributed storage.
Oracle Linux, Ansible, Terraform, Python, Bash, Prometheus, Grafana, GlusterFS, VMware, Kubernetes, Docker, Jenkins, PostgreSQL, Active Directory, LDAP, Kerberos, NFS, SMB, iSCSI, NVMe-oF
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer I
Boston or Seattle or Atlanta
$134k-$215k/yr HybridFull Time
Axon
AxonNASDAQ: AXON: Develops weapons and software for law enforcement and safety.
7+ YOEBachelor's in CS/Engineering, 7+ years software engineering experience, expertise in distributed systems, Kubernetes, cloud (Azure/AWS/GCP), observability, Kafka, Terraform/Pulumi, and experience with agentic AI/LLM tooling preferred.
Kubernetes, Terraform, Pulumi, Kafka, Grafana, Datadog, New Relic, MySQL, Cassandra, PostgreSQL, Azure, AWS, GCP
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - SDN
San Francisco or San Jose or Bellevue
$240k-$312k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Kubernetes, SmartNICs, Python, Ansible, Go, C, Helm, Terraform, GitOps, CI/CD, Linux, OpenStack Neutron, OVN, OVS, DPDK, SR-IOV
3d
Save
Mark Applied
Hide
Senior Site Reliability Engineer (Multiple Positions)
Seattle or Bellevue
$212k-$368k/yr OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOEMaster's degree plus 2 years, or bachelor's degree plus 5 years of related experience. Requires 2 years in reliability support, SDLC, runbooks, data operations, and Linux administration.
Linux, OS networking protocol stack
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Data Infrastructure
Seattle, Washington, United States
$202k-$368k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
5+ YOE5+ years SRE/production engineering experience, proficiency with Go/Python/Bash, deep Linux, networking and distributed systems knowledge, bachelor’s degree or equivalent.
Kubernetes, Redis, MySQL, Message Queue, Kafka, Flink, Go, Python, Bash
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Production Engineer - ThousandEyes
San Francisco or Seattle or Austin or New York City
$165k-$241k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE5+ years experience; proficiency in Python or Go; expertise with Kubernetes, cloud (AWS), Unix/Linux; strong SRE principles, incident response, and security-minded engineering.
Python, Go, Kubernetes, Service Mesh, Prometheus, OpenTelemetry, ArgoCD, CNCF, AWS, Unix, Linux
1mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer
Bellevue, Washington, United States
$120k-$150k/yr HybridFull Time
Practice by Numbers
Practice by Numbers: A software building reliable platform and infrastructure to support mission-critical healthcare workflows.
6+ YOEEngineering degree (BS/MS) required, 6+ years software/SRE experience, production-quality programming in Go/Python/Java/TypeScript, cloud experience (AWS preferred), on-call and incident leadership experience, SLO/SLI/observability skills.
AWS, Docker, Kubernetes, Terraform, Prometheus, Grafana, OpenTelemetry, Go, Python, TypeScript, GitHub Actions
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Azure Storage
United States or Redmond
$120k-$235k/yr OnsiteFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
4+ YOEBachelor's in CS/Engineering or equivalent experience,4+ years technical experience in service/performance/site reliability engineering,experience with large-scale distributed systems and performance analysis.
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer, AI Infrastructure
Seattle or Bellevue
$178k-$342k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOEBachelor's degree in computer science, software engineering, or related field; 3+ years of SRE, DevOps, or systems engineering experience; Linux, networking, distributed systems, programming, and automation skills.
Linux, Go, Python, C, C++, Java, Bash, Kubernetes, AWS, GCP, Azure, Terraform, Prometheus, Grafana
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Compute
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience owning large GPU/compute fleets, automation of repair/deployment pipelines, firmware/BMC/Redfish familiarity, incident response and paging, observability and metrics tooling, and proficiency with production automation.
Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, Kubernetes, Claude Code, Cursor, LLM APIs, MCP servers
1mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer (Starshield)
Hawthorne or Redmond or Washington
$165k-$230k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOE5+ years SRE/DevOps experience with Kubernetes and Linux, proficiency in Bash/Python and infrastructure automation, bachelor's in CS/IT/engineering or 7+ years experience, and ability to obtain/maintain Top Secret clearance.
Kubernetes, Linux, Bash, Python, Bazel, Makefiles, Terraform, Ansible, TCP/IP
2w
Save
Mark Applied
Hide
Senior Engineer 2 - Site Reliability Engineering (Hybrid, Seattle)
Seattle, Washington, United States
$166k-$258k/yr HybridFull Time
Nordstrom
Nordstrom: Operates luxury department stores and off-price retail outlets.
10+ YOEBachelor's in CS/Engineering or equivalent,10+ years software engineering experience in SRE/infrastructure,proficiency with Kubernetes,cloud providers,networking,strong problem-solving and communication skills.
Kubernetes, Java, Go, Python, AWS, GCP, Azure
2mo
Save
Mark Applied
Hide
Principal Software Engineer, Site Reliability
Bellevue, Washington, United States
$218k-$250k/yr OnsiteFull Time
UiPath
UiPathNYSE: PATH: Provides robotic process automation software for enterprise workflow automation.
10+ YOE10+ years building large-scale distributed systems; proficiency in OOP (C#, C++, Java, Python), cloud (Azure/AWS/GCP), Kubernetes, databases, CI/CD, SRE and incident management; strong mentoring and product-focused engineering experience.
C#, C++, Java, Python, Kubernetes, Azure, AWS, GCP, AKS, GKE, Azure SQL, CosmosDB, Azure Data Lake, Power BI, MongoDB, MySQL, DynamoDB, CI/CD, DevOps
3mo
Save
Mark Applied
Hide
Site Reliability Engineer, Frontend Performance
Seattle, Washington, United States
$167k-$204k/yr OnsiteFull Time
Rocket Companies
Rocket CompaniesNYSE: RKT: Provides digital mortgage, real estate, and personal finance services.
5+ YOE5+ years in SRE or systems performance; strong TypeScript & Java; observability and AI tooling experience; AWS/Kubernetes and Datadog a plus.
TypeScript, Java, React, Spring, Datadog, AWS, Kubernetes, CI/CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer (Copy) (Copy)
Seattle or San Francisco
$170k-$220k/yr HybridFull Time
Supio
Supio: AI-powered data analysis platform for personal injury law firms.
3+ YOE3+ years SRE/DevOps experience, software development background, Bash/Python/TypeScript/Postgres, AWS (EC2/Lambda/RDS/IAM/VPC), GitHub Actions, CI/CD, on-call experience, strong autonomy and problem-solving.
Bash, Python, TypeScript, Postgres SQL, AWS, EC2, Lambda, RDS, IAM, VPCs, GitHub, GitHub Actions, Claude, ChatGPT
1w
Save
Mark Applied
Hide
Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid)
New York City or Austin or Sunnyvale or Redmond
$140k-$215k/yr HybridFull Time
CrowdStrike
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
10+ YOE10+ years building distributed systems, 5+ years developing SaaS microservices, expert programming skills, distributed-systems expertise, architectural leadership, and a Computer Science degree or equivalent experience.
Go, Java, Scala, Kotlin, Python, Node.js, Kubernetes, AWS, Cassandra, Kafka, Elasticsearch, OpenSearch, Google Cloud Platform (GCP), Oracle Cloud Infrastructure (OCI), GitHub, Stack Overflow
5d
Save
Mark Applied
Hide
Staff+ Site Reliability Engineer, Safeguards ML Infra
San Francisco or Seattle or New York City
$405k-$485k/yr HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
8+ YOEProduction change-management experience, high-stakes release and on-call experience, AWS/GCP operations, Python proficiency, and a bachelor's degree or equivalent experience.
Python, Rust, AWS, GCP, AWS Bedrock, GCP Vertex, Claude
2mo
Save
Mark Applied
Hide
Senior Facilities Manager, Reliability - Site based Redmond, WA
Redmond, Washington, United States
$92k-$127k/yr OnsiteFull Time
Evotec
EvotecFrankfurt Stock Exchange: EVT: Provides drug discovery and development services through global research alliances.
10+ YOE5+ MgmtSenior facilities manager with 10+ years pharma facilities maintenance and leadership experience; CMMS proficiency; strong collaboration and problem-solving skills.
CMMS, Microsoft Office

Explore Jobs

Expand Your Job Search