44 infrastructure reliability engineer jobs at 26 companies in Covington, WA

3mo
Save
Mark Applied
Hide
Staff Infrastructure Reliability Engineer - Database & Storage
Seattle or San Francisco or Detroit or United States
$180k-$279k/yr HybridFull Time
Rocket Companies
Rocket CompaniesNYSE: RKT: Provides digital mortgage, real estate, and personal finance services.
7+ YOE7+ years AWS/cloud infra; 5+ years PostgreSQL/AWS services; Linux admin/scripting; mentoring; infrastructure as code and security; AI code generation tools; on-call readiness.
AWS, PostgreSQL, Aurora/RDS, S3, ElastiCache, OpenSearch, DynamoDB, Linux, Python, Infrastructure as Code, Security practices, AI code generation tools
2w
Save
Mark Applied
Hide
Reliability Engineer, R&D
Austin or New York City or San Francisco or Seattle
$203k-$232k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience in reliability engineering for infrastructure or complex hardware, building availability/RAM models, leading cross-discipline FMEAs, and mining field failure data.
2w
Save
Mark Applied
Hide
Senior Infrastructure Engineer
Seattle, Washington, United States
$160k-$220k/yr OnsiteFull Time
Aurelian
Aurelian: AI-powered voice automation for 9-1-1 non-emergency calls.
4+ YOE4+ years in infrastructure/platform/backend engineering, experience with reliability and scale, comfortable across backend and cloud, experience building analytics/observability/developer tooling.
ClickHouse
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Data Infrastructure
Seattle, Washington, United States
$202k-$368k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
5+ YOE5+ years SRE/production engineering experience, proficiency with Go/Python/Bash, deep Linux, networking and distributed systems knowledge, bachelor’s degree or equivalent.
Kubernetes, Redis, MySQL, Message Queue, Kafka, Flink, Go, Python, Bash
1mo
Save
Mark Applied
Hide
Sr. Hardware / Infrastructure Site Reliability Engineer (Starlink)
Redmond, Washington, United States
$165k-$230k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOEBachelor's in CS/IT/engineering plus 5+ years SRE/DevOps (or 7+ years without degree); 2+ years Linux; experience with Terraform/Ansible, Docker/Kubernetes; Bash, Python, Go/C++/C development; strong OS, networking, CI/CD skills.
Linux, Terraform, Ansible, Docker, Kubernetes, Bash, Python, Go, C++, C
1mo
Save
Mark Applied
Hide
Staff Software Engineer, Infrastructure
Canada or United States or Seattle or Paris or New York City
$238k-$382k/yr RemoteFull Time
Docker
Docker: Provides a platform for building, sharing, and running containerized applications.
8+ YOE8+ years backend/infrastructure engineering; strong Go; experience with Kubernetes/EKS, cloud platforms, networking, and reliability; Bachelor's or equivalent; experience leading cross-team technical initiatives and strong written/verbal communication.
Go, Terraform, Argo CD, EKS, Envoy Gateway, Grafana Cloud, OpenTelemetry, Prometheus, Grafana, GitHub Actions, Kubernetes, Linux, GitOps, Covey Scout
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
San Jose or Seattle or San Francisco
$159k-$302k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
5+ YOEBachelor's or equivalent, 5+ years SRE/infrastructure/backend experience; Kubernetes, Docker, Terraform, AWS, Postgres/Redis, observability, incident response, CI/CD, bash, Node.js/TypeScript experience; on-call participation.
Kubernetes, Docker, bash, CircleCI, Node.js, TypeScript, Postgres, Redis, AWS Aurora (Postgres-compatible), Terraform, AWS
3w
Save
Mark Applied
Hide
Tech Lead - Data Infrastructure Site Reliability
Seattle, Washington, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
5+ YOE5+ years SRE or software engineering experience designing, building, scaling, and operating cloud-based systems; expertise in databases, Kubernetes, distributed systems; strong communication and collaboration skills.
SQL, NoSQL, Kubernetes
2mo
Save
Mark Applied
Hide
Senior Network Reliability Engineer
Bellevue, Washington, United States
$139k-$195k/yr OnsiteFull Time
Robinhood
RobinhoodNASDAQ: HOOD: Provides a commission-free platform for investing and financial services.
1+ YOEMaster's in Computer Science or related field plus one year of experience; expertise in production network infrastructure, BGP/OSPF, network automation (Python or Golang), and monitoring.
Python, Golang, BGP, OSPF, Network Automation, Cloud Networking, Infrastructure as Code, Monitoring
3d
Save
Mark Applied
Hide
Senior Engineer 2 - Site Reliability Engineering (Hybrid, Seattle)
Seattle, Washington, United States
$166k-$258k/yr HybridFull Time
Nordstrom
Nordstrom: Operates luxury department stores and off-price retail outlets.
10+ YOEBachelor's in CS/Engineering or equivalent,10+ years software engineering experience in SRE/infrastructure,proficiency with Kubernetes,cloud providers,networking,strong problem-solving and communication skills.
Kubernetes, Java, Go, Python, AWS, GCP, Azure
2mo
Save
Mark Applied
Hide
SWE - Backend Infrastructure Engineer
San Francisco or Bellevue or New York
$175k-$280k/yr OnsiteFull Time
Sesame
Sesame: Designing wearable computers with lifelike voice-driven AI agents.
3+ YOEStrong systems thinker with reliability engineering experience; 3+ years in infrastructure, platform, or ML systems; Kubernetes production experience; strong communication.
Kubernetes, Terraform, CloudFormation, Pulumi, TorchServe, Seldon, KServe, Ray Serve, PyTorch, APIs, Database design
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
1mo
Save
Mark Applied
Hide
Senior Core Infrastructure Engineer
Seattle or Santa Clara
$79k-$210k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
3+ YOEDesign and implement scalable, reliable distributed systems; build fault-tolerant components; develop automation/IaC; apply security controls; participate in incident response and runbooks.
DevOps, Java, Infrastructure as Code (IaC)
4d
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Core Cloud Platform
San Francisco or San Jose or Bellevue
$240k-$356k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
7+ YOE7+ years SRE or production infrastructure experience, deep Kubernetes and Terraform knowledge, experience with observability and SLOs, proficiency in Go or Python, on-call and incident leadership experience.
Kubernetes, Terraform, Argo CD, Flux, Helm, Kustomize, OpenTelemetry, Prometheus, Grafana, Datadog, Go, Python, etcd, GitOps
1mo
Save
Mark Applied
Hide
Site Reliability Engineer - Software Ops and Scaling , One Material Handling System - Software, Controls and Science
Nashville or Arlington or Bellevue
$94k-$160k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
Experience automating, deploying, and supporting infrastructure; programming in Python, Ruby, Golang, Java, C++, C#, or Rust; Linux/Unix experience; CI/CD pipeline experience preferred.
Python, Ruby, Golang, Java, C++, C#, Rust, Linux/Unix, CI/CD
1mo
Save
Mark Applied
Hide
Senior Engineering Manager, Cloud Infrastructure
San Francisco or Seattle or New York
$207k-$362k/yr OnsiteFull Time
Rippling
Rippling: Unified platform managing workforce HR, IT, and finance operations
8+ YOE4+ Mgmt8+ years infrastructure or software engineering, 4+ years engineering management, deep AWS and Kubernetes experience, Terraform and Helm proficiency, production incident and reliability expertise, strong cross-functional leadership.
AWS, Kubernetes, Terraform, Helm
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Machine Learning Infrastructure - Generative AI
San Francisco or Sunnyvale or Seattle
$137k-$202k/yr OnsiteFull Time
DoorDash
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
6+ YOE6+ years software engineering experience; BS/MS/PhD in CS or equivalent; deep backend fundamentals in Python and distributed systems; experience with LLM inference/fine-tuning, production reliability, observability, and technical leadership.
Python, Claude Code, Codex, Cursor, vLLM, SGLang, TensorRT-LLM, Kubernetes, AWS, GCP, Modal
1mo
Save
Mark Applied
Hide
Senior System Architect, Infrastructure Reliability
Santa Clara or Westford or Austin or Durham or Redmond
$184k-$357k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOE6+ years systems programming experience, BS/MS/PhD in CS or EE (or equivalent), expertise in CPU/GPU diagnostics, C++ and Python proficiency, experience with RCA, cluster managers (Slurm/LSF/Kubernetes).
C++, Python, Slurm, LSF, Kubernetes, NVIDIA DCGM, NVIDIA Management Library (NVML), CRIU, CUDA, /dev/mcelog, dmesg, journald
6d
Save
Mark Applied
Hide
Software Engineer, Developer Infrastructure - Build Systems
New York City or Los Angeles or Austin or Seattle
$130k-$230k/yr OnsiteFull Time
Nominal
Nominal: Software platform for testing and operating complex hardware systems.
4+ YOE4+ years building and maintaining build systems, experience with Bazel and large monorepos, improving CI/CD performance and release reliability, strong engineering judgment.
Bazel, Datadog, Grafana, OpenTelemetry, Kubernetes
1mo
Save
Mark Applied
Hide
Operations Engineering Manager, Fleet Reliability
Dallas or Bellevue
$143k-$191k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
7+ YOE2+ Mgmt7+ years in software or infrastructure engineering with 2+ years leadership; SRE fundamentals, incident management, observability, change management; strong automation and people development skills.