Seattle or San Francisco or Detroit or United States
$180k-$279k/yrHybridFull Time
Rocket CompaniesNYSE: RKT: Provides digital mortgage, real estate, and personal finance services.
7+ YOE7+ years AWS/cloud infra; 5+ years PostgreSQL/AWS services; Linux admin/scripting; mentoring; infrastructure as code and security; AI code generation tools; on-call readiness.
AWS, PostgreSQL, Aurora/RDS, S3, ElastiCache, OpenSearch, DynamoDB, Linux, Python, Infrastructure as Code, Security practices, AI code generation tools
Austin or New York City or San Francisco or Seattle
$203k-$232k/yrOnsiteFull Time
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience in reliability engineering for infrastructure or complex hardware, building availability/RAM models, leading cross-discipline FMEAs, and mining field failure data.
Aurelian: AI-powered voice automation for 9-1-1 non-emergency calls.
4+ YOE4+ years in infrastructure/platform/backend engineering, experience with reliability and scale, comfortable across backend and cloud, experience building analytics/observability/developer tooling.
Senior Site Reliability Engineer - Data Infrastructure
Seattle, Washington, United States
$202k-$368k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
5+ YOE5+ years SRE/production engineering experience, proficiency with Go/Python/Bash, deep Linux, networking and distributed systems knowledge, bachelor’s degree or equivalent.
Sr. Hardware / Infrastructure Site Reliability Engineer (Starlink)
Redmond, Washington, United States
$165k-$230k/yrOnsiteFull Time
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOEBachelor's in CS/IT/engineering plus 5+ years SRE/DevOps (or 7+ years without degree); 2+ years Linux; experience with Terraform/Ansible, Docker/Kubernetes; Bash, Python, Go/C++/C development; strong OS, networking, CI/CD skills.
Linux, Terraform, Ansible, Docker, Kubernetes, Bash, Python, Go, C++, C
Canada or United States or Seattle or Paris or New York City
$238k-$382k/yrRemoteFull Time
Docker: Provides a platform for building, sharing, and running containerized applications.
8+ YOE8+ years backend/infrastructure engineering; strong Go; experience with Kubernetes/EKS, cloud platforms, networking, and reliability; Bachelor's or equivalent; experience leading cross-team technical initiatives and strong written/verbal communication.
ByteDance: Developing AI-driven content platforms and mobile applications.
5+ YOE5+ years SRE or software engineering experience designing, building, scaling, and operating cloud-based systems; expertise in databases, Kubernetes, distributed systems; strong communication and collaboration skills.
RobinhoodNASDAQ: HOOD: Provides a commission-free platform for investing and financial services.
1+ YOEMaster's in Computer Science or related field plus one year of experience; expertise in production network infrastructure, BGP/OSPF, network automation (Python or Golang), and monitoring.
Senior Engineer 2 - Site Reliability Engineering (Hybrid, Seattle)
Seattle, Washington, United States
$166k-$258k/yrHybridFull Time
Nordstrom: Operates luxury department stores and off-price retail outlets.
10+ YOEBachelor's in CS/Engineering or equivalent,10+ years software engineering experience in SRE/infrastructure,proficiency with Kubernetes,cloud providers,networking,strong problem-solving and communication skills.
Sesame: Designing wearable computers with lifelike voice-driven AI agents.
3+ YOEStrong systems thinker with reliability engineering experience; 3+ years in infrastructure, platform, or ML systems; Kubernetes production experience; strong communication.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
Senior Site Reliability Engineer - Core Cloud Platform
San Francisco or San Jose or Bellevue
$240k-$356k/yrHybridFull Time
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
7+ YOE7+ years SRE or production infrastructure experience, deep Kubernetes and Terraform knowledge, experience with observability and SLOs, proficiency in Go or Python, on-call and incident leadership experience.
Rippling: Unified platform managing workforce HR, IT, and finance operations
8+ YOE4+ Mgmt8+ years infrastructure or software engineering, 4+ years engineering management, deep AWS and Kubernetes experience, Terraform and Helm proficiency, production incident and reliability expertise, strong cross-functional leadership.
Senior Software Engineer, Machine Learning Infrastructure - Generative AI
San Francisco or Sunnyvale or Seattle
$137k-$202k/yrOnsiteFull Time
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
6+ YOE6+ years software engineering experience; BS/MS/PhD in CS or equivalent; deep backend fundamentals in Python and distributed systems; experience with LLM inference/fine-tuning, production reliability, observability, and technical leadership.
Senior System Architect, Infrastructure Reliability
Santa Clara or Westford or Austin or Durham or Redmond
$184k-$357k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOE6+ years systems programming experience, BS/MS/PhD in CS or EE (or equivalent), expertise in CPU/GPU diagnostics, C++ and Python proficiency, experience with RCA, cluster managers (Slurm/LSF/Kubernetes).
Software Engineer, Developer Infrastructure - Build Systems
New York City or Los Angeles or Austin or Seattle
$130k-$230k/yrOnsiteFull Time
Nominal: Software platform for testing and operating complex hardware systems.
4+ YOE4+ years building and maintaining build systems, experience with Bazel and large monorepos, improving CI/CD performance and release reliability, strong engineering judgment.
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
7+ YOE2+ Mgmt7+ years in software or infrastructure engineering with 2+ years leadership; SRE fundamentals, incident management, observability, change management; strong automation and people development skills.