127 systems reliability engineer jobs at 59 companies in Edmonds, WA
3w
Save
Mark Applied
Hide
3w
Senior Reliability Engineer – Hardware Systems
Seattle or California
$198k-$277k/yrOnsiteFull Time
Blue Origin: Develops reusable rockets and systems for human spaceflight.
7+ YOEBachelor's in engineering, 7+ years hardware reliability experience, HALT/HASS/ESS and environmental chamber experience, DFMEA and physics-of-failure modeling, proficiency with ReliaSoft/Minitab/JMP, and failure analysis skills.
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
6+ YOE6+ years software/network/systems engineering experience or equivalent degree+experience; experience with large-scale cloud services, observability, automation, and customer-facing reliability work.
Microsoft Fabric, Azure SQL DB, Azure Cosmos DB, Azure PostgreSQL, Azure Data Factory, Azure Synapse Analytics, Azure Service Bus, Azure Event Grid, Power BI, Logic Apps, Jupyter Notebooks
Pacific Northwest Defense Coalition: Trade association for the Northwest defense and security industry.
7+ YOE7+ years in safety and reliability engineering, experience with aerospace safety standards, FMEA/FTA/FRACAS, systems engineering, strong communication and customer-facing skills.
Grow Therapy: Platform connecting mental health providers with patients and insurance.
6+ YOE6+ years operating production systems; hands-on AWS, Kubernetes (EKS), Terraform; experience defining SLOs/SLAs and observability (DataDog); strong communication and systems-thinking skills; PostgreSQL experience a plus.
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
5+ YOETS/SCI with Polygraph and U.S. citizenship required. Bachelor's/Master's in CS or related, 5+ years SRE/Systems experience, expertise with Oracle Linux, Ansible, Terraform, Python, Bash, observability and distributed storage.
Senior Site Reliability Engineer - Data Infrastructure
Seattle, Washington, United States
$202k-$368k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
5+ YOE5+ years SRE/production engineering experience, proficiency with Go/Python/Bash, deep Linux, networking and distributed systems knowledge, bachelor’s degree or equivalent.
Sr Manager, AI Systems Quality & Reliability , Annapurna AI Servers and Systems
Austin or Seattle or Cupertino
$208k-$282k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
10+ YOE5+ Mgmt10+ years reliability/quality engineering experience with server or high-volume electronics, 5+ years people management, bachelor's degree in a relevant field, experience with root-cause analysis, quality systems, and multi-vendor manufacturing.
ByteDance: Developing AI-driven content platforms and mobile applications.
3+ YOEBachelor's degree in CS or related, 3+ years programming in C/C++/Java/Python/Go/Rust, familiarity with Unix/Linux, networking, and distributed systems.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
Site Reliability Engineer — HPC & Automation (Silicon Engineering)
Redmond, Washington, United States
$125k-$175k/yrOnsiteFull Time
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
2+ YOEBachelor's in CS/IS/engineering or 2+ years SRE/HPC/system administration experience; 1+ year dev experience with Bash/Python; 1+ year Linux experience; familiarity with containers, IaC, CI/CD, monitoring, databases, networking, and ASIC tool flows.
Senior Site Reliability Engineer, Platform Responsibility - USDS
Seattle, Washington, United States
$178k-$342k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
5+ YOE5+ years SRE/DevOps experience, BS in CS or related, strong Unix/Linux and distributed systems knowledge, experience with observability stacks and incident response, familiarity integrating AI/LLM into workflows.
Cowboy Space: Building satellite-based orbital infrastructure for AI data centers.
4+ YOE4+ years in software reliability or systems validation; strong Python/C++ skills; experience with distributed/embedded systems and automated validation frameworks.
Python, C++, Distributed Systems, Embedded Systems, CUDA
San Jose or San Francisco or Seattle or New York City or Chicago or California
$190k-$361k/yrOnsiteFull Time
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years professional software engineering with deep C++ expertise, production integration of AI/LLMs, cross‑platform systems design, reliability and observability, architecture leadership, and strong communication skills.
Airwallex: Global financial platform for business payments and money management.
5+ YOE5+ years software engineering; strong PostgreSQL and Redis experience; cloud (GCP) and IaC; building automation and platform tooling; strong system design.
Applications Engineer for High Reliability Compute Power
San Jose or Kirkland
OnsiteFull Time
Monolithic Power SystemsNasdaq: MPWR: Provides energy-efficient integrated power semiconductor solutions for electronics.
5+ YOE5-10 years in product definition, system engineering or applications for high reliability; BSEE or equivalent; PCB design and lab testing experience; travel up to 10%; experience with Space/Satellite PMICs is a plus.
PCB Design, Oscilloscopes, Multimeters, Power Management ICs, Lab Equipment
San Francisco or New York City or Seattle or Ann Arbor or United States or Canada
$151k-$206k/yrRemoteFull Time
Censys: Maps the internet to discover and manage security risks.
5+ YOE5+ years building distributed systems, strong Go and cloud experience (AWS/Azure/GCP), familiarity with messaging, databases, scalability and reliability concepts.