174 systems reliability engineer jobs at 125 companies in Vacaville, CA
1w
Save
Mark Applied
Hide
1w
Systems Reliability Engineer (SRE)
San Francisco or New York City
$150k-$170k/yrOnsiteFull Time
Claryo: AI-powered spatial software for optimizing warehouse operations
3+ YOE3+ years SRE/infrastructure experience, strong Linux and networking fundamentals, experience with Kubernetes, cloud platforms, observability tooling, and debugging distributed systems in production.
Lawrence Livermore National Laboratory: Develops science and technology for United States national security.
Bachelor’s degree in engineering or equivalent education and experience; engineering troubleshooting experience; complex systems analysis; DOE Q clearance eligibility requiring U.S. citizenship; SES.3 requires advanced technical and leadership experience.
Form Energy: Developing multi-day batteries for grid-scale energy storage.
4+ YOEBachelor's degree in relevant engineering field with 4+ years industry experience, expertise in designing reliability tests and accelerated lifetime testing, experience with liquid/gas handling systems, strong communication and cross-functional collaboration.
Eight Sleep: Develops smart mattresses and temperature-regulated sleep technology.
5+ YOE5+ years reliability engineering in electromechanical systems; experience with DOE, FMEA, Weibull, SPC; Python/MATLAB; hardware test equipment; strong collaboration.
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
4+ YOEAssociate degree in engineering, 4+ years in FDA/ISO-regulated environment, experience in product failure analysis, hardware/software integration, embedded systems and electrical design, strong problem-solving and project management skills.
Beast Industries: Produces digital media and consumer goods for MrBeast brands.
Expert in software quality engineering and site reliability for consumer-scale distributed systems; owns test strategy, SLOs/error budgets, CI/CD test gates, observability, incident response, and reliability tooling.
Dedalus Labs: Infrastructure for building and deploying AI agent applications.
Fluency in Rust/Go/C/C++; strong software engineering fundamentals and systems knowledge (OS, networking, distributed systems); performance-oriented debugging and reliability focus.
Thinking Machines: Building AI systems to extend human will and judgment.
Experience in distributed systems/cloud/site reliability, software automation for reliability, incident response and postmortems, strong communication and coordination skills.
OusterNASDAQ: OUST: Designs and manufactures high-resolution digital lidar sensors.
8+ YOE8+ years designing and executing reliability and verification strategies for rugged electromechanical and compute systems; FMEA/RCA experience; reliability physics, HALT, thermal and vibration testing; functional safety knowledge preferred.
Stuut: Automates business accounts receivable and collections through AI agents.
7+ YOE7+ years in SRE/infrastructure or backend engineering. Experience with AWS, Kubernetes/EKS, Docker, observability, SLOs/SLIs, Python or TypeScript, CI/CD, and production-grade distributed systems.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
5+ YOE5+ years systems and software engineering experience for large-scale internet services; expertise in SRE principles, containers, observability, incident management, Python and Go, and applying AI/ML to operations.
Ivo: AI-powered contract review and intelligence platform for legal teams.
5+ YOEMinimum 5 years experience; own uptime and reliability, define SLIs/SLOs/SLAs, design failover and disaster recovery, implement security controls, lead incident response; familiarity with cloud and LLM-driven systems.
Grow Therapy: Platform connecting mental health providers with patients and insurance.
6+ YOE6+ years operating production systems; hands-on AWS, Kubernetes (EKS), Terraform; experience defining SLOs/SLAs and observability (DataDog); strong communication and systems-thinking skills; PostgreSQL experience a plus.
Retool: Software platform for building custom internal business applications.
Experience operating production infrastructure (AWS), Kubernetes, Terraform, Postgres; programming in Go/Python/TypeScript/Java/Ruby; building observability and automation for customer-facing SaaS systems.
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yrRemoteFull Time
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
CircleNYSE: CRCL: Digital currency issuer and blockchain financial infrastructure provider.
6+ YOE6+ years SRE/DevOps experience preferred; strong Kubernetes, IaC (Terraform/Pulumi), cloud networking, Go or Python, CI/CD, observability, and distributed systems experience; blockchain node operation experience is a plus.
Scottsdale or San Francisco or Chicago or New York
$66k-$82k/yrHybridFull Time
Early Warning Services: Operates payment and risk solutions for the financial industry.
2+ YOEBachelor's degree in CS/IS; 2-5 years in IT/DevOps; ITIL/ITSM knowledge; Agile; strong problem solving; ability to relate business needs to system capabilities.
ITIL, ITSM, Agile, Networking, Distributed Systems