259 systems reliability engineer jobs at 172 companies in Ross, CA
2w
Save
Mark Applied
Hide
2w
Systems Reliability Engineer (SRE)
San Francisco or New York City
$150k-$170k/yrOnsiteFull Time
Claryo: AI-powered spatial software for optimizing warehouse operations
3+ YOE3+ years SRE/infrastructure experience, strong Linux and networking fundamentals, experience with Kubernetes, cloud platforms, observability tooling, and debugging distributed systems in production.
Lawrence Livermore National Laboratory: Develops science and technology for United States national security.
Bachelor’s degree in engineering or equivalent education and experience; engineering troubleshooting experience; complex systems analysis; DOE Q clearance eligibility requiring U.S. citizenship; SES.3 requires advanced technical and leadership experience.
ThoughtSpot: AI-powered analytics platform for enterprise business intelligence.
Experience troubleshooting Linux systems and cloud platforms, hands-on with monitoring tools, on-call/incident management experience, scripting in Python/Go/Bash/Java, B.S. in CS or equivalent preferred.
Mainspring Energy: Manufactures fuel-flexible linear generators for onsite power generation.
5+ YOEBachelor in Electrical or Mechanical Engineering; 5+ years in electrical hardware reliability testing; experience with FMEA, CAPA; data analysis skills (Python, Matlab, JMP); familiar with UL/IEC standards; strong communication and cross-functional collaboration.
Python, Matlab, Weibull++, JMP, ECad tools, Oscilloscope, Function Generator, DMM
Form Energy: Developing multi-day batteries for grid-scale energy storage.
4+ YOEBachelor's degree in relevant engineering field with 4+ years industry experience, expertise in designing reliability tests and accelerated lifetime testing, experience with liquid/gas handling systems, strong communication and cross-functional collaboration.
Eight Sleep: Develops smart mattresses and temperature-regulated sleep technology.
5+ YOE5+ years reliability engineering in electromechanical systems; experience with DOE, FMEA, Weibull, SPC; Python/MATLAB; hardware test equipment; strong collaboration.
Rivian and Volkswagen Group Technologies: Joint venture focused on automotive electronics and software-defined vehicles.
5+ YOE5–10 years in reliability engineering; automotive electronics or embedded systems preferred; BS in engineering; strong data and reliability methods; SQL/Python, Databricks; dashboard tools.
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
4+ YOEAssociate degree in engineering, 4+ years in FDA/ISO-regulated environment, experience in product failure analysis, hardware/software integration, embedded systems and electrical design, strong problem-solving and project management skills.
Beast Industries: Produces digital media and consumer goods for MrBeast brands.
Expert in software quality engineering and site reliability for consumer-scale distributed systems; owns test strategy, SLOs/error budgets, CI/CD test gates, observability, incident response, and reliability tooling.
Dedalus Labs: Infrastructure for building and deploying AI agent applications.
Fluency in Rust/Go/C/C++; strong software engineering fundamentals and systems knowledge (OS, networking, distributed systems); performance-oriented debugging and reliability focus.
Zipline: Operates an autonomous drone delivery system for medical supplies.
B.S. or M.S. in engineering, strong mechanical fundamentals, hands-on experience with safety-critical electromechanical systems, DFMEA/ALT/environmental testing/failure analysis experience, and strong technical communication.
Thinking Machines: Building AI systems to extend human will and judgment.
Experience in distributed systems/cloud/site reliability, software automation for reliability, incident response and postmortems, strong communication and coordination skills.
Parasail: Provides scalable cloud infrastructure for AI model inference.
5+ YOE5+ years production engineering experience operating customer-facing systems; strong SRE and production diagnostics skills; Kubernetes, Linux, distributed systems, and software engineering proficiency; ability to lead incident response and build observability.
OusterNASDAQ: OUST: Designs and manufactures high-resolution digital lidar sensors.
8+ YOE8+ years designing and executing reliability and verification strategies for rugged electromechanical and compute systems; FMEA/RCA experience; reliability physics, HALT, thermal and vibration testing; functional safety knowledge preferred.
Stuut: Automates business accounts receivable and collections through AI agents.
7+ YOE7+ years in SRE/infrastructure or backend engineering. Experience with AWS, Kubernetes/EKS, Docker, observability, SLOs/SLIs, Python or TypeScript, CI/CD, and production-grade distributed systems.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
5+ YOE5+ years systems and software engineering experience for large-scale internet services; expertise in SRE principles, containers, observability, incident management, Python and Go, and applying AI/ML to operations.
EarnIn: Provides immediate access to earned wages through a mobile app.
3+ YOE3+ years SRE or related experience; hands-on Python/Go coding; experience with observability, incident response, SLOs/SLIs, and distributed systems; strong communication and documentation skills.