82 reliability engineering manager jobs at 43 companies in Lathrop, CA
1mo
Save
Mark Applied
Hide
1mo
Reliability Engineering Manager - (M5)
Santa Clara, California, United States
$172k-$236k/yrOnsiteFull Time
Applied MaterialsNASDAQ: AMAT: Manufacturers of equipment for semiconductor and display production.
Managing reliability verification and qualification for semiconductor products, leading and developing reliability teams, subject-matter expertise in system-level reliability, familiarity with semiconductor fabrication equipment, and developing AI/Copilot automation.
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
5+ YOE3+ Mgmt5+ years network reliability engineering, 3+ years engineering/operations management, strong cloud networking and distributed systems expertise, proven people leadership, excellent communication and organizational skills.
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
18+ YOE10+ Mgmt18+ years in board and system reliability, 5+ years on data center equipment, 10+ years leading reliability management; deep reliability and physics-of-failure expertise; statistics and reliability modeling skills; bachelor’s in engineering or related (graduate preferred).
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
18+ YOE10+ Mgmt18+ years in board/system reliability with 5+ years on data center equipment and 10+ years leading reliability; deep reliability, testing, modeling, statistics, and physics-of-failure expertise.
5+ YOE3+ Mgmt5+ years software engineering, 3+ years managing engineering teams; hands-on Kubernetes, Terraform, AWS; experience with production reliability, SLOs, on-call and incident response.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
SandiskNasdaq: SNDK: Designs and manufactures flash memory and data storage products.
6+ YOEBachelor's degree in engineering/CS, 6+ years engineering experience in test infrastructure or reliability, hands-on with environmental chambers, failure analysis, vendor management, and global lab alignment.
Santa Clara or St. Louis or Bangalore or London or Paris or Melbourne or Taipei or Tokyo
OnsiteFull Time
NetskopeNASDAQ: NTSK: Cloud-native cybersecurity and data protection platform for enterprises.
3+ YOEBachelor's in CS/Engineering or equivalent; 3+ years building/managing complex systems (including 1-2 years SRE); experience with cloud services, microservices, availability/performance optimization, debugging, and strong communication.
Contract Lead, Site Reliability Engineering — AI Accelerator Infrastructure
Santa Clara, California, United States
$195k-$285k/yrHybridContract, Full Time
d-Matrix: Develops high-performance semiconductor chips for generative AI inference.
15+ YOE5+ MgmtBachelor's in CS/EE,15+ years SRE/infrastructure engineering,5+ years leading SRE teams,deep Linux,Terraform,Ansible,Kubernetes,Prometheus/Grafana/Datadog,Python or Go,cloud (AWS/Azure/GCP).
Altera: Manufacturer of field-programmable gate arrays and programmable logic devices.
15+ YOE5+ MgmtLead global hardware engineering; 15+ years in semiconductor hardware; manage US/Asia teams; expertise in packaging, board hardware, SI/PI, and reliability.
Ayar Labs: Develops optical interconnect technology for high-speed data movement.
7+ YOEMaster’s in EE/Physics/Materials; 7+ years semiconductor reliability with 4+ years in silicon photonics; foundry and packaging reliability expertise; PoF modeling and statistical analysis experience; familiarity with JEDEC, Telcordia, MIL-STD.
MarvellNASDAQ: MRVL: Designs and develops high-performance semiconductor and infrastructure solutions.
3+ YOEBachelor’s in CS/EE with 5-10 years in product qualification and reliability testing for semiconductors; HTOL/production burn-in experience; JEDEC/AEC standards; vendor management; data analysis.
Lam ResearchNASDAQ: LRCX: Designs and manufactures wafer fabrication equipment for the semiconductor industry.
10+ YOE5+ MgmtBachelor's in Electrical or RF Engineering; 10+ years RF engineering experience and 5+ years engineering management; expertise in RF architecture, modeling, simulation, test, reliability, and cross‑functional program execution; strong leadership and people development skills.
BS or MS in CS or related field; expertise in configuration management (Ansible, Terraform, Kubernetes); Python and/or Go; Kubernetes with autoscaling; production engineering/DevOps/SRE experience; public cloud (GCP/AWS); Linux networking; CI/CD with GitLab/GitHub; distributed systems; strong communication; ownership and monitoring as code.