73 reliability engineering manager jobs at 33 companies in Salinas, CA
1mo
Save
Mark Applied
Hide
1mo
Reliability Engineering Manager - (M5)
Santa Clara, California, United States
$172k-$236k/yrOnsiteFull Time
Applied MaterialsNASDAQ: AMAT: Manufacturers of equipment for semiconductor and display production.
Managing reliability verification and qualification for semiconductor products, leading and developing reliability teams, subject-matter expertise in system-level reliability, familiarity with semiconductor fabrication equipment, and developing AI/Copilot automation.
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
5+ YOE3+ Mgmt5+ years network reliability engineering, 3+ years engineering/operations management, strong cloud networking and distributed systems expertise, proven people leadership, excellent communication and organizational skills.
Site Reliability Engineering Manager, Storage - Apple Services Engineering
Cupertino, California, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Engineering manager for distributed storage systems and large-scale storage infrastructure; experience in distributed systems and site reliability engineering.
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
18+ YOE10+ Mgmt18+ years in board and system reliability, 5+ years on data center equipment, 10+ years leading reliability management; deep reliability and physics-of-failure expertise; statistics and reliability modeling skills; bachelor’s in engineering or related (graduate preferred).
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
12+ YOE5+ MgmtDegree in CS/ECE/Physics/Math/Engineering or equivalent, 12+ years SRE/ITSM experience, 5+ years managing global IT/service teams, expertise in incident/problem/change management, observability, AI/automation, and executive communication.
5+ YOE3+ Mgmt5+ years software engineering, 3+ years managing engineering teams; hands-on Kubernetes, Terraform, AWS; experience with production reliability, SLOs, on-call and incident response.
NetflixNASDAQ: NFLX: Global video streaming and media production service.
Lead a team of distributed systems engineers; build and operate large-scale payments services; collaborate with product, regional leads, and stakeholders; drive reliability, scalability, and innovation.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
NetflixNASDAQ: NFLX: Provider of global streaming entertainment and video content.
3+ Mgmt3+ years managing distributed engineering teams, platform mindset, experience with reliable 24x7 services, strong communication, hiring/coaching, and ability to drive adoption of data platforms and governance.
graph-based data modeling, knowledge graphs, RDF, ontologies
Santa Clara or St. Louis or Bangalore or London or Paris or Melbourne or Taipei or Tokyo
OnsiteFull Time
NetskopeNASDAQ: NTSK: Cloud-native cybersecurity and data protection platform for enterprises.
3+ YOEBachelor's in CS/Engineering or equivalent; 3+ years building/managing complex systems (including 1-2 years SRE); experience with cloud services, microservices, availability/performance optimization, debugging, and strong communication.
Contract Lead, Site Reliability Engineering — AI Accelerator Infrastructure
Santa Clara, California, United States
$195k-$285k/yrHybridContract, Full Time
d-Matrix: Develops high-performance semiconductor chips for generative AI inference.
15+ YOE5+ MgmtBachelor's in CS/EE,15+ years SRE/infrastructure engineering,5+ years leading SRE teams,deep Linux,Terraform,Ansible,Kubernetes,Prometheus/Grafana/Datadog,Python or Go,cloud (AWS/Azure/GCP).
Altera: Manufacturer of field-programmable gate arrays and programmable logic devices.
15+ YOE5+ MgmtLead global hardware engineering; 15+ years in semiconductor hardware; manage US/Asia teams; expertise in packaging, board hardware, SI/PI, and reliability.
Ayar Labs: Develops optical interconnect technology for high-speed data movement.
7+ YOEMaster’s in EE/Physics/Materials; 7+ years semiconductor reliability with 4+ years in silicon photonics; foundry and packaging reliability expertise; PoF modeling and statistical analysis experience; familiarity with JEDEC, Telcordia, MIL-STD.
MarvellNASDAQ: MRVL: Designs and develops high-performance semiconductor and infrastructure solutions.
3+ YOEBachelor’s in CS/EE with 5-10 years in product qualification and reliability testing for semiconductors; HTOL/production burn-in experience; JEDEC/AEC standards; vendor management; data analysis.
Sr Manager, AI Systems Quality & Reliability , Annapurna AI Servers and Systems
Austin or Seattle or Cupertino
$208k-$282k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
10+ YOE5+ Mgmt10+ years reliability/quality engineering experience with server or high-volume electronics, 5+ years people management, bachelor's degree in a relevant field, experience with root-cause analysis, quality systems, and multi-vendor manufacturing.