300 reliability engineering manager jobs at 168 companies in Albany, CA

2w
Save
Mark Applied
Hide
Director, Reliability Engineering
Menlo Park, California, United States
$200k-$265k/yr HybridFull Time
Mainspring Energy
Mainspring Energy: Private U.S. power infrastructure manufacturing and deploying fuel-flexible linear generators for utilities, data centers, and enterprises.
10+ YOE3+ MgmtBachelor's degree in electrical, mechanical, or related engineering; 10+ years in reliability, test, or quality engineering; 3+ years managing people; team-building and FMEA, HALT, RCA, CAPA expertise.
FMEA, HALT, CAPA, RCA, 8D, X-ray, CT, Python, MATLAB, Weibull++, JMP, UL, NFPA, IEC
1mo
Save
Mark Applied
Hide
Senior Director, Reliability Engineering
Santa Clara, California, United States
$332k-$500k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
18+ YOE10+ MgmtBachelor's degree in engineering, physics, materials science, or related field; 18+ years in board and system reliability; 5+ years in data center reliability; 10+ years leading reliability management.
NUDD, FMEA, finite element analysis
1mo
Save
Mark Applied
Hide
Senior Director, Reliability Engineering
Santa Clara, California, United States
$332k-$500k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
18+ YOE10+ Mgmt18+ years in board and system reliability, 5+ years on data center equipment, 10+ years leading reliability management; deep reliability and physics-of-failure expertise; statistics and reliability modeling skills; bachelor’s in engineering or related (graduate preferred).
1mo
Save
Mark Applied
Hide
Manager, Site Reliability Engineering
San Francisco, California, United States
$204k-$306k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Identity management and access control software provider.
3+ Mgmt3+ years technical leadership experience; experience with cloud-native architectures, Kubernetes, Terraform, CI/CD, observability platforms; strong software development and automation background; US Person status required.
Amazon Web Services (AWS), Kubernetes, Terraform, Grafana, Splunk, APM, CI/CD
1mo
Save
Mark Applied
Hide
Senior Director, Reliability Engineering
Santa Clara, California, United States
$332k-$500k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
18+ YOE10+ Mgmt18+ years in board/system reliability with 5+ years on data center equipment and 10+ years leading reliability; deep reliability, testing, modeling, statistics, and physics-of-failure expertise.
1d
Save
Mark Applied
Hide
Senior Manager, Reliability Engineering & AIOps
Fremont, California, United States
$137k-$287k/yr HybridFull Time
Lam Research
Lam ResearchNasdaq: LRCX: Global supplier of wafer fabrication equipment for semiconductors.
10+ YOEBachelor's degree with 10 years of related experience, master's with 8 years, or equivalent; reliability team leadership, incident command, disaster recovery, paging platforms, cloud infrastructure, observability, infrastructure as code, and Python or Go.
Azure, AWS, GCP, Microsoft Copilot, Cursor, GitHub Copilot, PagerDuty, Prometheus, Grafana, Loki, Tempo, Terraform, Python, Go
1w
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineering
Mountain View or Mountain View or California or United States
$222k-$301k/yr OnsiteFull Time
Intuit
IntuitNASDAQ: INTU: A global financial technology platform powering prosperity.
8+ YOE3+ Mgmt8+ years in systems, SRE, or infrastructure engineering; 3+ years managing engineering teams; AWS at scale; distributed systems, Kubernetes, IaC, observability, incident management, and AI Ops experience; bachelor's degree required.
AWS, Amazon EC2, Amazon EKS, Amazon ECS, Amazon VPC, Amazon RDS, Amazon DynamoDB, AWS IAM, Amazon CloudWatch, AWS Auto Scaling, Kubernetes, Terraform, AWS CloudFormation, Datadog, Splunk, PagerDuty, Prometheus, Grafana, AI Ops, AIOps platforms
3w
Save
Mark Applied
Hide
Reliability Engineering Technical Leader
San Jose, California, United States
$163k-$205k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Global leader in networking, cybersecurity, and cloud-native technology solutions.
8+ YOEBachelor's in Engineering with 12+ years or Master's with 8+ years; deep hardware reliability and PCBA knowledge; expertise in RAS, risk management, data-driven reliability, and executive influence; proven leadership and mentoring skills.
4d
Save
Mark Applied
Hide
Manager Systems Engineering 2 – Reliability
Sunnyvale, California, United States
$161k-$242k/yr OnsiteFull Time
Northrop Grumman
Northrop GrummanNYSE: NOC: Develops and manufactures advanced aerospace, defense, and space systems.
9+ YOEBachelor’s STEM degree with 9 years, master’s with 7 years, or PhD with 4 years of engineering experience; project management, reliability and system safety expertise, failure analysis, and active US Secret clearance required.
Web, Model Based Systems Engineering (MBSE), Agile, SCRUM
4d
Save
Mark Applied
Hide
Manager Systems Engineering 2 – Reliability
Sunnyvale, California, United States
$161k-$242k/yr OnsiteFull Time
Northrop Grumman
Northrop GrummanNYSE: NOC: Global aerospace and defense technology.
9+ YOESTEM bachelor's degree with 9 years, master's with 7, or PhD with 4 years of engineering experience; project management, reliability and system safety expertise, failure analysis, statistical analysis, Secret clearance, and US citizenship required.
Model Based Systems Engineering (MBSE), Agile, SCRUM
2mo
Save
Mark Applied
Hide
Manager, Software Engineering (Reliability Platform)
United States or California or Washington or New York or New Jersey or Connecticut or Los Angeles or San Francisco
$204k-$290k/yr RemoteFull Time
Affirm
AffirmNASDAQ: AFRM: Financial technology offering point-of-sale payment solutions.
7+ YOE2+ Mgmt7+ years backend/full-stack engineering experience with 2+ years engineering leadership; SRE/production engineering experience; observability and platform-building experience; strong programming (Python, Kotlin, Java); Bachelor\u0002s degree or equivalent experience.
Python, Kotlin, Java
6d
Save
Mark Applied
Hide
Director, Engineering, Maintenance & Reliability
Sacramento or Pittsburg
OnsiteFull Time
Nivagen Pharmaceuticals
Nivagen Pharmaceuticals: Specialty pharma manufacturer of generic and branded prescription drugs, sterile injectables, and OTC products for North America.
12+ YOE5+ MgmtBachelor's degree in engineering or related technical field, 12+ years regulated manufacturing experience, 5+ years leadership preferred, sterile manufacturing expertise, driver's license, and U.S. work authorization.
PLC, HMI, SCADA, data historians, Microsoft Excel
20h
Save
Mark Applied
Hide
Senior Manager, Reliability Engineering & AIOps
Fremont or Japan or Singapore or Malaysia or India or Korea
$137k-$287k/yr HybridFull Time
Lam Research
Lam ResearchNasdaq: LRCX: Global supplier of wafer fabrication equipment for semiconductors.
10+ YOEBachelor's degree with 10 years' related experience, master's with 8 years, or equivalent. Requires reliability leadership, incident command, disaster recovery, cloud infrastructure, observability, infrastructure as code, and Python or Go.
Microsoft Azure, AWS, GCP, Microsoft Copilot, Cursor, GitHub Copilot, PagerDuty, Prometheus, Grafana, Loki, Tempo, Terraform, Python, Go
1mo
Save
Mark Applied
Hide
Principal Tech Lead Manager - Data Platform & Reliability Engineering
Mountain View, California, United States
$215k-$275k/yr OnsiteFull Time
ID.me
ID.me: Private American digital identity wallet and identity-verification helping people securely access government, healthcare, and commercial services.
5+ YOE3+ Mgmt8+ years engineering experience with 3+ years managing teams,5+ years in data/platform/SRE; bachelor\u0002s or equivalent; deep PostgreSQL and data reliability expertise; strong communication and cloud/IaC experience.
PostgreSQL, Neo4j, Amazon Neptune, Kafka, Kinesis, Kubernetes, Terraform, Helm, AWS
3d
Save
Mark Applied
Hide
Site Reliability Engineering Manager, Vehicle Software
Sunnyvale, California, United States
$276k-$294k/yr HybridFull Time
Wayve
Wayve: British autonomous-driving software licensing vehicle-agnostic AI Driver technology to automakers and fleet owners.
8+ YOE3+ MgmtRequires 8+ years building production software systems, 3+ years of people leadership, SRE and reliability expertise, architecture experience, and hands-on coding in C++, Rust, Python, or Go.
C++, Rust, Python, Go, Linux, CI/CD
3d
Save
Mark Applied
Hide
Site Reliability Engineering Manager, Vehicle Software
Sunnyvale, California, United States
$276k-$294k/yr HybridFull Time
Wayve
Wayve: British autonomous-driving software licensing vehicle-agnostic AI Driver technology to automakers and fleet owners.
8+ YOE3+ MgmtRequires 8+ years building production software systems, 3+ years people leadership, SRE and reliability practices, architecture experience, and production coding in C++, Rust, Python, or Go.
Linux, C++, Rust, Python, Go, CI/CD
2d
Save
Mark Applied
Hide
RELIABILITY ENGINEER
Carolina or Caguas or Boston or San Francisco or United States or Puerto Rico or Dominican Republic or Mexico or Germany or Canada or South America
FieldFull Time
Mentor Technical Group
Mentor Technical Group: Technical, engineering, and compliance solutions for life sciences.
Bachelor's degree in engineering required, mechanical engineering preferred. Reliability, vibration analysis, thermography, or project management certifications are preferred or advantageous.
Computer Maintenance Management System (CMMS)
2mo
Save
Mark Applied
Hide
Director of Platform & Reliability Engineering
New York City or San Francisco
$235k-$245k/yr HybridFull Time
Forge Global
Forge GlobalNYSE: FRGE: Financial technology operating a private-market marketplace and data, custody, and investment solutions for companies and investors.
8+ YOE5+ Mgmt8+ years software engineering experience with infrastructure/platform focus, 5+ years people leadership, cloud and infrastructure as code experience, observability and incident response expertise, Bachelor's in CS or equivalent, strong communication.
Kubernetes, CI/CD
4w
Save
Mark Applied
Hide
Site Reliability Engineering (SRE) Manager, Apple Maps
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designing and manufacturing consumer electronics, software, and digital services.
Build, manage, and deliver highly available, automated infrastructure for Apple Maps at global scale; focus on reliability, scalability, and operational excellence.
1mo
Save
Mark Applied
Hide
Engineering Manager, Observability
Sunnyvale, California, United States
$182k-$242k/yr OnsiteFull Time
CoreWeave
CoreWeaveNasdaq: CRWV: Specialized cloud provider for large-scale AI and machine learning.
5+ YOE2+ Mgmt5+ years software engineering, 2+ years engineering management, experience with observability platforms, reliability engineering, scaling telemetry, and hiring/managing teams.
OpenTelemetry, Grafana, Prometheus, Kubernetes

Explore Jobs

Expand Your Job Search