225 reliability engineering manager jobs at 145 companies in El Cerrito, CA

1mo
Save
Mark Applied
Hide
Senior Director, Reliability Engineering
Santa Clara, California, United States
$332k-$500k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
18+ YOE10+ Mgmt18+ years in board and system reliability, 5+ years on data center equipment, 10+ years leading reliability management; deep reliability and physics-of-failure expertise; statistics and reliability modeling skills; bachelor’s in engineering or related (graduate preferred).
1mo
Save
Mark Applied
Hide
Manager, Site Reliability Engineering
San Francisco, California, United States
$204k-$306k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
3+ Mgmt3+ years technical leadership experience; experience with cloud-native architectures, Kubernetes, Terraform, CI/CD, observability platforms; strong software development and automation background; US Person status required.
Amazon Web Services (AWS), Kubernetes, Terraform, Grafana, Splunk, APM, CI/CD
1mo
Save
Mark Applied
Hide
Senior Director, Reliability Engineering
Santa Clara, California, United States
$332k-$500k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
18+ YOE10+ Mgmt18+ years in board/system reliability with 5+ years on data center equipment and 10+ years leading reliability; deep reliability, testing, modeling, statistics, and physics-of-failure expertise.
1w
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineering
Irving or San Leandro
OnsiteFull Time
Wells Fargo
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
7+ YOE3+ MgmtRequires 7+ years in systems engineering or technology architecture, 3+ years of management, 5+ years leading engineering or SRE teams, and experience with customer-facing platforms, incident management, SRE, DevOps, cloud, and production operations.
Splunk, Grafana, AppDynamics, Dynatrace, OpenTelemetry, Prometheus, Kubernetes, OpenShift, AWS, Microsoft Azure, Google Cloud Platform, Infrastructure as Code (IaC), CI/CD, AIOps, ITIL
1w
Save
Mark Applied
Hide
Reliability Engineering Technical Leader
San Jose, California, United States
$163k-$205k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
8+ YOEBachelor's in Engineering with 12+ years or Master's with 8+ years; deep hardware reliability and PCBA knowledge; expertise in RAS, risk management, data-driven reliability, and executive influence; proven leadership and mentoring skills.
1mo
Save
Mark Applied
Hide
Manager, Software Engineering (Reliability Platform)
United States or California or Washington or New York or New Jersey or Connecticut or Los Angeles or San Francisco
$204k-$290k/yr RemoteFull Time
Affirm
AffirmNasdaq: AFRM: Financial platform providing installment loans for consumer purchases.
7+ YOE2+ Mgmt7+ years backend/full-stack engineering experience with 2+ years engineering leadership; SRE/production engineering experience; observability and platform-building experience; strong programming (Python, Kotlin, Java); Bachelor\u0002s degree or equivalent experience.
Python, Kotlin, Java
1mo
Save
Mark Applied
Hide
Principal Tech Lead Manager - Data Platform & Reliability Engineering
Mountain View, California, United States
$215k-$275k/yr OnsiteFull Time
ID.me
ID.me: Provides secure digital identity verification and authentication services.
5+ YOE3+ Mgmt8+ years engineering experience with 3+ years managing teams,5+ years in data/platform/SRE; bachelor\u0002s or equivalent; deep PostgreSQL and data reliability expertise; strong communication and cloud/IaC experience.
PostgreSQL, Neo4j, Amazon Neptune, Kafka, Kinesis, Kubernetes, Terraform, Helm, AWS
2mo
Save
Mark Applied
Hide
Director of Platform & Reliability Engineering
New York City or San Francisco
$235k-$245k/yr HybridFull Time
Forge Global
Forge GlobalNYSE: FRGE: Marketplace for trading private shares and pre-IPO stock.
8+ YOE5+ Mgmt8+ years software engineering experience with infrastructure/platform focus, 5+ years people leadership, cloud and infrastructure as code experience, observability and incident response expertise, Bachelor's in CS or equivalent, strong communication.
Kubernetes, CI/CD
2w
Save
Mark Applied
Hide
Site Reliability Engineering (SRE) Manager, Apple Maps
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Build, manage, and deliver highly available, automated infrastructure for Apple Maps at global scale; focus on reliability, scalability, and operational excellence.
3w
Save
Mark Applied
Hide
Engineering Manager, Observability
Sunnyvale, California, United States
$182k-$242k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
5+ YOE2+ Mgmt5+ years software engineering, 2+ years engineering management, experience with observability platforms, reliability engineering, scaling telemetry, and hiring/managing teams.
OpenTelemetry, Grafana, Prometheus, Kubernetes
1mo
Save
Mark Applied
Hide
Engineering Manager
Berlin or Toronto or San Francisco or New York City or London or Paris or Montreal or Seoul or Germany
HybridFull Time
Cohere
Cohere: Provides enterprise-grade large language models and AI software platforms.
Proven experience managing engineering teams, building and shipping production AI-first full-stack applications, partnering with product and ML teams, and delivering secure, reliable systems.
2mo
Save
Mark Applied
Hide
Lead Site Reliability Engineering - Network
Palo Alto or Columbus
$152k-$215k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE10+ MgmtFormal network engineering training, 5+ years applied experience, 10+ years leading technologists, advanced network reliability skills, SD-WAN and cloud (AWS, Azure) proficiency, major network vendor experience, observability tooling and incident leadership.
SD-WAN, AWS, Azure, Palo Alto, Juniper, F5, Broadcom, Arista, Cisco, Grafana, SevOne, Prometheus, Kibana, ThousandEyes, Splunk, Jenkins, GitLab, Terraform, eBPF, TCP/IP, HTTPS, BGP
2mo
Save
Mark Applied
Hide
Engineering Manager
San Francisco, California, United States
$250k-$300k/yr OnsiteFull Time
LiteLLM
LiteLLM: Open-source AI gateway for standardizing LLM API access.
Proven engineering manager with experience shipping through teams, metrics-driven reliability/process design, incident/RCAs, hiring and scaling, customer-facing escalation handling, and familiarity with Rust migrations and startup/open-source environments.
Rust, Python
3d
Save
Mark Applied
Hide
Engineering Manager, Verifications
Denver or San Francisco or Nashville or Santiago
$197k-$274k/yr HybridFull Time
Checkr
Checkr: AI-powered platform for background checks and identity verification.
8+ YOE4+ MgmtRequires 4+ years managing engineering teams, 8+ years as a software engineer, service architecture and distributed systems expertise, integration reliability, incident management, regulated products, and strong stakeholder communication.
AI-assisted development, agentic coding tools
1mo
Save
Mark Applied
Hide
Engineering Manager (Platform)
San Francisco, California, United States
OnsiteFull Time
Rec Technologies
Rec Technologies: Modern software platform for parks and recreation departments.
5+ YOE4+ Mgmt5+ years as a software engineer with 4+ years managing engineering teams, AI fluency, experience with platform reliability or payments, strong coding and hiring skills, and ability to coach engineers and shape roadmap.
Next.js, Vercel, React, TypeScript, Node.js, Koa, Objection.js, PostgreSQL, AWS, Temporal, Twilio, Stripe, Claude Code, Codex, Cursor
1w
Save
Mark Applied
Hide
Senior Manager, Hardware Reliability & Test
San Jose, California, United States
$202k-$223k/yr HybridFull Time
Muon Space
Muon Space: Designs, builds, and operates low Earth orbit satellite constellations.
10+ YOE3+ Mgmt10+ years in hardware test/qualification or reliability engineering with 3+ years leading teams, BS in engineering or CS, hands-on spaceflight or environmental test experience, reliability analysis knowledge, and strong communication skills.
2mo
Save
Mark Applied
Hide
Engineering Manager, Platform
San Francisco or United States or Canada
$255k-$295k/yr RemoteFull Time
Render
Render: Cloud platform for building, deploying, and scaling AI-native applications.
8+ YOE4+ Mgmt8+ years building infrastructure or platform products for developers, 4+ years managing engineers, strong system design, reliability, CI/CD, observability, configuration management, and developer experience.
1mo
Save
Mark Applied
Hide
Engineering Manager
New York City or San Francisco or United States
$128k-$183k/yr RemoteFull Time
Pathstream
Pathstream: Online workforce development and professional certificate programs.
4+ YOE2+ Mgmt4+ years software engineering experience, 2+ years leading engineers, proficiency in modern web stacks and cloud (Ruby on Rails, React, JS/TS, Python, Docker, PostgreSQL, AWS), experience with production reliability, security, and AI-enabled development tools.
Ruby on Rails, React, JavaScript, TypeScript, Python, Docker, PostgreSQL, AWS, Claude Code
1d
Save
Mark Applied
Hide
Engineering Manager, Metadata
Oakland or Colorado
$187k-$233k/yr RemoteFull Time
Fivetran
Fivetran: Automates data movement into cloud data warehouses.
Experience managing software engineering teams, reviewing designs and code, delivering complex cloud projects, owning production services, driving reliability, planning with product partners, and using AI development tools.
GraphQL, GitHub Copilot, Claude Code, Claude, ChatGPT, TDDs, PRDs, FDLC, SLA
2mo
Save
Mark Applied
Hide
Engineering Manager, Agents
New York City or San Francisco or North America
$220k-$400k/yr RemoteFull Time
Hightouch
Hightouch: Syncs customer data from warehouses to business and marketing tools.
Lead an engineering team for an Agentic Marketing Platform: roadmap and ship features, hire and grow engineers, improve reliability and execution, and provide strong technical and people leadership.

Explore Jobs

Expand Your Job Search