99 sre manager jobs at 77 companies in Pacifica, CA
3mo
Save
Mark Applied
Hide
3mo
Engineering Lead – Platform & SRE
Santa Clara, California, United States
$175k-$215k/yrHybridFull Time
Kerrigan Robotics: AI-powered orchestration platform for factory automation systems.
5+ YOE5+ years SRE/DevOps experience; deep Kubernetes knowledge; expertise with Pulumi/Terraform, Helm; proficiency in Go, Linux networking, observability (Prometheus, Grafana), and CI/CD (GitHub Actions).
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
10+ YOE3+ Mgmt10+ years software engineering experience with SRE/platform focus, 3+ years managing SRE teams, proficiency in cloud (AWS/Azure/GCP), deep SRE principles, incident management, bachelor's in CS or equivalent, able to work 2+ days/week in Sunnyvale.
Pune or India or Milpitas or Seattle or Princeton or Cape Town or London or Zurich or Singapore or Mexico City
OnsiteFull Time
ZensarNational Stock Exchange of India: ZENSARTECH: Global technology firm providing digital transformation and infrastructure services.
15+ YOE15+ years IT experience in product owner/manager or technical leadership; strong knowledge of IT operations, monitoring/observability, ITSM (ServiceNow/Remedy), automation (StackStorm), cloud/hybrid, DevOps/SRE, Agile/SAFe, and AIOps initiatives.
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bengaluru or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Washington or Dallas
$200k-$275k/yrOnsiteFull Time
SingleStore: Provides a distributed SQL database for real-time analytics.
Product leadership experience with distributed databases, query processing, AI agents or MCP servers, observability systems, SRE workflows, complex technical initiatives, and cross-functional people leadership.
Plaud: Develops AI-powered voice recorders and automated transcription software.
8+ YOE8+ years in SRE/Infrastructure/Platform engineering, strong cloud (AWS/GCP/Azure) and Kubernetes experience, on-call/incident management experience, proficiency in Go/Python/Java, and experience building observability and SLO-driven systems.
8+ YOE8+ MgmtBachelor's in CS or equivalent,8+ years software engineering,8+ years people management,5+ years quality/reliability engineering,experience with distributed systems and cloud,strong leadership and partnership skills.
SkillzNYSE: SKLZ: Operates a platform for competitive multiplayer mobile gaming.
14+ YOE14+ years infrastructure engineering experience with public cloud (AWS), 5+ years running Kubernetes (EKS), leadership experience, observability/CI-CD expertise, cost optimization track record, and proficiency in Go, Python, or Java.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
7+ YOE7+ years product management experience including 3+ years on technical infrastructure or platform products; experience defining observability, SLIs/SLOs/SLA, partnering with SRE and fleet engineering; strong writing, influence, and execution skills.
Site Reliability Engineering (SRE) Manager, Apple Maps
Cupertino, California, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Build, manage, and deliver highly available, automated infrastructure for Apple Maps at global scale; focus on reliability, scalability, and operational excellence.
Sentry: Developer platform for error tracking and performance monitoring.
8+ YOE8+ years TPM experience in high-growth tech with platform/infrastructure/SRE/security exposure; proven cross-functional program leadership; incident management experience; strong analytical and communication skills.
WalmartNYSE: WMT: Multinational retail operating discount stores and supermarkets.
5+ YOEPlan and execute medium- to high-complexity technical programs, manage dependencies and risks, translate product needs into technical designs, apply SRE and cloud best practices, and align stakeholders across engineering and product teams.
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bangalore or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Dallas
$200k-$275k/yrOnsiteFull Time
SingleStore: Real-time distributed SQL database for transactions and analytics.
Requires deep distributed database, query processing, performance optimization, observability, and AI product expertise, plus multi-team leadership, customer engagement, execution, communication, and mentorship skills.
6+ YOEBS/MS in Engineering or CS (or equivalent),6+ years product management experience,experience with cloud operations,SRE,distributed systems,release processes,and using AI to augment product work.
NubankNYSE: NU: Digital financial platform offering banking, credit, and investment services.
Requires deep infrastructure, SRE, or backend engineering expertise; hands-on technical leadership; massive cloud-native systems experience; architectural mastery; AI/ML transformation experience; and executive communication skills.
Sr. Technical Program Manager, Product Escalations
Sunnyvale, California, United States
$157k-$180k/yrOnsiteFull Time
Illumio: Provides zero-trust segmentation software to contain cyberattacks.
7+ YOERequires 7+ years in technical support, escalation, account, incident, customer success engineering, or technical program management; enterprise escalation, SaaS/cloud, cross-functional, RCA, and executive communication experience.
Glean: AI platform for enterprise search and automated workplace agents
8+ YOE8+ years TPM/infrastructure or SRE experience with 3+ years leading infra/platform programs; BS/MS in CS/Engineering or related; strong cloud, ML/LLM, reliability, and cross-functional leadership skills.
Microsoft Teams, Zoom, ServiceNow, Zendesk, GitHub, AWS, GCP, Azure, LLM
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Experienced manager for cloud platform/SRE teams with strong Kubernetes and production infrastructure background, familiar with IaC and CI/CD tooling, recruiting and incident management skills.