86 sre manager jobs at 60 companies in Larkspur, CA
1w
Save
Mark Applied
Hide
1w
Senior Director – Observability | SRE
San Francisco or Dallas
OnsiteFull Time
Gap Inc.NYSE: GAP: Global specialty retailer of apparel and accessories.
10+ YOEStrategic technology leader with 10+ years driving operational transformation, expertise in ITIL, SRE, architecture, observability, infrastructure, cloud operations, automation, and service management, plus large-team leadership experience.
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
10+ YOE3+ Mgmt10+ years software engineering experience with SRE/platform focus, 3+ years managing SRE teams, proficiency in cloud (AWS/Azure/GCP), deep SRE principles, incident management, bachelor's in CS or equivalent, able to work 2+ days/week in Sunnyvale.
10+ YOEBachelor's degree in a technical field or equivalent experience, 10 years of program management experience, SQL experience, and preferred experience managing complex cross-functional projects.
San Francisco or Boston or New York City or Austin or Tokyo or London or Bangalore or United States or India or Europe
OnsiteFull Time
Postman: Platform for building, testing, and managing software APIs.
15+ YOE7+ MgmtRequires 15+ years in infrastructure, platform, or SRE engineering and 7+ years in leadership, with distributed-team management, Kubernetes ecosystem, Istio, AWS, Azure, and SRE expertise.
Sr Manager - Infrastructure, SRE, & AI Platforms - Services Special Projects
Cupertino or San Francisco
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
12+ YOE6+ MgmtMS in computer science or related field; 12+ years engineering leadership; 6+ years managing multi-layered engineering organizations; expertise in cloud infrastructure, Kubernetes, AI/ML systems, SRE, and global operations.
Pune or India or Milpitas or Seattle or Princeton or Cape Town or London or Zurich or Singapore or Mexico City
OnsiteFull Time
ZensarNational Stock Exchange of India: ZENSARTECH: Global technology firm providing digital transformation and infrastructure services.
15+ YOE15+ years IT experience in product owner/manager or technical leadership; strong knowledge of IT operations, monitoring/observability, ITSM (ServiceNow/Remedy), automation (StackStorm), cloud/hybrid, DevOps/SRE, Agile/SAFe, and AIOps initiatives.
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bengaluru or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Washington or Dallas
$200k-$275k/yrOnsiteFull Time
SingleStore: Provides a distributed SQL database for real-time analytics.
Product leadership experience with distributed databases, query processing, AI agents or MCP servers, observability systems, SRE workflows, complex technical initiatives, and cross-functional people leadership.
Plaud: Develops AI-powered voice recorders and automated transcription software.
8+ YOE8+ years in SRE/Infrastructure/Platform engineering, strong cloud (AWS/GCP/Azure) and Kubernetes experience, on-call/incident management experience, proficiency in Go/Python/Java, and experience building observability and SLO-driven systems.
SkillzNYSE: SKLZ: Operates a platform for competitive multiplayer mobile gaming.
14+ YOE14+ years infrastructure engineering experience with public cloud (AWS), 5+ years running Kubernetes (EKS), leadership experience, observability/CI-CD expertise, cost optimization track record, and proficiency in Go, Python, or Java.
WalmartNYSE: WMT: Multinational retail operating discount stores and supermarkets.
5+ YOEPlan and execute medium- to high-complexity technical programs, manage dependencies and risks, translate product needs into technical designs, apply SRE and cloud best practices, and align stakeholders across engineering and product teams.
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bangalore or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Dallas
$200k-$275k/yrOnsiteFull Time
SingleStore: Real-time distributed SQL database for transactions and analytics.
Requires deep distributed database, query processing, performance optimization, observability, and AI product expertise, plus multi-team leadership, customer engagement, execution, communication, and mentorship skills.
NubankNYSE: NU: Digital financial platform offering banking, credit, and investment services.
Requires deep infrastructure, SRE, or backend engineering expertise; hands-on technical leadership; massive cloud-native systems experience; architectural mastery; AI/ML transformation experience; and executive communication skills.
Sr. Technical Program Manager, Product Escalations
Sunnyvale, California, United States
$157k-$180k/yrOnsiteFull Time
Illumio: Provides zero-trust segmentation software to contain cyberattacks.
7+ YOERequires 7+ years in technical support, escalation, account, incident, customer success engineering, or technical program management; enterprise escalation, SaaS/cloud, cross-functional, RCA, and executive communication experience.
Technical Customer Success / Customer Program Manager
United States or Pleasanton or India
RemoteFull Time
Ciroos: AI platform for automated site reliability engineering.
Experience leading complex enterprise customer programs; technical fluency across cloud, SRE, observability, integrations, and security; excellent communication and customer relationship skills.
AWS, Microsoft Azure, Google Cloud Platform, Kubernetes, Splunk, Datadog, ServiceNow, Slack, Jira, Linear, SRE, ITSM
Glean: AI platform for enterprise search and automated workplace agents
8+ YOE8+ years TPM/infrastructure or SRE experience with 3+ years leading infra/platform programs; BS/MS in CS/Engineering or related; strong cloud, ML/LLM, reliability, and cross-functional leadership skills.
Microsoft Teams, Zoom, ServiceNow, Zendesk, GitHub, AWS, GCP, Azure, LLM
Principal Technical Program Manager (TPM) - AI Infrastructure Operations
Houston or New York City or San Francisco or Seattle
OnsiteFull Time
Nscale: Vertically integrated AI infrastructure provider for high-performance computing.
5+ YOE5+ years in technical program management for complex infrastructure or software programs; knowledge of data centers, distributed systems, Linux, networking, operational metrics, and Agile/Scrum. Technical degree preferred.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
Experience with system design, scalable cloud infrastructure, configuration management, programming in Python or Go, distributed systems, automation, documentation, and cross-functional collaboration.
AWS, GCP, Azure, Chef, Ansible, Terraform, GitHub Actions, Python, Go
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Experienced manager for cloud platform/SRE teams with strong Kubernetes and production infrastructure background, familiar with IaC and CI/CD tooling, recruiting and incident management skills.