86 sre manager jobs at 60 companies in Larkspur, CA

1w
Save
Mark Applied
Hide
Senior Director – Observability | SRE
San Francisco or Dallas
OnsiteFull Time
Gap Inc.
Gap Inc.NYSE: GAP: Global specialty retailer of apparel and accessories.
10+ YOEStrategic technology leader with 10+ years driving operational transformation, expertise in ITIL, SRE, architecture, observability, infrastructure, cloud operations, automation, and service management, plus large-team leadership experience.
ITIL, SRE, Live Sight Insights
2mo
Save
Mark Applied
Hide
Manager, Engineering - Dev Ops/SRE (Hybrid)
Sunnyvale, California, United States
$140k-$215k/yr HybridFull Time
CrowdStrike
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
10+ YOE3+ Mgmt10+ years software engineering experience with SRE/platform focus, 3+ years managing SRE teams, proficiency in cloud (AWS/Azure/GCP), deep SRE principles, incident management, bachelor's in CS or equivalent, able to work 2+ days/week in Sunnyvale.
Golang, Kubernetes, Istio, Linkerd, Apache Kafka, Apache Flink, Prometheus, Grafana, Jaeger, OpenTelemetry, ELK, Splunk, AWS, Azure, GCP
3d
Save
Mark Applied
Hide
Senior Technical Program Manager II, Workspace SRE
Sunnyvale, California, United States
$240k-$333k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
10+ YOEBachelor's degree in a technical field or equivalent experience, 10 years of program management experience, SQL experience, and preferred experience managing complex cross-functional projects.
SQL
6d
Save
Mark Applied
Hide
Head of Engineering, Infrastructure & SRE
San Francisco or Boston or New York City or Austin or Tokyo or London or Bangalore or United States or India or Europe
OnsiteFull Time
Postman
Postman: Platform for building, testing, and managing software APIs.
15+ YOE7+ MgmtRequires 15+ years in infrastructure, platform, or SRE engineering and 7+ years in leadership, with distributed-team management, Kubernetes ecosystem, Istio, AWS, Azure, and SRE expertise.
Kubernetes, Cluster API, Argo, Helm, Crossplane, Istio, AWS, Azure, GitOps, CI/CD, FinOps
3d
Save
Mark Applied
Hide
Sr Manager - Infrastructure, SRE, & AI Platforms - Services Special Projects
Cupertino or San Francisco
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
12+ YOE6+ MgmtMS in computer science or related field; 12+ years engineering leadership; 6+ years managing multi-layered engineering organizations; expertise in cloud infrastructure, Kubernetes, AI/ML systems, SRE, and global operations.
Kubernetes, AWS, GCP, Cassandra, FoundationDB, Kafka, Redis, PostgreSQL
2mo
Save
Mark Applied
Hide
Product Manager - AIOPS & Automation
Pune or India or Milpitas or Seattle or Princeton or Cape Town or London or Zurich or Singapore or Mexico City
OnsiteFull Time
Zensar
ZensarNational Stock Exchange of India: ZENSARTECH: Global technology firm providing digital transformation and infrastructure services.
15+ YOE15+ years IT experience in product owner/manager or technical leadership; strong knowledge of IT operations, monitoring/observability, ITSM (ServiceNow/Remedy), automation (StackStorm), cloud/hybrid, DevOps/SRE, Agile/SAFe, and AIOps initiatives.
ServiceNow, Remedy, StackStorm, Resolve.in, DRYiCE, BigPanda, Moogsoft, ServiceNow Event Management, SAFe, DevOps, SRE
3w
Save
Mark Applied
Hide
Principal Product Manager Lead
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bengaluru or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Washington or Dallas
$200k-$275k/yr OnsiteFull Time
SingleStore
SingleStore: Provides a distributed SQL database for real-time analytics.
Product leadership experience with distributed databases, query processing, AI agents or MCP servers, observability systems, SRE workflows, complex technical initiatives, and cross-functional people leadership.
Aura Analyst, AI Agents, MCP servers, Python UDFs, Cloud Functions, Container Services, SQrL, SRE, CPU, GPU
2mo
Save
Mark Applied
Hide
Senior SRE Engineer - San Francisco
San Francisco, California, United States
HybridFull Time
Plaud
Plaud: Develops AI-powered voice recorders and automated transcription software.
8+ YOE8+ years in SRE/Infrastructure/Platform engineering, strong cloud (AWS/GCP/Azure) and Kubernetes experience, on-call/incident management experience, proficiency in Go/Python/Java, and experience building observability and SLO-driven systems.
AWS, GCP, Azure, Kubernetes, Go, Python, Java, Cursor, GPT models, Gemini, Claude
2mo
Save
Mark Applied
Hide
Head - SRE
Bengaluru or Las Vegas or San Francisco
HybridFull Time
Skillz
SkillzNYSE: SKLZ: Operates a platform for competitive multiplayer mobile gaming.
14+ YOE14+ years infrastructure engineering experience with public cloud (AWS), 5+ years running Kubernetes (EKS), leadership experience, observability/CI-CD expertise, cost optimization track record, and proficiency in Go, Python, or Java.
EKS, EC2, VPC, IAM, Cost Explorer, Savings Plans, Kubernetes, Istio, Datadog, Prometheus, Jaeger, X-Ray, ArgoCD, GitHub Actions, Go, Python, Java
3mo
Save
Mark Applied
Hide
Senior Product Manager
Toronto or San Francisco
HybridFull Time
PagerDuty
PagerDutyNYSE: PD: Platform for real-time incident response and digital operations automation.
5+ YOE5+ years product management; SaaS B2B; SRE/DevOps domain experience; workflow automation and integration tooling; security and permissions expertise; strong communication; leadership in roadmap execution.
Workflow automation, APIs, Telemetry analysis, Security and authorization models, SaaS platforms
1mo
Save
Mark Applied
Hide
Staff, Technical Program Manager
Bentonville or Sunnyvale
$110k-$220k/yr OnsiteFull Time
Walmart
WalmartNYSE: WMT: Multinational retail operating discount stores and supermarkets.
5+ YOEPlan and execute medium- to high-complexity technical programs, manage dependencies and risks, translate product needs into technical designs, apply SRE and cloud best practices, and align stakeholders across engineering and product teams.
3w
Save
Mark Applied
Hide
Principal Product Manager Lead
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bangalore or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Dallas
$200k-$275k/yr OnsiteFull Time
SingleStore
SingleStore: Real-time distributed SQL database for transactions and analytics.
Requires deep distributed database, query processing, performance optimization, observability, and AI product expertise, plus multi-team leadership, customer engagement, execution, communication, and mentorship skills.
Aura Analyst, AI Agents, MCP servers, Python UDFs, Cloud Functions, Container Services, SQrL, SRE
1d
Save
Mark Applied
Hide
Director, Technical Program Manager (Resiliency and Reliability Engineering)
McLean or Richmond or New York City or Plano or San Francisco
$210k-$287k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: A diversified financial services providing banking and credit products.
7+ YOEBachelor's degree and 7+ years managing technical programs required. Preferred: distributed systems, cloud, SRE, resilience engineering, Agile delivery, complex program leadership, and regulated-environment experience.
Agile
2w
Save
Mark Applied
Hide
Senior Staff TPM - Core Infrastructure & Platform Evolution
Palo Alto or Virginia or Miami or São Paulo
HybridFull Time
Nubank
NubankNYSE: NU: Digital financial platform offering banking, credit, and investment services.
Requires deep infrastructure, SRE, or backend engineering expertise; hands-on technical leadership; massive cloud-native systems experience; architectural mastery; AI/ML transformation experience; and executive communication skills.
AI, ML, SRE, cloud-native, distributed databases, microservices
2w
Save
Mark Applied
Hide
Sr. Technical Program Manager, Product Escalations
Sunnyvale, California, United States
$157k-$180k/yr OnsiteFull Time
Illumio
Illumio: Provides zero-trust segmentation software to contain cyberattacks.
7+ YOERequires 7+ years in technical support, escalation, account, incident, customer success engineering, or technical program management; enterprise escalation, SaaS/cloud, cross-functional, RCA, and executive communication experience.
Jira, Salesforce, ServiceNow, PagerDuty, Zendesk, ITIL, SRE
2w
Save
Mark Applied
Hide
Technical Customer Success / Customer Program Manager
United States or Pleasanton or India
RemoteFull Time
Ciroos
Ciroos: AI platform for automated site reliability engineering.
Experience leading complex enterprise customer programs; technical fluency across cloud, SRE, observability, integrations, and security; excellent communication and customer relationship skills.
AWS, Microsoft Azure, Google Cloud Platform, Kubernetes, Splunk, Datadog, ServiceNow, Slack, Jira, Linear, SRE, ITSM
3mo
Save
Mark Applied
Hide
Senior Technical Program Manager, Infrastructure
Mountain View, California, United States
$198k-$236k/yr HybridFull Time
Glean
Glean: AI platform for enterprise search and automated workplace agents
8+ YOE8+ years TPM/infrastructure or SRE experience with 3+ years leading infra/platform programs; BS/MS in CS/Engineering or related; strong cloud, ML/LLM, reliability, and cross-functional leadership skills.
Microsoft Teams, Zoom, ServiceNow, Zendesk, GitHub, AWS, GCP, Azure, LLM
1w
Save
Mark Applied
Hide
Principal Technical Program Manager (TPM) - AI Infrastructure Operations
Houston or New York City or San Francisco or Seattle
OnsiteFull Time
Nscale
Nscale: Vertically integrated AI infrastructure provider for high-performance computing.
5+ YOE5+ years in technical program management for complex infrastructure or software programs; knowledge of data centers, distributed systems, Linux, networking, operational metrics, and Agile/Scrum. Technical degree preferred.
Linux, Agile, Scrum, NVIDIA GPUs, InfiniBand, RDMA, SRE, Continuous Integration/Continuous Deployment (CI/CD)
2w
Save
Mark Applied
Hide
IT Systems Engineer - Internal Platforms & SRE
San Francisco or San Jose
$206k-$275k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
Experience with system design, scalable cloud infrastructure, configuration management, programming in Python or Go, distributed systems, automation, documentation, and cross-functional collaboration.
AWS, GCP, Azure, Chef, Ansible, Terraform, GitHub Actions, Python, Go
2mo
Save
Mark Applied
Hide
Engineering Manager, Cloud Platform
San Francisco, California, United States
$165k-$330k/yr HybridFull Time
Baseten
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Experienced manager for cloud platform/SRE teams with strong Kubernetes and production infrastructure background, familiar with IaC and CI/CD tooling, recruiting and incident management skills.
Kubernetes, Terraform, CloudFormation, Pulumi, GitHub Actions, GitLab CI, CircleCI, Jenkins, Prometheus, ELK stack, Grafana, OpenTelemetry

Explore Jobs

Expand Your Job Search