54 sre manager jobs at 40 companies in Vacaville, CA
1w
Save
Mark Applied
Hide
1w
Principal Product Manager Lead
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bengaluru or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Washington or Dallas
$200k-$275k/yrOnsiteFull Time
SingleStore: Provides a distributed SQL database for real-time analytics.
Product leadership experience with distributed databases, query processing, AI agents or MCP servers, observability systems, SRE workflows, complex technical initiatives, and cross-functional people leadership.
Plaud: Develops AI-powered voice recorders and automated transcription software.
8+ YOE8+ years in SRE/Infrastructure/Platform engineering, strong cloud (AWS/GCP/Azure) and Kubernetes experience, on-call/incident management experience, proficiency in Go/Python/Java, and experience building observability and SLO-driven systems.
SkillzNYSE: SKLZ: Operates a platform for competitive multiplayer mobile gaming.
14+ YOE14+ years infrastructure engineering experience with public cloud (AWS), 5+ years running Kubernetes (EKS), leadership experience, observability/CI-CD expertise, cost optimization track record, and proficiency in Go, Python, or Java.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
7+ YOE7+ years product management experience including 3+ years on technical infrastructure or platform products; experience defining observability, SLIs/SLOs/SLA, partnering with SRE and fleet engineering; strong writing, influence, and execution skills.
Sentry: Developer platform for error tracking and performance monitoring.
8+ YOE8+ years TPM experience in high-growth tech with platform/infrastructure/SRE/security exposure; proven cross-functional program leadership; incident management experience; strong analytical and communication skills.
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bangalore or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Dallas
$200k-$275k/yrOnsiteFull Time
SingleStore: Real-time distributed SQL database for transactions and analytics.
Requires deep distributed database, query processing, performance optimization, observability, and AI product expertise, plus multi-team leadership, customer engagement, execution, communication, and mentorship skills.
Technical Customer Success / Customer Program Manager
United States or Pleasanton or India
RemoteFull Time
Ciroos: AI platform for automated site reliability engineering.
Experience leading complex enterprise customer programs; technical fluency across cloud, SRE, observability, integrations, and security; excellent communication and customer relationship skills.
AWS, Microsoft Azure, Google Cloud Platform, Kubernetes, Splunk, Datadog, ServiceNow, Slack, Jira, Linear, SRE, ITSM
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Experienced manager for cloud platform/SRE teams with strong Kubernetes and production infrastructure background, familiar with IaC and CI/CD tooling, recruiting and incident management skills.
Quincy or Princeton or Irvine or Austin or Atlanta or Boston or Sacramento
$120k-$203k/yrHybridFull Time
State StreetNYSE: STT: Provides investment servicing and management to institutional investors.
12+ YOE7+ Mgmt12+ years in cyber security or technology operations, 7+ years in production/ service management; expertise in public cloud, ITIL/ITSM, SRE/DevOps, operational governance, audit readiness, and AI-enabled automation.
Bellevue or Livingston or New York City or Sunnyvale or San Francisco
$182k-$242k/yrOnsiteFull Time
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
7+ YOE3+ Mgmt3+ years engineering management and 7+ years technical experience in cloud operations/SRE; strong knowledge of cloud platforms, K8S, observability, incident management, and systems programming (Go); experience defining SLAs/SLOs and running on-call.
Databricks: A unified platform for data analytics and artificial intelligence.
5+ YOE5+ years in incident management/SRE/production operations for cloud-native systems; lead high-severity incidents; cloud (AWS/Azure/GCP) and observability expertise; log analysis; scripting (Python/Go/Bash); BS/MS in CS/CE or related field.
FastlyNYSE: FSLY: Provides edge cloud platform for content delivery and cybersecurity.
10+ YOE4+ Mgmt4+ years managing engineering/SRE teams,10+ years technical experience with distributed systems/CDN/cloud infrastructure, hands-on development in Go/Rust/C/C++, Linux, roadmap and cross-functional leadership.
Anyscale: Cloud platform for scaling distributed machine learning applications.
Experienced engineering leader for infrastructure, SRE, and governance; deep distributed systems knowledge; Kubernetes and cloud provider experience; hiring, coaching, and execution skills.
Hinge HealthNYSE: HNGE: Digital provider of musculoskeletal care and physical therapy
10+ YOE4+ Mgmt10+ years in technology; 4+ years leading engineering teams; strong cloud infra (AWS, Kubernetes/EKS) and IaC (Terraform); proven platform reliability, SRE focus, and cost optimization; cross-functional leadership across geographies.
10+ YOE5+ Mgmt10+ years in IT infrastructure/operations, 5+ years senior leadership, experience with hybrid cloud (AWS/Azure/GCP), SRE, ITSM/ITIL, vendor management, budgeting and retail technology (POS) preferred.
Sr Director, Regional Sales Management- Security Sales (San Francisco, California, US)
San Francisco, California, United States
$240k-$300k/yrRemoteFull Time
DynatraceNYSE: DT: Provides an AI-powered observability platform for cloud monitoring.
10+ YOE5+ Mgmt10+ years in Presales/DevOps/SRE; 5+ years of management; CKA/CKAD preferred; Terraform, Helm, YAML; Go/Java/Node.js knowledge; Bachelor in CS/Engineering; MBA or equivalent a plus.
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
15+ YOE15+ years in SRE/infrastructure/platform engineering with large-scale fleets; experience in capacity management, orchestration, observability, SLOs/SLIs, incident response, and cross-team architecture.