54 sre manager jobs at 39 companies in Rio Vista, CA

23h
Save
Mark Applied
Hide
Senior Director – Observability | SRE
San Francisco or Dallas
OnsiteFull Time
Gap Inc.
Gap Inc.NYSE: GAP: Global specialty retailer of apparel and accessories.
10+ YOEStrategic technology leader with 10+ years driving operational transformation, expertise in ITIL, SRE, architecture, observability, infrastructure, cloud operations, automation, and service management, plus large-team leadership experience.
ITIL, SRE, Live Sight Insights
1w
Save
Mark Applied
Hide
Principal Product Manager Lead
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bengaluru or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Washington or Dallas
$200k-$275k/yr OnsiteFull Time
SingleStore
SingleStore: Provides a distributed SQL database for real-time analytics.
Product leadership experience with distributed databases, query processing, AI agents or MCP servers, observability systems, SRE workflows, complex technical initiatives, and cross-functional people leadership.
Aura Analyst, AI Agents, MCP servers, Python UDFs, Cloud Functions, Container Services, SQrL, SRE, CPU, GPU
2mo
Save
Mark Applied
Hide
Senior SRE Engineer - San Francisco
San Francisco, California, United States
HybridFull Time
Plaud
Plaud: Develops AI-powered voice recorders and automated transcription software.
8+ YOE8+ years in SRE/Infrastructure/Platform engineering, strong cloud (AWS/GCP/Azure) and Kubernetes experience, on-call/incident management experience, proficiency in Go/Python/Java, and experience building observability and SLO-driven systems.
AWS, GCP, Azure, Kubernetes, Go, Python, Java, Cursor, GPT models, Gemini, Claude
2mo
Save
Mark Applied
Hide
Head - SRE
Bengaluru or Las Vegas or San Francisco
HybridFull Time
Skillz
SkillzNYSE: SKLZ: Operates a platform for competitive multiplayer mobile gaming.
14+ YOE14+ years infrastructure engineering experience with public cloud (AWS), 5+ years running Kubernetes (EKS), leadership experience, observability/CI-CD expertise, cost optimization track record, and proficiency in Go, Python, or Java.
EKS, EC2, VPC, IAM, Cost Explorer, Savings Plans, Kubernetes, Istio, Datadog, Prometheus, Jaeger, X-Ray, ArgoCD, GitHub Actions, Go, Python, Java
2mo
Save
Mark Applied
Hide
Senior Product Manager
Toronto or San Francisco
HybridFull Time
PagerDuty
PagerDutyNYSE: PD: Platform for real-time incident response and digital operations automation.
5+ YOE5+ years product management; SaaS B2B; SRE/DevOps domain experience; workflow automation and integration tooling; security and permissions expertise; strong communication; leadership in roadmap execution.
Workflow automation, APIs, Telemetry analysis, Security and authorization models, SaaS platforms
1mo
Save
Mark Applied
Hide
Staff Product Manager - Observability
Bellevue or San Francisco
$291k-$430k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
7+ YOE7+ years product management experience including 3+ years on technical infrastructure or platform products; experience defining observability, SLIs/SLOs/SLA, partnering with SRE and fleet engineering; strong writing, influence, and execution skills.
NCCL, InfiniBand, PyTorch, DCGM, Datadog, Grafana, Prometheus
2mo
Save
Mark Applied
Hide
Staff Technical Program Manager
San Francisco, California, United States
$200k-$240k/yr HybridFull Time
Sentry
Sentry: Developer platform for error tracking and performance monitoring.
8+ YOE8+ years TPM experience in high-growth tech with platform/infrastructure/SRE/security exposure; proven cross-functional program leadership; incident management experience; strong analytical and communication skills.
1w
Save
Mark Applied
Hide
Principal Product Manager Lead
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bangalore or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Dallas
$200k-$275k/yr OnsiteFull Time
SingleStore
SingleStore: Real-time distributed SQL database for transactions and analytics.
Requires deep distributed database, query processing, performance optimization, observability, and AI product expertise, plus multi-team leadership, customer engagement, execution, communication, and mentorship skills.
Aura Analyst, AI Agents, MCP servers, Python UDFs, Cloud Functions, Container Services, SQrL, SRE
6d
Save
Mark Applied
Hide
Technical Customer Success / Customer Program Manager
United States or Pleasanton or India
RemoteFull Time
Ciroos
Ciroos: AI platform for automated site reliability engineering.
Experience leading complex enterprise customer programs; technical fluency across cloud, SRE, observability, integrations, and security; excellent communication and customer relationship skills.
AWS, Microsoft Azure, Google Cloud Platform, Kubernetes, Splunk, Datadog, ServiceNow, Slack, Jira, Linear, SRE, ITSM
2mo
Save
Mark Applied
Hide
Engineering Manager, Cloud Platform
San Francisco, California, United States
$165k-$330k/yr HybridFull Time
Baseten
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Experienced manager for cloud platform/SRE teams with strong Kubernetes and production infrastructure background, familiar with IaC and CI/CD tooling, recruiting and incident management skills.
Kubernetes, Terraform, CloudFormation, Pulumi, GitHub Actions, GitLab CI, CircleCI, Jenkins, Prometheus, ELK stack, Grafana, OpenTelemetry
1w
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineering
Irving or San Leandro
OnsiteFull Time
Wells Fargo
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
7+ YOE3+ MgmtRequires 7+ years in systems engineering or technology architecture, 3+ years of management, 5+ years leading engineering or SRE teams, and experience with customer-facing platforms, incident management, SRE, DevOps, cloud, and production operations.
Splunk, Grafana, AppDynamics, Dynatrace, OpenTelemetry, Prometheus, Kubernetes, OpenShift, AWS, Microsoft Azure, Google Cloud Platform, Infrastructure as Code (IaC), CI/CD, AIOps, ITIL
2w
Save
Mark Applied
Hide
Global Operational Excellence Lead
Quincy or Princeton or Irvine or Austin or Atlanta or Boston or Sacramento
$120k-$203k/yr HybridFull Time
State Street
State StreetNYSE: STT: Provides investment servicing and management to institutional investors.
12+ YOE7+ Mgmt12+ years in cyber security or technology operations, 7+ years in production/ service management; expertise in public cloud, ITIL/ITSM, SRE/DevOps, operational governance, audit readiness, and AI-enabled automation.
AWS, Azure, GCP, SIEM, Security Lakehouse, SOAR, AI, SRE, DevOps, ITOM, ITSM, ITIL
1mo
Save
Mark Applied
Hide
Incident Manager
United States or Texas or San Francisco
$104k-$146k/yr RemoteFull Time
Databricks
Databricks: A unified platform for data analytics and artificial intelligence.
5+ YOE5+ years in incident management/SRE/production operations for cloud-native systems; lead high-severity incidents; cloud (AWS/Azure/GCP) and observability expertise; log analysis; scripting (Python/Go/Bash); BS/MS in CS/CE or related field.
AWS, Azure, GCP, Datadog, Elasticsearch, Splunk, Cloud Logging, OpenTelemetry, Prometheus, Grafana, Python, Go, Bash, Apache Spark, Delta Lake, MLflow
4d
Save
Mark Applied
Hide
Engineering Manager, Kubernetes Infrastructure (Bare Metal)
Livingston or New York City or Sunnyvale or Bellevue or San Francisco
$182k-$242k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
Experience managing infrastructure, platform, or SRE engineering teams; strong Kubernetes, distributed systems, production infrastructure, incident response, communication, and cross-functional leadership skills.
Kubernetes, Go, Python
2w
Save
Mark Applied
Hide
Senior Engineering Manager - Containers at Edge
San Francisco or Denver or New York City
$228k-$274k/yr HybridFull Time
Fastly
FastlyNYSE: FSLY: Provides edge cloud platform for content delivery and cybersecurity.
10+ YOE4+ Mgmt4+ years managing engineering/SRE teams,10+ years technical experience with distributed systems/CDN/cloud infrastructure, hands-on development in Go/Rust/C/C++, Linux, roadmap and cross-functional leadership.
Go, Rust, C/C++, Linux
3w
Save
Mark Applied
Hide
Engineering Manager, Platform Infrastructure (Foundations)
San Francisco, California, United States
$270k-$320k/yr OnsiteFull Time
Anyscale
Anyscale: Cloud platform for scaling distributed machine learning applications.
Experienced engineering leader for infrastructure, SRE, and governance; deep distributed systems knowledge; Kubernetes and cloud provider experience; hiring, coaching, and execution skills.
Kubernetes, AWS, GCP, Azure, VMs
1mo
Save
Mark Applied
Hide
Manager, Software Engineering (Reliability Platform)
United States or California or Washington or New York or New Jersey or Connecticut or Los Angeles or San Francisco
$204k-$290k/yr RemoteFull Time
Affirm
AffirmNasdaq: AFRM: Financial platform providing installment loans for consumer purchases.
7+ YOE2+ Mgmt7+ years backend/full-stack engineering experience with 2+ years engineering leadership; SRE/production engineering experience; observability and platform-building experience; strong programming (Python, Kotlin, Java); Bachelor\u0002s degree or equivalent experience.
Python, Kotlin, Java
3mo
Save
Mark Applied
Hide
Senior Engineering Manager, Service Enablement
San Francisco, California, United States
$212k-$318k/yr HybridFull Time
Hinge Health
Hinge HealthNYSE: HNGE: Digital provider of musculoskeletal care and physical therapy
10+ YOE4+ Mgmt10+ years in technology; 4+ years leading engineering teams; strong cloud infra (AWS, Kubernetes/EKS) and IaC (Terraform); proven platform reliability, SRE focus, and cost optimization; cross-functional leadership across geographies.
AWS, Kubernetes, EKS, Terraform, NestJS, TypeScript, GitHub Actions, NX, Okteto, Datadog, CI/CD, Helm, Infisical, Vault
1mo
Save
Mark Applied
Hide
VP – IT Operations
Emeryville, California, United States
$220k-$250k/yr HybridFull Time
Grocery Outlet
Grocery OutletNASDAQ: GO: Discount grocery retailer selling overstocked and closeout name-brand products.
10+ YOE5+ Mgmt10+ years in IT infrastructure/operations, 5+ years senior leadership, experience with hybrid cloud (AWS/Azure/GCP), SRE, ITSM/ITIL, vendor management, budgeting and retail technology (POS) preferred.
AWS, Azure, GCP, POS, SRE, ITSM, ITIL
2mo
Save
Mark Applied
Hide
Director of Site Reliability Engineering
San Francisco, California, United States
$210k-$310k/yr HybridFull Time
Stellar Development Foundation
Stellar Development Foundation: Developing and maintaining the open-source Stellar blockchain network.
3+ YOE3+ MgmtLead SRE team; define OKRs; coordinate with dev teams; coach; 3+ years SRE and 3+ years managing; Kubernetes; IaC; strong communication.
Terraform, Ansible, Puppet, Kubernetes

Explore Jobs

Expand Your Job Search