United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bengaluru or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Washington or Dallas
$200k-$275k/yrOnsiteFull Time
SingleStore: Provides a distributed SQL database for real-time analytics.
Product leadership experience with distributed databases, query processing, AI agents or MCP servers, observability systems, SRE workflows, complex technical initiatives, and cross-functional people leadership.
Plaud: Develops AI-powered voice recorders and automated transcription software.
8+ YOE8+ years in SRE/Infrastructure/Platform engineering, strong cloud (AWS/GCP/Azure) and Kubernetes experience, on-call/incident management experience, proficiency in Go/Python/Java, and experience building observability and SLO-driven systems.
8+ YOE8+ MgmtBachelor's in CS or equivalent,8+ years software engineering,8+ years people management,5+ years quality/reliability engineering,experience with distributed systems and cloud,strong leadership and partnership skills.
SkillzNYSE: SKLZ: Operates a platform for competitive multiplayer mobile gaming.
14+ YOE14+ years infrastructure engineering experience with public cloud (AWS), 5+ years running Kubernetes (EKS), leadership experience, observability/CI-CD expertise, cost optimization track record, and proficiency in Go, Python, or Java.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
7+ YOE7+ years product management experience including 3+ years on technical infrastructure or platform products; experience defining observability, SLIs/SLOs/SLA, partnering with SRE and fleet engineering; strong writing, influence, and execution skills.
Sentry: Developer platform for error tracking and performance monitoring.
8+ YOE8+ years TPM experience in high-growth tech with platform/infrastructure/SRE/security exposure; proven cross-functional program leadership; incident management experience; strong analytical and communication skills.
United States or Seattle or San Francisco or Sunnyvale or Raleigh or Boston or London or Lisbon or Bangalore or Dublin or Kyiv or Chicago or New York City or Austin or Phoenix or Las Vegas or Dallas
$200k-$275k/yrOnsiteFull Time
SingleStore: Real-time distributed SQL database for transactions and analytics.
Requires deep distributed database, query processing, performance optimization, observability, and AI product expertise, plus multi-team leadership, customer engagement, execution, communication, and mentorship skills.
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Experienced manager for cloud platform/SRE teams with strong Kubernetes and production infrastructure background, familiar with IaC and CI/CD tooling, recruiting and incident management skills.
Bellevue or Livingston or New York City or Sunnyvale or San Francisco
$182k-$242k/yrOnsiteFull Time
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
7+ YOE3+ Mgmt3+ years engineering management and 7+ years technical experience in cloud operations/SRE; strong knowledge of cloud platforms, K8S, observability, incident management, and systems programming (Go); experience defining SLAs/SLOs and running on-call.
Databricks: A unified platform for data analytics and artificial intelligence.
5+ YOE5+ years in incident management/SRE/production operations for cloud-native systems; lead high-severity incidents; cloud (AWS/Azure/GCP) and observability expertise; log analysis; scripting (Python/Go/Bash); BS/MS in CS/CE or related field.
FastlyNYSE: FSLY: Provides edge cloud platform for content delivery and cybersecurity.
10+ YOE4+ Mgmt4+ years managing engineering/SRE teams,10+ years technical experience with distributed systems/CDN/cloud infrastructure, hands-on development in Go/Rust/C/C++, Linux, roadmap and cross-functional leadership.
Anyscale: Cloud platform for scaling distributed machine learning applications.
Experienced engineering leader for infrastructure, SRE, and governance; deep distributed systems knowledge; Kubernetes and cloud provider experience; hiring, coaching, and execution skills.
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
5+ YOE5+ Mgmt5+ years leading engineering teams and 5+ years in infrastructure/platform/backend roles; strong platform mindset, SRE experience (SLOs/error budgets), AWS and cloud fundamentals, influence and hiring experience, experience with incident/incident response processes.
Hinge HealthNYSE: HNGE: Digital provider of musculoskeletal care and physical therapy
10+ YOE4+ Mgmt10+ years in technology; 4+ years leading engineering teams; strong cloud infra (AWS, Kubernetes/EKS) and IaC (Terraform); proven platform reliability, SRE focus, and cost optimization; cross-functional leadership across geographies.
10+ YOE5+ Mgmt10+ years in IT infrastructure/operations, 5+ years senior leadership, experience with hybrid cloud (AWS/Azure/GCP), SRE, ITSM/ITIL, vendor management, budgeting and retail technology (POS) preferred.
Charlotte or Chandler or San Francisco or Columbus
$119k-$224k/yrHybridFull Time
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
5+ YOERequires 5+ years in systems engineering or architecture, 5+ years SRE, cloud observability experience, hosting platforms, and familiarity with DevOps, Agile, and IT service management.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
10+ YOE5+ MgmtBachelor's in a technical field,10+ years engineering experience with 5+ years leading SRE/Platform teams; experience with observability, incident management, distributed systems, and cloud architecture.
AWS, New Relic, Splunk, Datadog, Sentry, Honeycomb, Grafana, Prometheus, OpenTelemetry