22 staff site reliability engineer jobs at 21 companies in Concord, CA
3mo
Save
Mark Applied
Hide
3mo
Staff Site Reliability Engineer
Mountain View, California, United States
$252k-$308k/yrHybridFull Time
EarnIn: Provides immediate access to earned wages through a mobile app.
7+ YOE7+ years in SRE or related field; experience applying AI/LLMs to operations; strong SLO/SLI and incident management; software engineering in Python or Go; observability and IaC proficiency; AI-assisted development tools; fintech/regulated environment experience.
CircleNYSE: CRCL: Digital currency issuer and blockchain financial infrastructure provider.
6+ YOE6+ years SRE/DevOps experience preferred; strong Kubernetes, IaC (Terraform/Pulumi), cloud networking, Go or Python, CI/CD, observability, and distributed systems experience; blockchain node operation experience is a plus.
Skydio: Develops autonomous AI drones for defense and industrial inspection.
8+ YOE8+ years in SRE, platform, DevOps, production engineering, or equivalent; strong Kubernetes and AWS experience; Terraform, CI/CD, networking, observability, and production reliability expertise.
Staff Site Reliability Engineer-Production Operations
Palo Alto, California, United States
$186k-$233k/yrHybridFull Time
Rivian and Volkswagen Group Technologies: A joint venture creating software-defined vehicle technology and connected services for electric vehicles.
Senior SRE with incident command experience, strong systems engineering for distributed systems, hands-on coding (Python or Go), observability expertise (Datadog or comparable), and experience with blameless post-incident practices.
LiveRampNew York Stock Exchange: RAMP: Provides a platform for secure data collaboration and identity.
10+ YOESenior SRE with 10+ years in production engineering; Kubernetes, Terraform, Python/Go; globally distributed systems; CI/CD; FinOps; strong communication.
San Francisco or San Jose or New York City or Milpitas or Mountain View or Holmdel or Goleta or Redwood City or Fremont or Sunnyvale or Brooklyn or Palo Alto
$187k-$268k/yrHybridFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
6+ YOERequires 8+ years with a bachelor's, 6+ with a master's, or 3+ with a PhD; 6+ years in SRE or infrastructure engineering, 5+ years operating Kubernetes, cloud, CI/CD, and Python or Go.
RubrikNYSE: RBRK: Secures enterprise data across cloud and on-premises environments.
8+ YOEUS citizen; 8-12+ years software engineering with SRE/DevOps; BS/MS/PhD in CS/CE or related field; proficient in Go/Python/Java; distributed systems; Unix/Linux; on-call; leadership experience.
8+ YOEBachelor's degree or equivalent practical experience; 8 years building infrastructure or distributed systems; 5 years programming in C++ or Go and reliability engineering; distributed systems experience required.
C++, Go, Java, GoogleSQL, Software Development Kit (SDK)
Member of Technical Staff, Site Reliability Engineer
San Francisco, California, United States
$200k-$400k/yrOnsiteFull Time
Inferact: An artificial intelligence infrastructure advancing open-source inference technology to make model serving faster and more affordable.
Bachelor's degree or equivalent experience, production systems expertise, SRE fundamentals, incident response, Linux, networking, observability, distributed systems, and Python, Go, or Bash scripting.
Staff+ Site Reliability Engineer, Safeguards ML Infra
San Francisco or Seattle or New York City
$405k-$485k/yrHybridFull Time
Anthropic: Developing safe and reliable artificial intelligence systems.
8+ YOEProduction change-management experience, high-stakes release and on-call experience, AWS/GCP operations, Python proficiency, and a bachelor's degree or equivalent experience.
Python, Rust, AWS, GCP, AWS Bedrock, GCP Vertex, Claude
Member of Technical Staff, Site Reliablity Engineer
San Francisco, California, United States
$200k-$270k/yrHybridFull Time
Vapi: Build and deploy AI-powered voice agents via flexible APIs.
Experience running incident command and postmortems, operating SLOs/error budgets, capacity planning and load testing, Kubernetes production ops, KEDA autoscaling, and shipping services in Go or TypeScript.
Senior Staff Service Reliability and Operational Intelligence Engineer
Santa Clara, California, United States
$188k-$270k/yrOnsiteFull Time
IonQNYSE: IONQ: Develops and sells trapped-ion quantum computers and cloud services.
12+ YOERequires 12+ years in production, site reliability, platform engineering, or cloud operations; AWS or GCP; distributed systems, Kubernetes, observability, incident response, automation, and cross-team technical leadership.
Langan: Provides integrated engineering, environmental, and site development consulting services.
3+ YOEBachelor's in Civil/Geotechnical Engineering, 3+ years geotechnical experience, fieldwork and report writing, ability to perform site inspections, reliable transportation and valid driver’s license; FE/EIT preferred.