104 site operations engineer jobs at 83 companies in California

1mo
Save
Mark Applied
Hide
Site Operations Engineer, Bay Area
California, United States
$90k-$120k/yr OnsiteFull Time
Glimpse
Glimpse: Private battery-quality using high-throughput X-ray CT and software to help cell producers and buyers detect defects.
3+ YOE3+ years experience establishing lab or production operations, machine operation with safety protocols, DOT/IATA and radiation training, strong project management, attention to detail, US citizenship or green card required.
Glimpse Portal, CT scanner, X-ray
1w
Save
Mark Applied
Hide
Senior Site Engineer
Los Angeles or United States
$100k-$175k/yr OnsiteFull Time
Northwood Space
Northwood Space: Northwood is an end-to-end ground infrastructure provider for space missions, delivering hardware, software, and network services.
5+ YOERequires 5+ years of relevant experience, a bachelor's degree in mechanical, civil, electrical, or related engineering, and willingness to travel domestically and internationally. Construction, facilities, operations, and dashboard experience preferred.
Grafana, SQL
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer-Production Operations
Palo Alto, California, United States
$186k-$233k/yr HybridFull Time
Rivian and Volkswagen Group Technologies
Rivian and Volkswagen Group Technologies: Joint-venture automotive technology developing software-defined vehicle architecture and software for Rivian and Volkswagen Group electric vehicles.
Senior SRE with incident command experience, strong systems engineering for distributed systems, hands-on coding (Python or Go), observability expertise (Datadog or comparable), and experience with blameless post-incident practices.
Python, Go, Datadog, LLM
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
San Francisco, California, United States
$149k-$224k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: The #1 AI CRM driving customer success together.
5+ YOE5+ years systems and software engineering experience for large-scale internet services; expertise in SRE principles, containers, observability, incident management, Python and Go, and applying AI/ML to operations.
Temporal, Airflow, Argo Workflows, Docker, Kubernetes, DNS, HTTP, Grafana, Prometheus, ELK, Splunk, Datadog, Python, Go, Linux, Claude Code, GitHub Copilot, Codex, Cursor, AWS, GCP, MCP
4w
Save
Mark Applied
Hide
ZfG Operations Engineer
San Jose, California, United States
$99k-$229k/yr HybridFull Time
Zoom
ZoomNasdaq Global Select Market: ZM: American publicly traded communications platform serving businesses and individuals with AI-assisted video, voice, chat, and phone services.
Proven experience in site reliability, DevOps, or infrastructure operations; proficiency with Python, Go, or Bash; knowledge of distributed systems, monitoring/observability, incident management, and automation; experience supporting production systems and on-call rotations.
Python, Go, Bash, BrightHire
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
San Francisco or New York City
$164k-$306k/yr HybridFull Time
Retool
Retool: Private software providing an internal-tools development platform for business and enterprise teams.
Experience operating production infrastructure (AWS), Kubernetes, Terraform, Postgres; programming in Go/Python/TypeScript/Java/Ruby; building observability and automation for customer-facing SaaS systems.
Kubernetes, Helm, Docker Compose, Terraform, AWS, Postgres, Go, Python, TypeScript, Java, Ruby
4w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
San Francisco, California, United States
$195k-$258k/yr RemoteFull Time
Circle Internet Group, Inc.
Circle Internet Group, Inc.NYSE: CRCL: Public financial technology providing stablecoin, digital-asset, payments, and blockchain infrastructure to businesses and developers.
6+ YOE6+ years SRE/DevOps experience preferred; strong Kubernetes, IaC (Terraform/Pulumi), cloud networking, Go or Python, CI/CD, observability, and distributed systems experience; blockchain node operation experience is a plus.
Kubernetes, Helm, Terraform, Pulumi, Go, Python, CI/CD, RBAC, VPCs, DNS, SQL
2mo
Save
Mark Applied
Hide
Staff Engineer, Site Reliability
New York City or Los Angeles or United States or Canada
RemoteFull Time
Babylist
Babylist: Private baby-registry and e-commerce platform helping expecting parents plan, shop, and prepare.
Hands-on Terraform expertise, proven AWS experience (EKS, RDS, networking, CDNs), production Kubernetes operation, CI/CD design, observability and alerting, on-call/incident management, cross-functional developer support, familiarity with AI tooling.
Ruby on Rails, AWS, Sidekiq, MySQL, Redis, Terraform, EKS, RDS, Kubernetes, CircleCI, GitHub Actions, Datadog, Sentry, PagerDuty, Cronitor, Claude, ChatGPT
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer Platform Private Cloud Engineer
San Jose, California, United States
$94k-$130k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesBSE: 532540: Global leader in IT services, consulting, and business solutions.
7+ YOERequires 7+ years designing and operating enterprise or cloud environments, private cloud and Kubernetes expertise, scripting, IaC tools, Unix/Linux knowledge, and a CS or engineering degree.
VMware, AWS, GCP, Kubernetes, Helm, ArgoCD, Python, Bash, Ruby, Scala, Ansible, Terraform, Unix, Linux
6d
Save
Mark Applied
Hide
Launch Operations Engineer
Torrance, California, United States
$88k-$155k/yr OnsiteFull Time
Parsons
ParsonsNYSE: PSN: Technology and engineering solutions for defense and infrastructure.
Bachelor’s degree or equivalent experience, 5+ years in integration or launch engineering, launch campaign experience, Secret clearance, technical project management, and willingness to travel to remote sites.
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer (FedRAMP)
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Identity management and access control software provider.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
AWS, Terraform, Chef, Ansible, Puppet, Apache httpd, nginx, Apache Tomcat, Bash, Python, Golang, git, gdb, strace, ltrace, tcpdump, Wireshark, Docker, Kubernetes
2w
Save
Mark Applied
Hide
Site Reliability Engineer (Multiple Positions)
San Jose, California, United States
$226k-$317k/yr OnsiteFull Time
TikTok
TikTok: Short-form mobile video and social media platform.
1+ YOEMaster's degree and 1 year of related experience, or bachelor's degree and 3 years; requires 1 year supporting critical systems, monitoring, troubleshooting, data operations, error analysis, and runbook creation.
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Cloud
Santa Clara, California, United States
$168k-$265k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
8+ YOEBS/MS in computer science, engineering, or equivalent experience; 8+ years in live-site production operations; advanced Python; Kubernetes, cloud automation, incident management, and SRE on-call experience.
Akamai Edge Redirector Cloudlets, Akamai Forward Rewrite Cloudlets, Akamai CDN, Akamai Cloudlets Policy Manager, AWS, Kubernetes, Python, WAF
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Sunnyvale, California, United States
$160k-$240k/yr OnsiteFull Time
Fiserv
FiservNYSE: FI: Global leader in financial services technology solutions.
Requires mid-to-senior site reliability, operations, or DevOps experience; shell scripting; GCP, GKE, Kubernetes, IaC, monitoring tools, HAProxy, GitHub Actions, and strong troubleshooting skills.
Google Cloud Platform (GCP), GKE, Kubernetes, Terraform, Ansible, Puppet, Prometheus, Grafana, Datadog, HAProxy, GitHub, GitHub Actions, Python, Go, Java
1mo
Save
Mark Applied
Hide
Cloud Site Reliability Engineer - DCS Cloud
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Global technology specializing in AI-powered content platforms.
2+ YOEBachelor's degree in CS or related,2+ years in Linux operations/SRE/DevOps,programming in Go/Python/C++,cloud and reliability practices experience,strong troubleshooting and communication skills.
Go, Python, C++, Linux, OCI, AWS, Azure, GCP, KVM, QEMU, Docker, Kubernetes, containerd, cgroups, namespaces, CUDA, MIG
2mo
Save
Mark Applied
Hide
Member of Technical Staff, Site Reliablity Engineer
San Francisco, California, United States
$200k-$270k/yr HybridFull Time
Vapi
Vapi: Private voice AI platform that lets developers and enterprises build, deploy, and manage conversational voice agents.
Experience running incident command and postmortems, operating SLOs/error budgets, capacity planning and load testing, Kubernetes production ops, KEDA autoscaling, and shipping services in Go or TypeScript.
Go, TypeScript, Bash, Chronosphere, Prometheus, Grafana, Datadog, OpenTelemetry, Kubernetes, EKS, KEDA
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
San Francisco, California, United States
$117k-$209k/yr OnsiteFull Time
Autodesk
AutodeskNASDAQ: ADSK: Global provider of software for design, engineering, and manufacturing.
7+ YOEBachelor's degree or equivalent experience and 7+ years in SRE, software, platform, cloud infrastructure, or production operations; experience with cloud platforms, automation, IaC, CI/CD, and reliability engineering.
AWS, Azure, Python, Go, Java, PowerShell, Bash, Infrastructure as Code, CI/CD, AWS GovCloud, Kubernetes, Splunk, Dynatrace, Datadog, CloudWatch
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Lincoln or San Francisco
$125k-$165k/yr RemoteFull Time
TELCOR
TELCOR: Private healthcare software provider serving laboratories, pathology practices, and hospitals with point-of-care and revenue cycle management tools.
2+ YOEExperience with distributed systems, Redis, queuing systems, Kubernetes (2+ yrs), AWS, Terraform (2+ yrs), observability, and production operations.
Redis, Kubernetes, AWS, Terraform
6d
Save
Mark Applied
Hide
Manager, Site Reliability Engineer
New York City or San Francisco or California or New York
$150k-$220k/yr OnsiteFull Time
Forge Global
Forge GlobalNYSE: FRGE: Financial technology operating a private-market marketplace and data, custody, and investment solutions for companies and investors.
10+ YOE5+ MgmtRequires 5+ years leading SRE, DevOps, cloud operations, or similar functions; 10+ years in engineering or operations; bachelor's degree or equivalent; cloud infrastructure and distributed systems experience.
AWS, Azure, Kubernetes, Terraform, Ansible, Datadog, CloudWatch, CI/CD
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Foster City, California, United States
$250k-$300k/yr HybridFull Time
Zoox
Zoox: Autonomous mobility developing a fully electric robotaxi fleet.
5+ YOE5+ years operating GitHub Enterprise at scale, monorepo management, CI/CD integration, infrastructure-as-code (Terraform/Pulumi), cloud platform experience, technical leadership and migration planning.
Git, GitHub Enterprise, GitHub Cloud, Buildkite, GitHub Actions, Jenkins, GitLab CI, Terraform, Pulumi, Bazel, Buck, Reviewable, Gerrit

Explore Jobs

Expand Your Job Search