34 staff site reliability engineer jobs at 25 companies in Daly City, CA

3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Mountain View, California, United States
$252k-$308k/yr HybridFull Time
EarnIn
EarnIn: Fintech helping workers access earned wages in real time and manage finances without interest or mandatory fees.
7+ YOE7+ years in SRE or related field; experience applying AI/LLMs to operations; strong SLO/SLI and incident management; software engineering in Python or Go; observability and IaC proficiency; AI-assisted development tools; fintech/regulated environment experience.
Datadog, CloudWatch, OpenTelemetry, Terraform, Kubernetes, AWS, Python, Go, Cursor, Claude Code, Copilot
3w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
San Francisco, California, United States
$195k-$258k/yr RemoteFull Time
Circle Internet Group, Inc.
Circle Internet Group, Inc.NYSE: CRCL: Public financial technology providing stablecoin, digital-asset, payments, and blockchain infrastructure to businesses and developers.
6+ YOE6+ years SRE/DevOps experience preferred; strong Kubernetes, IaC (Terraform/Pulumi), cloud networking, Go or Python, CI/CD, observability, and distributed systems experience; blockchain node operation experience is a plus.
Kubernetes, Helm, Terraform, Pulumi, Go, Python, CI/CD, RBAC, VPCs, DNS, SQL
1d
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Ads
San Francisco or United States
$217k-$304k/yr RemoteFull Time, Contract
Reddit
RedditNYSE: RDDT: Social news aggregation, web content rating, and discussion platform.
8+ YOE8+ years in site reliability or infrastructure engineering, distributed systems, cloud-native architecture, observability, automation, incident management, performance optimization, and backend software engineering.
Go, Kubernetes, Kafka, ClickHouse, Spark, Flink, BigQuery
2w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
San Mateo or United States
$240k-$300k/yr RemoteFull Time
Skydio
Skydio: American autonomous-drone manufacturer serving public safety, government, utility, and enterprise customers.
8+ YOE8+ years in SRE, platform, DevOps, production engineering, or equivalent; strong Kubernetes and AWS experience; Terraform, CI/CD, networking, observability, and production reliability expertise.
Kubernetes, Amazon Web Services (AWS), Amazon Elastic Kubernetes Service (EKS), Terraform, Argo CD, Spinnaker, GitHub Actions, GitLab CI/CD, Jenkins, Linux, Python, Go, Helm, GitOps, Datadog, PostgreSQL
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
New York or Mountain View
$200k-$220k/yr HybridFull Time
ASAPP
ASAPP: ASAPP is a private enterprise AI software providing agentic customer-service platforms for contact centers.
10+ YOE10+ years production software experience; distributed systems; automation; AWS; Terraform; Python/Go; Kubernetes; on-call; strong communication.
Terraform, Python, Go, AWS, Kubernetes
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Mountain View, California, United States
$218k-$260k/yr OnsiteFull Time
ID.me
ID.me: Private American digital identity wallet and identity-verification helping people securely access government, healthcare, and commercial services.
10+ YOEBachelor's in CS; 10+ years coding; 5+ years cloud (GCP preferred, AWS acceptable).
GCP, AWS, CI/CD, Observability, Infrastructure as Code
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
San Jose, California, United States
$119k-$170k/yr HybridFull Time
Zscaler
ZscalerNASDAQ: ZS: Cloud-native Zero Trust cybersecurity platform for digital transformation.
5+ YOE5+ years Linux/UNIX admin, Kubernetes/Docker, automation (Ansible), network/security fundamentals, and strong security practices.
Docker, Kubernetes, Ansible, Python, Golang, BASH, Openstack, CEPH, HashiCorp Vault, nftables
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer (FedRAMP)
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Identity management and access control software provider.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
AWS, Terraform, Chef, Ansible, Puppet, Apache httpd, nginx, Apache Tomcat, Bash, Python, Golang, git, gdb, strace, ltrace, tcpdump, Wireshark, Docker, Kubernetes
3mo
Save
Mark Applied
Hide
Senior Staff Site Reliability Engineer
San Francisco or New York City or Chicago
$245k-$270k/yr HybridFull Time
Ironclad
Ironclad: Private software providing AI contract lifecycle management tools for legal and business teams.
8+ YOE8+ years DevOps/SRE; 5+ years coding; Kubernetes and GCP expertise; build resilient infra; GitOps with Terraform/Pulumi, CircleCI, ArgoCD; AI tooling experience; strong communication; cross-functional collaboration.
Kubernetes, Google Cloud Platform, Terraform, Pulumi, CircleCI, ArgoCD, Claude Code, Cursor, Zed
5d
Save
Mark Applied
Hide
Staff Site Reliability Engineer
San Francisco or Chicago or Dallas or United States
$121k-$151k/yr OnsiteFull Time
WEX
WEXNYSE: WEX: Public fintech and payments helping businesses manage fleet fuel, employee benefits, and corporate payments.
8+ YOE8+ years in large-scale system reliability; expertise in architecture, cloud platforms, automation, Kubernetes, service meshes, distributed tracing, observability, AI agents, and production AI security and governance.
Kubernetes, Grafana, ELK, Splunk, Docker, MySQL, PostgreSQL, OpenTelemetry, Jaeger, Prometheus, APIs, CI/CD
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer-Production Operations
Palo Alto, California, United States
$186k-$233k/yr HybridFull Time
Rivian and Volkswagen Group Technologies
Rivian and Volkswagen Group Technologies: Joint-venture automotive technology developing software-defined vehicle architecture and software for Rivian and Volkswagen Group electric vehicles.
Senior SRE with incident command experience, strong systems engineering for distributed systems, hands-on coding (Python or Go), observability expertise (Datadog or comparable), and experience with blameless post-incident practices.
Python, Go, Datadog, LLM
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Foster City, California, United States
$250k-$300k/yr HybridFull Time
Zoox
Zoox: Autonomous mobility developing a fully electric robotaxi fleet.
5+ YOE5+ years operating GitHub Enterprise at scale, monorepo management, CI/CD integration, infrastructure-as-code (Terraform/Pulumi), cloud platform experience, technical leadership and migration planning.
Git, GitHub Enterprise, GitHub Cloud, Buildkite, GitHub Actions, Jenkins, GitLab CI, Terraform, Pulumi, Bazel, Buck, Reviewable, Gerrit
3mo
Save
Mark Applied
Hide
Senior Staff Site Reliability Engineer
San Francisco, California, United States
$181k-$263k/yr OnsiteFull Time
LiveRamp
LiveRampNYSE: RAMP: Public data collaboration platform helping brands, publishers, retailers, and media platforms unify, activate, and measure data.
10+ YOESenior SRE with 10+ years in production engineering; Kubernetes, Terraform, Python/Go; globally distributed systems; CI/CD; FinOps; strong communication.
Terraform, Kubernetes, Python, Go, Jenkins, CircleCI, SingleStore, ScyllaDB, Cassandra, DynamoDB, AWS, GCP
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer - Paze
San Francisco or Chicago or Scottsdale or New York City or Phoenix or Washington
$116k-$174k/yr HybridFull Time
Early Warning Services
Early Warning Services: U.S. bank-owned fintech and consumer reporting agency providing identity, fraud-risk, and real-time payment solutions to financial institutions.
8+ YOE8+ years SRE/engineering experience, bachelor’s degree preferred, hands-on with Python/Go/Java, Docker, CI/CD, Linux, cloud and messaging/datastore technologies; on-call rotation and incident leadership experience.
Python, Go, Java, Docker, Kafka, SQS, JMS, Oracle, Dynamo DB, Aurora, Redis, memcached, Linux, GIT, Chef, Maven, Jenkins, Kubernetes, Swarm, AWS
2d
Save
Mark Applied
Hide
Senior Staff Site Reliability Engineer - Compute Core Engineering
United States or Santa Clara
$200k-$322k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
12+ YOEBachelor's degree or equivalent experience and 12+ years in compute platform engineering focused on automation. Requires containerization, distributed systems, Terraform, Linux kernel internals, Go or Python, and networking expertise.
DNS, NTP/PTP, DHCP, LDAP, SR-IOV, DPU, eBPF, XDP, Terraform, Go, Python, Linux, VLAN, VxLAN, SDN, BGP, Anycast
2d
Save
Mark Applied
Hide
Senior Staff Site Reliability Engineer - Compute Core Engineering
United States or Santa Clara
$200k-$322k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
12+ YOEBachelor’s degree or equivalent experience, 12+ years in compute platform engineering and automation, containerization and distributed systems expertise, Linux kernel proficiency, Terraform, configuration management, Go or Python, and network architecture knowledge.
DNS, NTP, PTP, DHCP, LDAP, SR-IOV, DPU, eBPF, XDP, Terraform, Go, Python, Linux, Kernel Internals, BareMetal, VLAN, VxLAN, SDN, BGP, Anycast, containers, microservices, IaC
3w
Save
Mark Applied
Hide
Staff Platform Site Reliability Engineer
London or Toronto or New York City or Montreal or Kitchener or San Francisco
HybridFull Time
Index Exchange
Index Exchange: Independent ad-tech supply-side platform helping media owners monetize digital content and enabling brands to buy programmatic advertising.
8+ YOERequires 8+ years in platform engineering, SRE, infrastructure engineering, or DevOps; deep Linux and Kubernetes expertise; IaC at scale; Go or Python; networking; and cross-team technical strategy.
Kubernetes, Terraform, Ansible, GitOps, ArgoCD, Go, Python, Linux, EKS, GKE, Ceph, Hadoop, Spark, HBase, Kafka, Prometheus, Grafana, ELK, Mimir, Loki, Tempo, Vault, AWS, GCP
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Semantic Understanding
San Jose, California, United States
$207k-$300k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOG, GOOGL: Global technology specializing in internet-related services and products.
8+ YOE3+ MgmtBachelor’s degree or equivalent experience; 8 years software development, 4 years Design for Reliability, 3 years SRE, project leadership, and distributed systems experience.
Google Infrastructure
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Waymo Fleet
San Francisco or Mountain View or United States or North America
$251k-$310k/yr OnsiteFull Time
Waymo
Waymo: Autonomous driving technology and robotaxi service provider.
8+ YOERequires 8+ years architecting mission-critical systems in C++, Java, or Python; reliability leadership; cross-functional influence; and a bachelor's degree or 10+ years of similar experience. Advanced degree preferred.
C++, Java, Python
3mo
Save
Mark Applied
Hide
Staff Software Engineer - Reliability
Palo Alto, California, United States
$218k-$328k/yr OnsiteFull Time
Rubrik
RubrikNYSE: RBRK: Public cybersecurity and AI operations software helping organizations protect, monitor, and recover data, identities, and workloads.
8+ YOEUS citizen; 8-12+ years software engineering with SRE/DevOps; BS/MS/PhD in CS/CE or related field; proficient in Go/Python/Java; distributed systems; Unix/Linux; on-call; leadership experience.
Go, Python, Java, Kubernetes, MySQL, Terraform, Pulumi, Prometheus, Grafana, OpenTelemetry