63 service reliability engineer jobs at 46 companies in Mill Valley, CA
1mo
Save
Mark Applied
Hide
1mo
Site Reliability Engineer
Santa Clara or St. Louis or Bangalore or London or Paris or Melbourne or Taipei or Tokyo
OnsiteFull Time
NetskopeNASDAQ: NTSK: Cloud-native cybersecurity and data protection platform for enterprises.
3+ YOEBachelor's in CS/Engineering or equivalent; 3+ years building/managing complex systems (including 1-2 years SRE); experience with cloud services, microservices, availability/performance optimization, debugging, and strong communication.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
5+ YOE5+ years systems and software engineering experience for large-scale internet services; expertise in SRE principles, containers, observability, incident management, Python and Go, and applying AI/ML to operations.
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Work on globally scaled, revenue-critical internet services (App Store, Music, Books, Podcasts, Fitness+); ensure reliability and scalability of services used by billions of devices.
5+ YOEBS/MS or equivalent experience,5+ years software engineering, experience with distributed databases/SaaS deployments, proficiency in Python/Golang/Bash, Kubernetes and cloud platform experience preferred.
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOEBS in CS or equivalent with 5+ years supporting critical services; experience with large-scale HPC clusters (Slurm, LSF, Kubernetes), IaC, CI/CD, multi‑cloud (AWS/GCP/OCI), and 2+ languages such as Python or Go.
Saviynt: Provides AI-powered identity governance and cloud security platforms.
9+ YOE9+ years in platform/infra/SRE roles, deep Kubernetes and GCP expertise, strong Go and Python skills, experience with CI/CD, event-driven systems, observability, distributed systems, and building shared platform services.
Site Reliability Engineer for Linux administration
Ontario or Greenville or Clearwater or Fremont
OnsiteFull Time
Hyve SolutionsNYSE: SNX: Designs and manufactures custom hardware for hyperscale data centers.
3+ YOE3+ years Linux production administration (RHEL/Ubuntu/Rocky/CentOS); knowledge of core services, storage, backups, monitoring, security hardening, and basic scripting/automation. Bachelor's in CS/IT or equivalent experience; RHCSA/CompTIA Linux+/LPIC-1 are nice-to-have.
1X: Manufacturing safe, general-purpose humanoid robots for home and work.
Experienced software engineer with reliability and testing expertise across cloud, mobile, on-robot services, and embedded systems; strong DevOps and Python/Linux skills; statistical and systems reasoning.
Seattle or San Francisco or Detroit or United States
$180k-$279k/yrHybridFull Time
Rocket CompaniesNYSE: RKT: Provides digital mortgage, real estate, and personal finance services.
7+ YOE7+ years AWS/cloud infra; 5+ years PostgreSQL/AWS services; Linux admin/scripting; mentoring; infrastructure as code and security; AI code generation tools; on-call readiness.
AWS, PostgreSQL, Aurora/RDS, S3, ElastiCache, OpenSearch, DynamoDB, Linux, Python, Infrastructure as Code, Security practices, AI code generation tools
Sacramento or San Francisco or San Jose or Modesto or Stockton
$80k-$90k/yrFieldFull Time
MGC Diagnostics: Sells non-invasive cardiorespiratory diagnostic systems and respiratory software.
2+ YOEAssociate or bachelor's in electronics, biomedical engineering, or related; minimum 2 years field service experience servicing medical equipment; valid driver’s license, reliable transportation; strong troubleshooting, communication, and computer skills.
Member of Technical Staff, Site Reliablity Engineer
San Francisco, California, United States
$200k-$270k/yrHybridFull Time
Vapi: Build and deploy AI-powered voice agents via flexible APIs.
Experience running incident command and postmortems, operating SLOs/error budgets, capacity planning and load testing, Kubernetes production ops, KEDA autoscaling, and shipping services in Go or TypeScript.
Senior Software Engineer - Observability and Reliability
New York City or San Francisco or London or Sydney
$170k-$240k/yrOnsiteFull Time
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
5+ YOE5+ years building high-quality software, strong CS fundamentals, experience building observability tools, proficiency with Go, OpenTelemetry, Kubernetes, participation in on-call rotations, cloud service administration (GCP/AWS/Azure) preferred.
Polymr: AI-native procurement and manufacturing workflow automation platform.
Experience building multi-tenant systems, data layers, integrations, deployments, and reliable production services; customer-facing collaboration; familiarity with AI coding tools preferred.
8+ YOEBA/BS in engineering; 8+ years infrastructure design/operation experience; 3+ years campus environment experience; proficiency with BAS and CMMS (e.g., SAP, Siemens); knowledge of NFPA, ASHRAE, OSHA; reliability engineering and capital project leadership.
Noble Thermodynamic Systems: Developing zero-emissions, ultra-efficient power generation technology.
Owns piping systems across power projects from conceptual design through construction, commissioning, and operations support, ensuring reliability, safety, accessibility, and serviceability.
DocuSignNASDAQ: DOCU: Provider of e-signature and intelligent agreement management software.
8+ YOE8+ years full-stack development experience with frontend and backend technologies; experience designing backend APIs, CI/CD, service reliability, and strong communication skills.
Applied MaterialsNASDAQ: AMAT: Manufacturers of equipment for semiconductor and display production.
Experience developing service procedures and troubleshooting complex equipment, conducting failure analysis, creating technical manuals and training, and participating in quality/reliability and CAPA processes.
San Francisco or Sacramento or West Sacramento or United States
$7k-$11k/moHybridFull Time
State Controller's Office: California's fiscal controller managing state financial operations and assets.
Software engineering experience across applications, databases, APIs, integrations, testing, cloud services, delivery pipelines, and production operations, with strong judgment in security, privacy, accessibility, and reliability.