73 site operations engineer jobs at 55 companies in Fairview, CA

2w
Save
Mark Applied
Hide
Staff Site Reliability Engineer-Production Operations
Palo Alto, California, United States
$186k-$233k/yr HybridFull Time
Rivian and Volkswagen Group Technologies
Rivian and Volkswagen Group Technologies: A joint venture creating software-defined vehicle technology and connected services for electric vehicles.
Senior SRE with incident command experience, strong systems engineering for distributed systems, hands-on coding (Python or Go), observability expertise (Datadog or comparable), and experience with blameless post-incident practices.
Python, Go, Datadog, LLM
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
San Francisco, California, United States
$149k-$224k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
5+ YOE5+ years systems and software engineering experience for large-scale internet services; expertise in SRE principles, containers, observability, incident management, Python and Go, and applying AI/ML to operations.
Temporal, Airflow, Argo Workflows, Docker, Kubernetes, DNS, HTTP, Grafana, Prometheus, ELK, Splunk, Datadog, Python, Go, Linux, Claude Code, GitHub Copilot, Codex, Cursor, AWS, GCP, MCP
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Mountain View, California, United States
$252k-$308k/yr HybridFull Time
EarnIn
EarnIn: Provides immediate access to earned wages through a mobile app.
7+ YOE7+ years in SRE or related field; experience applying AI/LLMs to operations; strong SLO/SLI and incident management; software engineering in Python or Go; observability and IaC proficiency; AI-assisted development tools; fintech/regulated environment experience.
Datadog, CloudWatch, OpenTelemetry, Terraform, Kubernetes, AWS, Python, Go, Cursor, Claude Code, Copilot
2mo
Save
Mark Applied
Hide
Lab Operations Site Supervisor
Santa Clara, California, United States
$60k-$121k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
2+ YOE1+ MgmtBachelor's degree in a technical field; 2+ years lab engineering/technical experience; strong lab equipment knowledge; junior management experience; ability to travel within the Bay Area.
MS Visio, MS SharePoint, Jira
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
San Francisco or New York City
$164k-$306k/yr HybridFull Time
Retool
Retool: Software platform for building custom internal business applications.
Experience operating production infrastructure (AWS), Kubernetes, Terraform, Postgres; programming in Go/Python/TypeScript/Java/Ruby; building observability and automation for customer-facing SaaS systems.
Kubernetes, Helm, Docker Compose, Terraform, AWS, Postgres, Go, Python, TypeScript, Java, Ruby
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
San Francisco, California, United States
$195k-$258k/yr RemoteFull Time
Circle
CircleNYSE: CRCL: Digital currency issuer and blockchain financial infrastructure provider.
6+ YOE6+ years SRE/DevOps experience preferred; strong Kubernetes, IaC (Terraform/Pulumi), cloud networking, Go or Python, CI/CD, observability, and distributed systems experience; blockchain node operation experience is a plus.
Kubernetes, Helm, Terraform, Pulumi, Go, Python, CI/CD, RBAC, VPCs, DNS, SQL
2mo
Save
Mark Applied
Hide
Lab Operations Site Supervisor
Santa Clara, California, United States
$60k-$121k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
2+ YOEBachelor's in a technical field or equivalent experience, 2+ years experience, lab equipment and engineering background, junior management experience, ability to assemble and move equipment, debug PCBs, run tests on Windows/Linux, and occasionally travel within the Bay Area.
Microsoft Visio, Microsoft SharePoint, Jira, Windows, Linux, oscilloscopes
13h
Save
Mark Applied
Hide
Senior Site Reliability Engineer Platform Private Cloud Engineer
San Jose, California, United States
$94k-$130k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
7+ YOERequires 7+ years designing and operating enterprise or cloud environments, private cloud and Kubernetes expertise, scripting, IaC tools, Unix/Linux knowledge, and a CS or engineering degree.
VMware, AWS, GCP, Kubernetes, Helm, ArgoCD, Python, Bash, Ruby, Scala, Ansible, Terraform, Unix, Linux
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Palo Alto, California, United States
$200k-$400k/yr HybridFull Time
Nectar Social
Nectar Social: AI platform for social commerce and community management.
5+ YOE5+ years operating production systems; cloud (AWS); infrastructure as code; programming; startup environment; reliability-focused with cost awareness.
AWS, Pulumi, Postgres, ClickHouse, Turbopuffer, Temporal
3w
Save
Mark Applied
Hide
Staff Site Reliability Engineer (FedRAMP)
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
AWS, Terraform, Chef, Ansible, Puppet, Apache httpd, nginx, Apache Tomcat, Bash, Python, Golang, git, gdb, strace, ltrace, tcpdump, Wireshark, Docker, Kubernetes
1mo
Save
Mark Applied
Hide
Cloud Site Reliability Engineer - DCS Cloud
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOEBachelor's degree in CS or related,2+ years in Linux operations/SRE/DevOps,programming in Go/Python/C++,cloud and reliability practices experience,strong troubleshooting and communication skills.
Go, Python, C++, Linux, OCI, AWS, Azure, GCP, KVM, QEMU, Docker, Kubernetes, containerd, cgroups, namespaces, CUDA, MIG
2mo
Save
Mark Applied
Hide
Member of Technical Staff, Site Reliablity Engineer
San Francisco, California, United States
$200k-$270k/yr HybridFull Time
Vapi
Vapi: Build and deploy AI-powered voice agents via flexible APIs.
Experience running incident command and postmortems, operating SLOs/error budgets, capacity planning and load testing, Kubernetes production ops, KEDA autoscaling, and shipping services in Go or TypeScript.
Go, TypeScript, Bash, Chronosphere, Prometheus, Grafana, Datadog, OpenTelemetry, Kubernetes, EKS, KEDA
1mo
Save
Mark Applied
Hide
Site Reliability Engineer - Global SRE, Monetization Technology
San Jose, California, United States
$156k-$317k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
3+ YOEBachelor's or equivalent, 3+ years programming experience (C,C++,Java,Python,Perl,Go), Unix/Linux and IP networking expertise, production operations and automation experience, strong communication and ownership.
C, C++, Java, Python, Perl, Go, Unix, Linux, IP networking
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Lincoln or San Francisco
$125k-$165k/yr RemoteFull Time
TELCOR
TELCOR: Provides healthcare software for laboratory and point-of-care operations.
2+ YOEExperience with distributed systems, Redis, queuing systems, Kubernetes (2+ yrs), AWS, Terraform (2+ yrs), observability, and production operations.
Redis, Kubernetes, AWS, Terraform
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Foster City, California, United States
$250k-$300k/yr HybridFull Time
Zoox
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
5+ YOE5+ years operating GitHub Enterprise at scale, monorepo management, CI/CD integration, infrastructure-as-code (Terraform/Pulumi), cloud platform experience, technical leadership and migration planning.
Git, GitHub Enterprise, GitHub Cloud, Buildkite, GitHub Actions, Jenkins, GitLab CI, Terraform, Pulumi, Bazel, Buck, Reviewable, Gerrit
2mo
Save
Mark Applied
Hide
Site Reliability Engineer - USDS (Multiple Positions)
San Jose, California, United States
$188k-$259k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
1+ YOERequires degree in CS/Engineering/Information Systems/Data Science/Mathematics plus relevant experience; experience with Linux administration, monitoring, troubleshooting, SDLC and cloud-native operations.
Linux
3mo
Save
Mark Applied
Hide
Assistant Site Manager (Electrical Operations)
San Francisco or New York or Denver or Austin or Calgary or Toronto
$118k-$135k/yr HybridFull Time
Intersect
IntersectNASDAQ: GOOGL: Develops-located data centers and renewable energy infrastructure.
1+ YOEB.S. in Electrical Engineering; 1–3 years field experience; interpret electrical diagrams; analyze data from SCADA/CMMS; strong communication; safety-focused.
SCADA, Inverter platforms, CMMS, Electrical Schematics
5d
Save
Mark Applied
Hide
Staff Site Reliability Engineer (SRE) (Hybrid)
San Francisco or San Jose or New York City or Milpitas or Mountain View or Holmdel or Goleta or Redwood City or Fremont or Sunnyvale or Brooklyn or Palo Alto
$187k-$268k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
6+ YOERequires 8+ years with a bachelor's, 6+ with a master's, or 3+ with a PhD; 6+ years in SRE or infrastructure engineering, 5+ years operating Kubernetes, cloud, CI/CD, and Python or Go.
Kubernetes, AWS, GCP, Python, Go, Terraform, MLOps, CI/CD
1mo
Save
Mark Applied
Hide
ASE Senior Site Reliability Engineer
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Design and build secure end-to-end solutions, develop server-side systems, APIs and tooling to operate large-scale services while upholding privacy and high performance.
1d
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Storage
San Francisco or San Jose
$267k-$356k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOERequires 5+ years operating Linux systems in production or HPC environments, large-scale storage experience, incident response, monitoring, Kubernetes, CI/CD, Python or Go, and Terraform or Ansible.
Linux, CEPH, Lustre, GPFS, Prometheus, Grafana, Alertmanager, Datadog, SumoLogic, Kubernetes, ArgoCD, Helm, Kustomize, GitHub Actions, Jenkins, BuildKite, Docker, Podman, Python, Go, Terraform, Ansible, NFS, SMB, S3, NVMe-oF/TCP, VAST, Weka, NetApp, Dell PowerScale, KVM/QEMU, GPUDirect Storage, RDMA, InfiniBand, RoCE, ethtool, mlxlink, clush