59 site operations engineer jobs at 42 companies in Morgan Hill, CA

2w
Save
Mark Applied
Hide
Staff Site Reliability Engineer-Production Operations
Palo Alto, California, United States
$186k-$233k/yr HybridFull Time
Rivian and Volkswagen Group Technologies
Rivian and Volkswagen Group Technologies: A joint venture creating software-defined vehicle technology and connected services for electric vehicles.
Senior SRE with incident command experience, strong systems engineering for distributed systems, hands-on coding (Python or Go), observability expertise (Datadog or comparable), and experience with blameless post-incident practices.
Python, Go, Datadog, LLM
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Mountain View, California, United States
$252k-$308k/yr HybridFull Time
EarnIn
EarnIn: Provides immediate access to earned wages through a mobile app.
7+ YOE7+ years in SRE or related field; experience applying AI/LLMs to operations; strong SLO/SLI and incident management; software engineering in Python or Go; observability and IaC proficiency; AI-assisted development tools; fintech/regulated environment experience.
Datadog, CloudWatch, OpenTelemetry, Terraform, Kubernetes, AWS, Python, Go, Cursor, Claude Code, Copilot
1w
Save
Mark Applied
Hide
ZfG Operations Engineer
San Jose, California, United States
$99k-$229k/yr HybridFull Time
Zoom
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
Proven experience in site reliability, DevOps, or infrastructure operations; proficiency with Python, Go, or Bash; knowledge of distributed systems, monitoring/observability, incident management, and automation; experience supporting production systems and on-call rotations.
Python, Go, Bash, BrightHire
3mo
Save
Mark Applied
Hide
Lab Operations Site Supervisor
Santa Clara, California, United States
$60k-$121k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
2+ YOE1+ MgmtBachelor's degree in a technical field; 2+ years lab engineering/technical experience; strong lab equipment knowledge; junior management experience; ability to travel within the Bay Area.
MS Visio, MS SharePoint, Jira
3mo
Save
Mark Applied
Hide
Lab Operations Site Supervisor
Santa Clara, California, United States
$60k-$121k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
2+ YOEBachelor's in a technical field or equivalent experience, 2+ years experience, lab equipment and engineering background, junior management experience, ability to assemble and move equipment, debug PCBs, run tests on Windows/Linux, and occasionally travel within the Bay Area.
Microsoft Visio, Microsoft SharePoint, Jira, Windows, Linux, oscilloscopes
5d
Save
Mark Applied
Hide
Senior Site Reliability Engineer Platform Private Cloud Engineer
San Jose, California, United States
$94k-$130k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
7+ YOERequires 7+ years designing and operating enterprise or cloud environments, private cloud and Kubernetes expertise, scripting, IaC tools, Unix/Linux knowledge, and a CS or engineering degree.
VMware, AWS, GCP, Kubernetes, Helm, ArgoCD, Python, Bash, Ruby, Scala, Ansible, Terraform, Unix, Linux
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Palo Alto, California, United States
$200k-$400k/yr HybridFull Time
Nectar Social
Nectar Social: AI platform for social commerce and community management.
5+ YOE5+ years operating production systems; cloud (AWS); infrastructure as code; programming; startup environment; reliability-focused with cost awareness.
AWS, Pulumi, Postgres, ClickHouse, Turbopuffer, Temporal
4d
Save
Mark Applied
Hide
Site Reliability Engineer (Multiple Positions)
San Jose, California, United States
$226k-$317k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
1+ YOEMaster's degree and 1 year of related experience, or bachelor's degree and 3 years; requires 1 year supporting critical systems, monitoring, troubleshooting, data operations, error analysis, and runbook creation.
1mo
Save
Mark Applied
Hide
Cloud Site Reliability Engineer - DCS Cloud
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOEBachelor's degree in CS or related,2+ years in Linux operations/SRE/DevOps,programming in Go/Python/C++,cloud and reliability practices experience,strong troubleshooting and communication skills.
Go, Python, C++, Linux, OCI, AWS, Azure, GCP, KVM, QEMU, Docker, Kubernetes, containerd, cgroups, namespaces, CUDA, MIG
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Foster City, California, United States
$250k-$300k/yr HybridFull Time
Zoox
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
5+ YOE5+ years operating GitHub Enterprise at scale, monorepo management, CI/CD integration, infrastructure-as-code (Terraform/Pulumi), cloud platform experience, technical leadership and migration planning.
Git, GitHub Enterprise, GitHub Cloud, Buildkite, GitHub Actions, Jenkins, GitLab CI, Terraform, Pulumi, Bazel, Buck, Reviewable, Gerrit
5d
Save
Mark Applied
Hide
Engineering Operations Manager, on-site
Sunnyvale, California, United States
$145k-$155k/yr OnsiteFull Time
CBRE
CBRENYSE: CBRE: Provides global commercial real estate services and investment management.
5+ YOEBachelor's degree preferred or equivalent experience; 5+ years in engineering operations, facilities management, or building systems; leadership, budgeting, reporting, vendor management, and Microsoft Office expertise required.
Microsoft Office, Microsoft Excel, Microsoft Word, Microsoft Outlook, Microsoft PowerPoint
2mo
Save
Mark Applied
Hide
Site Reliability Engineer - USDS (Multiple Positions)
San Jose, California, United States
$188k-$259k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
1+ YOERequires degree in CS/Engineering/Information Systems/Data Science/Mathematics plus relevant experience; experience with Linux administration, monitoring, troubleshooting, SDLC and cloud-native operations.
Linux
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE) (Hybrid)
San Francisco or San Jose or New York City or Milpitas or Mountain View or Holmdel or Goleta or Redwood City or Fremont or Sunnyvale or Brooklyn or Palo Alto
$168k-$245k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
4+ YOERequires 7+ years with a bachelor's, 4+ with a master's, or 1 with a PhD; 4+ years in SRE or related engineering, 3+ years operating Kubernetes, Helm, CI/CD, cloud platforms, and Python or Go.
Kubernetes, Helm, CI/CD, AWS, GCP, Python, Go, Terraform, MLOps, VPC, DNS
1mo
Save
Mark Applied
Hide
ASE Senior Site Reliability Engineer
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Design and build secure end-to-end solutions, develop server-side systems, APIs and tooling to operate large-scale services while upholding privacy and high performance.
6d
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Storage
San Francisco or San Jose
$267k-$356k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOERequires 5+ years operating Linux systems in production or HPC environments, large-scale storage experience, incident response, monitoring, Kubernetes, CI/CD, Python or Go, and Terraform or Ansible.
Linux, CEPH, Lustre, GPFS, Prometheus, Grafana, Alertmanager, Datadog, SumoLogic, Kubernetes, ArgoCD, Helm, Kustomize, GitHub Actions, Jenkins, BuildKite, Docker, Podman, Python, Go, Terraform, Ansible, NFS, SMB, S3, NVMe-oF/TCP, VAST, Weka, NetApp, Dell PowerScale, KVM/QEMU, GPUDirect Storage, RDMA, InfiniBand, RoCE, ethtool, mlxlink, clush
1mo
Save
Mark Applied
Hide
On-site Support Engineer
San Carlos, California, United States
FieldFull Time
Genesis AI
Genesis AI: Building general-purpose robots with human-level physical intelligence.
Hands-on experience with robotics, hardware, teleoperation, or data-systems operations; comfortable debugging across hardware, software, and data; Linux and Python proficiency; customer-facing and travel-ready.
Linux, Python
1mo
Save
Mark Applied
Hide
Infrastructure Engineer (Data Center Operations)
Sunnyvale, California, United States
OnsiteFull Time
Cerebras Systems
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
3+ YOE3+ years in data center or infrastructure engineering; strong Linux administration, x86 hardware, enterprise networking, BIOS/firmware and remote management; scripting with Bash or Python; on-site work and ability to lift/move servers.
ip, dmesg, netstat, ping, Bash, Python, IPMI, iDRAC, iLO, PXE, NFS, RAID, Ansible, BIOS
3w
Save
Mark Applied
Hide
Sr. IAM Site Reliability Engineer
Santa Clara, California, United States
$186k-$279k/yr OnsiteFull Time
Pure Storage
Pure StorageNYSE: PSTG: Provides all-flash enterprise data storage and management solutions.
Operate and automate enterprise IAM platforms, implement IAC and automation, build observability and SLIs/SLOs, lead incident response and documentation; hands-on Terraform/Ansible and scripting experience required.
Terraform, Ansible, Tines, Python, PowerShell, Bash, Datadog, Prometheus, Grafana, Splunk
1w
Save
Mark Applied
Hide
IT & Security Operations Engineer
Santa Clara, California, United States
$88k-$140k/yr OnsiteFull Time
Bambu Lab
Bambu Lab: Manufacturer of desktop 3D printers and 3D printing accessories.
2+ YOEBachelor's or equivalent experience,2+ years IT on-site support/operations,proficient with Windows/macOS/Linux,networking fundamentals,experience with SCCM/Intune/GPO/MDM,SSO/MFA/VPN,Shell/PowerShell/Python,AWS/GCP,English communication.
Windows, macOS, Linux, TCP/IP, Wi-Fi, Wireshark, SCCM, Intune, GPO, MDM, SSO, MFA, VPN, Zero-trust, Shell, PowerShell, Python, AWS, GCP
2mo
Save
Mark Applied
Hide
Site Assessment and Remediation Specialist
Torrance or Oakland or San Mateo
$90k-$110k/yr FieldFull Time
NV5
NV5NASDAQ: NVEE: Technical engineering and consulting for infrastructure and energy projects.
1+ YOEDegree in physical sciences/environmental science/engineering,1–3 years site/remediation experience,soil/soil-vapor sampling experience,ability to oversee SVE installation/operation,MS Office proficiency,GIS/AutoCAD a plus,California driver’s license required.
Microsoft Office Suite, GIS, AutoCAD