87 sre engineer jobs at 40 companies in Patterson, CA

PromotedHiringCafe
Founding Backend / Infra Engineer
Cupertino, CA, US
$160k-$300k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Own the crawlers, pipelines, and infrastructure powering a real-time job search engine. Strong Node.js and Python fundamentals; bonus points for security and reverse-engineering chops.
Node.js, Python, Elasticsearch, Redis
2w
Save
Mark Applied
Hide
Senior SRE Engineer
Santa Clara, California, United States
$148k-$276k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOE5+ years SRE/platform experience, deep Kubernetes and CI (GitLab/GitHub) expertise, scripting in Python/Go/bash, IaC with Terraform/Helm/Ansible, observability tooling experience, BS/MS in CS or equivalent.
GitLab CI, GitHub Actions, GitLab-runner, Kubernetes, Python, Go, bash, Terraform, Helm, Ansible, Argo CD, Flux, Prometheus, Grafana, Loki, ELK, OpenTelemetry
1mo
Save
Mark Applied
Hide
Senior DevOps/SRE Engineer
San Jose, California, United States
$179k-$306k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Experience building and maintaining CI/CD pipelines (GitHub Actions, Jenkins), automation using IaC, Python and Bash, SRE/leadership experience, strong CI/CD and build-system knowledge, BA/BS in CS/CE/EE or equivalent.
GitHub Actions, Jenkins, GitHub, Infrastructure as Code (IaC), Python, Bash, Windows, Linux
2w
Save
Mark Applied
Hide
Senior SRE Engineer
Santa Clara, California, United States
$148k-$276k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years SRE/platform experience, deep Kubernetes administration, GitLab/GitHub CI at scale, scripting in Python/Go/bash, IaC with Terraform/Helm/Ansible, observability tooling, BS/MS in CS or equivalent experience.
Kubernetes, GitLab CI, GitHub Actions, GitLab-runner, Python, Go, bash, Terraform, Helm, Ansible, Argo CD, Flux, Prometheus, Grafana, Loki, ELK, OpenTelemetry, Linux
2mo
Save
Mark Applied
Hide
Engineering Lead – Platform & SRE
Santa Clara, California, United States
$175k-$215k/yr HybridFull Time
Kerrigan Robotics
Kerrigan Robotics: AI-powered orchestration platform for factory automation systems.
5+ YOE5+ years SRE/DevOps experience; deep Kubernetes knowledge; expertise with Pulumi/Terraform, Helm; proficiency in Go, Linux networking, observability (Prometheus, Grafana), and CI/CD (GitHub Actions).
Pulumi, Kubernetes, AWS, GCP, Azure, GitHub Actions, Prometheus, Grafana, OIDC, CNIs, Terraform, Helm, Go, Linux
3w
Save
Mark Applied
Hide
Sr. SRE Platform Architect
San Jose or Austin
HybridFull Time
Bitdeer
BitdeerNASDAQ: BTDR: Operates cryptocurrency mining and high-performance computing data centers.
10+ YOE10+ years production SRE/platform engineering or infra-architecture (including ≥3 years architect-level). Hands-on GPU/AI compute, multi-region observability, Kubernetes and cluster platforms, data-center operations, DDD and plugin framework experience; BS/MS CS.
NVIDIA, DCGM, MIG, vGPU, NVLink, NVSwitch, XID, NCCL, InfiniBand, RoCE, Lustre, NetApp, Pure, DDN, VAST, NVMe-oF, Kubernetes, GPU Operator, Slurm, Volcano, Kueue, Ray, KubeRay, ZTP, BMC, IPMI, Redfish, GitOps
4w
Save
Mark Applied
Hide
DevOps Engineer / Site Reliability Engineer (SRE)
Bangalore or India or Milpitas or Seattle or Princeton or Cape Town or London or Zurich or Singapore or Mexico City
OnsiteFull Time
Zensar
ZensarNational Stock Exchange of India: ZENSARTECH: Global technology firm providing digital transformation and infrastructure services.
4+ YOEHands-on GCP, Kubernetes (GKE), Docker, Terraform, Jenkins, Git, Linux administration, CI/CD, monitoring (Prometheus/Grafana), SRE practices; 4+ years experience; Bachelor's or equivalent experience.
Google Cloud Platform (GCP), Compute Engine, GKE, Cloud Storage, IAM, VPC, Terraform, Deployment Manager, Docker, Kubernetes (K8s), Jenkins (Pipeline as Code), ArgoCD, Spinnaker, Git (branching, merging strategies, pull requests), GitHub, GitLab, Bitbucket, Ubuntu, RHEL, CentOS, Bash, Python, Go, Prometheus, Grafana, Google Cloud Monitoring / Logging, Stackdriver (Cloud Monitoring), Elasticsearch, Logstash, Kibana, Istio, Linkerd, Vault, Jboss/Wildfly
1w
Save
Mark Applied
Hide
Machine Learning Ops Engineer, Global SRE
San Jose, California, United States
$245k-$450k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Bachelor's in CS or equivalent; expertise in Linux, networking, storage; programming in Python, Go, C, C++, or Java; troubleshooting and production operations experience; SRE of ML systems preferred.
Linux, Python, Go, C, C++, Java
4d
Save
Mark Applied
Hide
Job Posting Title AI/ DevOps Engineer
San Jose, California, United States
$174k-$331k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
6+ YOE6–10 years SRE/infrastructure experience; strong datastore, Kubernetes, cloud (AWS/Azure/GCP), observability and incident response skills; interest in AI/ML ops; automation and reliability focus.
Aerospike, FoundationDB, Postgres, CosmosDB, DynamoDB, Kubernetes, AWS, Azure, GCP, Prometheus, Grafana, OpenTelemetry, Copilot, Claude Code, Codex
3mo
Save
Mark Applied
Hide
Observability Lead - Cloud SRE & Network Reliability
Fremont, California, United States
$114k-$253k/yr HybridFull Time
Lam Research
Lam ResearchNASDAQ: LRCX: Designs and manufactures wafer fabrication equipment for the semiconductor industry.
12+ YOE6+ MgmtSenior SRE leader with 12+ years infrastructure/SRE/DevOps experience and 6+ years leading teams; deep multi-cloud networking, observability, DR/BCP, backup/restore, automation (Ansible/Terraform/Python), Kubernetes, and incident management experience.
Prometheus, Grafana, Datadog, PagerDuty, ThousandEyes, Azure Monitor, CloudWatch, Google Cloud Operations, Splunk, Ansible, Terraform, Python, Kubernetes, AKS, EKS, GKE, ServiceNow
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE) – CloudVision as a Service (CVaaS)
Santa Clara, California, United States
$101k-$161k/yr RemoteFull Time
Arista Networks
Arista NetworksNYSE: ANET: Provides cloud networking solutions and high-speed multilayer Ethernet switches.
5+ YOEBS/MS or equivalent experience,5+ years software engineering, experience with distributed databases/SaaS deployments, proficiency in Python/Golang/Bash, Kubernetes and cloud platform experience preferred.
Golang, Python, Ansible, Pulumi, Bash, Kubernetes, GKE, GCP
3w
Save
Mark Applied
Hide
Principal Systems Design Engineer
Mountain View or Santa Clara
$193k-$337k/yr OnsiteFull Time
ServiceNow
ServiceNowNYSE: NOW: Provides a cloud platform for automating enterprise digital workflows.
12+ YOE12+ years in infrastructure architecture with experience across networking, cloud (AWS/Azure/GCP), IAM (Okta/Entra ID), SRE/IaC/CI-CD, M&A integrations, and strong cross-functional collaboration.
Okta, Entra ID, AWS, Azure, GCP, IaC, CI/CD, O365, Slack, ServiceNow, SD-WAN, VPN, SRE
2w
Save
Mark Applied
Hide
Platform Engineer I
Pleasanton, California, United States
$32-$41/hr HybridFull Time
Blackhawk Network
Blackhawk Network: Provider of global branded payment and gift card solutions.
Bachelor's in CS/Engineering or equivalent; experience in platform/DevOps/SRE or similar; strong Linux, AWS, Git; scripting with Python/Bash; production support and incident management; experience with AI-assisted engineering tools.
AWS, Git, Python, Bash, Kubernetes, Docker, Jenkins, Splunk, New Relic, Prometheus, Grafana, OpenTelemetry, ServiceNow, Terraform, CloudFormation, GitHub Copilot, Cursor, Claude
1mo
Save
Mark Applied
Hide
Senior Platform Engineer
Pleasanton or California or United States
HybridFull Time
STN
STN: Provides high-performance GPU infrastructure, cloud, and managed IT services.
6+ YOE6+ years in platform/SRE/cloud engineering, deep Kubernetes expertise, Go and/or Python programming, experience operating GPU/AI infrastructure, bachelor's degree or equivalent experience.
Kubernetes, Slurm, Run:ai, Go, Python, NVIDIA GPU Operator, MIG, MPS, NCCL, KubeRay, Istio, Linkerd, CRDs
2w
Save
Mark Applied
Hide
Platform Engineer I
Pleasanton, California, United States
$41/hr HybridFull Time
Blackhawk Network
Blackhawk Network: Provider of branded payment and gift card solutions.
Bachelor's in CS/Engineering or equivalent experience; platform/DevOps/SRE experience; strong Linux, AWS, Git; scripting in Python/Bash; experience with major incident management and AI-assisted development tools.
AWS, Kubernetes, Git, Python, Bash, GitHub Copilot, Cursor, Claude, Docker, Jenkins, Splunk, New Relic, Prometheus, Grafana, OpenTelemetry, ServiceNow, Terraform, CloudFormation, Linux
1mo
Save
Mark Applied
Hide
Senior Production Engineer
San Francisco or Sunnyvale
$209k-$253k/yr OnsiteFull Time
Crusoe
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
Senior production engineer with distributed systems experience, hands-on with large language models, SRE mindset, and strong programming skills (Python, Go, Java, or C++), Kubernetes knowledge, and collaborative skills.
Python, Go, Java, C++, Kubernetes
2w
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara or St. Louis or Bangalore or London or Paris or Melbourne or Taipei or Tokyo
OnsiteFull Time
Netskope
NetskopeNASDAQ: NTSK: Cloud-native cybersecurity and data protection platform for enterprises.
3+ YOEBachelor's in CS/Engineering or equivalent; 3+ years building/managing complex systems (including 1-2 years SRE); experience with cloud services, microservices, availability/performance optimization, debugging, and strong communication.
Python, C, C++, Go, Rust, Docker, Kubernetes, AWS, GCP, KVM, OpenNebula, OpenStack, TCP/IP
1w
Save
Mark Applied
Hide
Site Reliability Engineer - Video Infrastructure
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
SRE for global multimedia transport, storage and processing; strong SRE, networking, OS, database, container and distributed systems troubleshooting skills; bachelor's in CS or equivalent.
C, C++, Java, Python, Go, Linux, MySQL, MongoDB, Redis, ELK, AWS, Google Cloud, Azures
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Pleasanton or Austin or San Francisco or United States
OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOESenior SRE with strong infrastructure, automation, and programming experience (Terraform, Chef, Ansible, Python, Java, Bash). Minimum multi-year experience in software engineering or equivalent; participates in on-call and incident response.
Terraform, Chef, Ansible, Python, Java, Bash, Kubernetes, Helm, Jenkins, Grafana, Prometheus, OCI - DevOps, Oracle Cloud Guard, Oracle Observability and Management
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer - SDN
San Francisco or San Jose or Bellevue
$240k-$312k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Kubernetes, SmartNICs, Python, Ansible, Go, C, Helm, Terraform, GitOps, CI/CD, Linux, OpenStack Neutron, OVN, OVS, DPDK, SR-IOV
1mo
Save
Mark Applied
Hide
Senior Platform Engineer (Cloud Workloads)
San Jose or Ohio or Seattle or United States
$173k-$321k/yr RemoteFull Time
Veeam
Veeam: Data resilience and security for hybrid cloud environments
5+ YOE5+ years in cloud platform engineering/SRE supporting SaaS; deep Elastic Stack experience, Azure/AWS operations, incident response/runbooks, IaC (Bicep/Terraform/Pulumi), CI/CD, and scripting (Bash/Python/PowerShell).
Elastic Stack, Elasticsearch, Kibana, Elastic Fleet, KQL, Query DSL, Azure Kubernetes Service (AKS), Azure Container Apps, Azure, AWS, Entra ID, Managed Identities, App Registrations, Key Vault, Azure Bicep, Terraform, Pulumi, Azure DevOps, GitHub Actions, ServiceNow, Salesforce, Jira, Incident.io, Service Bus, Cosmos DB, Bash, Python, PowerShell