181 sre engineer jobs at 95 companies in Tracy, CA
🚀PromotedHiringCafe
Founding Backend / Infra Engineer
Cupertino, CA, US
$160k-$300k/yrOn-SiteFull Time
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Own the crawlers, pipelines, and infrastructure powering a real-time job search engine. Strong Node.js and Python fundamentals; bonus points for security and reverse-engineering chops.
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOE5+ years SRE/platform experience, deep Kubernetes and CI (GitLab/GitHub) expertise, scripting in Python/Go/bash, IaC with Terraform/Helm/Ansible, observability tooling experience, BS/MS in CS or equivalent.
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Experience building and maintaining CI/CD pipelines (GitHub Actions, Jenkins), automation using IaC, Python and Bash, SRE/leadership experience, strong CI/CD and build-system knowledge, BA/BS in CS/CE/EE or equivalent.
GitHub Actions, Jenkins, GitHub, Infrastructure as Code (IaC), Python, Bash, Windows, Linux
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years SRE/platform experience, deep Kubernetes administration, GitLab/GitHub CI at scale, scripting in Python/Go/bash, IaC with Terraform/Helm/Ansible, observability tooling, BS/MS in CS or equivalent experience.
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
8+ YOE8+ years in DevOps/SRE or platform engineering; production Kubernetes; CI/CD with GitHub Actions/Jenkins/Tekton; IaC (Terraform/Pulumi); GitOps (ArgoCD/Flux); Observability (Prometheus/Grafana); multi-cloud or multi-region experience; able to work in Sunnyvale office 2+ days.
RingCentralNYSE: RNG: Sells cloud-based business phone and video conferencing software.
2+ YOEMaintain 24x7 production availability, implement automation/orchestration, partner with development, perform root cause analysis; required experience with cloud, containers, scripting, and monitoring.
Kerrigan Robotics: AI-powered orchestration platform for factory automation systems.
5+ YOE5+ years SRE/DevOps experience; deep Kubernetes knowledge; expertise with Pulumi/Terraform, Helm; proficiency in Go, Linux networking, observability (Prometheus, Grafana), and CI/CD (GitHub Actions).
BitdeerNASDAQ: BTDR: Operates cryptocurrency mining and high-performance computing data centers.
10+ YOE10+ years production SRE/platform engineering or infra-architecture (including ≥3 years architect-level). Hands-on GPU/AI compute, multi-region observability, Kubernetes and cluster platforms, data-center operations, DDD and plugin framework experience; BS/MS CS.
8+ YOEBachelor's degree or equivalent,8 years software/systems engineering experience,5 years SRE experience,5 years software design experience,EMR not mentioned; strong troubleshooting and stakeholder management skills.
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
Proven SRE/DevOps experience in hybrid cloud, Linux administration, networking (DHCP/PXE/NTP), IaC/configuration management, automation, and mentoring; growth mindset and independent execution.
GEICO: Provides vehicle and property insurance services to consumers.
8+ YOE8+ years in software or site reliability engineering; 5+ years in SRE/DevOps; strong Python; Golang preferred; AWS/Azure/GCP experience; CI/CD and IaC; observability and incident response; security tooling familiarity.
TikTok: Global short-form video hosting and social media platform.
Bachelor's in CS or equivalent; expertise in Linux, networking, storage; programming in Python, Go, C, C++, or Java; troubleshooting and production operations experience; SRE of ML systems preferred.
Observability Lead - Cloud SRE & Network Reliability
Fremont, California, United States
$114k-$253k/yrHybridFull Time
Lam ResearchNASDAQ: LRCX: Designs and manufactures wafer fabrication equipment for the semiconductor industry.
12+ YOE6+ MgmtSenior SRE leader with 12+ years infrastructure/SRE/DevOps experience and 6+ years leading teams; deep multi-cloud networking, observability, DR/BCP, backup/restore, automation (Ansible/Terraform/Python), Kubernetes, and incident management experience.
Kody: An agentic commerce platform providing integrated in-person payment solutions.
Senior SRE with deep AWS and GitHub experience, strong monitoring/logging and scripting skills, incident leadership, and absolute fluency in Mandarin and English.
5+ YOEBS/MS or equivalent experience,5+ years software engineering, experience with distributed databases/SaaS deployments, proficiency in Python/Golang/Bash, Kubernetes and cloud platform experience preferred.