49 senior site reliability engineer jobs at 26 companies in Lathrop, CA
1mo
Save
Mark Applied
Hide
1mo
Senior Site Reliability Engineer
Pleasanton or Austin or San Francisco or United States
OnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOESenior SRE with strong infrastructure, automation, and programming experience (Terraform, Chef, Ansible, Python, Java, Bash). Minimum multi-year experience in software engineering or equivalent; participates in on-call and incident response.
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
8+ YOE8+ years supporting live-site production environments; BS/MS or equivalent; strong Kubernetes, AWS, Python; Akamai/CDN and SRE on-call experience required.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Research Triangle Park or San Jose or Milpitas or Richardson or Santa Clara
$127k-$182k/yrHybridFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE5+ years SRE/Cloud Ops experience, Docker and Kubernetes proficiency, scripting in Python/Go/Bash, monitoring and incident response experience, Linux and networking knowledge, CI/CD and IaC familiarity, bachelor’s degree or equivalent.
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOERequires a master's degree and 2 years of related experience, or a bachelor's degree and 5 years of progressive experience. Requires cloud systems, Linux, Docker, Kubernetes, software lifecycle, observability, and reliability engineering expertise.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
Senior Site Reliability Engineer Kubernetes Platform
San Jose, California, United States
$64k-$130k/yrOnsiteFull Time
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
10+ YOE10+ years SRE/DevOps experience; production Kubernetes; FedRAMP High/DoD IL5 experience; Terraform, CI/CD, Python/Go; observability platforms; compliance and ATO support.
BS or MS in CS or related field; expertise in configuration management (Ansible, Terraform, Kubernetes); Python and/or Go; Kubernetes with autoscaling; production engineering/DevOps/SRE experience; public cloud (GCP/AWS); Linux networking; CI/CD with GitLab/GitHub; distributed systems; strong communication; ownership and monitoring as code.
Senior Site Reliability Engineer, Global E-Commerce
San Jose, California, United States
$213k-$388k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
5+ YOEBachelor's or equivalent,5+ years SRE/infra experience,proficiency in Go/Python/Java,strong Linux,networking and distributed systems knowledge,cloud-native production experience.
Senior Software Engineer, Site Reliability Engineering
New York or San Ramon or Reno
$153k-$210k/yrHybridFull Time
Ridgeline: Cloud-native platform for investment management operations.
3+ YOE3–6 years SRE/DevOps experience, 2+ years on AWS, proficiency with Terraform, observability, CI/CD, Python/Go/Bash, incident response, and strong communication and troubleshooting skills.
5+ YOEBS/MS or equivalent experience,5+ years software engineering, experience with distributed databases/SaaS deployments, proficiency in Python/Golang/Bash, Kubernetes and cloud platform experience preferred.
Saviynt: Provides AI-powered identity governance and cloud security platforms.
9+ YOE9+ years in platform/infra/SRE roles, deep Kubernetes and GCP expertise, strong Go and Python skills, experience with CI/CD, event-driven systems, observability, distributed systems, and building shared platform services.
Senior Site Reliability Engineer, AI Infrastructure
San Jose, California, United States
$187k-$360k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOEBachelor's degree or equivalent experience, 3+ years in SRE, DevOps, or systems engineering, Linux and networking expertise, programming, distributed systems, automation, and CI/CD experience.
8+ YOEBachelor's degree or equivalent practical experience; 8 years building infrastructure or distributed systems; 5 years programming in C++ or Go and reliability engineering; distributed systems experience required.
C++, Go, Java, GoogleSQL, Software Development Kit (SDK)
Pure StorageNYSE: PSTG: Provides all-flash enterprise data storage and management solutions.
Operate and automate enterprise IAM platforms, implement IAC and automation, build observability and SLIs/SLOs, lead incident response and documentation; hands-on Terraform/Ansible and scripting experience required.
BitdeerNASDAQ: BTDR: Operates cryptocurrency mining and high-performance computing data centers.
5+ YOERequires 5+ years of Kubernetes operations, 2+ years managing GPU workloads, Terraform, Helm, GitOps, SRE practices, monitoring, Go or Python, and multi-tenant platform experience.