30 cloud reliability engineer jobs at 18 companies in Manteca, CA
1w
Save
Mark Applied
Hide
1w
Cloud Site Reliability Engineer - DCS Cloud
San Jose, California, United States
OnsiteFull Time
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOEBachelor's degree in CS or related,2+ years in Linux operations/SRE/DevOps,programming in Go/Python/C++,cloud and reliability practices experience,strong troubleshooting and communication skills.
SambaNova Systems: Develops custom AI hardware and software for enterprise computing.
3+ YOE3-5+ years SRE/DevOps in public cloud; Bachelor's degree or equivalent; Python/Go/Java; Docker/Kubernetes; monitoring/observability tools; IaC; CI/CD.
Saviynt: Provides AI-powered identity governance and cloud security platforms.
9+ YOE9+ years in platform/infra/SRE roles, deep Kubernetes and GCP expertise, strong Go and Python skills, experience with CI/CD, event-driven systems, observability, distributed systems, and building shared platform services.
Pune or San Jose or Durham or Mexico City or Bangalore or Hoofddorp or Belgrade or Barcelona or Singapore or Sydney or Tokyo
HybridFull Time
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
7+ YOE7+ years SRE experience with networking, virtualization (VMware ESXi), Linux, cloud and strong customer-facing troubleshooting and communication skills.
VMware ESXi, VMware, Linux, DevOps, Cloud, Citrix, Microsoft
Research Triangle Park or San Jose or Milpitas or Richardson or Santa Clara
$127k-$182k/yrHybridFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE5+ years SRE/Cloud Ops experience, Docker and Kubernetes proficiency, scripting in Python/Go/Bash, monitoring and incident response experience, Linux and networking knowledge, CI/CD and IaC familiarity, bachelor’s degree or equivalent.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Observability Lead - Cloud SRE & Network Reliability
Fremont, California, United States
$114k-$253k/yrHybridFull Time
Lam ResearchNASDAQ: LRCX: Designs and manufactures wafer fabrication equipment for the semiconductor industry.
12+ YOE6+ MgmtSenior SRE leader with 12+ years infrastructure/SRE/DevOps experience and 6+ years leading teams; deep multi-cloud networking, observability, DR/BCP, backup/restore, automation (Ansible/Terraform/Python), Kubernetes, and incident management experience.
8+ YOEBachelor's in CS or equivalent,8+ years building infrastructure/distributed systems,5+ years programming in C++ or Go,5+ years reliability engineering,EMR not mentioned,experience with distributed systems and stakeholder collaboration.
Senior Site Reliability Engineer, Global E-Commerce
San Jose, California, United States
$213k-$388k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
5+ YOEBachelor's or equivalent,5+ years SRE/infra experience,proficiency in Go/Python/Java,strong Linux,networking and distributed systems knowledge,cloud-native production experience.
Lead Software Engineer - Cloud Storage (GO , Python automation)
San Jose, California, United States
$215k-$245k/yrOnsiteFull Time
NetAppNASDAQ: NTAP: Sells enterprise data storage and cloud management software.
8+ YOE8+ years software/systems engineering experience; 3+ years in data management/storage; strong Go and Python; experience with Kubernetes, cloud (Azure/AWS/GCP), distributed systems, APIs, observability, and production reliability.
San Francisco or San Jose or Seattle or Sacramento
RemoteFull Time
Unstructured: Enterprise data transformation for LLM and AI applications.
8+ YOE7-10+ years in production systems; cloud-native and distributed architectures; auth systems (OAuth/OIDC/SAML/JWT/RBAC/ABAC); enterprise identity integration; mentoring; focus on performance, reliability, and scalable design.
eBayNASDAQ: EBAY: Global online marketplace for buying and selling diverse products.
3+ YOE3+ years software engineering experience; strong Java and Spring Boot skills; GraphQL and API design experience; CI/CD (Maven/Jenkins); Kubernetes and cloud-native deployments; focus on reliability, scalability, and developer experience.
Java, Spring Boot, GraphQL, REST, CI/CD, Maven, Jenkins, Kubernetes
XperiNYSE: XPER: Develops entertainment and audio technology for consumer electronics and automotive.
Design, develop, and maintain scalable Java/goLang/Elixir microservices on AWS; collaborate with cross-functional teams to build cloud-native, performant, reliable, and secure applications.
Director / Senior Director, Platform Software Engineering
Hong Kong or San Jose
HybridFull Time
Nex: Gaming system that turns body movement into interactive play.
Experience leading teams delivering OS and cloud services at scale; strong people leadership; hands-on with device software and cloud; focus on reliability, security, privacy, and compliance; ability to work across global time zones.
Android, Cloud services, DevOps, OTA pipelines, Observability, Play Pass, Security, Privacy, Compliance
Oakland or Washington or Ohio or District of Columbia or California or Long Beach or Arizona or Colorado or Connecticut or Florida or Georgia or Maryland or Minnesota or Nevada or Oregon or El Dorado Hills or San Diego or Woodland Hills or Alabama or Illinois or Virginia or Wisconsin or Texas or New York
$205k-$308k/yrHybridFull Time
Ascendiun: Nonprofit parent overseeing health insurance and clinical service organizations.
10+ YOE6+ Mgmt10+ years engineering experience, 6+ years people management, experience leading multiple engineering teams, production ownership, cloud-native and reliability practices, strong communication and cross-functional partnership.