63 platform reliability engineer jobs at 37 companies in Watsonville, CA
5h
Save
Mark Applied
Hide
5h
Staff Reliability Engineer
Santa Clara, California, United States
$167k-$291k/yrRemoteFull Time
ServiceNowNYSE: NOW: Provides a cloud platform for automating enterprise digital workflows.
8+ YOE8+ years SRE/Platform/DevOps experience, strong Kubernetes and cloud-native platform skills, automation and CI/CD expertise, software engineering with Python/Go/Java/Ruby, observability and reliability knowledge.
Staff Site Reliability Engineer- Developer Platform
Palo Alto, California, United States
$186k-$233k/yrOnsiteFull Time
Rivian and Volkswagen Group Technologies: A joint venture creating cloud, connectivity and software-defined vehicle solutions for electric vehicles.
5+ YOE5+ years in Platform/DevOps/SRE; Terraform, Kubernetes, GitOps (ArgoCD or Flux), cloud (AWS/Azure/GCP), scripting (Python, Bash, Go); strong communication and mentoring skills.
Senior Site Reliability Engineer Platform Cloud Foundations Engineer
San Jose, California, United States
$64k-$130k/yrOnsiteFull Time
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
8+ YOE8+ years SRE/platform engineering experience with AWS multi-account, Terraform, automation (Python/Go/Ruby), cloud governance, and strong documentation and communication skills.
AWS Organizations, IAM, Terraform, Python, Go, Ruby, Control Tower, Account Factory for Terraform, CloudFormation, EventBridge, Lambda, SQS, IAM Identity Center, GCP
AbbottNYSE: ABT: Manufactures medical devices, diagnostics, and nutritional health products.
Ensure reliability, scalability, and performance of a medical-device remote monitoring platform; expertise in cloud (Azure), Kubernetes, observability, automation, and incident management; bachelor's in a technical discipline.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK/EFK, Datadog, Linux
Senior Site Reliability Engineer, Platform Infrastructure (Foundations)
San Francisco or Palo Alto
OnsiteFull Time
Anyscale: Cloud platform for scaling distributed machine learning applications.
3+ YOE3+ years writing production code; experience with distributed systems, Kubernetes, cloud (AWS/Azure/GCP); proficiency in Go and Python; familiarity with observability (Prometheus, Grafana); on-call experience.
Senior Site Reliability Engineer - Core Cloud Platform
San Francisco or San Jose or Bellevue
$240k-$356k/yrHybridFull Time
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
7+ YOE7+ years SRE or production infrastructure experience, deep Kubernetes and Terraform knowledge, experience with observability and SLOs, proficiency in Go or Python, on-call and incident leadership experience.
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
Senior SRE with strong distributed systems, cloud (Azure), Kubernetes, observability, automation, incident management, and cross-functional communication skills for a medical device remote monitoring platform.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK, EFK, Datadog, Linux
Principal Tech Lead Manager - Data Platform & Reliability Engineering
Mountain View, California, United States
$215k-$275k/yrOnsiteFull Time
ID.me: Provides secure digital identity verification and authentication services.
5+ YOE3+ Mgmt8+ years engineering experience with 3+ years managing teams,5+ years in data/platform/SRE; bachelor\u0002s or equivalent; deep PostgreSQL and data reliability expertise; strong communication and cloud/IaC experience.
Senior Site Reliability Engineer, Apple Data Platform SRE / Apple Services Engineering
Cupertino, California, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Apply SRE principles to mentor teams, ensure reliability for large-scale analytics infrastructure across Hadoop, HBase, Spark, Data Lakes, and Airflow; participate in production on-call.
ThoughtSpot: AI-powered analytics platform for enterprise business intelligence.
Experience troubleshooting Linux systems and cloud platforms, hands-on with monitoring tools, on-call/incident management experience, scripting in Python/Go/Bash/Java, B.S. in CS or equivalent preferred.
Saviynt: Provides AI-powered identity governance and cloud security platforms.
9+ YOE9+ years in platform/infra/SRE roles, deep Kubernetes and GCP expertise, strong Go and Python skills, experience with CI/CD, event-driven systems, observability, distributed systems, and building shared platform services.
United States or Houston or Santa Clara or Korea or Germany
RemoteFull Time
Qcells: Provider of solar modules, energy storage, and EPC services.
8+ YOE8+ years in SRE/DevOps or software engineering; experience with cloud platforms, distributed systems, observability, CI/CD, and incident response; willingness to travel up to 10%.
5+ YOEBS/MS or equivalent experience,5+ years software engineering, experience with distributed databases/SaaS deployments, proficiency in Python/Golang/Bash, Kubernetes and cloud platform experience preferred.
Wayve: Develops end-to-end artificial intelligence for autonomous driving systems.
10+ YOE10+ years building large-scale distributed systems or ML infrastructure, 3+ years at staff/principal level, experience with Spark, Ray, Kubernetes, Airflow, MLflow, web frameworks, reliability engineering, and mentoring engineers.
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOEBachelor's in CS/engineering/math +5 years software development or 7+ years experience; Linux experience required; SRE/DevOps, Kubernetes/Istio, streaming data platforms, and programming in Python/Go/Java/C#/Scala preferred.
eBayNASDAQ: EBAY: Global online marketplace for buying and selling diverse products.
3+ YOE3+ years software engineering experience; strong Java and Spring Boot skills; GraphQL and API design experience; CI/CD (Maven/Jenkins); Kubernetes and cloud-native deployments; focus on reliability, scalability, and developer experience.
Java, Spring Boot, GraphQL, REST, CI/CD, Maven, Jenkins, Kubernetes
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
5+ YOE5+ Mgmt5+ years leading engineering teams and 5+ years in infrastructure/platform/backend roles; strong platform mindset, SRE experience (SLOs/error budgets), AWS and cloud fundamentals, influence and hiring experience, experience with incident/incident response processes.