160 platform reliability engineer jobs at 118 companies in California

1mo
Save
Mark Applied
Hide
Senior Platform Reliability Engineer
San Francisco or New York City or Seattle
$182k-$250k/yr HybridFull Time
Grow Therapy
Grow Therapy: Platform connecting mental health providers with patients and insurance.
6+ YOE6+ years operating production systems; hands-on AWS, Kubernetes (EKS), Terraform; experience defining SLOs/SLAs and observability (DataDog); strong communication and systems-thinking skills; PostgreSQL experience a plus.
AWS, Kubernetes, EKS, Terraform, DataDog, PostgreSQL, Gem
3d
Save
Mark Applied
Hide
Staff AI Platform & Reliability Engineer
Santa Ana, California, United States
$172k-$220k/yr OnsiteFull Time
eJam
eJam: Builds and scales direct-to-consumer e-commerce brands.
Senior Python/GCP engineer to design and operate AI generation services, provider integrations, billing/usage metering, tenant security and platform reliability.
Python 3.12, FastAPI, Pydantic, asyncio, Cloud Run, Pub/Sub, Cloud Tasks, GCS, Firestore, Cloud SQL/Postgres, SQLAlchemy, Alembic, Terraform, IAM/OIDC, Secret Manager, CI/CD, NestJS, TypeScript, Claude Code, Codex, Cursor
2d
Save
Mark Applied
Hide
Staff Reliability Engineer
Santa Clara, California, United States
$167k-$291k/yr RemoteFull Time
ServiceNow
ServiceNowNYSE: NOW: Provides a cloud platform for automating enterprise digital workflows.
8+ YOE8+ years SRE/Platform/DevOps experience, strong Kubernetes and cloud-native platform skills, automation and CI/CD expertise, software engineering with Python/Go/Java/Ruby, observability and reliability knowledge.
Kubernetes, Python, Go, Java, Ruby, GitLab CI/CD, Argo CD, Flux, Playwright, Selenium, Cypress, REST Assured, PyTest, JUnit, TestNG, Ansible, Terraform, Helm, Argo Workflows, Kustomize, Istio, Linkerd, Gateway API, Ingress, Prometheus, OpenTelemetry, AWS (EKS), Azure (AKS), Google Cloud (GKE), GitOps
1w
Save
Mark Applied
Hide
Senior Reliability Engineer
San Francisco or Oakland
$139k-$205k/yr OnsiteFull Time
DoorDash
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
5+ YOE5+ years reliability validation or hardware testing experience for robotics/unmanned platforms, Bachelor's in engineering, proficiency with environmental test equipment, Python and CAD, strong communication and analytical skills.
Python, CAD, DAQ, FRACAS, HIL
3w
Save
Mark Applied
Hide
Staff Site Reliability Engineer- Developer Platform
Palo Alto, California, United States
$186k-$233k/yr OnsiteFull Time
Rivian and Volkswagen Group Technologies
Rivian and Volkswagen Group Technologies: A joint venture creating cloud, connectivity and software-defined vehicle solutions for electric vehicles.
5+ YOE5+ years in Platform/DevOps/SRE; Terraform, Kubernetes, GitOps (ArgoCD or Flux), cloud (AWS/Azure/GCP), scripting (Python, Bash, Go); strong communication and mentoring skills.
Terraform, Kubernetes, ArgoCD, Flux, Python, Bash, GoLang, AWS, Azure, GCP
3w
Save
Mark Applied
Hide
Site Reliability Engineer, Compute Platform
San Jose, California, United States
$156k-$388k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Bachelor's in CS/Engineering, strong Linux, networking, databases, Kubernetes, SRE/DevOps toolset knowledge, experience with ClickHouse/Spark/Presto/Doris/Hadoop, coding in Python/Shell/Java/Go, strong problem-solving and communication.
ClickHouse, Spark, Presto, Doris, Hadoop, Kubernetes, Python, Shell, Java, Go
4d
Save
Mark Applied
Hide
Senior Site Reliability Engineer Platform Cloud Foundations Engineer
San Jose, California, United States
$64k-$130k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
8+ YOE8+ years SRE/platform engineering experience with AWS multi-account, Terraform, automation (Python/Go/Ruby), cloud governance, and strong documentation and communication skills.
AWS Organizations, IAM, Terraform, Python, Go, Ruby, Control Tower, Account Factory for Terraform, CloudFormation, EventBridge, Lambda, SQS, IAM Identity Center, GCP
2mo
Save
Mark Applied
Hide
Director of Platform & Reliability Engineering
San Francisco or New York City
$235k-$245k/yr HybridFull Time
Forge Global
Forge GlobalNYSE: FRGE: Marketplace for trading private shares and pre-IPO stock.
8+ YOE5+ Mgmt8+ years software engineering experience with infrastructure/platform focus, 5+ years people leadership, deep cloud/observability/incident response experience, strong distributed systems judgment, and ability to set platform strategy.
Kubernetes, CI/CD
2mo
Save
Mark Applied
Hide
Senior Reliability Engineer
San Diego, California, United States
$107k-$120k/yr OnsiteFull Time
PCI Pharma Services
PCI Pharma Services: Integrated pharmaceutical development, manufacturing, and global packaging services.
7+ YOEBachelor’s in chemical, electrical, industrial, mechanical engineering or computer science; 7-10+ years engineering experience with at least 4 years cGMP; strong GMP and aseptic knowledge; automation and control platforms familiarity; proficient in MS Office; able to work autonomously and collaboratively.
Allen Bradley, Siemens, Wonderware, iFix, PI, Microsoft Office
2mo
Save
Mark Applied
Hide
Senior Platform Engineer
San Francisco, California, United States
$170k-$230k/yr HybridFull Time
Authorium
Authorium: Cloud-based administrative operations platform for government agencies.
6+ YOE6+ years in platform/infrastructure/DevOps engineering with distributed systems, cloud architecture, security, and reliability; strong collaboration in an in-person SF office (Mon–Thu).
Cursor, Claude, AWS ECS, AWS EKS, CloudWatch, IAM, VPC, Parameter Store, Terraform, Pulumi, CDK, Datadog, OpenTelemetry
3w
Save
Mark Applied
Hide
Site Reliability Engineer, Compute Platform
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Experience with Linux, networking, databases, Kubernetes, ClickHouse/Hadoop/Doris/Spark/Presto, scripting or programming (Python, Shell, Java, Go), incident management, and capacity planning.
ClickHouse, Spark, Presto, Doris, Hadoop, Kubernetes, Linux, Python, Shell, Java, Go
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Kubernetes Platform (Starshield)
Hawthorne or Redmond
$125k-$175k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
1+ YOEBachelor's in CS/IT/engineering + 1+ year SRE/DevOps experience (or 3+ years experience), Linux, Terraform/Ansible, Kubernetes and OCI containers, Bash/Python scripting, development in Python/C++/Go, willingness to obtain Top Secret clearance.
Terraform, Ansible, OCI containers, Kubernetes, Bash, Python, C++, Go, Bazel, Makefiles, TCP/IP, Linux
2w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Los Angeles, California, United States
$140k-$199k/yr HybridFull Time
Green Dot
Green DotNYSE: GDOT: Provides mobile banking and payment solutions to consumers and businesses.
7+ YOE7+ years in release/reliability engineering, cloud platform experience (AWS/Azure/GCP), automated deployment and observability proficiency, scripting with PowerShell/Bash/Python, excellent troubleshooting and communication skills.
AWS, Azure, GCP, PowerShell, Bash, Python
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Oakland, California, United States
$175k-$210k/yr HybridFull Time
Fivetran
Fivetran: Automates data movement into cloud data warehouses.
5+ YOE5+ years SaaS experience; managed Kubernetes, cloud platforms (AWS/GCP/Azure), Terraform/Ansible/ArgoCD; Python/Shell scripting, Linux admin, PostgreSQL; incident response and reliability engineering experience.
Kubernetes, EKS, AKS, GKE, PostgreSQL, ArgoCD, Terraform, Ansible, Python, Shell, Go, Java, AWS, GCP, Azure, Grafana, Buildkite, Temporal, Pulumi, Linux, VPN, PrivateLink, Private Service Connect (GCP)
6d
Save
Mark Applied
Hide
Platform Engineer, Billing Systems
San Francisco or United States
$220k-$350k/yr HybridFull Time
Wispr Flow
Wispr Flow: Provides AI-powered voice dictation software for computers and mobile devices.
Experienced engineer who has built billing systems, run payments migrations, reconciled billing across platforms, and built testing for high-reliability billing.
RevenueCat, Stripe, Sequence
1mo
Save
Mark Applied
Hide
Platform Engineer
San Francisco, California, United States
OnsiteFull Time
Phonic
Phonic: Building a platform for lifelike, reliable voice AI agents.
Strong experience with cloud infrastructure, containerized deployments, infrastructure-as-code, reliability engineering, SLOs, observability, and incident response; strong ownership and developer-experience focus.
GCP, AWS, Azure, Docker, Kubernetes, Terraform, WebSockets, WebRTC, TypeScript, Python
2mo
Save
Mark Applied
Hide
Platform Engineer (SRE) - AI Control Plane
San Francisco, California, United States
OnsiteFull Time
Speakeasy
Speakeasy: Automates API SDK and documentation generation for developers.
Platform Engineer (SRE) to own reliability, design deployments, and participate in on-call; strong systems and software engineering.
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Client Platform
Los Angeles, California, United States
$164k-$270k/yr OnsiteFull Time
Hadrian
Hadrian: Building autonomous factories for aerospace and defense manufacturing.
Experience building scalable automation, strong scripting (Python, Bash, PowerShell), IaC (Ansible, Terraform), MDM platform administration, patch/vulnerability remediation, endpoint hardening, and translating CMMC/compliance into code-managed baselines. ITAR eligibility required.
Python, Bash, PowerShell, Ansible, Terraform, Chef, Salt, Puppet, Pulumi, Fleet DM, Microsoft Intune, Workspace ONE, Jamf, osquery
1mo
Save
Mark Applied
Hide
Senior Platform Engineer
United States or San Mateo or Provo
RemoteFull Time
GC AI
GC AI: AI-powered legal platform for in-house legal teams.
Production experience with Google Cloud Platform, IAM, infrastructure as code (Terraform/Pulumi), CI/CD, observability, and backend development (TypeScript preferred). Strong reliability, operational excellence, and developer experience focus.
Google Cloud Platform, Terraform, Pulumi, CI/CD, TypeScript, IAM
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Sunnyvale, California, United States
$90k-$180k/yr OnsiteFull Time
Abbott
AbbottNYSE: ABT: Manufactures medical devices, diagnostics, and nutritional health products.
Ensure reliability, scalability, and performance of a medical-device remote monitoring platform; expertise in cloud (Azure), Kubernetes, observability, automation, and incident management; bachelor's in a technical discipline.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK/EFK, Datadog, Linux

Explore Jobs

Expand Your Job Search