44 devops reliability engineer jobs at 28 companies in Santa Cruz, CA

1d
Save
Mark Applied
Hide
Staff Reliability Engineer
Santa Clara, California, United States
$167k-$291k/yr RemoteFull Time
ServiceNow
ServiceNowNYSE: NOW: Provides a cloud platform for automating enterprise digital workflows.
8+ YOE8+ years SRE/Platform/DevOps experience, strong Kubernetes and cloud-native platform skills, automation and CI/CD expertise, software engineering with Python/Go/Java/Ruby, observability and reliability knowledge.
Kubernetes, Python, Go, Java, Ruby, GitLab CI/CD, Argo CD, Flux, Playwright, Selenium, Cypress, REST Assured, PyTest, JUnit, TestNG, Ansible, Terraform, Helm, Argo Workflows, Kustomize, Istio, Linkerd, Gateway API, Ingress, Prometheus, OpenTelemetry, AWS (EKS), Azure (AKS), Google Cloud (GKE), GitOps
2d
Save
Mark Applied
Hide
Site Reliability Engineer
San Mateo or Arizona or California or Colorado or Florida or Georgia or Illinois or Nevada or North Carolina or Oregon or Texas or Utah or Washington
$140k-$150k/yr RemoteFull Time
VyncaCare
VyncaCare: Offers palliative care services and advance care planning technology.
3+ YOE3+ years SRE/DevOps experience, strong AWS and Terraform skills, Kubernetes and Helm experience, observability and incident response knowledge, bachelor's or equivalent, on-call participation, East Coast hours.
AWS, Terraform, Kubernetes, Helm, Prometheus, Grafana, Datadog, CloudWatch, SigNoz, OpenTelemetry, ArgoCD, Flux, PostgreSQL, MySQL, Redshift, ClickHouse, AWS Secrets Manager, HashiCorp Vault, Snowflake, Python, Go, Linux
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Palo Alto or Newport Beach
$165k-$190k/yr OnsiteFull Time
Obsidian Security
Obsidian Security: Provides cybersecurity and threat detection for enterprise SaaS applications.
3+ YOE3+ years DevOps/SRE experience on GCP and/or AWS, Bachelor's in CS or related, proficiency with Kubernetes, Helm, GitLab CI/CD, ArgoCD, Prometheus, Grafana; programming in Golang or Python; strong communication and critical thinking.
Kubernetes, Helm, GitLab CI/CD, ArgoCD, Prometheus, Grafana, Golang, Python, Kafka, Elasticsearch, PostgreSQL, ScyllaDB, Databricks, Dagster, Sentry, Kong, AWS, GCP
2w
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara, California, United States
$230k-$250k/yr OnsiteFull Time
Forward Networks
Forward Networks: Provides a digital twin platform for enterprise network management.
6+ YOE6+ years SRE/DevOps experience in SaaS/cloud, strong networking fundamentals, Kubernetes, observability (Prometheus/Grafana/Datadog/Splunk), Python/Bash automation, cloud and IaC (AWS/GCP/Azure, Terraform/Ansible), and incident response ownership.
Kubernetes, Prometheus, Grafana, Datadog, Splunk, Python, Bash, AWS, GCP, Azure, Terraform, Ansible
2w
Save
Mark Applied
Hide
Site Reliability Engineer II
Sunnyvale, California, United States
$141k-$162k/yr OnsiteFull Time
Illumio
Illumio: Provides zero-trust segmentation software to contain cyberattacks.
2+ YOE2+ years SRE/DevOps experience with hands-on Azure experience; experience with AWS/GCP, IaC, CI/CD, scripting (PowerShell, Python, Go), and security best practices.
Azure, AWS, GCP, Terraform, ARM templates, CloudFormation, Azure DevOps, AWS CodePipeline, Jenkins, PowerShell, Python, Go
3w
Save
Mark Applied
Hide
Site Reliability Engineer, Compute Platform
San Jose, California, United States
$156k-$388k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Bachelor's in CS/Engineering, strong Linux, networking, databases, Kubernetes, SRE/DevOps toolset knowledge, experience with ClickHouse/Spark/Presto/Doris/Hadoop, coding in Python/Shell/Java/Go, strong problem-solving and communication.
ClickHouse, Spark, Presto, Doris, Hadoop, Kubernetes, Python, Shell, Java, Go
1mo
Save
Mark Applied
Hide
Systems Reliability Engineer II (Virtualization & Networking)
Bangalore or Pune or San Jose or Durham or Mexico City or Hoofddorp or Belgrade or Barcelona or Singapore or Sydney or Tokyo
HybridFull Time
Nutanix
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
4+ YOE4+ years SRE experience; expertise in virtualization (VMware ESXi), L2/L3 networking, Linux, DevOps and cloud; strong troubleshooting, customer interaction, mentoring, and collaboration skills; degree preferred.
VMware ESXi, VMware, Citrix, Microsoft, Linux, DevOps
4d
Save
Mark Applied
Hide
Sr. Site Reliability Engineer (Starlink)
Hawthorne or Palo Alto or Redmond
$165k-$270k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOEBachelor's in CS/engineering/math with 5 years software experience or 7+ years SRE/DevOps experience; Linux experience required; Kubernetes, Kafka, cloud-native tooling, and programming in Python/Go/Java/C#/Scala preferred.
Linux, Kubernetes, Istio, Apache Kafka, Apache Spark, HBase, HDFS, Apache Flink, Python, C#, Java, Scala, Go
3w
Save
Mark Applied
Hide
Cloud Site Reliability Engineer - DCS Cloud
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOEBachelor's degree in CS or related,2+ years in Linux operations/SRE/DevOps,programming in Go/Python/C++,cloud and reliability practices experience,strong troubleshooting and communication skills.
Go, Python, C++, Linux, OCI, AWS, Azure, GCP, KVM, QEMU, Docker, Kubernetes, containerd, cgroups, namespaces, CUDA, MIG
2mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Santa Clara, California, United States
$152k-$245k/yr OnsiteFull Time
Palo Alto Networks
Palo Alto NetworksNASDAQ: PANW: Provides enterprise-grade network, cloud, and endpoint security software.
BS or MS in CS or related field; expertise in configuration management (Ansible, Terraform, Kubernetes); Python and/or Go; Kubernetes with autoscaling; production engineering/DevOps/SRE experience; public cloud (GCP/AWS); Linux networking; CI/CD with GitLab/GitHub; distributed systems; strong communication; ownership and monitoring as code.
Kubernetes, Docker, GCP, AWS, Ansible, Terraform, Vault, GitLab, Spinnaker, Pub/Sub, Bigtable, Memorystore, BigQuery, RabbitMQ, Kafka, MySQL, Python, Go, Shell scripting, Golang
3w
Save
Mark Applied
Hide
Staff Site Reliability Engineer- Developer Platform
Palo Alto, California, United States
$186k-$233k/yr OnsiteFull Time
Rivian and Volkswagen Group Technologies
Rivian and Volkswagen Group Technologies: A joint venture creating cloud, connectivity and software-defined vehicle solutions for electric vehicles.
5+ YOE5+ years in Platform/DevOps/SRE; Terraform, Kubernetes, GitOps (ArgoCD or Flux), cloud (AWS/Azure/GCP), scripting (Python, Bash, Go); strong communication and mentoring skills.
Terraform, Kubernetes, ArgoCD, Flux, Python, Bash, GoLang, AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
Staff Database Reliability Engineer, DBRE
Palo Alto, California, United States
$165k-$185k/yr RemoteFull Time
Assured
Assured: Software platform for automating insurance claims processing
8+ YOE8+ years in SRE/DevOps/DBA roles; deep PostgreSQL and Amazon Aurora experience; proficiency with JavaScript/TypeScript and Node.js; experience optimizing production databases and building automation; Terraform, Docker/Kubernetes, Prisma, Redshift familiarity a plus.
PostgreSQL, Amazon Aurora, JavaScript, TypeScript, Node.js, Terraform, Terragrunt, Prisma, Docker, Kubernetes, Redshift, CI/CD
4d
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE)
Palo Alto, California, United States
$175k-$229k/yr HybridFull Time
Instrumental
Instrumental: AI-powered software for electronics manufacturing quality and optimization.
5+ YOE5+ years DevOps/SRE experience on public cloud (AWS preferred); expertise in Linux, shell, containers, Kubernetes, terraform, monitoring/logging/APM; strong automation, KPI measurement, and security awareness; U.S. citizenship required for access-controlled work.
AWS, Linux, shell, containerization, Kubernetes, terraform, APM
2mo
Save
Mark Applied
Hide
Staff Cyber Site Reliability Engineer (SRE)
Bethesda or Palo Alto or Dallas or Seattle
$110k-$230k/yr HybridFull Time
GEICO
GEICO: Provides vehicle and property insurance services to consumers.
8+ YOE8+ years in software or site reliability engineering; 5+ years in SRE/DevOps; strong Python; Golang preferred; AWS/Azure/GCP experience; CI/CD and IaC; observability and incident response; security tooling familiarity.
Python, Golang, Grafana, Prometheus, GitHub Actions, Jenkins, Terraform, Ansible
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Cupertino, California, United States
$135k-$180k/yr OnsiteFull Time
VITURE
VITURE: An innovative startup building AI-powered wearable technology.
2+ YOE2+ years in SRE/DevOps, cloud infra, IaC, and observability; strong containerization and CI/CD experience.
Terraform, Ansible, Docker, Kubernetes, Prometheus, Grafana, ELK Stack, Python, Go, Shell
4d
Save
Mark Applied
Hide
Senior Site Reliability Engineer Kubernetes Platform
San Jose, California, United States
$64k-$130k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
10+ YOE10+ years SRE/DevOps experience; production Kubernetes; FedRAMP High/DoD IL5 experience; Terraform, CI/CD, Python/Go; observability platforms; compliance and ATO support.
Kubernetes, EKS, AKS, GKE, Terraform, GitHub Actions, GitLab CI, Jenkins, ArgoCD, Python, Go, Prometheus, Grafana, OpenTelemetry, ELK, Istio, Linkerd, OPA/Gatekeeper, Kyverno
3w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
United States or Houston or Santa Clara or Korea or Germany
RemoteFull Time
Qcells
Qcells: Provider of solar modules, energy storage, and EPC services.
8+ YOE8+ years in SRE/DevOps or software engineering; experience with cloud platforms, distributed systems, observability, CI/CD, and incident response; willingness to travel up to 10%.
AWS, Azure, GCP, Docker, Kubernetes, CI/CD, AI/ML
2mo
Save
Mark Applied
Hide
Staff Software Engineer - Reliability
Palo Alto, California, United States
$218k-$328k/yr OnsiteFull Time
Rubrik
RubrikNYSE: RBRK: Secures enterprise data across cloud and on-premises environments.
8+ YOEUS citizen; 8-12+ years software engineering with SRE/DevOps; BS/MS/PhD in CS/CE or related field; proficient in Go/Python/Java; distributed systems; Unix/Linux; on-call; leadership experience.
Go, Python, Java, Kubernetes, MySQL, Terraform, Pulumi, Prometheus, Grafana, OpenTelemetry
2mo
Save
Mark Applied
Hide
Sr Staff Site Reliability Engineer
San Jose, California, United States
$133k-$200k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
3+ YOE3+ years in SRE/DevOps with AWS/EKS, observability, CI/CD, and security focus.
Amazon EKS, Prometheus, Grafana, ELK, Kafka, Airflow, Spark, Jenkins, GitLab CI, ArgoCD, AWS, Docker, Python, Bash, PowerShell
1mo
Save
Mark Applied
Hide
SRE/DevOps Engineer- Palo Alto, the US
Palo Alto, California, United States
HybridFull Time
Kody
Kody: An agentic commerce platform providing integrated in-person payment solutions.
Deep AWS and Git/GitHub experience, strong monitoring/logging and scripting skills, production incident leadership, bilingual Mandarin and English, ownership of CI/CD and deployment practices.
AWS, GitHub, Git

Explore Jobs

Expand Your Job Search