42 devops reliability engineer jobs at 26 companies in Aromas, CA

1mo
Save
Mark Applied
Hide
Site Reliability Engineer - Hardware Infrastructure
Santa Clara, California, United States
$184k-$357k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOEDegree in CS or related field (or equivalent experience), 8+ years SRE/DevOps/Production Engineering, SRE principles, infrastructure automation, production reliability, Python/Go/Perl/Ruby, Prometheus and Grafana, strong communication.
Python, Go, Perl, Ruby, Prometheus, Grafana
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Palo Alto or Newport Beach
$165k-$190k/yr OnsiteFull Time
Obsidian Security
Obsidian Security: Provides cybersecurity and threat detection for enterprise SaaS applications.
3+ YOE3+ years DevOps/SRE experience on GCP and/or AWS, Bachelor's in CS or related, proficiency with Kubernetes, Helm, GitLab CI/CD, ArgoCD, Prometheus, Grafana; programming in Golang or Python; strong communication and critical thinking.
Kubernetes, Helm, GitLab CI/CD, ArgoCD, Prometheus, Grafana, Golang, Python, Kafka, Elasticsearch, PostgreSQL, ScyllaDB, Databricks, Dagster, Sentry, Kong, AWS, GCP
3w
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara, California, United States
$230k-$250k/yr OnsiteFull Time
Forward Networks
Forward Networks: Provides a digital twin platform for enterprise network management.
6+ YOE6+ years SRE/DevOps experience in SaaS/cloud, strong networking fundamentals, Kubernetes, observability (Prometheus/Grafana/Datadog/Splunk), Python/Bash automation, cloud and IaC (AWS/GCP/Azure, Terraform/Ansible), and incident response ownership.
Kubernetes, Prometheus, Grafana, Datadog, Splunk, Python, Bash, AWS, GCP, Azure, Terraform, Ansible
3w
Save
Mark Applied
Hide
Site Reliability Engineer II
Sunnyvale, California, United States
$141k-$162k/yr OnsiteFull Time
Illumio
Illumio: Provides zero-trust segmentation software to contain cyberattacks.
2+ YOE2+ years SRE/DevOps experience with hands-on Azure experience; experience with AWS/GCP, IaC, CI/CD, scripting (PowerShell, Python, Go), and security best practices.
Azure, AWS, GCP, Terraform, ARM templates, CloudFormation, Azure DevOps, AWS CodePipeline, Jenkins, PowerShell, Python, Go
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Compute Platform
San Jose, California, United States
$156k-$388k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Bachelor's in CS/Engineering, strong Linux, networking, databases, Kubernetes, SRE/DevOps toolset knowledge, experience with ClickHouse/Spark/Presto/Doris/Hadoop, coding in Python/Shell/Java/Go, strong problem-solving and communication.
ClickHouse, Spark, Presto, Doris, Hadoop, Kubernetes, Python, Shell, Java, Go
2mo
Save
Mark Applied
Hide
Systems Reliability Engineer II (Virtualization & Networking)
Bangalore or Pune or San Jose or Durham or Mexico City or Hoofddorp or Belgrade or Barcelona or Singapore or Sydney or Tokyo
HybridFull Time
Nutanix
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
4+ YOE4+ years SRE experience; expertise in virtualization (VMware ESXi), L2/L3 networking, Linux, DevOps and cloud; strong troubleshooting, customer interaction, mentoring, and collaboration skills; degree preferred.
VMware ESXi, VMware, Citrix, Microsoft, Linux, DevOps
1w
Save
Mark Applied
Hide
Sr. Site Reliability Engineer (Starlink)
Hawthorne or Palo Alto or Redmond
$165k-$270k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOEBachelor's in CS/engineering/math with 5 years software experience or 7+ years SRE/DevOps experience; Linux experience required; Kubernetes, Kafka, cloud-native tooling, and programming in Python/Go/Java/C#/Scala preferred.
Linux, Kubernetes, Istio, Apache Kafka, Apache Spark, HBase, HDFS, Apache Flink, Python, C#, Java, Scala, Go
1mo
Save
Mark Applied
Hide
Cloud Site Reliability Engineer - DCS Cloud
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOEBachelor's degree in CS or related,2+ years in Linux operations/SRE/DevOps,programming in Go/Python/C++,cloud and reliability practices experience,strong troubleshooting and communication skills.
Go, Python, C++, Linux, OCI, AWS, Azure, GCP, KVM, QEMU, Docker, Kubernetes, containerd, cgroups, namespaces, CUDA, MIG
2mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Santa Clara, California, United States
$152k-$245k/yr OnsiteFull Time
Palo Alto Networks
Palo Alto NetworksNASDAQ: PANW: Provides enterprise-grade network, cloud, and endpoint security software.
BS or MS in CS or related field; expertise in configuration management (Ansible, Terraform, Kubernetes); Python and/or Go; Kubernetes with autoscaling; production engineering/DevOps/SRE experience; public cloud (GCP/AWS); Linux networking; CI/CD with GitLab/GitHub; distributed systems; strong communication; ownership and monitoring as code.
Kubernetes, Docker, GCP, AWS, Ansible, Terraform, Vault, GitLab, Spinnaker, Pub/Sub, Bigtable, Memorystore, BigQuery, RabbitMQ, Kafka, MySQL, Python, Go, Shell scripting, Golang
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer- Developer Platform
Palo Alto, California, United States
$186k-$233k/yr OnsiteFull Time
Rivian and Volkswagen Group Technologies
Rivian and Volkswagen Group Technologies: A joint venture creating cloud, connectivity and software-defined vehicle solutions for electric vehicles.
5+ YOE5+ years in Platform/DevOps/SRE; Terraform, Kubernetes, GitOps (ArgoCD or Flux), cloud (AWS/Azure/GCP), scripting (Python, Bash, Go); strong communication and mentoring skills.
Terraform, Kubernetes, ArgoCD, Flux, Python, Bash, GoLang, AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
Staff Database Reliability Engineer, DBRE
Palo Alto, California, United States
$165k-$185k/yr RemoteFull Time
Assured
Assured: Software platform for automating insurance claims processing
8+ YOE8+ years in SRE/DevOps/DBA roles; deep PostgreSQL and Amazon Aurora experience; proficiency with JavaScript/TypeScript and Node.js; experience optimizing production databases and building automation; Terraform, Docker/Kubernetes, Prisma, Redshift familiarity a plus.
PostgreSQL, Amazon Aurora, JavaScript, TypeScript, Node.js, Terraform, Terragrunt, Prisma, Docker, Kubernetes, Redshift, CI/CD
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE)
Palo Alto, California, United States
$175k-$229k/yr HybridFull Time
Instrumental
Instrumental: AI-powered software for electronics manufacturing quality and optimization.
5+ YOE5+ years DevOps/SRE experience on public cloud (AWS preferred); expertise in Linux, shell, containers, Kubernetes, terraform, monitoring/logging/APM; strong automation, KPI measurement, and security awareness; U.S. citizenship required for access-controlled work.
AWS, Linux, shell, containerization, Kubernetes, terraform, APM
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer Kubernetes Platform
San Jose, California, United States
$64k-$130k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
10+ YOE10+ years SRE/DevOps experience; production Kubernetes; FedRAMP High/DoD IL5 experience; Terraform, CI/CD, Python/Go; observability platforms; compliance and ATO support.
Kubernetes, EKS, AKS, GKE, Terraform, GitHub Actions, GitLab CI, Jenkins, ArgoCD, Python, Go, Prometheus, Grafana, OpenTelemetry, ELK, Istio, Linkerd, OPA/Gatekeeper, Kyverno
2mo
Save
Mark Applied
Hide
Staff Software Engineer - Reliability
Palo Alto, California, United States
$218k-$328k/yr OnsiteFull Time
Rubrik
RubrikNYSE: RBRK: Secures enterprise data across cloud and on-premises environments.
8+ YOEUS citizen; 8-12+ years software engineering with SRE/DevOps; BS/MS/PhD in CS/CE or related field; proficient in Go/Python/Java; distributed systems; Unix/Linux; on-call; leadership experience.
Go, Python, Java, Kubernetes, MySQL, Terraform, Pulumi, Prometheus, Grafana, OpenTelemetry
2mo
Save
Mark Applied
Hide
Sr Staff Site Reliability Engineer
San Jose, California, United States
$133k-$200k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
3+ YOE3+ years in SRE/DevOps with AWS/EKS, observability, CI/CD, and security focus.
Amazon EKS, Prometheus, Grafana, ELK, Kafka, Airflow, Spark, Jenkins, GitLab CI, ArgoCD, AWS, Docker, Python, Bash, PowerShell
4w
Save
Mark Applied
Hide
Staff Software Engineer, Reliability Engineering & SDLC Governance
Carlsbad or Germantown or San Jose or San Francisco or New York City
$165k-$261k/yr OnsiteFull Time
Viasat
ViasatNASDAQ: VSAT: Provides global satellite broadband and secure networking communication services.
8+ YOE8+ years software engineering experience with SDLC governance, SRE/DevOps knowledge, observability, incident analysis, and cross-team influence to improve reliability and customer experience.
Prometheus, Grafana, OpenTelemetry, Datadog, AS9115
1mo
Save
Mark Applied
Hide
SRE/DevOps Engineer- Palo Alto, the US
Palo Alto, California, United States
HybridFull Time
Kody
Kody: An agentic commerce platform providing integrated in-person payment solutions.
Deep AWS and Git/GitHub experience, strong monitoring/logging and scripting skills, production incident leadership, bilingual Mandarin and English, ownership of CI/CD and deployment practices.
AWS, GitHub, Git
1mo
Save
Mark Applied
Hide
Senior Kubernetes DevOps Engineer - K0rdent AI Apps/Core Services
San Jose, California, United States
HybridFull Time
Mirantis
Mirantis: Develops software for managing cloud and AI infrastructure.
Kubernetes operator expertise; experience with Linux, virtualization, networking, and storage; system architecture, scalability, and reliability; debugging production systems; microservices and enterprise IT experience; excellent English communication.
Kubernetes, k0rdent, k0s, k0smotron, Kubeflow, Kserve, vLLM, NVIDIA AI Enterprise, KubeVirt, Harbor, ArgoCD, Grafana, OpenStack, Linux
2mo
Save
Mark Applied
Hide
Observability Lead - Cloud SRE & Network Reliability (193698)
Fremont or San Francisco or Oakland
$114k-$253k/yr HybridFull Time
Lam Research
Lam ResearchNASDAQ: LRCX: Manufacturing equipment used to fabricate advanced semiconductor microchips.
12+ YOE6+ MgmtBS/MS/PhD or equivalent, 12+ years in infrastructure/SRE/DevOps/network engineering, 6+ years leading SRE/observability teams; multi-cloud networking, DR/BCP, observability platforms, IaC, automation, Python/Go experience.
Azure, AWS, GCP, Prometheus, Grafana, Datadog, PagerDuty, ThousandEyes, Azure Monitor, CloudWatch, Google Cloud Operations, Splunk, Ansible, Terraform, Python, Go, Kubernetes, AKS, EKS, GKE, ServiceNow
3d
Save
Mark Applied
Hide
Software Engineer III – Data Platform
Mountain View or McLean or New York City or Tampa
$173k-$201k/yr OnsiteFull Time
ID.me
ID.me: Provides secure digital identity verification and authentication services.
3+ YOEBachelor's degree in a technical field; 3–5 years in SRE, DevOps, or infrastructure engineering; cloud experience; and 1+ year with a modern programming language.
Prometheus, Grafana, OpenTelemetry, AWS, GCP, Azure, Java, Go, Python, Ruby, JavaScript, Docker, Kubernetes, Terraform, Pulumi, Ansible, GitOps, CI/CD, FedRAMP, NIST 800-53, SOC2