33 devops reliability engineer jobs at 18 companies in Prunedale, CA

1mo
Save
Mark Applied
Hide
Site Reliability Engineer - Hardware Infrastructure
Santa Clara, California, United States
$184k-$357k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOEDegree in CS or related field (or equivalent experience), 8+ years SRE/DevOps/Production Engineering, SRE principles, infrastructure automation, production reliability, Python/Go/Perl/Ruby, Prometheus and Grafana, strong communication.
Python, Go, Perl, Ruby, Prometheus, Grafana
3w
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara, California, United States
$230k-$250k/yr OnsiteFull Time
Forward Networks
Forward Networks: Provides a digital twin platform for enterprise network management.
6+ YOE6+ years SRE/DevOps experience in SaaS/cloud, strong networking fundamentals, Kubernetes, observability (Prometheus/Grafana/Datadog/Splunk), Python/Bash automation, cloud and IaC (AWS/GCP/Azure, Terraform/Ansible), and incident response ownership.
Kubernetes, Prometheus, Grafana, Datadog, Splunk, Python, Bash, AWS, GCP, Azure, Terraform, Ansible
3w
Save
Mark Applied
Hide
Site Reliability Engineer II
Sunnyvale, California, United States
$141k-$162k/yr OnsiteFull Time
Illumio
Illumio: Provides zero-trust segmentation software to contain cyberattacks.
2+ YOE2+ years SRE/DevOps experience with hands-on Azure experience; experience with AWS/GCP, IaC, CI/CD, scripting (PowerShell, Python, Go), and security best practices.
Azure, AWS, GCP, Terraform, ARM templates, CloudFormation, Azure DevOps, AWS CodePipeline, Jenkins, PowerShell, Python, Go
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Compute Platform
San Jose, California, United States
$156k-$388k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Bachelor's in CS/Engineering, strong Linux, networking, databases, Kubernetes, SRE/DevOps toolset knowledge, experience with ClickHouse/Spark/Presto/Doris/Hadoop, coding in Python/Shell/Java/Go, strong problem-solving and communication.
ClickHouse, Spark, Presto, Doris, Hadoop, Kubernetes, Python, Shell, Java, Go
2mo
Save
Mark Applied
Hide
Systems Reliability Engineer II (Virtualization & Networking)
Bangalore or Pune or San Jose or Durham or Mexico City or Hoofddorp or Belgrade or Barcelona or Singapore or Sydney or Tokyo
HybridFull Time
Nutanix
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
4+ YOE4+ years SRE experience; expertise in virtualization (VMware ESXi), L2/L3 networking, Linux, DevOps and cloud; strong troubleshooting, customer interaction, mentoring, and collaboration skills; degree preferred.
VMware ESXi, VMware, Citrix, Microsoft, Linux, DevOps
1mo
Save
Mark Applied
Hide
Cloud Site Reliability Engineer - DCS Cloud
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOEBachelor's degree in CS or related,2+ years in Linux operations/SRE/DevOps,programming in Go/Python/C++,cloud and reliability practices experience,strong troubleshooting and communication skills.
Go, Python, C++, Linux, OCI, AWS, Azure, GCP, KVM, QEMU, Docker, Kubernetes, containerd, cgroups, namespaces, CUDA, MIG
2mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Santa Clara, California, United States
$152k-$245k/yr OnsiteFull Time
Palo Alto Networks
Palo Alto NetworksNASDAQ: PANW: Provides enterprise-grade network, cloud, and endpoint security software.
BS or MS in CS or related field; expertise in configuration management (Ansible, Terraform, Kubernetes); Python and/or Go; Kubernetes with autoscaling; production engineering/DevOps/SRE experience; public cloud (GCP/AWS); Linux networking; CI/CD with GitLab/GitHub; distributed systems; strong communication; ownership and monitoring as code.
Kubernetes, Docker, GCP, AWS, Ansible, Terraform, Vault, GitLab, Spinnaker, Pub/Sub, Bigtable, Memorystore, BigQuery, RabbitMQ, Kafka, MySQL, Python, Go, Shell scripting, Golang
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer Kubernetes Platform
San Jose, California, United States
$64k-$130k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
10+ YOE10+ years SRE/DevOps experience; production Kubernetes; FedRAMP High/DoD IL5 experience; Terraform, CI/CD, Python/Go; observability platforms; compliance and ATO support.
Kubernetes, EKS, AKS, GKE, Terraform, GitHub Actions, GitLab CI, Jenkins, ArgoCD, Python, Go, Prometheus, Grafana, OpenTelemetry, ELK, Istio, Linkerd, OPA/Gatekeeper, Kyverno
3d
Save
Mark Applied
Hide
Site Reliability Engineer, AI Infrastructure
San Jose, California, United States
$123k-$259k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
1+ YOEBachelor's degree or equivalent experience, 1+ year in SRE, DevOps, or systems engineering, Linux and networking knowledge, distributed systems experience, programming, scripting, CI/CD, and automation skills.
Linux, Go, Python, C, C++, Java, Bash, Kubernetes, AWS, GCP, Azure, Terraform, Prometheus, Grafana, Distributed Tracing, LLMs, Agentic AI
2mo
Save
Mark Applied
Hide
Sr Staff Site Reliability Engineer
San Jose, California, United States
$133k-$200k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
3+ YOE3+ years in SRE/DevOps with AWS/EKS, observability, CI/CD, and security focus.
Amazon EKS, Prometheus, Grafana, ELK, Kafka, Airflow, Spark, Jenkins, GitLab CI, ArgoCD, AWS, Docker, Python, Bash, PowerShell
1mo
Save
Mark Applied
Hide
Staff Software Engineer, Reliability Engineering & SDLC Governance
Carlsbad or Germantown or San Jose or San Francisco or New York City
$165k-$261k/yr OnsiteFull Time
Viasat
ViasatNASDAQ: VSAT: Provides global satellite broadband and secure networking communication services.
8+ YOE8+ years software engineering experience with SDLC governance, SRE/DevOps knowledge, observability, incident analysis, and cross-team influence to improve reliability and customer experience.
Prometheus, Grafana, OpenTelemetry, Datadog, AS9115
1mo
Save
Mark Applied
Hide
SRE/Devops Engineer- San Jose, the US
San Jose, California, United States
HybridFull Time
Kody
Kody: An agentic commerce platform providing integrated in-person payment solutions.
Deep AWS and GitHub experience, strong monitoring/logging and scripting skills, incident management ownership, and absolute fluency in Mandarin and English.
AWS, GitHub
1mo
Save
Mark Applied
Hide
Senior Kubernetes DevOps Engineer - K0rdent AI Apps/Core Services
San Jose, California, United States
HybridFull Time
Mirantis
Mirantis: Develops software for managing cloud and AI infrastructure.
Kubernetes operator expertise; experience with Linux, virtualization, networking, and storage; system architecture, scalability, and reliability; debugging production systems; microservices and enterprise IT experience; excellent English communication.
Kubernetes, k0rdent, k0s, k0smotron, Kubeflow, Kserve, vLLM, NVIDIA AI Enterprise, KubeVirt, Harbor, ArgoCD, Grafana, OpenStack, Linux
5d
Save
Mark Applied
Hide
Software Engineer III – Data Platform
Mountain View or McLean or New York City or Tampa
$173k-$201k/yr OnsiteFull Time
ID.me
ID.me: Provides secure digital identity verification and authentication services.
3+ YOEBachelor's degree in a technical field; 3–5 years in SRE, DevOps, or infrastructure engineering; cloud experience; and 1+ year with a modern programming language.
Prometheus, Grafana, OpenTelemetry, AWS, GCP, Azure, Java, Go, Python, Ruby, JavaScript, Docker, Kubernetes, Terraform, Pulumi, Ansible, GitOps, CI/CD, FedRAMP, NIST 800-53, SOC2
3w
Save
Mark Applied
Hide
Job Posting Title AI/ DevOps Engineer
San Jose, California, United States
$174k-$331k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
6+ YOE6–10 years SRE/infrastructure experience; strong datastore, Kubernetes, cloud (AWS/Azure/GCP), observability and incident response skills; interest in AI/ML ops; automation and reliability focus.
Aerospike, FoundationDB, Postgres, CosmosDB, DynamoDB, Kubernetes, AWS, Azure, GCP, Prometheus, Grafana, OpenTelemetry, Copilot, Claude Code, Codex
6d
Save
Mark Applied
Hide
ZfG Operations Engineer
San Jose, California, United States
$99k-$229k/yr HybridFull Time
Zoom
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
Proven experience in site reliability, DevOps, or infrastructure operations; proficiency with Python, Go, or Bash; knowledge of distributed systems, monitoring/observability, incident management, and automation; experience supporting production systems and on-call rotations.
Python, Go, Bash, BrightHire
1mo
Save
Mark Applied
Hide
(USA) Distinguished, Software Engineer
Bentonville or Sunnyvale
$130k-$338k/yr OnsiteFull Time
Walmart
WalmartNYSE: WMT: Multinational retail operating discount stores and supermarkets.
6+ YOE6+ years software engineering experience or 8 years alternative; strong system architecture, cloud-native and AI integration, reliability/DevOps, CI/CD, and technical leadership skills.
DevOps, CI/CD, WCAG 2.2 AA
5d
Save
Mark Applied
Hide
Senior Director of Engineering – F5 Distributed Cloud
San Jose, California, United States
OnsiteFull Time
F5
F5NASDAQ: FFIV: Provides application delivery networking and multi-cloud security solutions.
10+ YOE10+ years leading engineering teams; required experience operating and scaling cloud services, with expertise in AWS, Azure, or GCP, distributed systems, AI traffic, reliability, security, metrics, and global team leadership.
AWS, Azure, GCP, CI/CD, DevOps, Agile