171 sre engineer jobs at 78 companies in Aromas, CA

6d
Save
Mark Applied
Hide
SRE Engineer - Cloud Environment
Santa Clara, California, United States
$88k-$112k/yr OnsiteFull Time
Bambu Lab
Bambu Lab: Manufacturer of desktop 3D printers and 3D printing accessories.
3+ YOE3+ years SRE/DevOps experience, proficiency with Linux, TCP/IP networking, Kubernetes, cloud (AWS/GCP), IaC (Terraform), scripting (Python/Go/Shell), and ability to support on-call rotation.
Kubernetes, Terraform, Python, Go, Shell, AWS, GCP, Linux, TCP/IP
2w
Save
Mark Applied
Hide
AI-First SRE/DevOps Engineer
San Jose, California, United States
$120k-$160k/yr HybridFull Time
Axiad
Axiad: Identity security platform for passwordless authentication and credential management.
5+ YOE5–8 years SRE/DevOps experience with production Kubernetes, infrastructure-as-code, GitOps, CI/CD, cloud provider experience, observability, Docker, AI-First tooling; Go or Python preferred.
Kubernetes, Docker, GitOps, Go, Python, Claude Code, Cursor, Windsurf
2mo
Save
Mark Applied
Hide
SRE/Dev Ops Engineer (Hybrid, Sunnyvale)
Sunnyvale, California, United States
$120k-$180k/yr HybridFull Time
CrowdStrike
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
8+ YOE8+ years in DevOps/SRE or platform engineering; production Kubernetes; CI/CD with GitHub Actions/Jenkins/Tekton; IaC (Terraform/Pulumi); GitOps (ArgoCD/Flux); Observability (Prometheus/Grafana); multi-cloud or multi-region experience; able to work in Sunnyvale office 2+ days.
Kubernetes, GitHub Actions, Jenkins, Tekton, Terraform, Pulumi, Crossplane, ArgoCD, Flux, Prometheus, Grafana, Jaeger, OpenTelemetry, Temporal, Argo Workflows, Istio, Linkerd, Go
3mo
Save
Mark Applied
Hide
Engineering Lead – Platform & SRE
Santa Clara, California, United States
$175k-$215k/yr HybridFull Time
Kerrigan Robotics
Kerrigan Robotics: AI-powered orchestration platform for factory automation systems.
5+ YOE5+ years SRE/DevOps experience; deep Kubernetes knowledge; expertise with Pulumi/Terraform, Helm; proficiency in Go, Linux networking, observability (Prometheus, Grafana), and CI/CD (GitHub Actions).
Pulumi, Kubernetes, AWS, GCP, Azure, GitHub Actions, Prometheus, Grafana, OIDC, CNIs, Terraform, Helm, Go, Linux
1w
Save
Mark Applied
Hide
Senior/Lead SRE Platform Services Engineer Technical Leader
San Jose, California, United States
$64k-$130k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
8+ years SRE/platform engineering experience, strong software skills (Python/Go/Ruby), infrastructure-as-code, Kubernetes, AWS, CI/CD, production reliability, security/compliance translation, and technical leadership.
Python, Go, Ruby, Kubernetes, AWS, CI/CD
1mo
Save
Mark Applied
Hide
Sr. SRE Platform Architect
San Jose or Austin
HybridFull Time
Bitdeer
BitdeerNASDAQ: BTDR: Operates cryptocurrency mining and high-performance computing data centers.
10+ YOE10+ years production SRE/platform engineering or infra-architecture (including ≥3 years architect-level). Hands-on GPU/AI compute, multi-region observability, Kubernetes and cluster platforms, data-center operations, DDD and plugin framework experience; BS/MS CS.
NVIDIA, DCGM, MIG, vGPU, NVLink, NVSwitch, XID, NCCL, InfiniBand, RoCE, Lustre, NetApp, Pure, DDN, VAST, NVMe-oF, Kubernetes, GPU Operator, Slurm, Volcano, Kueue, Ray, KubeRay, ZTP, BMC, IPMI, Redfish, GitOps
2mo
Save
Mark Applied
Hide
Evaluation Reliability SRE
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Work on Siri and Apple Intelligence to build AI-driven assistant capabilities with strong focus on privacy and cross-platform impact.
3w
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Quota SRE
Sunnyvale, California, United States
$207k-$301k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
8+ YOEBachelor's degree or equivalent,8 years software/systems engineering experience,5 years SRE experience,5 years software design experience,EMR not mentioned; strong troubleshooting and stakeholder management skills.
Google Cloud, Quotaserver, Bouncer, Slicer
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE)
Palo Alto, California, United States
$175k-$229k/yr HybridFull Time
Instrumental
Instrumental: AI-powered software for electronics manufacturing quality and optimization.
5+ YOE5+ years DevOps/SRE experience on public cloud (AWS preferred); expertise in Linux, shell, containers, Kubernetes, terraform, monitoring/logging/APM; strong automation, KPI measurement, and security awareness; U.S. citizenship required for access-controlled work.
AWS, Linux, shell, containerization, Kubernetes, terraform, APM
1mo
Save
Mark Applied
Hide
Machine Learning Ops Engineer, Global SRE
San Jose, California, United States
$245k-$450k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Bachelor's in CS or equivalent; expertise in Linux, networking, storage; programming in Python, Go, C, C++, or Java; troubleshooting and production operations experience; SRE of ML systems preferred.
Linux, Python, Go, C, C++, Java
3w
Save
Mark Applied
Hide
Job Posting Title AI/ DevOps Engineer
San Jose, California, United States
$174k-$331k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
6+ YOE6–10 years SRE/infrastructure experience; strong datastore, Kubernetes, cloud (AWS/Azure/GCP), observability and incident response skills; interest in AI/ML ops; automation and reliability focus.
Aerospike, FoundationDB, Postgres, CosmosDB, DynamoDB, Kubernetes, AWS, Azure, GCP, Prometheus, Grafana, OpenTelemetry, Copilot, Claude Code, Codex
1w
Save
Mark Applied
Hide
Senior AI Tools Engineer, SRE Operations - GeForce NOW
United States or Santa Clara
$144k-$230k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOEB.S. or equivalent and 5+ years experience; strong Python; experience with AI/LLM systems, Kubernetes, AWS, large-scale data pipelines, monitoring/visualization; strong automation and SRE knowledge.
Python, Go, Kubernetes, AWS, Grafana
4d
Save
Mark Applied
Hide
Staff Site Reliability Engineer (SRE) (Hybrid)
San Francisco or San Jose or New York City or Milpitas or Mountain View or Holmdel or Goleta or Redwood City or Fremont or Sunnyvale or Brooklyn or Palo Alto
$187k-$268k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
6+ YOERequires 8+ years with a bachelor's, 6+ with a master's, or 3+ with a PhD; 6+ years in SRE or infrastructure engineering, 5+ years operating Kubernetes, cloud, CI/CD, and Python or Go.
Kubernetes, AWS, GCP, Python, Go, Terraform, MLOps, CI/CD
1mo
Save
Mark Applied
Hide
SRE/Devops Engineer- Sunnyvale, CA, the US
Sunnyvale, California, United States
OnsiteFull Time
Kody
Kody: An agentic commerce platform providing integrated in-person payment solutions.
Senior SRE with deep AWS and GitHub experience, strong monitoring/logging and scripting skills, incident leadership, and absolute fluency in Mandarin and English.
AWS, GitHub
2w
Save
Mark Applied
Hide
SRE III (L3 Tech Support+Virtualization+Linux+Network)8-12yrs
Pune or San Jose or Durham or Mexico City or Bangalore or Hoofddorp or Belgrade or Barcelona or Singapore or Sydney or Tokyo
HybridFull Time
Nutanix
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
7+ YOE7+ years SRE experience with networking, virtualization (VMware ESXi), Linux, cloud/DevOps, strong customer-support skills and degree in engineering/computer science preferred.
VMware ESXi, Linux, DevOps, Cloud, VMware, Citrix, Microsoft
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE) – CloudVision as a Service (CVaaS)
Santa Clara, California, United States
$101k-$161k/yr RemoteFull Time
Arista Networks
Arista NetworksNYSE: ANET: Provides cloud networking solutions and high-speed multilayer Ethernet switches.
5+ YOEBS/MS or equivalent experience,5+ years software engineering, experience with distributed databases/SaaS deployments, proficiency in Python/Golang/Bash, Kubernetes and cloud platform experience preferred.
Golang, Python, Ansible, Pulumi, Bash, Kubernetes, GKE, GCP
2mo
Save
Mark Applied
Hide
Senior Platform Engineer
Menlo Park or New York or Seattle
$180k-$220k/yr RemoteFull Time
Verantos
Verantos: Generates high-accuracy real-world evidence for clinical and regulatory use
5+ YOE5+ years in DevOps/platform engineering or SRE; strong AWS, Terraform, CI/CD; Python/TypeScript; healthcare or regulated environments.
AWS, Terraform, Docker, Kubernetes, ECS, EKS, Python, TypeScript, Snowflake, GitHub Actions
1mo
Save
Mark Applied
Hide
Site Reliability Engineer - Hardware Infrastructure
Santa Clara, California, United States
$184k-$357k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOEDegree in CS or related field (or equivalent experience), 8+ years SRE/DevOps/Production Engineering, SRE principles, infrastructure automation, production reliability, Python/Go/Perl/Ruby, Prometheus and Grafana, strong communication.
Python, Go, Perl, Ruby, Prometheus, Grafana
3mo
Save
Mark Applied
Hide
Senior Production Engineer, Oeprational Excellence
San Francisco or Sunnyvale
$172k-$209k/yr OnsiteFull Time
Crusoe
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
5+ YOE5+ years in Production Engineering or SRE; GPU workloads; Linux; IaC; Kubernetes; Go/Python; strong communication.
Prometheus, Grafana, OpenTelemetry, Terraform, Ansible, Kubernetes, AWS, GCP, Linux
3d
Save
Mark Applied
Hide
IT Systems Engineer - Internal Platforms & SRE
San Francisco or San Jose
$206k-$275k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
Experience with system design, scalable cloud infrastructure, configuration management, programming in Python or Go, distributed systems, automation, documentation, and cross-functional collaboration.
AWS, GCP, Azure, Chef, Ansible, Terraform, GitHub Actions, Python, Go