207 sre engineer jobs at 100 companies in Santa Cruz, CA
2w
Save
Mark Applied
Hide
2w
SRE Engineer - Cloud Environment
Santa Clara, California, United States
$88k-$112k/yrOnsiteFull Time
Bambu Lab: Manufacturer of desktop 3D printers and 3D printing accessories.
3+ YOE3+ years SRE/DevOps experience, proficiency with Linux, TCP/IP networking, Kubernetes, cloud (AWS/GCP), IaC (Terraform), scripting (Python/Go/Shell), and ability to support on-call rotation.
Axiad: Identity security platform for passwordless authentication and credential management.
5+ YOE5–8 years SRE/DevOps experience with production Kubernetes, infrastructure-as-code, GitOps, CI/CD, cloud provider experience, observability, Docker, AI-First tooling; Go or Python preferred.
Kubernetes, Docker, GitOps, Go, Python, Claude Code, Cursor, Windsurf
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
8+ YOE8+ years in DevOps/SRE or platform engineering; production Kubernetes; CI/CD with GitHub Actions/Jenkins/Tekton; IaC (Terraform/Pulumi); GitOps (ArgoCD/Flux); Observability (Prometheus/Grafana); multi-cloud or multi-region experience; able to work in Sunnyvale office 2+ days.
Kerrigan Robotics: AI-powered orchestration platform for factory automation systems.
5+ YOE5+ years SRE/DevOps experience; deep Kubernetes knowledge; expertise with Pulumi/Terraform, Helm; proficiency in Go, Linux networking, observability (Prometheus, Grafana), and CI/CD (GitHub Actions).
BitdeerNASDAQ: BTDR: Operates cryptocurrency mining and high-performance computing data centers.
10+ YOE10+ years production SRE/platform engineering or infra-architecture (including ≥3 years architect-level). Hands-on GPU/AI compute, multi-region observability, Kubernetes and cluster platforms, data-center operations, DDD and plugin framework experience; BS/MS CS.
Instrumental: AI-powered software for electronics manufacturing quality and optimization.
5+ YOE5+ years DevOps/SRE experience on public cloud (AWS preferred); expertise in Linux, shell, containers, Kubernetes, terraform, monitoring/logging/APM; strong automation, KPI measurement, and security awareness; U.S. citizenship required for access-controlled work.
ServiceNowNYSE: NOW: Enterprise cloud platform for digital workflow automation.
3+ YOEBachelor's degree in computer science or related field; 3+ years in site reliability engineering; 2+ years with AWS and cloud automation; Kubernetes, Linux, Terraform, networking, GitOps, monitoring, and customer support experience.
AWS, Kubernetes, Helm, Linux, Terraform, GitOps, Prometheus, Grafana, Bazel, CueLang, Version Control, Okta, Snowflake, Google
TikTok: Global short-form video hosting and social media platform.
Bachelor's in CS or equivalent; expertise in Linux, networking, storage; programming in Python, Go, C, C++, or Java; troubleshooting and production operations experience; SRE of ML systems preferred.
Kody: An agentic commerce platform providing integrated in-person payment solutions.
Senior SRE with deep AWS and GitHub experience, strong monitoring/logging and scripting skills, incident leadership, and absolute fluency in Mandarin and English.
Alibaba Cloud-Cloud Infrastructure – Site Reliability Engineer (SRE)-Sunnyvale
Sunnyvale, California, United States
$104k-$171k/yrOnsiteFull Time
Alibaba CloudNYSE: BABA: Global cloud computing infrastructure and services provider
2+ YOE2+ years in distributed systems reliability engineering; high-availability architecture, Kafka/RocketMQ, Kubernetes, automation, and proficiency in Python, Go, or Java required. Bachelor's degree listed.
5+ YOEBS/MS or equivalent experience,5+ years software engineering, experience with distributed databases/SaaS deployments, proficiency in Python/Golang/Bash, Kubernetes and cloud platform experience preferred.
WiproNYSE: WIT: Global technology services and consulting for digital transformation.
5+ YOERequires 5–8 years of experience in data engineering and SRE, with capacity planning, forecasting, scaling, pipeline development, data modeling, automation, monitoring, and resilience testing skills.
Blackhawk Network: Provider of global branded payment and gift card solutions.
Bachelor's in CS/Engineering or equivalent; experience in platform/DevOps/SRE or similar; strong Linux, AWS, Git; scripting with Python/Bash; production support and incident management; experience with AI-assisted engineering tools.
AWS, Git, Python, Bash, Kubernetes, Docker, Jenkins, Splunk, New Relic, Prometheus, Grafana, OpenTelemetry, ServiceNow, Terraform, CloudFormation, GitHub Copilot, Cursor, Claude
STN: Provides high-performance GPU infrastructure, cloud, and managed IT services.
6+ YOE6+ years in platform/SRE/cloud engineering, deep Kubernetes expertise, Go and/or Python programming, experience operating GPU/AI infrastructure, bachelor's degree or equivalent experience.
Site Reliability Engineer - Hardware Infrastructure
Santa Clara, California, United States
$184k-$357k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOEDegree in CS or related field (or equivalent experience), 8+ years SRE/DevOps/Production Engineering, SRE principles, infrastructure automation, production reliability, Python/Go/Perl/Ruby, Prometheus and Grafana, strong communication.