35 service reliability engineer jobs at 21 companies in Soquel, CA

1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara or St. Louis or Bangalore or London or Paris or Melbourne or Taipei or Tokyo
OnsiteFull Time
Netskope
NetskopeNASDAQ: NTSK: Cloud-native cybersecurity and data protection platform for enterprises.
3+ YOEBachelor's in CS/Engineering or equivalent; 3+ years building/managing complex systems (including 1-2 years SRE); experience with cloud services, microservices, availability/performance optimization, debugging, and strong communication.
Python, C, C++, Go, Rust, Docker, Kubernetes, AWS, GCP, KVM, OpenNebula, OpenStack, TCP/IP
3w
Save
Mark Applied
Hide
Site Reliability Engineer - System Service Global
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Bachelor's in related field and strong experience with large-scale Linux host management, core data-center services (DNS, NTP, DHCP, NAT, APT, Kerberos), DevOps tooling, SRE practices, and troubleshooting.
BIND, PowerDNS, NTP, DHCP, NAT, APT, Kerberos, Ansible, Salt, Puppet, CI/CD, Python, Go, Bash, Linux
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, ASE
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Work on globally scaled, revenue-critical internet services (App Store, Music, Books, Podcasts, Fitness+); ensure reliability and scalability of services used by billions of devices.
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE) – CloudVision as a Service (CVaaS)
Santa Clara, California, United States
$101k-$161k/yr RemoteFull Time
Arista Networks
Arista NetworksNYSE: ANET: Provides cloud networking solutions and high-speed multilayer Ethernet switches.
5+ YOEBS/MS or equivalent experience,5+ years software engineering, experience with distributed databases/SaaS deployments, proficiency in Python/Golang/Bash, Kubernetes and cloud platform experience preferred.
Golang, Python, Ansible, Pulumi, Bash, Kubernetes, GKE, GCP
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - HPC
Santa Clara or Austin or Durham
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years building/supporting critical services; experience with large-scale HPC clusters (Slurm, LSF, Kubernetes); IaC and CI/CD proficiency; coding in Python/Go/Perl/Ruby; monitoring, capacity planning, and incident response skills.
Slurm, LSF, Kubernetes, Infrastructure as Code (IaC), CI/CD, AWS, GCP, OCI, Python, Go, Perl, Ruby
1mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer, Google Cloud
Atlanta or Milpitas
$240k-$250k/yr HybridFull Time
Saviynt
Saviynt: Provides AI-powered identity governance and cloud security platforms.
9+ YOE9+ years in platform/infra/SRE roles, deep Kubernetes and GCP expertise, strong Go and Python skills, experience with CI/CD, event-driven systems, observability, distributed systems, and building shared platform services.
Go (Golang), Python, Kubernetes, GCP, AWS, Azure, Kafka, RMQ, NATS, Google Pub/Sub, GitLab CI, ArgoCD, Prometheus, Grafana, ELK stack, Datadog, Envoy, Istio, MySQL, PostgresSQL
1mo
Save
Mark Applied
Hide
Site Reliability Engineer for Linux administration
Ontario or Greenville or Clearwater or Fremont
OnsiteFull Time
Hyve Solutions
Hyve SolutionsNYSE: SNX: Designs and manufactures custom hardware for hyperscale data centers.
3+ YOE3+ years Linux production administration (RHEL/Ubuntu/Rocky/CentOS); knowledge of core services, storage, backups, monitoring, security hardening, and basic scripting/automation. Bachelor's in CS/IT or equivalent experience; RHCSA/CompTIA Linux+/LPIC-1 are nice-to-have.
RHEL, Ubuntu, Rocky, CentOS, SSH, DNS, DHCP, NTP, LDAP, SSSD, Postfix, NFS, SMB, systemd, cron, yum, dnf, apt, Prometheus, Zabbix, Veeam, Bacula, NetBackup, LVM, mdadm, multipath, iSCSI, FC, ext4, xfs, btrfs, CIS, STIG, iptables, nftables, firewalld, SELinux, AppArmor, Shell, Python, Ansible, Nagios, VMware, KVM, Nutanix, Docker, Git, journalctl, syslog, auditd
3d
Save
Mark Applied
Hide
Software Engineer - NEO Reliability
San Carlos, California, United States
$200k-$300k/yr OnsiteFull Time
1X
1X: Manufacturing safe, general-purpose humanoid robots for home and work.
Experienced software engineer with reliability and testing expertise across cloud, mobile, on-robot services, and embedded systems; strong DevOps and Python/Linux skills; statistical and systems reasoning.
Python, Linux, CI/CD, APIs, HIL
1mo
Save
Mark Applied
Hide
Field Service Engineer MGC
Sacramento or San Francisco or San Jose or Modesto or Stockton
$80k-$90k/yr FieldFull Time
MGC Diagnostics
MGC Diagnostics: Sells non-invasive cardiorespiratory diagnostic systems and respiratory software.
2+ YOEAssociate or bachelor's in electronics, biomedical engineering, or related; minimum 2 years field service experience servicing medical equipment; valid driver’s license, reliable transportation; strong troubleshooting, communication, and computer skills.
CRM
1mo
Save
Mark Applied
Hide
Associate Field Service Engineer (Fixed Term – 12-month assignment)
Massachusetts or Connecticut or New Hampshire or New York or New Jersey or Pennsylvania or San Jose or Canada
$55k-$60k/yr FieldFull Time, Temporary
Sony
SonyNYSE: SONY: Sells consumer electronics, video games, and entertainment media.
0+ YOERecent college graduates (within 6 months). Must hold an Associate (required) or Bachelor’s (preferred) degree in engineering, biomedical, electronics, mechanical, life sciences or related field. Valid driver’s license and reliable vehicle, valid U.S. passport, U.S. work authorization, ability to travel (including overnight), basic computer skills,...
Salesforce
1mo
Save
Mark Applied
Hide
Production Engineer
San Jose or United States
$102k-$128k/yr HybridFull Time
Zscaler
ZscalerNASDAQ: ZS: Provides cloud-native cybersecurity solutions through zero trust architecture.
1+ YOE1-3 years managing reliability/availability for large-scale services; strong programming (Python, Go, C/C++), Linux/RHEL, networking, incident management, observability, and public cloud experience.
Python, Go, C/C++, AWS, GCP, Azure, Prometheus, Grafana, OpenTelemetry, Ansible, Terraform, Helm, Temporal, HAProxy, BGP, GRE, IPSec, Linux, RHEL, ITIL
3mo
Save
Mark Applied
Hide
Lead Systems Engineer - Traffic Management
Durham or Miami or Palo Alto or Washington
$15k-$23k/yr HybridFull Time
Nubank
NubankNYSE: NU: Digital financial platform offering banking, credit, and investment services.
Experience operating large-scale cloud distributed systems, Kubernetes, service mesh (Istio/Linkerd/Envoy), AWS networking, IaC with Pulumi or Terraform, strong reliability/observability and AI-assisted workflows.
Kubernetes, Istio, Linkerd, Envoy, Finagle, AWS, ALB, NLB, VPC, Route 53, IAM, Pulumi, Terraform
5d
Save
Mark Applied
Hide
Senior/Lead SRE Platform Services Engineer Technical Leader
San Jose, California, United States
$64k-$130k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
8+ years SRE/platform engineering experience, strong software skills (Python/Go/Ruby), infrastructure-as-code, Kubernetes, AWS, CI/CD, production reliability, security/compliance translation, and technical leadership.
Python, Go, Ruby, Kubernetes, AWS, CI/CD
1mo
Save
Mark Applied
Hide
Software Engineer III – Trust Service Team
Mountain View or McLean or New York City or Tampa
$173k-$201k/yr OnsiteFull Time
ID.me
ID.me: Provides secure digital identity verification and authentication services.
4+ YOEBachelor's degree or equivalent, 4+ years backend software experience (Java/JVM preferred), strong SQL/PostgreSQL skills, REST/OpenAPI design experience, familiarity with AI-assisted development tooling, and production reliability practices.
Java, JVM, Claude Code, Cursor, OpenAPI, PostgreSQL, SQL, OAuth2, OpenID Connect
1mo
Save
Mark Applied
Hide
Staff Software Engineer, Storage Platform
Bellevue or Menlo Park or Toronto
$230k-$270k/yr HybridFull Time
Robinhood
RobinhoodNASDAQ: HOOD: Provides a commission-free platform for investing and financial services.
5+ YOEDeep expertise in PostgreSQL/Aurora, distributed systems (sharding, replication, transactions), proficiency in Go or Rust, experience with Kubernetes and AWS services, and strong reliability/performance engineering skills.
PostgreSQL, Aurora PostgreSQL, Go, Rust, Kubernetes, RDS, DynamoDB
4d
Save
Mark Applied
Hide
Staff Backend Software Engineer
Sunnyvale, California, United States
$210k-$314k/yr OnsiteFull Time
Carbon
Carbon: Manufacturer of industrial 3D printers and advanced polymer materials.
8+ YOE8+ years building and operating production backend services; deep Linux and networking fundamentals; experience with OS provisioning, reliability, and mentoring; bachelor's in CS/Engineering or equivalent.
Linux, systemd, Bazel, MQTT, Python, Go, Node, TypeScript, C++
1mo
Save
Mark Applied
Hide
Software Engineer, AGI Platform
Los Altos or California
OnsiteFull Time
JazzX AI
JazzX AI: Building autonomous AI agents for enterprise mortgage workflows.
3+ YOE3+ years building production software; strong fundamentals in Java, Python, Go, JavaScript/TypeScript, or C++; experience with backend services, distributed systems, cloud, containers, testing, and operational reliability.
Java, Python, Go, JavaScript, TypeScript, C++, LangChain, LlamaIndex, CrewAI, AutoGen, Kubernetes, Docker
2d
Save
Mark Applied
Hide
Lead Backend Software Engineer
San Carlos, California, United States
$185k-$260k/yr HybridFull Time
Beacon AI
Beacon AI: Developing an AI-powered-pilot for safer flight operations.
5+ YOE5+ years building backend services, APIs, and distributed systems; experience with Python and JavaScript/TypeScript (Node.js); data pipeline and production reliability experience; strong ownership and communication.
Python, JavaScript, TypeScript, Node.js
1mo
Save
Mark Applied
Hide
Senior Kubernetes DevOps Engineer - K0rdent AI Apps/Core Services
San Jose, California, United States
HybridFull Time
Mirantis
Mirantis: Develops software for managing cloud and AI infrastructure.
Kubernetes operator expertise; experience with Linux, virtualization, networking, and storage; system architecture, scalability, and reliability; debugging production systems; microservices and enterprise IT experience; excellent English communication.
Kubernetes, k0rdent, k0s, k0smotron, Kubeflow, Kserve, vLLM, NVIDIA AI Enterprise, KubeVirt, Harbor, ArgoCD, Grafana, OpenStack, Linux
2mo
Save
Mark Applied
Hide
Engineering Manager - Payments
Los Gatos, California, United States
$436k-$791k/yr OnsiteFull Time
Netflix
NetflixNASDAQ: NFLX: Global video streaming and media production service.
Lead a team of distributed systems engineers; build and operate large-scale payments services; collaborate with product, regional leads, and stakeholders; drive reliability, scalability, and innovation.