62 reliability automation engineer jobs at 43 companies in Live Oak, CA

3mo
Save
Mark Applied
Hide
Reliability Engineer
Santa Clara, California, United States
$45k-$121k/yr OnsiteFull Time
Wipro
WiproNYSE: WIT: Global technology services and consulting for digital transformation.
2+ YOEBachelor's in electrical engineering, 2+ years electronics or lab reliability experience, knowledge of failure analysis (PFA, CSAM, X-ray), VLSI board design, test automation scripting, and lab instruments; familiar with Windows/Linux/CentOS and Microsoft Office.
Windows, Linux, CentOS, Microsoft Office, Physical Failure Analysis (PFA), C-Scan Acoustic Microscopy (CSAM), X-ray imaging, oscilloscopes, multimeters, curve tracers, VLSI Board Design
3mo
Save
Mark Applied
Hide
Software Reliability Engineer
Mountain View, California, United States
$146k-$219k/yr OnsiteFull Time
Nuro
Nuro: Builds autonomous driving software and electric delivery robots.
Production software experience; build automation/tools; strong debugging; reliability engineering interest.
Python, Go, Bash, C++, Observability, Telemetry
4d
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Santa Clara, California, United States
$50-$60/hr RemoteContract
ServiceNow
ServiceNowNYSE: NOW: Enterprise cloud platform for digital workflow automation.
3+ YOEBachelor's degree in computer science or related field; 3+ years in site reliability engineering; 2+ years with AWS and cloud automation; Kubernetes, Linux, Terraform, networking, GitOps, monitoring, and customer support experience.
AWS, Kubernetes, Helm, Linux, Terraform, GitOps, Prometheus, Grafana, Bazel, CueLang, Version Control, Okta, Snowflake, Google
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Sunnyvale, California, United States
$90k-$180k/yr OnsiteFull Time
Abbott
AbbottNYSE: ABT: Manufactures medical devices, diagnostics, and nutritional health products.
Ensure reliability, scalability, and performance of a medical-device remote monitoring platform; expertise in cloud (Azure), Kubernetes, observability, automation, and incident management; bachelor's in a technical discipline.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK/EFK, Datadog, Linux
2mo
Save
Mark Applied
Hide
Site Reliability Engineer - Hardware Infrastructure
Santa Clara, California, United States
$184k-$357k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOEDegree in CS or related field (or equivalent experience), 8+ years SRE/DevOps/Production Engineering, SRE principles, infrastructure automation, production reliability, Python/Go/Perl/Ruby, Prometheus and Grafana, strong communication.
Python, Go, Perl, Ruby, Prometheus, Grafana
13h
Save
Mark Applied
Hide
Site Reliability Engineer - rednote
Palo Alto, California, United States
OnsiteFull Time
Rednote
Rednote: A lifestyle-focused social media and e-commerce discovery platform.
Experience with large-scale reliability, high-availability architecture, incident response, cross-region disaster recovery, Linux, networking, middleware, cloud-native infrastructure, automation, and Python, Go, or Java.
Linux, MySQL, Redis, Kafka, Kubernetes, Service Mesh, Python, Go, Java
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Pleasanton or Austin or San Francisco or United States
OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOESenior SRE with strong infrastructure, automation, and programming experience (Terraform, Chef, Ansible, Python, Java, Bash). Minimum multi-year experience in software engineering or equivalent; participates in on-call and incident response.
Terraform, Chef, Ansible, Python, Java, Bash, Kubernetes, Helm, Jenkins, Grafana, Prometheus, OCI - DevOps, Oracle Cloud Guard, Oracle Observability and Management
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara, California, United States
$230k-$250k/yr OnsiteFull Time
Forward Networks
Forward Networks: Provides a digital twin platform for enterprise network management.
6+ YOE6+ years SRE/DevOps experience in SaaS/cloud, strong networking fundamentals, Kubernetes, observability (Prometheus/Grafana/Datadog/Splunk), Python/Bash automation, cloud and IaC (AWS/GCP/Azure, Terraform/Ansible), and incident response ownership.
Kubernetes, Prometheus, Grafana, Datadog, Splunk, Python, Bash, AWS, GCP, Azure, Terraform, Ansible
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - SDN
San Francisco or San Jose or Bellevue
$240k-$312k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Kubernetes, SmartNICs, Python, Ansible, Go, C, Helm, Terraform, GitOps, CI/CD, Linux, OpenStack Neutron, OVN, OVS, DPDK, SR-IOV
3w
Save
Mark Applied
Hide
Principal Design Automation Engineer
San Jose, California, United States
$176k-$298k/yr OnsiteFull Time
Micron Technology
Micron TechnologyNASDAQ: MU: Designs and manufactures semiconductor memory and data storage solutions.
8+ YOEMS in electrical engineering required, 8+ years NAND design experience, proficiency in analog/mixed-signal design and layout, circuit verification, Python automation, Cadence tools, HSPICE/Fast SPICE/Verilog, and semiconductor reliability understanding.
HSPICE, Fast SPICE, Verilog, Python, Cadence, UNIX, LVS, DRC
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Sunnyvale or Sylmar
$90k-$180k/yr OnsiteFull Time
Abbott
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
Senior SRE with strong distributed systems, cloud (Azure), Kubernetes, observability, automation, incident management, and cross-functional communication skills for a medical device remote monitoring platform.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK, EFK, Datadog, Linux
1mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer
Palo Alto or Palo Alto or Washington
$165k-$230k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOE5+ years experience with Kubernetes and Linux, proficiency in Bash/Python, experience with infrastructure automation and large-scale server management; Top Secret/SCI clearance required or obtainable.
Kubernetes, Linux, Bash, Python, Bazel, Makefiles, Terraform, Ansible, TCP/IP
1mo
Save
Mark Applied
Hide
Alibaba Cloud-Cloud Infrastructure – Site Reliability Engineer (SRE)-Sunnyvale
Sunnyvale, California, United States
$104k-$171k/yr OnsiteFull Time
Alibaba Cloud
Alibaba CloudNYSE: BABA: Global cloud computing infrastructure and services provider
2+ YOE2+ years in distributed systems reliability engineering; high-availability architecture, Kafka/RocketMQ, Kubernetes, automation, and proficiency in Python, Go, or Java required. Bachelor's degree listed.
RocketMQ, Kafka, Kubernetes, K8s, Java, Go, Python, Shell, Terraform, Helm, Operator
17h
Save
Mark Applied
Hide
Database Reliability Engineer
Livermore, California, United States
$146k-$223k/yr HybridFull Time, Contract
Lawrence Livermore National Laboratory
Lawrence Livermore National Laboratory: Develops science and technology for United States national security.
Bachelor's degree or equivalent experience; broad database reliability, automation, administration, monitoring, application server, operating system, storage, backup, and recovery experience; U.S. citizenship and ability to obtain DOE Q clearance.
Oracle, MySQL, Microsoft SQL Server, MongoDB, Cassandra, PostgreSQL, DynamoDB, WebLogic, Tomcat, SQL, Oracle Forms/Reports, PL/SQL, Java, Python, Windows, Linux, Oracle Enterprise Manager, Datadog, Grafana, SQL Server Management Studio, Icinga, Ansible, SCCM, PowerShell
1w
Save
Mark Applied
Hide
Staff Engineer, Hardware Reliability
Sunnyvale, California, United States
$156k-$255k/yr HybridFull Time
LinkedInNASDAQ: MSFT: Professional social network for career development and job recruitment.
6+ YOEBS in CS/CE or equivalent experience,6+ years in Linux-based infrastructure,4+ years hardware troubleshooting,experience with hardware qualification/integration,firmware/BMC knowledge,and automation for infrastructure at scale.
Linux, BMC, BIOS, IPMI, Redfish, SPEC, SPECpower, fio, unixbench, NCCL, CUDA, InfiniBand, RDMA, RoCE, GPFS, HDFS, Kubernetes
1mo
Save
Mark Applied
Hide
Staff Database Reliability Engineer, DBRE
Palo Alto, California, United States
$165k-$185k/yr RemoteFull Time
Assured
Assured: Software platform for automating insurance claims processing
8+ YOE8+ years in SRE/DevOps/DBA roles; deep PostgreSQL and Amazon Aurora experience; proficiency with JavaScript/TypeScript and Node.js; experience optimizing production databases and building automation; Terraform, Docker/Kubernetes, Prisma, Redshift familiarity a plus.
PostgreSQL, Amazon Aurora, JavaScript, TypeScript, Node.js, Terraform, Terragrunt, Prisma, Docker, Kubernetes, Redshift, CI/CD
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE)
Palo Alto, California, United States
$175k-$229k/yr HybridFull Time
Instrumental
Instrumental: AI-powered software for electronics manufacturing quality and optimization.
5+ YOE5+ years DevOps/SRE experience on public cloud (AWS preferred); expertise in Linux, shell, containers, Kubernetes, terraform, monitoring/logging/APM; strong automation, KPI measurement, and security awareness; U.S. citizenship required for access-controlled work.
AWS, Linux, shell, containerization, Kubernetes, terraform, APM
3d
Save
Mark Applied
Hide
Quality & Reliability Engineer
San Francisco or San Mateo
OnsiteFull Time
Beast Industries
Beast Industries: Produces digital media and consumer goods for MrBeast brands.
2+ YOERequires 2+ years building automated test infrastructure for consumer mobile or web products, with test automation frameworks, CI/CD pipelines, monitoring tools, and strong collaboration skills.
Playwright, Appium, Espresso, XCTest, CI/CD, SLO, AI tools
2d
Save
Mark Applied
Hide
Site Reliability Engineer Spring Co-op 2027
Lowell or Durham or San Jose or Austin
$76k-$166k/yr HybridMultiple Commitments Available
IBM
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Actively enrolled in a bachelor's program, available for a 16-week full-time co-op, and knowledgeable in Linux, monitoring, troubleshooting, automation, scripting, cloud platforms, and production support.
Linux, Python, Go, Bash, IBM Cloud, AWS, Microsoft Azure, Google Cloud Platform, Kubernetes, OpenShift, Ansible, Terraform, Jenkins, IBM Continuous Delivery, ArgoCD, Instana, New Relic, Grafana, Prometheus, PostgreSQL, CouchDB, Redis, Kafka, Spark, SQL, NoSQL, CI/CD
1mo
Save
Mark Applied
Hide
Site Reliability Engineer - Global SRE, Monetization Technology
San Jose, California, United States
$156k-$317k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
3+ YOEBachelor's or equivalent, 3+ years programming experience (C,C++,Java,Python,Perl,Go), Unix/Linux and IP networking expertise, production operations and automation experience, strong communication and ownership.
C, C++, Java, Python, Perl, Go, Unix, Linux, IP networking