141 reliability automation engineer jobs at 96 companies in California

3mo
Save
Mark Applied
Hide
Reliability Engineer
Santa Clara, California, United States
$45k-$121k/yr OnsiteFull Time
Wipro
WiproNYSE: WIT: Global technology services and consulting for digital transformation.
2+ YOEBachelor's in electrical engineering, 2+ years electronics or lab reliability experience, knowledge of failure analysis (PFA, CSAM, X-ray), VLSI board design, test automation scripting, and lab instruments; familiar with Windows/Linux/CentOS and Microsoft Office.
Windows, Linux, CentOS, Microsoft Office, Physical Failure Analysis (PFA), C-Scan Acoustic Microscopy (CSAM), X-ray imaging, oscilloscopes, multimeters, curve tracers, VLSI Board Design
3mo
Save
Mark Applied
Hide
Software Reliability Engineer
Mountain View, California, United States
$146k-$219k/yr OnsiteFull Time
Nuro
Nuro: Builds autonomous driving software and electric delivery robots.
Production software experience; build automation/tools; strong debugging; reliability engineering interest.
Python, Go, Bash, C++, Observability, Telemetry
1mo
Save
Mark Applied
Hide
Reliability Engineer, Supercomputing
San Francisco, California, United States
$350k-$475k/yr OnsiteFull Time
Thinking Machines
Thinking Machines: Building AI systems to extend human will and judgment.
Ensure reliability of GPU supercomputing fleet across hardware, firmware, and OS; debug kernel/driver/hardware issues; engage vendors; automate monitoring and runroot-cause analysis.
Python, Rust, Kubernetes, Slurm, Linux, BMC, iDRAC, IPMI, Redfish, DCGM, NVLink, NVSwitch, Linux kernel
2w
Save
Mark Applied
Hide
Senior Reliability Engineer, Labs
San Francisco or Oakland or United States
$139k-$205k/yr OnsiteFull Time
DoorDash
DoorDashNYSE: DASH: Local food delivery and on-demand logistics platform.
5+ YOEFive years of reliability validation or hardware testing experience in robotics or automated vehicles; bachelor's or higher in engineering; experience with test equipment, CAD, shop tools, Python, and hardware-software validation.
Python, CAD, DAQ, FRACAS, FMEA, FTA, HIL
2mo
Save
Mark Applied
Hide
Senior Reliability Engineer
San Diego, California, United States
$107k-$120k/yr OnsiteFull Time
PCI Pharma Services
PCI Pharma Services: Integrated pharmaceutical development, manufacturing, and global packaging services.
7+ YOEBachelor’s in chemical, electrical, industrial, mechanical engineering or computer science; 7-10+ years engineering experience with at least 4 years cGMP; strong GMP and aseptic knowledge; automation and control platforms familiarity; proficient in MS Office; able to work autonomously and collaboratively.
Allen Bradley, Siemens, Wonderware, iFix, PI, Microsoft Office
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Los Angeles, California, United States
$140k-$199k/yr HybridFull Time
Green Dot
Green DotNYSE: GDOT: Provides mobile banking and payment solutions to consumers and businesses.
7+ YOE7+ years in release/reliability engineering, cloud platform experience (AWS/Azure/GCP), automated deployment and observability proficiency, scripting with PowerShell/Bash/Python, excellent troubleshooting and communication skills.
AWS, Azure, GCP, PowerShell, Bash, Python
5d
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Santa Clara, California, United States
$50-$60/hr RemoteContract
ServiceNow
ServiceNowNYSE: NOW: Enterprise cloud platform for digital workflow automation.
3+ YOEBachelor's degree in computer science or related field; 3+ years in site reliability engineering; 2+ years with AWS and cloud automation; Kubernetes, Linux, Terraform, networking, GitOps, monitoring, and customer support experience.
AWS, Kubernetes, Helm, Linux, Terraform, GitOps, Prometheus, Grafana, Bazel, CueLang, Version Control, Okta, Snowflake, Google
3w
Save
Mark Applied
Hide
Site Reliability Engineer
San Francisco or Alpharetta or Arlington or Augusta or Ashburn or Allentown or Appleton or Atlanta or Annapolis Junction or Ann Arbor or Herndon or Allen
$165k-$241k/yr RemoteFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
7+ YOE7+ years SRE or related experience; BS/MS/PhD with corresponding years; U.S. Person required for FedRAMP/IL-5 work; on-call participation; strong coding, automation, reliability, and security skills.
1w
Save
Mark Applied
Hide
Principal Site Reliability Engineer (Hybrid)
Merrimack or San Diego
$118k-$201k/yr HybridFull Time
BAE Systems
BAE SystemsLondon Stock Exchange: BA: Provides advanced defense, aerospace, and security technology solutions.
4+ YOERequires 4–6+ years of site reliability engineering, Juniper networking, cloud technologies, automation, storage, virtualization, and security clearance eligibility; Security+ required or obtainable within 90 days.
Juniper, Ansible, Helm Charts, NFS, JDFS, Ceph, S3, VMware, Open Stack, Azure Stack, Kubernetes, Terraform
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Sunnyvale, California, United States
$90k-$180k/yr OnsiteFull Time
Abbott
AbbottNYSE: ABT: Manufactures medical devices, diagnostics, and nutritional health products.
Ensure reliability, scalability, and performance of a medical-device remote monitoring platform; expertise in cloud (Azure), Kubernetes, observability, automation, and incident management; bachelor's in a technical discipline.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK/EFK, Datadog, Linux
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (Raptor)
Hawthorne, California, United States
$125k-$175k/yr OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
1+ YOE1+ years hands-on experience with client/server hardware, networking, Linux/Windows, scripting and automation; bachelor's in CS/engineering/math or 2+ years software experience in lieu; HPC and systems engineering experience preferred.
Infiniband, ANSYS, StarCCM+, Bash, Python, Puppet, Ansible, Kubernetes, Docker, Linux, Windows
2mo
Save
Mark Applied
Hide
Site Reliability Engineer - Hardware Infrastructure
Santa Clara, California, United States
$184k-$357k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOEDegree in CS or related field (or equivalent experience), 8+ years SRE/DevOps/Production Engineering, SRE principles, infrastructure automation, production reliability, Python/Go/Perl/Ruby, Prometheus and Grafana, strong communication.
Python, Go, Perl, Ruby, Prometheus, Grafana
1d
Save
Mark Applied
Hide
Site Reliability Engineer - rednote
Palo Alto, California, United States
OnsiteFull Time
Rednote
Rednote: A lifestyle-focused social media and e-commerce discovery platform.
Experience with large-scale reliability, high-availability architecture, incident response, cross-region disaster recovery, Linux, networking, middleware, cloud-native infrastructure, automation, and Python, Go, or Java.
Linux, MySQL, Redis, Kafka, Kubernetes, Service Mesh, Python, Go, Java
2mo
Save
Mark Applied
Hide
Senior Database Reliability Engineer
San Francisco or New York City or Seattle or Boston or Los Angeles or Chicago or Washington or United States
$145k-$230k/yr HybridFull Time
Scribe
Scribe: Automatically documents digital workflows into step-by-step process guides.
Deep PostgreSQL and ORM expertise, experience with CDC pipelines (AWS DMS), OpenSearch, Redis, message brokers, observability tools, Python/Go automation, Terraform/IaC, and building reliability/scale for data tiers.
Django, PostgreSQL, Aurora Serverless V2, OpenSearch, Redis, ElastiCache, SQS, RabbitMQ, DMS, S3, Parquet, Snowflake, pganalyze, CloudWatch, Honeycomb, OpenTelemetry, Datadog DBM, pg_stat_statements, Kafka, Python, Go, Terraform, Debezium, Fivetran, Airbyte, pgbouncer, RDS Proxy, Snowpipe, BigQuery, Redshift, SQLAlchemy, ActiveRecord
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Pleasanton or Austin or San Francisco or United States
OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOESenior SRE with strong infrastructure, automation, and programming experience (Terraform, Chef, Ansible, Python, Java, Bash). Minimum multi-year experience in software engineering or equivalent; participates in on-call and incident response.
Terraform, Chef, Ansible, Python, Java, Bash, Kubernetes, Helm, Jenkins, Grafana, Prometheus, OCI - DevOps, Oracle Cloud Guard, Oracle Observability and Management
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara, California, United States
$230k-$250k/yr OnsiteFull Time
Forward Networks
Forward Networks: Provides a digital twin platform for enterprise network management.
6+ YOE6+ years SRE/DevOps experience in SaaS/cloud, strong networking fundamentals, Kubernetes, observability (Prometheus/Grafana/Datadog/Splunk), Python/Bash automation, cloud and IaC (AWS/GCP/Azure, Terraform/Ansible), and incident response ownership.
Kubernetes, Prometheus, Grafana, Datadog, Splunk, Python, Bash, AWS, GCP, Azure, Terraform, Ansible
3mo
Save
Mark Applied
Hide
Automation Engineer, NDT
Los Angeles, California, United States
$145k-$230k/yr OnsiteFull Time
Hadrian
Hadrian: Building autonomous factories for aerospace and defense manufacturing.
3+ YOE3+ years designing, deploying, or commissioning automated inspection/NDT systems in production/high-reliability environments; direct ownership of at least one automated inspection station; experience with UT, VT, RT, CT; ITAR eligible
Robotics, Industrial Controls, Fanuc, ABB, Kuka, Yaskawa, Beckhoff TwinCAT, Siemens TIA, RoboDK, Process Simulate, Visual Components, Python, C#
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Compute
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience owning large GPU/compute fleets, automation of repair/deployment pipelines, firmware/BMC/Redfish familiarity, incident response and paging, observability and metrics tooling, and proficiency with production automation.
Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, Kubernetes, Claude Code, Cursor, LLM APIs, MCP servers
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - SDN
San Francisco or San Jose or Bellevue
$240k-$312k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Kubernetes, SmartNICs, Python, Ansible, Go, C, Helm, Terraform, GitOps, CI/CD, Linux, OpenStack Neutron, OVN, OVS, DPDK, SR-IOV
1mo
Save
Mark Applied
Hide
Staff Reliability/FA Engineer (Irvine, CA, US)
Irvine, California, United States
$114k-$220k/yr OnsiteFull Time
Skyworks Solutions
Skyworks SolutionsNASDAQ: SWKS: Designs and manufactures analog and mixed-signal semiconductor solutions.
8+ YOEBS degree, minimum 8 years' experience (5+ in silicon semiconductor industry), knowledge of IC reliability and JEDEC standards, Python and ML/AI familiarity, Microsoft Office proficiency, strong analytical and communication skills.
Python, Microsoft Excel, Microsoft Word, Microsoft PowerPoint, machine learning, AI, automation tools

Explore Jobs

Expand Your Job Search