51 reliability automation engineer jobs at 33 companies in Gilroy, CA
3mo
Save
Mark Applied
Hide
3mo
Reliability Engineer
Santa Clara, California, United States
$45k-$121k/yrOnsiteFull Time
WiproNYSE: WIT: Global technology services and consulting for digital transformation.
2+ YOEBachelor's in electrical engineering, 2+ years electronics or lab reliability experience, knowledge of failure analysis (PFA, CSAM, X-ray), VLSI board design, test automation scripting, and lab instruments; familiar with Windows/Linux/CentOS and Microsoft Office.
AbbottNYSE: ABT: Manufactures medical devices, diagnostics, and nutritional health products.
Ensure reliability, scalability, and performance of a medical-device remote monitoring platform; expertise in cloud (Azure), Kubernetes, observability, automation, and incident management; bachelor's in a technical discipline.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK/EFK, Datadog, Linux
Site Reliability Engineer - Hardware Infrastructure
Santa Clara, California, United States
$184k-$357k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOEDegree in CS or related field (or equivalent experience), 8+ years SRE/DevOps/Production Engineering, SRE principles, infrastructure automation, production reliability, Python/Go/Perl/Ruby, Prometheus and Grafana, strong communication.
Pleasanton or Austin or San Francisco or United States
OnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOESenior SRE with strong infrastructure, automation, and programming experience (Terraform, Chef, Ansible, Python, Java, Bash). Minimum multi-year experience in software engineering or equivalent; participates in on-call and incident response.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
Senior SRE with strong distributed systems, cloud (Azure), Kubernetes, observability, automation, incident management, and cross-functional communication skills for a medical device remote monitoring platform.
Python, Go, Bash, PowerShell, Microsoft Azure, Azure Kubernetes Service (AKS), Azure Monitor, Azure DevOps, Azure Policy, Kubernetes, Docker, Prometheus, Grafana, ELK, EFK, Datadog, Linux
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOE5+ years experience with Kubernetes and Linux, proficiency in Bash/Python, experience with infrastructure automation and large-scale server management; Top Secret/SCI clearance required or obtainable.
LinkedInNASDAQ: MSFT: Professional social network for career development and job recruitment.
6+ YOEBS in CS/CE or equivalent experience,6+ years in Linux-based infrastructure,4+ years hardware troubleshooting,experience with hardware qualification/integration,firmware/BMC knowledge,and automation for infrastructure at scale.
Assured: Software platform for automating insurance claims processing
8+ YOE8+ years in SRE/DevOps/DBA roles; deep PostgreSQL and Amazon Aurora experience; proficiency with JavaScript/TypeScript and Node.js; experience optimizing production databases and building automation; Terraform, Docker/Kubernetes, Prisma, Redshift familiarity a plus.
Instrumental: AI-powered software for electronics manufacturing quality and optimization.
5+ YOE5+ years DevOps/SRE experience on public cloud (AWS preferred); expertise in Linux, shell, containers, Kubernetes, terraform, monitoring/logging/APM; strong automation, KPI measurement, and security awareness; U.S. citizenship required for access-controlled work.
Site Reliability Engineer - Global SRE, Monetization Technology
San Jose, California, United States
$156k-$317k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
3+ YOEBachelor's or equivalent, 3+ years programming experience (C,C++,Java,Python,Perl,Go), Unix/Linux and IP networking expertise, production operations and automation experience, strong communication and ownership.
C, C++, Java, Python, Perl, Go, Unix, Linux, IP networking
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
1+ YOEBachelor's degree or equivalent experience, 1+ year in SRE, DevOps, or systems engineering, Linux and networking knowledge, distributed systems experience, programming, scripting, CI/CD, and automation skills.
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
5+ YOE5+ years PostgreSQL and production data-system experience; 3+ years Linux engineering and infrastructure automation; 2+ years Python, Bash, Go, Ruby, or Perl; cloud, Kubernetes, and distributed data-system experience.
Site Reliability Engineering (SRE) Manager, Apple Maps
Cupertino, California, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Build, manage, and deliver highly available, automated infrastructure for Apple Maps at global scale; focus on reliability, scalability, and operational excellence.