36 reliability automation engineer jobs at 21 companies in Patterson, CA
2mo
Save
Mark Applied
Hide
2mo
Reliability Engineer
Santa Clara, California, United States
$45k-$121k/yrOnsiteFull Time
WiproNYSE: WIT: Global technology services and consulting for digital transformation.
2+ YOEBachelor's in electrical engineering, 2+ years electronics or lab reliability experience, knowledge of failure analysis (PFA, CSAM, X-ray), VLSI board design, test automation scripting, and lab instruments; familiar with Windows/Linux/CentOS and Microsoft Office.
Site Reliability Engineer - Hardware Infrastructure
Santa Clara, California, United States
$184k-$357k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOEDegree in CS or related field (or equivalent experience), 8+ years SRE/DevOps/Production Engineering, SRE principles, infrastructure automation, production reliability, Python/Go/Perl/Ruby, Prometheus and Grafana, strong communication.
Pleasanton or Austin or San Francisco or United States
OnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOESenior SRE with strong infrastructure, automation, and programming experience (Terraform, Chef, Ansible, Python, Java, Bash). Minimum multi-year experience in software engineering or equivalent; participates in on-call and incident response.
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Senior Site Reliability Engineer Platform Cloud Foundations Engineer
San Jose, California, United States
$64k-$130k/yrOnsiteFull Time
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
8+ YOE8+ years SRE/platform engineering experience with AWS multi-account, Terraform, automation (Python/Go/Ruby), cloud governance, and strong documentation and communication skills.
AWS Organizations, IAM, Terraform, Python, Go, Ruby, Control Tower, Account Factory for Terraform, CloudFormation, EventBridge, Lambda, SQS, IAM Identity Center, GCP
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
8+ YOE3+ Mgmt8+ years with distributed databases, 5+ years AWS and infra automation experience, CI/CD and IaC expertise, Python/Java skills, production SRE experience, and 3+ years leading technical projects.
Lawrence Livermore National Laboratory: Develops science and technology for United States national security.
Ability to obtain DOE Q clearance (U.S. citizenship), bachelor’s degree or equivalent, broad production DB administration and automation experience across Oracle, MySQL, MSSQL, PostgreSQL, MongoDB, Cassandra, DynamoDB; observability, scripting, and enterprise app/server experience.
Site Reliability Engineer - Global SRE, Monetization Technology
San Jose, California, United States
$156k-$317k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
3+ YOEBachelor's or equivalent, 3+ years programming experience (C,C++,Java,Python,Perl,Go), Unix/Linux and IP networking expertise, production operations and automation experience, strong communication and ownership.
C, C++, Java, Python, Perl, Go, Unix, Linux, IP networking
Pure StorageNYSE: PSTG: Provides all-flash enterprise data storage and management solutions.
Operate and automate enterprise IAM platforms, implement IAC and automation, build observability and SLIs/SLOs, lead incident response and documentation; hands-on Terraform/Ansible and scripting experience required.
Applied MaterialsNASDAQ: AMAT: Manufacturers of equipment for semiconductor and display production.
Managing reliability verification and qualification for semiconductor products, leading and developing reliability teams, subject-matter expertise in system-level reliability, familiarity with semiconductor fabrication equipment, and developing AI/Copilot automation.
Site Reliability Engineer for Linux administration
Ontario or Greenville or Clearwater or Fremont
OnsiteFull Time
Hyve SolutionsNYSE: SNX: Designs and manufactures custom hardware for hyperscale data centers.
3+ YOE3+ years Linux production administration (RHEL/Ubuntu/Rocky/CentOS); knowledge of core services, storage, backups, monitoring, security hardening, and basic scripting/automation. Bachelor's in CS/IT or equivalent experience; RHCSA/CompTIA Linux+/LPIC-1 are nice-to-have.
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years in network operations; strong TCP/IP, BGP, OSPF, MPLS, EVPN; experience with AWS/Azure/GCP; hands-on with automation; Bachelor’s in CS or related field.
DataDirect Networks: High-performance storage and data management for AI and HPC.
10+ YOE10+ years in quality engineering for distributed systems, hands-on Python/Bash automation, CI/CD (Git,Jenkins), performance and reliability testing, observability and debugging skills.
Microchip TechnologyNasdaq: MCHP: Manufacturer of microcontrollers, analog, and mixed-signal integrated circuits.
5+ YOEBachelor's in Electrical Engineering, 5+ years experience, JEDEC qualification knowledge, silicon reliability understanding, hands-on software automation and hardware development, Verilog/VHDL/C/Python/Perl, ATE and wafer test experience.
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
Proven experience in site reliability, DevOps, or infrastructure operations; proficiency with Python, Go, or Bash; knowledge of distributed systems, monitoring/observability, incident management, and automation; experience supporting production systems and on-call rotations.
IntelNasdaq: INTC: Designs and manufactures microprocessors and semiconductor components.
6+ YOEBachelor's in CE/EE/CS (or related) plus 6+ years (4+ with MS, 2+ with PhD). Experience with EM/IR reliability concepts, Redhawk/Totem, SPICE, parasitic extraction, transistor-level power integrity, and scripting/automation in Linux.