T-MobileNASDAQ: TMUS: The Un-carrier providing wireless and home internet services.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, scripting, APIs, cybersecurity, and reliability engineering experience, plus U.S. work authorization.
C, C#, Java, Perl, Python, Go, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Linux, Shell
Site Reliability Engineer — HPC & Automation (Silicon Engineering)
Redmond, Washington, United States
$125k-$175k/yrOnsiteFull Time
SpaceXNasdaq: SPCX: Designing, manufacturing, and launching advanced rockets and spacecraft.
2+ YOEBachelor's in CS/IS/engineering or 2+ years SRE/HPC/system administration experience; 1+ year dev experience with Bash/Python; 1+ year Linux experience; familiarity with containers, IaC, CI/CD, monitoring, databases, networking, and ASIC tool flows.
Lockbourne or Columbus or California or Colorado or Illinois or Massachusetts or Montana or Nebraska or Washington
$105k-$157k/yrOnsiteFull Time
OnTrac: National last-mile delivery and e-commerce logistics provider.
5+ YOEBachelor's degree or equivalent experience, 5+ years in reliability and maintenance engineering for facilities, conveyor, automated handling, or sortation systems, plus occupational safety knowledge.
San Francisco or New York City or Seattle or Boston or Los Angeles or Chicago or Washington or United States
$145k-$230k/yrHybridFull Time
Scribe: Workflow AI software that helps organizations capture, improve, and scale how work gets done.
Deep PostgreSQL and ORM expertise, experience with CDC pipelines (AWS DMS), OpenSearch, Redis, message brokers, observability tools, Python/Go automation, Terraform/IaC, and building reliability/scale for data tiers.
New Technology Introduction Automation Controls Engineer , Global Central Reliability Team
Boston or Austin or Nashville or Bellevue or Arlington or New York
$69k-$115k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Multinational technology focused on e-commerce and cloud computing.
3+ YOEBachelor's in engineering, 3+ years PLC programming and automation controls engineering experience, experience with complex control system design, cross-functional program coordination, and familiarity with industrial control standards and safety.
OpenEye: Private commercial cloud video surveillance serving businesses with AI-driven analytics, business intelligence, and loss-prevention tools.
1+ YOERequires 1–5 years of related experience, cloud, CI/CD, infrastructure automation, monitoring, scripting or development experience, TCP/IP knowledge, Agile familiarity, and strong problem-solving and communication skills.
Fluidstack: Building and operating civilization-scale data center infrastructure for AI.
Experience owning large GPU/compute fleets, automation of repair/deployment pipelines, firmware/BMC/Redfish familiarity, incident response and paging, observability and metrics tooling, and proficiency with production automation.
T-MobileNASDAQ: TMUS: The Un-carrier providing wireless and home internet services.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, Python, APIs, Power Platform, identity governance, and reliability engineering experience.
C, C#, Java, Perl, Python, Go, Shell, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Microsoft Power Platform, Linux, VMs
TikTok USDS Joint Venture LLC: Ensuring U.S. data security and content integrity for TikTok.
1+ YOEBachelor's degree or equivalent practical experience; 1+ year SRE, DevOps, or systems engineering experience; Linux, networking, distributed systems, programming, Bash, CI/CD, and automation skills.
Lambda: AI infrastructure building GPU cloud services and supercomputers for researchers, enterprises, and hyperscalers.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
Site Reliability Engineer (SRE) - Hybrid Cloud Storage
Seattle, Washington, United States
$140k-$210k/yrHybridFull Time
Qumulo: The world's most advanced file system – any data, any location, total control.
3+ YOE3+ years building/operating automated testing for complex software; strong C and Python skills; experience with on‑prem and cloud (AWS/GCP/Azure); Linux/Ubuntu fluency; familiarity with Kubernetes, Ansible, Terraform, and observability tooling.
Senior Site Reliability Engineer - CTJ - Top Secret
Redmond or Reston
$120k-$235k/yrHybridFull Time
MicrosoftNASDAQ: MSFT: Multinational technology providing software, cloud, and AI solutions.
2+ YOEDegree in CS/IT (MS+2yrs or BS+4yrs) or equivalent experience; active Top Secret clearance with ability/eligibility to maintain TS/SCI (with polygraph); experience with large-scale cloud/distributed systems; proficiency in C#, Go, Java, or Python; CI/CD and automation experience.
AppleNASDAQ: AAPL: Designing and manufacturing consumer electronics, software, and digital services.
Experience developing automation and tooling for distributed production systems, strong collaboration skills, and experience with Kafka and large-scale, low-latency platforms.
HP Inc.NYSE: HPQ: Global leader in personal computing, printing, and technology solutions.
15+ YOE8+ MgmtBachelor’s or master’s degree in computer science, engineering, or related field; 15+ years in software, infrastructure, platform, or reliability engineering; 8+ years leading engineering organizations; VP-level experience required.
Reliability Maintenance & Engineering – Tech School Partnerships (USA)
Seattle or Matteson or Minneapolis or Milwaukee or Des Moines or Houston or Miami or Baton Rouge or Boise or Windsor or Somerset or Niagara Falls or Orlando
OnsiteFull Time
JLLNYSE: JLL: Global professional services firm specializing in real estate and investment management.
Certificate or associate degree in industrial maintenance, reliability, or automation; willingness to network with recruiters; eligible to work in the United States.
Cowboy Space Corporation: Space infrastructure building orbital AI data centers and a solar-powered power grid for AI workloads.
4+ YOE4+ years in software reliability or systems validation; strong Python/C++ skills; experience with distributed/embedded systems and automated validation frameworks.
Python, C++, Distributed Systems, Embedded Systems, CUDA
Director, Performance Engineering and Test Automation
Oakland or Washington or Ohio or District of Columbia or California or Long Beach or Arizona or Colorado or Connecticut or Florida or Georgia or Maryland or Minnesota or Nevada or Oregon or El Dorado Hills or San Diego or Woodland Hills or Alabama or Illinois or Virginia or Wisconsin or Texas or New York
$182k-$273k/yrHybridFull Time
Stellarus: Nonprofit California health plan providing insurance coverage to individuals, families, employers, and public-program members.
10+ YOE6+ MgmtBachelor’s degree or equivalent experience, 10 years of relevant technical experience, and 6 years leading teams or technical programs in performance, automation, reliability, platform quality, or production readiness.
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
8+ YOEBachelor's, master's, Ph.D., or equivalent experience; 8+ years in infrastructure or reliability engineering; production Linux, Kubernetes, cloud, observability, networking, automation, and Python, Go, or shell scripting experience.