T-MobileNASDAQ: TMUS: Provides wireless network services and mobile communication devices.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, scripting, APIs, cybersecurity, and reliability engineering experience, plus U.S. work authorization.
C, C#, Java, Perl, Python, Go, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Linux, Shell
Site Reliability Engineer — HPC & Automation (Silicon Engineering)
Redmond, Washington, United States
$125k-$175k/yrOnsiteFull Time
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
2+ YOEBachelor's in CS/IS/engineering or 2+ years SRE/HPC/system administration experience; 1+ year dev experience with Bash/Python; 1+ year Linux experience; familiarity with containers, IaC, CI/CD, monitoring, databases, networking, and ASIC tool flows.
New Technology Introduction Automation Controls Engineer , Global Central Reliability Team
Boston or Austin or Nashville or Bellevue or Arlington or New York
$69k-$115k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
3+ YOEBachelor's in engineering, 3+ years PLC programming and automation controls engineering experience, experience with complex control system design, cross-functional program coordination, and familiarity with industrial control standards and safety.
San Francisco or New York City or Seattle or Boston or Los Angeles or Chicago or Washington or United States
$145k-$230k/yrHybridFull Time
Scribe: Automatically documents digital workflows into step-by-step process guides.
Deep PostgreSQL and ORM expertise, experience with CDC pipelines (AWS DMS), OpenSearch, Redis, message brokers, observability tools, Python/Go automation, Terraform/IaC, and building reliability/scale for data tiers.
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience owning large GPU/compute fleets, automation of repair/deployment pipelines, firmware/BMC/Redfish familiarity, incident response and paging, observability and metrics tooling, and proficiency with production automation.
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, Python, APIs, Power Platform, identity governance, and reliability engineering experience.
C, C#, Java, Perl, Python, Go, Shell, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Microsoft Power Platform, Linux, VMs
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
1+ YOEBachelor's degree or equivalent practical experience; 1+ year SRE, DevOps, or systems engineering experience; Linux, networking, distributed systems, programming, Bash, CI/CD, and automation skills.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
5+ YOE5+ years in systems/site reliability/DevOps/AV automation, proficiency in Python and REST API integrations, experience with automation/configuration tools and networking concepts, related technical degree required.
Splunk, Grafana, Slack, NetBox, Python, Ansible, Terraform, Puppet, Chef, Salt, New Relic, Kentik, Google Meet, Logitech, Neat, Cisco, Q-SYS, Google Workspace, Zoom, WebEx, Git, REST, AWS
Senior Principal Network Reliability Engineer - Network Region Build (NRB)
Seattle or United States
$126k-$264k/yrOnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
6+ YOEExpertise in hyperscale cloud networking, routing and transport protocols, reliability engineering, automation, and cross-organizational technical leadership; 6+ years experience preferred; English required.
Site Reliability Engineer (SRE) - Hybrid Cloud Storage
Seattle, Washington, United States
$140k-$210k/yrHybridFull Time
Qumulo: Unified file and object storage for hybrid cloud environments.
3+ YOE3+ years building/operating automated testing for complex software; strong C and Python skills; experience with on‑prem and cloud (AWS/GCP/Azure); Linux/Ubuntu fluency; familiarity with Kubernetes, Ansible, Terraform, and observability tooling.
Senior Site Reliability Engineer - CTJ - Top Secret
Redmond or Reston
$120k-$235k/yrHybridFull Time
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
2+ YOEDegree in CS/IT (MS+2yrs or BS+4yrs) or equivalent experience; active Top Secret clearance with ability/eligibility to maintain TS/SCI (with polygraph); experience with large-scale cloud/distributed systems; proficiency in C#, Go, Java, or Python; CI/CD and automation experience.
Engineering License and Workload Automation Developer
Seattle, Washington, United States
$165k-$231k/yrOnsiteFull Time
Blue Origin: Develops reusable rockets and systems for human spaceflight.
5+ YOE5+ years supporting engineering applications or license infrastructure, experience with FlexNet Publisher (FLEXlm), scripting/automation (Python preferred), distributed-systems troubleshooting, and ability to convert operational needs into reliable systems. Must be a U.S. person.
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience developing automation and tooling for distributed production systems, strong collaboration skills, and experience with Kafka and large-scale, low-latency platforms.
Reliability Maintenance & Engineering – Tech School Partnerships (USA)
Seattle or Matteson or Minneapolis or Milwaukee or Des Moines or Houston or Miami or Baton Rouge or Boise or Windsor or Somerset or Niagara Falls or Orlando
OnsiteFull Time
JLLNYSE: JLL: Global commercial real estate and investment management services.
Certificate or associate degree in industrial maintenance, reliability, or automation; willingness to network with recruiters; eligible to work in the United States.
Cowboy Space: Building satellite-based orbital infrastructure for AI data centers.
4+ YOE4+ years in software reliability or systems validation; strong Python/C++ skills; experience with distributed/embedded systems and automated validation frameworks.
Python, C++, Distributed Systems, Embedded Systems, CUDA
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
7+ YOE2+ Mgmt7+ years in software or infrastructure engineering with 2+ years leadership; SRE fundamentals, incident management, observability, change management; strong automation and people development skills.