T-MobileNASDAQ: TMUS: Provides wireless network services and mobile communication devices.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, scripting, APIs, cybersecurity, and reliability engineering experience, plus U.S. work authorization.
C, C#, Java, Perl, Python, Go, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Linux, Shell
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
6+ YOE6+ years software/network/systems engineering experience or equivalent degree+experience; experience with large-scale cloud services, observability, automation, and customer-facing reliability work.
Microsoft Fabric, Azure SQL DB, Azure Cosmos DB, Azure PostgreSQL, Azure Data Factory, Azure Synapse Analytics, Azure Service Bus, Azure Event Grid, Power BI, Logic Apps, Jupyter Notebooks
Site Reliability Engineer — HPC & Automation (Silicon Engineering)
Redmond, Washington, United States
$125k-$175k/yrOnsiteFull Time
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
2+ YOEBachelor's in CS/IS/engineering or 2+ years SRE/HPC/system administration experience; 1+ year dev experience with Bash/Python; 1+ year Linux experience; familiarity with containers, IaC, CI/CD, monitoring, databases, networking, and ASIC tool flows.
San Francisco or New York City or Seattle or Boston or Los Angeles or Chicago or Washington or United States
$145k-$230k/yrHybridFull Time
Scribe: Automatically documents digital workflows into step-by-step process guides.
Deep PostgreSQL and ORM expertise, experience with CDC pipelines (AWS DMS), OpenSearch, Redis, message brokers, observability tools, Python/Go automation, Terraform/IaC, and building reliability/scale for data tiers.
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience owning large GPU/compute fleets, automation of repair/deployment pipelines, firmware/BMC/Redfish familiarity, incident response and paging, observability and metrics tooling, and proficiency with production automation.
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, Python, APIs, Power Platform, identity governance, and reliability engineering experience.
C, C#, Java, Perl, Python, Go, Shell, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Microsoft Power Platform, Linux, VMs
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years SRE/production engineering experience; Kubernetes, Linux networking, observability, on-call/incident response, automation with Python/Ansible; experience with multi-datacenter and hybrid cloud environments.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
5+ YOE5+ years in systems/site reliability/DevOps/AV automation, proficiency in Python and REST API integrations, experience with automation/configuration tools and networking concepts, related technical degree required.
Splunk, Grafana, Slack, NetBox, Python, Ansible, Terraform, Puppet, Chef, Salt, New Relic, Kentik, Google Meet, Logitech, Neat, Cisco, Q-SYS, Google Workspace, Zoom, WebEx, Git, REST, AWS
RobinhoodNASDAQ: HOOD: Provides a commission-free platform for investing and financial services.
1+ YOEMaster's in Computer Science or related field plus one year of experience; expertise in production network infrastructure, BGP/OSPF, network automation (Python or Golang), and monitoring.
Reliability Engineer, PSO, Amazon Prime Air, Amazon Prime Air
Seattle, Washington, United States
$129k-$175k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
5+ YOEBachelor's in STEM, 5+ years experience in robotics/automation or related product development, strong data analysis and troubleshooting skills, systems reliability knowledge, and ability to lead cross-functional investigations.
Senior Principal Network Reliability Engineer - Network Region Build (NRB)
Seattle or United States
$126k-$264k/yrOnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
6+ YOEExpertise in hyperscale cloud networking, routing and transport protocols, reliability engineering, automation, and cross-organizational technical leadership; 6+ years experience preferred; English required.
Site Reliability Engineer (SRE) - Hybrid Cloud Storage
Seattle, Washington, United States
$140k-$210k/yrHybridFull Time
Qumulo: Unified file and object storage for hybrid cloud environments.
3+ YOE3+ years building/operating automated testing for complex software; strong C and Python skills; experience with on‑prem and cloud (AWS/GCP/Azure); Linux/Ubuntu fluency; familiarity with Kubernetes, Ansible, Terraform, and observability tooling.
Engineering License and Workload Automation Developer
Seattle, Washington, United States
$165k-$231k/yrOnsiteFull Time
Blue Origin: Develops reusable rockets and systems for human spaceflight.
5+ YOE5+ years supporting engineering applications or license infrastructure, experience with FlexNet Publisher (FLEXlm), scripting/automation (Python preferred), distributed-systems troubleshooting, and ability to convert operational needs into reliable systems. Must be a U.S. person.
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience developing automation and tooling for distributed production systems, strong collaboration skills, and experience with Kafka and large-scale, low-latency platforms.
Reliability Maintenance & Engineering – Tech School Partnerships (USA)
Seattle or Matteson or Minneapolis or Milwaukee or Des Moines or Houston or Miami or Baton Rouge or Boise or Windsor or Somerset or Niagara Falls or Orlando
OnsiteFull Time
JLLNYSE: JLL: Global commercial real estate and investment management services.
Certificate or associate degree in industrial maintenance, reliability, or automation; willingness to network with recruiters; eligible to work in the United States.
Cowboy Space: Building satellite-based orbital infrastructure for AI data centers.
4+ YOE4+ years in software reliability or systems validation; strong Python/C++ skills; experience with distributed/embedded systems and automated validation frameworks.
Python, C++, Distributed Systems, Embedded Systems, CUDA
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
7+ YOE2+ Mgmt7+ years in software or infrastructure engineering with 2+ years leadership; SRE fundamentals, incident management, observability, change management; strong automation and people development skills.