53 reliability engineering manager jobs at 36 companies in Round Rock, TX

1mo
Save
Mark Applied
Hide
Reliability Engineer, MHE Reliability Engineering Team
Tempe or Austin or Nashville or Bellevue or Atlanta or Santa Monica or Irving or Houston
$69k-$115k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
3+ YOE3+ years engineering or maintenance experience, experience with material handling equipment or reliability programs, Associate degree or equivalent Amazon RME experience, flexible schedule with travel, project management and cross-functional leadership skills.
CMMS, Overall Equipment Effectiveness (OEE), AutoCAD, VBA, SQL, Failure Mode and Effects Analysis (FMEA), Gantt Charts
1d
Save
Mark Applied
Hide
Manager, Site Reliability Engineering
Reston or Austin
OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOE1+ Mgmt8+ years in software engineering or infrastructure (or fewer with relevant degrees), 3–5 years automation/programming experience, data analysis skills, 1 year leadership experience preferred.
2w
Save
Mark Applied
Hide
Site Reliability Engineering (SRE) Manager
Research Triangle Park or Austin
$131k-$245k/yr OnsiteFull Time
IBM
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Bachelor's degree; proven experience leading SRE/DevOps/engineering teams; ownership of SLI/SLOs, capacity planning, incident response, automation, and compliance with security and regulatory standards.
Kubernetes, OpenShift, AWS, Azure, GCP, IBM Cloud, Terraform, Ansible, Jira, Python
1mo
Save
Mark Applied
Hide
Principal Architect, Site Reliability Engineering
Southlake or Austin
$221k-$252k/yr OnsiteFull Time
Charles Schwab
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
5+ YOE3+ Mgmt5+ years in SRE with 3+ years in architect/leadership; design scalable, fault-tolerant systems; strong observability; CI/CD; postmortems; SRE leadership.
Prometheus, Grafana, Datadog, Splunk
4d
Save
Mark Applied
Hide
Engineering Manager, Support & Stability
Columbus or Boston or New York City or Chicago or Austin or Los Angeles
$150k-$190k/yr RemoteFull Time
Loop Returns
Loop Returns: Software platform automating e-commerce returns and post-purchase experiences.
3+ YOE3+ years engineering management experience, experience with platform reliability/system health, AI agent adoption, technical depth in high-risk domains, ability to manage and grow engineers.
PHP, Laravel, Vue.js, MySQL, DynamoDB, Kubernetes, AWS, Jira, Claude, Cursor, Datadog, Snowflake
3w
Save
Mark Applied
Hide
Manager, Engineering Observability
Austin, Texas, United States
$169k-$232k/yr HybridFull Time
Procore
ProcoreNYSE: PCOR: Cloud-based construction management software for projects and teams.
7+ YOE2+ Mgmt7+ years as a software engineer, 2+ years managing teams; hands-on backend/distributed systems experience; experience with observability tools (Datadog, OpenTelemetry, Prometheus/Grafana, Honeycomb, SumoLogic, Bugsnag); SRE/reliability background; strong leadership and roadmap skills.
Datadog, Honeycomb, SumoLogic, OpenTelemetry, Bugsnag, Prometheus, Grafana
3d
Save
Mark Applied
Hide
Engineering Manager, Data Feeds
New York City or Boston or Chicago or Salt Lake City or Austin or Montreal or Canada or Washington or Dallas or United States or Vancouver or Toronto or Charlotte or Denver
$129k-$304k/yr RemoteFull Time
Chainlink Labs
Chainlink Labs: Building decentralized oracle networks for blockchain smart contracts.
Deep experience building and operating production blockchain infrastructure, strong engineering judgment, leadership and coaching experience, stakeholder management, and ability to deliver reliable, scalable systems.
4d
Save
Mark Applied
Hide
Sr. Manager, Site Reliability
Austin, Texas, United States
HybridFull Time
Omnicell
OmnicellNASDAQ: OMCL: Provider of automated medication management and pharmacy solutions.
8+ YOE8+ years platform/SRE experience with 4+ years in SRE/DevOps roles, experience setting SLOs and incident command, public cloud, Terraform, Kubernetes, Python, CI/CD, and regulated environments (HIPAA,SOC2).
DataDog, IBM/Instana, Prometheus, Grafana, OpenTelemetry, Terraform, Chef, Puppet, CodeFresh, TeamCity, GitHub Actions, Octopus Deploy, Kubernetes, Docker, Helm, Istio, Linkerd, Elasticsearch/Kibana, Chaos Monkey, LitmusChaos, Databricks, Team Foundation Server, Jenkins, Python, Kafka, RabbitMQ
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin, Texas, United States
HybridFull Time
2K
2KNASDAQ: TTWO: Publishes and develops global video game franchises and entertainment.
5+ YOE5+ years SRE/platform engineering experience, deep Kubernetes (EKS/GKE), Terraform/Pulumi and GitOps, observability with Prometheus/Grafana/Datadog, production coding in Go/Python/TypeScript, Linux and networking expertise, incident management.
Terraform, Pulumi, ArgoCD, Flux, Kubernetes, EKS, GKE, Istio, Cilium, Helm, Terragrunt, Prometheus, Grafana, Datadog, OpenTelemetry, GitHub Actions, Jenkins, Go, Python, TypeScript, PasswordState, 1Password, AWS Secrets Manager, OPA/Gatekeeper, AWS, GCP, VMware, Ansible, Puppet, AWS Systems Manager
1mo
Save
Mark Applied
Hide
Lead Hardware Reliability Engineer (Starlink)
Bastrop, Texas, United States
OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
2+ YOE2+ MgmtBachelor's in engineering, 2+ years executing complex projects, 2+ years leading teams of 5+, experience with high-volume manufacturing, PCB/PCBA/SMT processes, AS9100/ISO9001 quality systems, ability to work extended hours/weekends, onsite in Bastrop, TX, and must meet ITAR eligibility.
ISO9001, AS9100, PCB, PCBA, SMT, MRB, EIS
2mo
Save
Mark Applied
Hide
Director, Site Reliability Engineering & Cloud Operations (SRE)
Austin or Golden Valley
$198k-$295k/yr HybridFull Time
Resideo
ResideoNYSE: REZI: Manufacturing and distributing home comfort and security solutions.
15+ YOE8+ Mgmt15+ years CS/EE; 15+ years Cloud Operations/SRE; 8+ years leadership; 5+ years Azure; Kubernetes; Terraform; CI/CD; IaC; large-scale IoT; observability tools
Azure, Kubernetes, Terraform, Ansible, CI/CD, Prometheus, Grafana, ELK, Pulumi, CDK
1mo
Save
Mark Applied
Hide
Advanced Technology Program Manager (Distribution Reliability)
Austin, Texas, United States
$113k-$145k/yr HybridFull Time
City of Austin
City of Austin: Providing municipal services and public infrastructure to the Austin community.
5+ YOE2+ MgmtBachelor's in Business or Engineering plus 5 years related experience (including 2 years program/project management). Requires utility operations knowledge, budget and contract management, data analysis, communications, and leadership skills.
1w
Save
Mark Applied
Hide
Director of Maintenance & Reliability
Elgin or Ferris or Amarillo
OnsiteFull Time
JC Davis Power
JC Davis Power: Provides temporary power and climate control for worksites.
Proven experience running multi-site maintenance/repair shops for fleets or heavy equipment; strong reliability engineering, PM programs, failure analysis, and people development; data-driven with telematics experience.
2mo
Save
Mark Applied
Hide
Team Lead, Site Reliability Engineer
Clearwater or Austin
HybridFull Time
TeamViewer
TeamViewerFrankfurt Stock Exchange: TMV: Provides remote connectivity and digital workplace software solutions.
3+ YOELead SRE team with hands-on Azure, containers, CI/CD; deliver reliability for global SaaS; manage people.
Microsoft Azure, Azure App Services, Application Gateway, Docker, Kubernetes, GitLab CI/CD, Azure DevOps, Terraform, Argo CD, PowerShell, Bash, Python, Grafana, Prometheus, Datadog, Keycloak, Entra, PostgreSQL, MS SQL
1d
Save
Mark Applied
Hide
Senior AI Ops & Incident/Site Reliability Engineer
Austin or Plano or Fort Mill or Boston or New York City or Tempe or San Diego
$94k-$171k/yr HybridFull Time
Perficient
Perficient: Provides digital transformation and AI consulting for global enterprises.
8+ YOEBachelor's degree or equivalent,8+ years in IT operations/SRE or production support,5+ years leading incident management,experience with Dynatrace,ServiceNow,cloud platforms and AIOps implementations.
Dynatrace, ServiceNow, AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
Engineering Project Manager (SAP FI - Financial Accounting), IS&T Enterprise Systems
Austin, Texas, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience with SAP technologies and enterprise systems, strong analytical mindset, collaboration with engineers and business teams to deliver scalable, reliable solutions.
SAP
1mo
Save
Mark Applied
Hide
Senior System Architect, Infrastructure Reliability
Santa Clara or Westford or Austin or Durham or Redmond
$184k-$357k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOE6+ years systems programming experience, BS/MS/PhD in CS or EE (or equivalent), expertise in CPU/GPU diagnostics, C++ and Python proficiency, experience with RCA, cluster managers (Slurm/LSF/Kubernetes).
C++, Python, Slurm, LSF, Kubernetes, NVIDIA DCGM, NVIDIA Management Library (NVML), CRIU, CUDA, /dev/mcelog, dmesg, journald
1w
Save
Mark Applied
Hide
Senior Engineer, Hybrid Services & Reliability (SRE)
Sunnyvale or Austin
$148k-$222k/yr HybridFull Time
General Motors
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
Proven SRE/DevOps experience in hybrid cloud, Linux administration, networking (DHCP/PXE/NTP), IaC/configuration management, automation, and mentoring; growth mindset and independent execution.
Python, Go, Linux, DHCP, PXE, NTP, Chef, Ansible, Terraform, Kubernetes (k8s)
1mo
Save
Mark Applied
Hide
AI Infrastructure Manager
Bengaluru or San Francisco or Boston or New York City or Austin or Tokyo or London
HybridFull Time
Postman
Postman: Platform for building, testing, and managing software APIs.
Experience leading engineering teams building GenAI or AI infrastructure and distributed systems; strong cloud, accelerator, and performance optimization knowledge; proficiency in Python or Go; architecture and reliability experience.
Python, Go