50 reliability engineering manager jobs at 33 companies in Blanco, TX
2w
Save
Mark Applied
Hide
2w
Manager, Site Reliability Engineering
Reston or Austin
OnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOE1+ Mgmt8+ years in software engineering or infrastructure (or fewer with relevant degrees), 3–5 years automation/programming experience, data analysis skills, 1 year leadership experience preferred.
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
5+ YOE3+ Mgmt5+ years in SRE with 3+ years in architect/leadership; design scalable, fault-tolerant systems; strong observability; CI/CD; postmortems; SRE leadership.
Columbus or Boston or New York City or Chicago or Austin or Los Angeles
$150k-$190k/yrRemoteFull Time
Loop Returns: Software platform automating e-commerce returns and post-purchase experiences.
3+ YOE3+ years engineering management experience, experience with platform reliability/system health, AI agent adoption, technical depth in high-risk domains, ability to manage and grow engineers.
ProcoreNYSE: PCOR: Cloud-based construction management software for projects and teams.
7+ YOE2+ Mgmt7+ years as a software engineer, 2+ years managing teams; hands-on backend/distributed systems experience; experience with observability tools (Datadog, OpenTelemetry, Prometheus/Grafana, Honeycomb, SumoLogic, Bugsnag); SRE/reliability background; strong leadership and roadmap skills.
New York City or Boston or Chicago or Salt Lake City or Austin or Montreal or Canada or Washington or Dallas or United States or Vancouver or Toronto or Charlotte or Denver
$129k-$304k/yrRemoteFull Time
Chainlink Labs: Building decentralized oracle networks for blockchain smart contracts.
Deep experience building and operating production blockchain infrastructure, strong engineering judgment, leadership and coaching experience, stakeholder management, and ability to deliver reliable, scalable systems.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
San Antonio or Ashburn or North America or Europe or Asia
$125k-$135k/yrHybridFull Time
Vantage Data Centers: Provides hyperscale data center campuses for cloud and AI providers.
2+ YOEBachelor's degree in electrical or mechanical engineering preferred, 2–3 years in critical facility operations and maintenance, strong project management and communication skills, and willingness to travel up to 25%.
Apptronik: Designs and manufactures humanoid robots for industrial automation.
8+ YOE2+ MgmtHands-on technical manager to build and lead production engineering for robot deployment, commissioning, and fleet reliability; 8+ years engineering or 4+ years humanoid experience, 2+ years managerial experience, willing to travel up to 25%.
2KNASDAQ: TTWO: Publishes and develops global video game franchises and entertainment.
5+ YOE5+ years SRE/platform engineering experience, deep Kubernetes (EKS/GKE), Terraform/Pulumi and GitOps, observability with Prometheus/Grafana/Datadog, production coding in Go/Python/TypeScript, Linux and networking expertise, incident management.
Sr Manager, AI Systems Quality & Reliability , Annapurna AI Servers and Systems
Austin or Seattle or Cupertino
$208k-$282k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
10+ YOE5+ Mgmt10+ years reliability/quality engineering experience with server or high-volume electronics, 5+ years people management, bachelor's degree in a relevant field, experience with root-cause analysis, quality systems, and multi-vendor manufacturing.
iHeartMediaNASDAQ: IHRT: Provides radio broadcasting, podcasting, and digital audio streaming services.
2+ YOEMaster's in CS/CE/EE/IS (or equivalent) plus 24 months as a Software Engineer or related; experience leading SRE/DevOps teams; strong communication, delegation, troubleshooting, and improvement skills.
AtlanticusNASDAQ: ATLC: Provides credit cards and lending solutions for underserved consumers.
5+ YOERequires 5+ years supporting production applications, Java, AWS, Kubernetes, Docker, Datadog or Splunk, CI/CD, Python or Bash, Linux, cloud troubleshooting, and incident management experience.
CrowdStrikeNASDAQ: CRWD: Provides cloud-native endpoint protection and cybersecurity services.
10+ YOE10+ years building distributed systems, 5+ years developing SaaS microservices, expert programming skills, distributed-systems expertise, architectural leadership, and a Computer Science degree or equivalent experience.
Electric Reliability Compliance Analyst Senior - Operations & Planning
Austin, Texas, United States
HybridFull Time
City of Austin: Providing municipal services and public infrastructure to the Austin community.
4+ YOEBachelor's in Business, Engineering, or related plus 4 years energy/electric utility experience; knowledge of NERC/FERC/ERCOT/PUCT reliability requirements; audit, reporting and training experience; ability to travel and obtain required clearances.
Site Reliability Engineer, Apple Data Platform / Multi-Cloud Infrastructure
Austin, Texas, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Manage and operate a massive multi-cloud data platform, run incident response, provide hands-on support to internal teams, and partner with developers to keep services reliable across AWS, GCP, and on-prem Kubernetes.
Senior AI Ops & Incident/Site Reliability Engineer
Austin or Plano or Fort Mill or Boston or New York City or Tempe or San Diego
$94k-$171k/yrHybridFull Time
Perficient: Provides digital transformation and AI consulting for global enterprises.
8+ YOEBachelor's degree or equivalent,8+ years in IT operations/SRE or production support,5+ years leading incident management,experience with Dynatrace,ServiceNow,cloud platforms and AIOps implementations.
Senior System Architect, Infrastructure Reliability
Santa Clara or Westford or Austin or Durham or Redmond
$184k-$357k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOE6+ years systems programming experience, BS/MS/PhD in CS or EE (or equivalent), expertise in CPU/GPU diagnostics, C++ and Python proficiency, experience with RCA, cluster managers (Slurm/LSF/Kubernetes).