20 infrastructure reliability engineer jobs at 12 companies in Bastrop, TX
🚀PromotedHiringCafe
Founding Backend / Infra Engineer
Cupertino, CA, US
$160k-$300k/yrOn-SiteFull Time
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Own the crawlers, pipelines, and infrastructure powering a real-time job search engine. Strong Node.js and Python fundamentals; bonus points for security and reverse-engineering chops.
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yrRemoteFull Time
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
Austin or New York City or San Francisco or Seattle
$203k-$232k/yrOnsiteFull Time
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience in reliability engineering for infrastructure or complex hardware, building availability/RAM models, leading cross-discipline FMEAs, and mining field failure data.
Pleasanton or Austin or San Francisco or United States
OnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOESenior SRE with strong infrastructure, automation, and programming experience (Terraform, Chef, Ansible, Python, Java, Bash). Minimum multi-year experience in software engineering or equivalent; participates in on-call and incident response.
Brivo: Cloud-native platform for physical security and video surveillance management.
2+ YOE2+ years of SRE/infrastructure experience; strong Linux in production; Kubernetes or similar; Python or Bash (Golang a plus); incident response experience; ability to implement scalable reliability improvements; familiarity with LLM-based tooling for automation.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
Zello: A voice-first push-to-talk communication platform for frontline workers.
7+ YOESeasoned SRE with 7+ years in production databases and cloud infrastructure; strong expertise with MySQL and MongoDB; skilled in observability, on-call, and incident response; proficient in Python/Go/bash for automation.
Bengaluru or San Francisco or Boston or New York City or Austin or Tokyo or London
HybridFull Time
Postman: Platform for building, testing, and managing software APIs.
Experience leading engineering teams building GenAI or AI infrastructure and distributed systems; strong cloud, accelerator, and performance optimization knowledge; proficiency in Python or Go; architecture and reliability experience.
H-E-B: A retail grocery operating digital technology platforms to deliver food and household products to customers.
5+ YOE5+ years platform or SRE experience managing cloud-native data infrastructure; Databricks, AWS, Terraform, Python, and observability experience; bachelor\u0002s degree or equivalent; strong troubleshooting and communication skills.
Senior System Architect, Infrastructure Reliability
Santa Clara or Austin or Westford or Durham or Redmond
$184k-$357k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
6+ YOE6+ years systems programming experience; BS/MS/PhD in CS or EE (or equivalent); experience building RCA pipelines for HPC/cloud; deep CPU/GPU architecture knowledge; strong C++ and Python; familiarity with Slurm/LSF/Kubernetes.
Senior System Architect, Infrastructure Reliability
Santa Clara or Westford or Austin or Durham or Redmond
$184k-$357k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOE6+ years systems programming experience, BS/MS/PhD in CS or EE (or equivalent), expertise in CPU/GPU diagnostics, C++ and Python proficiency, experience with RCA, cluster managers (Slurm/LSF/Kubernetes).
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
5+ YOE5+ years technical product management, 3+ years technical (software/network) experience, bachelor's degree in CS/engineering/math/finance/economics, experience with full product lifecycle and enterprise/cloud products.
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience designing and implementing efficient, reliable, scalable software and cloud infrastructure; deep expertise optimizing applications and platform architecture.
New York City or Boston or Chicago or Salt Lake City or Austin or Montreal or Canada or Washington or Dallas or United States or Vancouver or Toronto or Charlotte or Denver
$129k-$304k/yrRemoteFull Time
Chainlink Labs: Building decentralized oracle networks for blockchain smart contracts.
Deep experience building and operating production blockchain infrastructure, strong engineering judgment, leadership and coaching experience, stakeholder management, and ability to deliver reliable, scalable systems.
Manager III, Software Development - Content Data Platform
Austin or Bozeman or Denver or Minneapolis or Missoula or Portland or Salt Lake City or Seattle or Charlotte or Kalispell or Boise or Charleston or Dallas or Fort Worth or Phoenix or Richmond or Spokane or Vermont
$150k-$188k/yrRemoteFull Time
onXmaps: Digital mapping and navigation apps for outdoor recreation.
5+ Mgmt5+ years managing software engineers; experience building data platform/infrastructure and migration off legacy systems; hiring and team development skills; experience with data quality, SLOs, and operational reliability; US work authorization required; ability to travel.