20 infrastructure reliability engineer jobs at 11 companies in Austin, TX

PromotedHiringCafe
Founding Backend / Infra Engineer
Cupertino, CA, US
$160k-$300k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Own the crawlers, pipelines, and infrastructure powering a real-time job search engine. Strong Node.js and Python fundamentals; bonus points for security and reverse-engineering chops.
Node.js, Python, Elasticsearch, Redis
1mo
Save
Mark Applied
Hide
Customer Reliability Engineer - Infrastructure
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yr RemoteFull Time
Astronomer
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
Apache Airflow, AWS, Azure, CI/CD, GCP, Infrastructure as Code (IaC), Kubernetes, Linux, Python
1w
Save
Mark Applied
Hide
Reliability Engineer, R&D
Austin or New York City or San Francisco or Seattle
$203k-$232k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience in reliability engineering for infrastructure or complex hardware, building availability/RAM models, leading cross-discipline FMEAs, and mining field failure data.
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Pleasanton or Austin or San Francisco or United States
OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
8+ YOESenior SRE with strong infrastructure, automation, and programming experience (Terraform, Chef, Ansible, Python, Java, Bash). Minimum multi-year experience in software engineering or equivalent; participates in on-call and incident response.
Terraform, Chef, Ansible, Python, Java, Bash, Kubernetes, Helm, Jenkins, Grafana, Prometheus, OCI - DevOps, Oracle Cloud Guard, Oracle Observability and Management
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Austin, Texas, United States
OnsiteFull Time
Brivo
Brivo: Cloud-native platform for physical security and video surveillance management.
2+ YOE2+ years of SRE/infrastructure experience; strong Linux in production; Kubernetes or similar; Python or Bash (Golang a plus); incident response experience; ability to implement scalable reliability improvements; familiarity with LLM-based tooling for automation.
Kubernetes, Linux, Python, Bash, Golang, Prometheus, Grafana
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin, Texas, United States
HybridFull Time
Zello
Zello: A voice-first push-to-talk communication platform for frontline workers.
7+ YOESeasoned SRE with 7+ years in production databases and cloud infrastructure; strong expertise with MySQL and MongoDB; skilled in observability, on-call, and incident response; proficient in Python/Go/bash for automation.
Prometheus, OpenTelemetry, Docker, Kubernetes, Google Cloud, AWS, Azure, MySQL, MongoDB, ScyllaDB, Cassandra, Elasticsearch, Redis, Git, Python, Go, Bash
7h
Save
Mark Applied
Hide
Site Reliability Engineer, AiDP Production Engineering
Austin, Texas, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Design and operate scalable, maintainable data-platform infrastructure using Kafka, Spark, Iceberg, and Airflow for on-premises and cloud environments.
Kafka, Spark, Iceberg, Airflow
1mo
Save
Mark Applied
Hide
AI Infrastructure Manager
Bengaluru or San Francisco or Boston or New York City or Austin or Tokyo or London
HybridFull Time
Postman
Postman: Platform for building, testing, and managing software APIs.
Experience leading engineering teams building GenAI or AI infrastructure and distributed systems; strong cloud, accelerator, and performance optimization knowledge; proficiency in Python or Go; architecture and reliability experience.
Python, Go
1mo
Save
Mark Applied
Hide
Senior System Architect, Infrastructure Reliability
Santa Clara or Austin or Westford or Durham or Redmond
$184k-$357k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
6+ YOE6+ years systems programming experience; BS/MS/PhD in CS or EE (or equivalent); experience building RCA pipelines for HPC/cloud; deep CPU/GPU architecture knowledge; strong C++ and Python; familiarity with Slurm/LSF/Kubernetes.
C++, Python, Slurm, LSF, Kubernetes, CUDA, DCGM, NVML, CRIU, /dev/mcelog, dmesg, journald, Linux kernel
1mo
Save
Mark Applied
Hide
Senior System Architect, Infrastructure Reliability
Santa Clara or Westford or Austin or Durham or Redmond
$184k-$357k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOE6+ years systems programming experience, BS/MS/PhD in CS or EE (or equivalent), expertise in CPU/GPU diagnostics, C++ and Python proficiency, experience with RCA, cluster managers (Slurm/LSF/Kubernetes).
C++, Python, Slurm, LSF, Kubernetes, NVIDIA DCGM, NVIDIA Management Library (NVML), CRIU, CUDA, /dev/mcelog, dmesg, journald
1w
Save
Mark Applied
Hide
Senior Product Manger - Tech, Infrastructure Reliability
Austin or Nashville or Arlington or North Reading
$145k-$206k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
5+ YOE5+ years technical product management, 3+ years technical (software/network) experience, bachelor's degree in CS/engineering/math/finance/economics, experience with full product lifecycle and enterprise/cloud products.
3d
Save
Mark Applied
Hide
Engineering Manager, Data Feeds
New York City or Boston or Chicago or Salt Lake City or Austin or Montreal or Canada or Washington or Dallas or United States or Vancouver or Toronto or Charlotte or Denver
$129k-$304k/yr RemoteFull Time
Chainlink Labs
Chainlink Labs: Building decentralized oracle networks for blockchain smart contracts.
Deep experience building and operating production blockchain infrastructure, strong engineering judgment, leadership and coaching experience, stakeholder management, and ability to deliver reliable, scalable systems.
1mo
Save
Mark Applied
Hide
Manager III, Software Development - Content Data Platform
Austin or Bozeman or Denver or Minneapolis or Missoula or Portland or Salt Lake City or Seattle or Charlotte or Kalispell or Boise or Charleston or Dallas or Fort Worth or Phoenix or Richmond or Spokane or Vermont
$150k-$188k/yr RemoteFull Time
onXmaps
onXmaps: Digital mapping and navigation apps for outdoor recreation.
5+ Mgmt5+ years managing software engineers; experience building data platform/infrastructure and migration off legacy systems; hiring and team development skills; experience with data quality, SLOs, and operational reliability; US work authorization required; ability to travel.
GCP, Managed Spark, BigQuery, BigLake, Pub/Sub, Managed Airflow, Apache Iceberg, Spark, PySpark, DuckDB, dbt, Airflow, GDAL, PostGIS, Apache Sedona, DuckDB spatial extensions, Knowledge Graph, AI-assisted tools

Explore Jobs

Expand Your Job Search