49 site reliability engineer jobs at 24 companies in Georgetown, TX
2mo
Save
Mark Applied
Hide
2mo
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Future Secure AI: Deploys secure, persona-based AI-Workers for large enterprises.
5+ YOEHands-on Kubernetes, Terraform, and Helm experience; programming in Python/Go/Java/Bash/PowerShell/Ruby; SRE experience with on-call, incident response, SLIs/SLOs; cloud and CI/CD experience; 5+ years preferred.
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
3+ YOE3+ years systems support/scripting in enterprise environments, Windows/Linux proficiency, cloud experience, automation and CI/CD familiarity, bachelor's or equivalent, strong communication and problem-solving skills.
2KNASDAQ: TTWO: Publishes and develops global video game franchises and entertainment.
5+ YOE5+ years SRE/platform engineering experience, deep Kubernetes (EKS/GKE), Terraform/Pulumi and GitOps, observability with Prometheus/Grafana/Datadog, production coding in Go/Python/TypeScript, Linux and networking expertise, incident management.
VisaNYSE: V: Global payment technology facilitating electronic funds transfers.
2+ YOE2+ years with a Bachelor's or 5+ years experience; hands-on Azure, Kubernetes, Terraform, IaC/GitOps, CI/CD, observability, service mesh (Istio preferred); strong SRE, troubleshooting, documentation, and English (B2+).
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience owning large GPU/compute fleets, automation of repair/deployment pipelines, firmware/BMC/Redfish familiarity, incident response and paging, observability and metrics tooling, and proficiency with production automation.
Cox Enterprises: Providing global communications, automotive services, and media solutions.
6+ YOEBachelor's in CS or related and 6 years experience (or alternate degree/experience combos). Experience with observability (New Relic, CloudWatch, Grafana, Datadog), AWS and CI/CD, Terraform or AWS CloudFormation, C#/Java/Python, and AppSec tools (Veracode, CloudSploit, Data Theorem).
Infrastructure as Code (IaC), CI/CD, New Relic, CloudWatch, Grafana, Datadog, AWS, Terraform, AWS CloudFormation, C#, Java, Python, Veracode, CloudSploit, Data Theorem
Realtor.comNasdaq: NWSA: Online marketplace for buying, selling, and renting homes.
5+ YOE5+ years SRE/DevOps experience, 3+ years with AWS and Kubernetes, proficiency in Python/Go/Java, IaC (Terraform/CloudFormation), observability tools, CI/CD and on-call/incident response experience.
SecurityScorecard: Provides continuous cybersecurity ratings and risk monitoring for organizations.
6+ YOE6+ years in SRE/DevOps with production Kubernetes, CI/CD pipeline expertise, IaC (Terraform/Helm/Pulumi), Python/Bash/Go proficiency, observability tooling, and experience with Kafka/Flink/ClickHouse and AI/LLM tooling integration.
Brivo: Cloud-native platform for physical security and video surveillance management.
2+ YOE2+ years of SRE/infrastructure experience; strong Linux in production; Kubernetes or similar; Python or Bash (Golang a plus); incident response experience; ability to implement scalable reliability improvements; familiarity with LLM-based tooling for automation.
Senior Site Reliability Engineer, Production Engineer - ThousandEyes
San Francisco or Seattle or Austin or New York City
$165k-$241k/yrHybridFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE5+ years experience; proficiency in Python or Go; expertise with Kubernetes, cloud (AWS), Unix/Linux; strong SRE principles, incident response, and security-minded engineering.
Python, Go, Kubernetes, Service Mesh, Prometheus, OpenTelemetry, ArgoCD, CNCF, AWS, Unix, Linux
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOEBS in CS or equivalent with 5+ years supporting critical services; experience with large-scale HPC clusters (Slurm, LSF, Kubernetes), IaC, CI/CD, multi‑cloud (AWS/GCP/OCI), and 2+ languages such as Python or Go.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
Senior Site Reliability Engineer - Workflow Automation
Austin, Texas, United States
HybridFull Time
Dimensional Fund Advisors: Provides systematic investment solutions based on financial science.
5+ YOE5+ years SRE/DevOps experience, deep Airflow and enterprise scheduler experience, strong Linux/Windows and cloud (AWS) skills, Python and shell scripting, observability and automation focus.
Symphony: Secure collaboration and communication platform for financial institutions
Strong experience with IaC, Kubernetes, Linux, cloud (GCP/AWS), observability, and on-call production support; proficient in Python/Go; solid networking and communication.
Site Reliability Engineer, Apple Data Platform - AI/ML Platform
Austin, Texas, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience operating large multi-cloud data platforms, incident response, supporting internal engineering teams, and running services such as Spark, Flink, Airflow, Ray and notebook/LLM agent platforms.
Zello: A voice-first push-to-talk communication platform for frontline workers.
7+ YOESeasoned SRE with 7+ years in production databases and cloud infrastructure; strong expertise with MySQL and MongoDB; skilled in observability, on-call, and incident response; proficient in Python/Go/bash for automation.