51 site reliability engineer jobs at 26 companies in Manor, TX

2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Austin, Texas, United States
$140k-$200k/yr OnsiteFull Time
Future Secure AI
Future Secure AI: Deploys secure, persona-based AI-Workers for large enterprises.
5+ YOEHands-on Kubernetes, Terraform, and Helm experience; programming in Python/Go/Java/Bash/PowerShell/Ruby; SRE experience with on-call, incident response, SLIs/SLOs; cloud and CI/CD experience; 5+ years preferred.
Kubernetes, EKS, AKS, GKE, Terraform, Helm, SLIs, SLOs, SLAs, Python, Go, Java, Bash, PowerShell, Ruby, ArgoCD, CI/CD, GitOps, AWS, Azure, Google Cloud
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Reston or Austin
$81k-$187k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
3+ YOEProvide SRE escalation, automate tasks, manage complex change requests, mentor SREs, and support large-scale production reliability.
Linux, Unix, Docker, Kubernetes, Terraform, Bash, Perl, Python, Ruby, JavaScript, Java, Chef, Puppet
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin, Texas, United States
HybridFull Time
2K
2KNASDAQ: TTWO: Publishes and develops global video game franchises and entertainment.
5+ YOE5+ years SRE/platform engineering experience, deep Kubernetes (EKS/GKE), Terraform/Pulumi and GitOps, observability with Prometheus/Grafana/Datadog, production coding in Go/Python/TypeScript, Linux and networking expertise, incident management.
Terraform, Pulumi, ArgoCD, Flux, Kubernetes, EKS, GKE, Istio, Cilium, Helm, Terragrunt, Prometheus, Grafana, Datadog, OpenTelemetry, GitHub Actions, Jenkins, Go, Python, TypeScript, PasswordState, 1Password, AWS Secrets Manager, OPA/Gatekeeper, AWS, GCP, VMware, Ansible, Puppet, AWS Systems Manager
1mo
Save
Mark Applied
Hide
LEAD SITE RELIABILITY ENGINEER
Austin, Texas, United States
$167k-$204k/yr HybridFull Time
Cox Enterprises
Cox Enterprises: Providing global communications, automotive services, and media solutions.
6+ YOEBachelor's in CS or related and 6 years experience (or alternate degree/experience combos). Experience with observability (New Relic, CloudWatch, Grafana, Datadog), AWS and CI/CD, Terraform or AWS CloudFormation, C#/Java/Python, and AppSec tools (Veracode, CloudSploit, Data Theorem).
Infrastructure as Code (IaC), CI/CD, New Relic, CloudWatch, Grafana, Datadog, AWS, Terraform, AWS CloudFormation, C#, Java, Python, Veracode, CloudSploit, Data Theorem
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin or New York City
$152k-$195k/yr HybridFull Time
SecurityScorecard
SecurityScorecard: Provides continuous cybersecurity ratings and risk monitoring for organizations.
6+ YOE6+ years in SRE/DevOps with production Kubernetes, CI/CD pipeline expertise, IaC (Terraform/Helm/Pulumi), Python/Bash/Go proficiency, observability tooling, and experience with Kafka/Flink/ClickHouse and AI/LLM tooling integration.
Kubernetes, MCP servers, CI/CD, GitHub Actions, Jenkins, GitLab CI, EKS, GKE, AKS, Terraform, Helm, Argo CD, Pulumi, GitOps, Python, Bash, Go, Prometheus, Grafana, Datadog, OpenTelemetry, Kafka, Flink, ClickHouse, Langsmith, Langfuse
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin, Texas, United States
$111k-$172k/yr HybridFull Time
Visa
VisaNYSE: V: Global payment technology facilitating electronic funds transfers.
2+ YOE2+ years with a Bachelor's or 5+ years experience; hands-on Azure, Kubernetes, Terraform, IaC/GitOps, CI/CD, observability, service mesh (Istio preferred); strong SRE, troubleshooting, documentation, and English (B2+).
Azure, AWS, Kubernetes, Terraform, GitOps, CI/CD, Istio, App Mesh, Linkerd, Service Mesh
1mo
Save
Mark Applied
Hide
Sr Site Reliability Engineer
Austin, Texas, United States
HybridFull Time
Realtor.com
Realtor.comNasdaq: NWSA: Online marketplace for buying, selling, and renting homes.
5+ YOE5+ years SRE/DevOps experience, 3+ years with AWS and Kubernetes, proficiency in Python/Go/Java, IaC (Terraform/CloudFormation), observability tools, CI/CD and on-call/incident response experience.
AWS, EKS, Fargate, ECS, EC2, RDS, S3, CloudWatch, IAM, VPC, Route53, CloudFront, Lambda, Kubernetes, Docker, Istio, Argo CD, CircleCI, Jenkins, GitHub Actions, New Relic, Datadog, Prometheus, Grafana, Splunk, Terraform, CloudFormation, Helm, Kustomize, Python, Go, Java, Bash, Tyk, Kong, Apollo GraphQL, AWS Secrets Manager, Vault, OpsGenie, PagerDuty, ServiceNow, Skyway, Frontdoor, Pantheon
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Compute
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience owning large GPU/compute fleets, automation of repair/deployment pipelines, firmware/BMC/Redfish familiarity, incident response and paging, observability and metrics tooling, and proficiency with production automation.
Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, Kubernetes, Claude Code, Cursor, LLM APIs, MCP servers
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Austin, Texas, United States
OnsiteFull Time
Brivo
Brivo: Cloud-native platform for physical security and video surveillance management.
2+ YOE2+ years of SRE/infrastructure experience; strong Linux in production; Kubernetes or similar; Python or Bash (Golang a plus); incident response experience; ability to implement scalable reliability improvements; familiarity with LLM-based tooling for automation.
Kubernetes, Linux, Python, Bash, Golang, Prometheus, Grafana
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Production Engineer - ThousandEyes
San Francisco or Seattle or Austin or New York City
$165k-$241k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE5+ years experience; proficiency in Python or Go; expertise with Kubernetes, cloud (AWS), Unix/Linux; strong SRE principles, incident response, and security-minded engineering.
Python, Go, Kubernetes, Service Mesh, Prometheus, OpenTelemetry, ArgoCD, CNCF, AWS, Unix, Linux
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer ( SRE)
Austin, Texas, United States
$130k-$175k/yr OnsiteFull Time
Charles Schwab
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
8+ YOE8+ years supporting enterprise-scale applications; strong automation, SDLC, Linux/Windows, cloud, networking, monitoring, programming (Python/Java/PowerShell/Bash/.NET), databases, messaging, observability, and AI/ML ops; bachelor’s degree.
Python, Java, PowerShell, Bash, .NET, SQL Server, Oracle, MongoDB, Kafka, RabbitMQ, IBM MQ, Solace, Splunk, AppDynamics, Jenkins, Harness, GitHub Actions, Kubernetes, OpenShift, Google Cloud Platform (GCP), AWS, Microsoft Azure
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - HPC
Santa Clara or Durham or Austin
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOEBS in CS or equivalent with 5+ years supporting critical services; experience with large-scale HPC clusters (Slurm, LSF, Kubernetes), IaC, CI/CD, multi‑cloud (AWS/GCP/OCI), and 2+ languages such as Python or Go.
Slurm, LSF, Kubernetes, AWS, GCP, OCI, Infrastructure as Code (IaC), CI/CD, Python, Go, Perl, Ruby, AIOps
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - HPC
Santa Clara or Austin or Durham
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years building/supporting critical services; experience with large-scale HPC clusters (Slurm, LSF, Kubernetes); IaC and CI/CD proficiency; coding in Python/Go/Perl/Ruby; monitoring, capacity planning, and incident response skills.
Slurm, LSF, Kubernetes, Infrastructure as Code (IaC), CI/CD, AWS, GCP, OCI, Python, Go, Perl, Ruby
1mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer - Austin, Texas
Austin, Texas, United States
HybridFull Time
ShipperHQ
ShipperHQ: Provides shipping rate management and checkout optimization software for e-commerce.
10+ YOE10+ years SRE/platform/devops experience; proven AWS cloud experience; expert Terraform and GitLab CI/CD; Kubernetes and container expertise; strong software engineering, observability, SLO/SLI, networking, Linux, and mentoring skills.
AWS, Terraform, GitLab, Kubernetes, Linux
1d
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin, Texas, United States
HybridFull Time
TeamViewer
TeamViewerFrankfurt Stock Exchange: TMV: Provides remote connectivity and digital workplace software solutions.
5+ YOEDegree in computer science, software engineering, IT, or equivalent experience; 5+ years in SRE, DevOps, or software development; Azure, IaC, automation, containers, monitoring, databases, security, and scripting expertise.
Microsoft Azure, Kubernetes, GitOps, Terraform, Argo CD, PowerShell, Docker, MS SQL, Postgres, Datadog, Grafana, Prometheus
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
3d
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Workflow Automation
Austin, Texas, United States
HybridFull Time
Dimensional Fund Advisors: Provides systematic investment solutions based on financial science.
5+ YOE5+ years SRE/DevOps experience, deep Airflow and enterprise scheduler experience, strong Linux/Windows and cloud (AWS) skills, Python and shell scripting, observability and automation focus.
Apache Airflow, Automic/UC4, Control-M, Linux, Windows, AWS, Python, shell scripting, Kubernetes, Docker, Terraform, Helm, Ansible, ELK, Grafana, Prometheus, dbt, Kafka, Snowflake, AWS MWAA, Cloud Composer, Astronomer, CI/CD
4d
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Austin or Atlanta
$100k-$115k/yr OnsiteFull Time
Atlanticus
AtlanticusNASDAQ: ATLC: Provides credit cards and lending solutions for underserved consumers.
5+ YOERequires 5+ years supporting production applications, Java, AWS, Kubernetes, Docker, Datadog or Splunk, CI/CD, Python or Bash, Linux, cloud troubleshooting, and incident management experience.
AWS, Amazon EKS, Amazon EC2, ALB/NLB, Amazon RDS, IAM, Amazon Route 53, Amazon CloudWatch, Amazon S3, VPC, Datadog, Splunk, Docker, Kubernetes, Jenkins, GitHub Actions, Argo CD, MySQL, Oracle, Python, Bash, Linux, Helm, Terraform, Prometheus, Grafana, OpenTelemetry, Karpenter, Cluster Autoscaler, Java, JVM
4d
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Atlanta or Austin
$100k-$115k/yr OnsiteFull Time
Atlanticus
AtlanticusNASDAQ: ATLC: Provide credit products and financial services to underserved consumers.
5+ YOERequires 5+ years supporting production applications, Java, AWS, Kubernetes, Docker, Datadog, Splunk, CI/CD, databases, Python or Bash, Linux, networking, and incident management experience.
Amazon Web Services (AWS), Amazon EKS, Amazon EC2, ALB, NLB, Amazon RDS, IAM, Amazon Route 53, Amazon CloudWatch, Amazon S3, Amazon VPC, Java, Datadog, Splunk, Docker, Kubernetes, Jenkins, GitHub Actions, Argo CD, MySQL, Oracle, Python, Bash, Linux, Helm, Terraform, Prometheus, Grafana, OpenTelemetry, Karpenter, Cluster Autoscaler, AI-assisted development tools, agentic AI systems