169 site reliability engineer jobs at 80 companies in Texas

1w
Save
Mark Applied
Hide
Site Reliability Engineer
Austin or Reston
HybridFull Time
Seekr Technologies
Seekr Technologies: Private American enterprise AI providing explainable, secure AI software and hardware to government and critical-infrastructure customers.
5+ YOERequires 5+ years in site reliability engineering and Linux systems, monitoring and logging expertise, programming or scripting skills, Docker, Kubernetes, configuration automation, and incident response experience.
Linux, ELK, Prometheus, InfluxDB, Grafana, Python, Ruby, Bash, Java, Docker, Kubernetes, Puppet, Chef, Ansible, Terraform, Elasticsearch, Kafka, Aerospike, Git, GitHub, GitLab, ArgoCD
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Customer engagement platform for cross-channel marketing and analytics.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
San Francisco or Alpharetta or Arlington or Augusta or Ashburn or Allentown or Appleton or Atlanta or Annapolis Junction or Ann Arbor or Herndon or Allen
$165k-$241k/yr RemoteFull Time
Cisco
CiscoNASDAQ: CSCO: Global leader in networking, cybersecurity, and cloud-native technology solutions.
7+ YOE7+ years SRE or related experience; BS/MS/PhD with corresponding years; U.S. Person required for FedRAMP/IL-5 work; on-call participation; strong coding, automation, reliability, and security skills.
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Global financial services and banking institution.
4+ YOE4+ years cloud/platform engineering experience with GCP exposure, Terraform/IaC, CI/CD, observability, DevSecOps practices, incident response, and automation for reliability and resiliency.
GCP, Azure, Terraform, Terraform Enterprise, Log Analytics, Dynatrace, Resource Graph, CI/CD, IAM
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Idaho or Plano
$117k-$209k/yr RemoteFull Time
Autodesk
AutodeskNASDAQ: ADSK: Global provider of software for design, engineering, and manufacturing.
7+ YOEU.S. citizen with 7+ years SRE/platform/cloud experience, strong reliability engineering skills (SLOs/SLIs, observability, incident management), cloud experience (AWS/Azure), automation using Python/Go/Java/PowerShell/Bash, and ability to operate production services in regulated GovCloud environments.
AWS, Azure, Python, Go, Java, PowerShell, Bash, Infrastructure as Code, CI/CD, Splunk, Dynatrace, Datadog, CloudWatch, Kubernetes
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Houston, Texas, United States
HybridFull Time
NOV
NOVNYSE: NOV: Provider of technology and equipment for the energy industry.
5+ YOE5+ years SRE/DevOps experience; expertise in Kubernetes, AKKA.NET, cloud platforms (AWS/Azure/GCP), scripting (Bash/PowerShell/Python), observability stacks, and PostgreSQL tuning; proven incident management skills.
Prometheus, Grafana, Datadog, OpenTelemetry, ELK, Phobos, AKKA.NET, PostgreSQL, GitHub Actions, Azure Pipelines, GitLab CI, Bash, PowerShell, Python, AWS, Azure, GCP, Kubernetes, Docker, Terraform, GitHub, GitLab, Azure DevOps, C#
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Austin, Texas, United States
$140k-$200k/yr OnsiteFull Time
Future Secure AI
Future Secure AI: Private enterprise AI building and operating bespoke AI-Workers that automate complex, high-stakes workflows.
5+ YOEHands-on Kubernetes, Terraform, and Helm experience; programming in Python/Go/Java/Bash/PowerShell/Ruby; SRE experience with on-call, incident response, SLIs/SLOs; cloud and CI/CD experience; 5+ years preferred.
Kubernetes, EKS, AKS, GKE, Terraform, Helm, SLIs, SLOs, SLAs, Python, Go, Java, Bash, PowerShell, Ruby, ArgoCD, CI/CD, GitOps, AWS, Azure, Google Cloud
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Global financial services and banking institution.
4+ YOE4+ years cloud/platform engineering experience with Terraform and GCP; strong IaC, observability, automation, incident response, and DevSecOps skills; ability to define SLIs/SLOs and mentor engineers.
Terraform, Terraform Enterprise, Google Cloud Platform (GCP), Azure, Log Analytics, Dynatrace, Resource Graph, CI/CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Reston or Austin
$81k-$187k/yr OnsiteFull Time
Oracle Corporation
Oracle CorporationNYSE: ORCL: Cloud infrastructure and enterprise software solutions provider.
3+ YOEProvide SRE escalation, automate tasks, manage complex change requests, mentor SREs, and support large-scale production reliability.
Linux, Unix, Docker, Kubernetes, Terraform, Bash, Perl, Python, Ruby, JavaScript, Java, Chef, Puppet
2mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Plano, Texas, United States
OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services and investment banking firm.
5+ YOE5+ years applied SRE experience, formal SRE training/certification, deep knowledge of reliability/scalability/security, observability tooling experience, Python/Ansible fluency, SDLC experience, and strong communication/mentoring skills.
Grafana, Dynatrace, Prometheus, CloudWatch, Splunk, Python, Ansible
2w
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Westlake, Texas, United States
OnsiteFull Time
Fidelity Investments
Fidelity Investments: Provider of investment, retirement, and financial planning services.
3+ YOEBachelor’s degree and 5 years, or master’s degree and 3 years, in site reliability engineering or related work. Requires Kubernetes, cloud, infrastructure-as-code, monitoring, automation, Python, and distributed systems expertise.
Kubernetes, Power BI, Grafana, Azure ARM, Terraform, Datadog, Splunk, Jenkins, Azure DevOps, Team Foundation Version Control, Cloud Formation Template, Amazon Web Services (AWS), LAMBDA, API Gateway, Fault Injection Service (FIS), Azure Chaos Studio, Python, Windows, Linux
2mo
Save
Mark Applied
Hide
Sr Site Reliability Engineer
Austin, Texas, United States
HybridFull Time
Move, Inc.
Move, Inc.: Online real estate services connecting buyers, sellers, renters, and agents through listings, tools, and software.
5+ YOE5+ years SRE/DevOps experience, 3+ years with AWS and Kubernetes, proficiency in Python/Go/Java, IaC (Terraform/CloudFormation), observability tools, CI/CD and on-call/incident response experience.
AWS, EKS, Fargate, ECS, EC2, RDS, S3, CloudWatch, IAM, VPC, Route53, CloudFront, Lambda, Kubernetes, Docker, Istio, Argo CD, CircleCI, Jenkins, GitHub Actions, New Relic, Datadog, Prometheus, Grafana, Splunk, Terraform, CloudFormation, Helm, Kustomize, Python, Go, Java, Bash, Tyk, Kong, Apollo GraphQL, AWS Secrets Manager, Vault, OpsGenie, PagerDuty, ServiceNow, Skyway, Frontdoor, Pantheon
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin or New York City
$152k-$195k/yr HybridFull Time
SecurityScorecard
SecurityScorecard: Private cybersecurity ratings and third-party risk platform serving organizations managing supply-chain risk.
6+ YOE6+ years in SRE/DevOps with production Kubernetes, CI/CD pipeline expertise, IaC (Terraform/Helm/Pulumi), Python/Bash/Go proficiency, observability tooling, and experience with Kafka/Flink/ClickHouse and AI/LLM tooling integration.
Kubernetes, MCP servers, CI/CD, GitHub Actions, Jenkins, GitLab CI, EKS, GKE, AKS, Terraform, Helm, Argo CD, Pulumi, GitOps, Python, Bash, Go, Prometheus, Grafana, Datadog, OpenTelemetry, Kafka, Flink, ClickHouse, Langsmith, Langfuse
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Compute
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Building and operating civilization-scale data center infrastructure for AI.
Experience owning large GPU/compute fleets, automation of repair/deployment pipelines, firmware/BMC/Redfish familiarity, incident response and paging, observability and metrics tooling, and proficiency with production automation.
Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, Kubernetes, Claude Code, Cursor, LLM APIs, MCP servers
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Copenhagen or Dallas or Dubai or London or Miami or Pune or Buenos Aires or Bogota or Mexico or Singapore
RemoteFull Time
CellPoint Digital
CellPoint Digital: Private fintech providing payment orchestration and digital commerce platforms for airlines, travel businesses, hospitality, and merchants.
6+ YOE6+ years SRE/DevOps experience, deep GCP and Kubernetes knowledge, Terraform expertise, incident leadership, automation, security collaboration, strong communication.
GCP, Kubernetes (GKE), CloudSQL, Spanner, Terraform, IAM
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin, Texas, United States
$111k-$172k/yr HybridFull Time
Visa
VisaNYSE: V: Global leader in digital payments and transaction technology.
2+ YOE2+ years with a Bachelor's or 5+ years experience; hands-on Azure, Kubernetes, Terraform, IaC/GitOps, CI/CD, observability, service mesh (Istio preferred); strong SRE, troubleshooting, documentation, and English (B2+).
Azure, AWS, Kubernetes, Terraform, GitOps, CI/CD, Istio, App Mesh, Linkerd, Service Mesh
2mo
Save
Mark Applied
Hide
LEAD SITE RELIABILITY ENGINEER
Austin, Texas, United States
$167k-$204k/yr HybridFull Time
Cox Automotive
Cox Automotive: Privately held automotive services and software serving dealers, fleets, lenders, automakers, and car shoppers.
6+ YOEBachelor's in CS or related and 6 years experience (or alternate degree/experience combos). Experience with observability (New Relic, CloudWatch, Grafana, Datadog), AWS and CI/CD, Terraform or AWS CloudFormation, C#/Java/Python, and AppSec tools (Veracode, CloudSploit, Data Theorem).
Infrastructure as Code (IaC), CI/CD, New Relic, CloudWatch, Grafana, Datadog, AWS, Terraform, AWS CloudFormation, C#, Java, Python, Veracode, CloudSploit, Data Theorem
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Dallas or Charlotte or Wayne or Plano
HybridFull Time
Vanguard
Vanguard: Global investment management firm owned by its client funds.
Experience with observability, monitoring, reliability metrics, alerting, automation, resilience engineering, incident response, and production troubleshooting; Python-based automation and chaos engineering experience are mentioned.
Splunk, Honeycomb, Amazon CloudWatch, Dynatrace, AppDynamics, Python, Blue Prism, UiPath
3w
Save
Mark Applied
Hide
Site Reliability Engineer
Austin, Texas, United States
HybridFull Time
Thales
ThalesEuronext Paris: HO: Global technology leader in aerospace, defense, and security.
5+ YOEEngineer or equivalent with at least 5 years of experience, Java development, public cloud, containers, microservices, CI/CD, automation, monitoring, and observability. U.S. or dual citizenship required.
Terraform, Ansible, Kubernetes, GitLab, Datadog, Java, GCP, AWS, Docker, Jenkins, Helm, NoSQL
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Dallas, Texas, United States
HybridFull Time
Longbridge Group
Longbridge Group: Singapore-headquartered fintech group operating online brokerage, institutional trading technology, and AI financial infrastructure for investors and institutions.
5+ YOE5+ years in SRE, DevOps, or production engineering; AWS/GCP/Azure, Docker, Kubernetes, Linux, CI/CD, incident management, distributed systems, and programming experience required.
Terraform, Ansible, Helm, Kubernetes, Prometheus, AWS, GCP, Docker, Python, Go, Linux, CI/CD