35 cloud reliability engineer jobs at 20 companies in San Marcos, TX
1mo
Save
Mark Applied
Hide
1mo
Site Reliability Engineer
Austin, Texas, United States
$140k-$200k/yrOnsiteFull Time
Future Secure AI: Deploys secure, persona-based AI-Workers for large enterprises.
5+ YOEHands-on Kubernetes, Terraform, and Helm experience; programming in Python/Go/Java/Bash/PowerShell/Ruby; SRE experience with on-call, incident response, SLIs/SLOs; cloud and CI/CD experience; 5+ years preferred.
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
3+ YOEUS citizenship and TS/SCI w/Poly required; 3+ years experience in operations/engineering; proficiency with Linux, scripting, cloud, change management, and monitoring; bachelor’s or equivalent preferred.
Senior Site Reliability Engineer, Production Engineer - ThousandEyes
San Francisco or Seattle or Austin or New York City
$165k-$241k/yrHybridFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE5+ years experience; proficiency in Python or Go; expertise with Kubernetes, cloud (AWS), Unix/Linux; strong SRE principles, incident response, and security-minded engineering.
Python, Go, Kubernetes, Service Mesh, Prometheus, OpenTelemetry, ArgoCD, CNCF, AWS, Unix, Linux
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yrRemoteFull Time
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
ArmNASDAQ: ARM: Designs and licenses processor architectures and semiconductor intellectual property.
Build and operate a secure OpenStack private cloud, automate infrastructure with IaC, develop platform tooling, debug compute/network/storage, and contribute to reliability and CI/CD pipelines.
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
8+ YOE8+ years software engineering experience with 4+ years SRE, experience with cloud-native containers, IaC, CI/CD, observability, on-call rotations, and deploying/operating LLM-powered applications.
Terraform, Google Cloud Platform, Gemini, Claude, OpenAI
Site Reliability Engineer, Apple Data Platform / Multi-Cloud Infrastructure
Austin, Texas, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Manage and operate a massive multi-cloud data platform, run incident response, provide hands-on support to internal teams, and partner with developers to keep services reliable across AWS, GCP, and on-prem Kubernetes.
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOEBS in CS or equivalent with 5+ years supporting critical services; experience with large-scale HPC clusters (Slurm, LSF, Kubernetes), IaC, CI/CD, multi‑cloud (AWS/GCP/OCI), and 2+ languages such as Python or Go.
Senior Site Reliability Engineer - Workflow Automation
Austin, Texas, United States
HybridFull Time
Dimensional Fund Advisors: Provides systematic investment solutions based on financial science.
5+ YOE5+ years SRE/DevOps experience, deep Airflow and enterprise scheduler experience, strong Linux/Windows and cloud (AWS) skills, Python and shell scripting, observability and automation focus.
Guidehouse: Provides management and technology consulting services to diverse organizations.
4+ YOEBA/BS or equivalent experience, 4+ years IT/admin/software/platform experience with AWS, 1+ years cloud deployment experience, proficiency with CI/CD and IaC tools (Terraform, Ansible, GitLab, Artifactory, Packer), scripting (Python, PowerShell, Bash), Windows/Linux, Agile, and ability to obtain Public Trust.
Zello: A voice-first push-to-talk communication platform for frontline workers.
7+ YOESeasoned SRE with 7+ years in production databases and cloud infrastructure; strong expertise with MySQL and MongoDB; skilled in observability, on-call, and incident response; proficient in Python/Go/bash for automation.
Symphony: Secure collaboration and communication platform for financial institutions
Strong experience with IaC, Kubernetes, Linux, cloud (GCP/AWS), observability, and on-call production support; proficient in Python/Go; solid networking and communication.
AtlanticusNASDAQ: ATLC: Provides credit cards and lending solutions for underserved consumers.
5+ YOERequires 5+ years supporting production applications, Java, AWS, Kubernetes, Docker, Datadog or Splunk, CI/CD, Python or Bash, Linux, cloud troubleshooting, and incident management experience.
Senior AI Ops & Incident/Site Reliability Engineer
Austin or Plano or Fort Mill or Boston or New York City or Tempe or San Diego
$94k-$171k/yrHybridFull Time
Perficient: Provides digital transformation and AI consulting for global enterprises.
8+ YOEBachelor's degree or equivalent,8+ years in IT operations/SRE or production support,5+ years leading incident management,experience with Dynatrace,ServiceNow,cloud platforms and AIOps implementations.
DISCONYSE: LAW: Provides AI-powered software for ediscovery and legal document review.
15+ YOE15+ years experience building data-intensive distributed systems; expertise with big data technologies, DDD, reliability, CI/CD, cloud providers, APIs, and security-minded design.