22 aws site reliability engineer jobs at 15 companies in Dallas, TX

2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Idaho or Plano
$117k-$209k/yr RemoteFull Time
Autodesk
AutodeskNASDAQ: ADSK: Developing software for architecture, engineering, and entertainment industries.
7+ YOEU.S. citizen with 7+ years SRE/platform/cloud experience, strong reliability engineering skills (SLOs/SLIs, observability, incident management), cloud experience (AWS/Azure), automation using Python/Go/Java/PowerShell/Bash, and ability to operate production services in regulated GovCloud environments.
AWS, Azure, Python, Go, Java, PowerShell, Bash, Infrastructure as Code, CI/CD, Splunk, Dynatrace, Datadog, CloudWatch, Kubernetes
5d
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Westlake, Texas, United States
OnsiteFull Time
Fidelity Investments
Fidelity Investments: Provides investment management, retirement planning, and brokerage services.
3+ YOEBachelor’s degree and 5 years, or master’s degree and 3 years, in site reliability engineering or related work. Requires Kubernetes, cloud, infrastructure-as-code, monitoring, automation, Python, and distributed systems expertise.
Kubernetes, Power BI, Grafana, Azure ARM, Terraform, Datadog, Splunk, Jenkins, Azure DevOps, Team Foundation Version Control, Cloud Formation Template, Amazon Web Services (AWS), LAMBDA, API Gateway, Fault Injection Service (FIS), Azure Chaos Studio, Python, Windows, Linux
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin or Southlake
$129k-$175k/yr OnsiteFull Time
Charles Schwab
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
10+ YOEBachelor's in CS or related; 10+ years software development/SRE experience (8+ years DevOps/SRE), 8+ years CI/CD and observability, 5+ years leading reliability practices; strong automation, scripting (Python/shell), cloud and distributed systems experience; must be authorized to work in the U.S. without sponsorship.
Python, shell scripting, CI/CD, Splunk, Kubernetes, Terraform, AWS, GCP, Azure
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer - 3 Month Contract
Dallas or Minot
RemoteContract
Orion Health
Orion HealthToronto Stock Exchange: AIDX: Developing population-scale health platforms and healthcare data interoperability software.
4+ YOERequires 4–6 years of site reliability or equivalent experience, systems/application support or development experience, scripting, cloud production support, infrastructure automation, networking, and a technical bachelor's degree or equivalent.
Amazon Web Services (AWS), Windows, Linux, Active Directory, Group Policy Object (GPO), DNS, PowerShell, Python, Bash, Puppet, Ansible, Kubernetes, CloudFormation, Terraform, Splunk, TCP/IP, DHCP, VLANs, VPNs, firewall, Load Balancers, Continuous Integration/Continuous Delivery (CI/CD), Red Hat, Oracle, SQL, HIPAA, HITRUST
2mo
Save
Mark Applied
Hide
Site Reliability Engineer I
Arlington or Irving
HybridFull Time
GM Financial: Provides automotive financing and leasing services for dealers and consumers.
3+ YOE3-5 years cloud DevOps/SRE, Azure/AWS, Kubernetes, Terraform, CI/CD, Linux/Windows, automation, scripting (Python/PowerShell), OpenShift knowledge preferred.
Azure, AWS, GCP, Kubernetes, Terraform, Arm Templates, Powershell, Python, Git, Jenkins, Ansible, CI/CD, Azure CLI
2mo
Save
Mark Applied
Hide
Site Reliability Engineer III
Jersey City or Dallas
$133k-$185k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
3+ YOE3+ years SRE experience, proficiency with SLI/SLO concepts, Python/Java/Spring Boot/.Net or PySpark, observability tools, CI/CD, cloud platforms (AWS), and experience implementing infrastructure-as-code.
Databricks, Snowflake, AWS, Kubernetes, Python, PySpark, Java, Spring Boot, .Net, Grafana, Dynatrace, Prometheus, Datadog, Splunk
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Columbus or Chicago or Charlotte or Indianapolis or Dallas or Pittsburgh or Detroit or Minnetonka or Cincinnati or Akron or Cleveland
$57k-$113k/yr OnsiteFull Time
Huntington
HuntingtonNASDAQ: HBAN: Provides regional commercial, consumer, and mortgage banking services.
5+ YOEExperience supporting mainframe and API ecosystems, troubleshooting production incidents, batch scheduling, API modernization, monitoring and RCA; bachelor’s degree or several years' related experience often required.
z/OS, MVS, CICS, MQ, DB2, COBOL, JCL, TSO/ISPF, Easytrieve, SORT, FileAid, Xpediter, ChangeMan, AbendAid, Zeke, Zena, CA7, Control-M, JSON, XML, REST, SOAP, API Gateway, Postman, SoapUI, Swagger/OpenAPI, Dynatrace, Splunk, Grafana, Azure Monitor, AppDynamics, CloudWatch, ServiceNow, AWS, Azure, GCP, IBM MQ, Kafka, Event Hub, SQL
1mo
Save
Mark Applied
Hide
Compliance Engineering, Site Reliability Engineer SRE, Associate, Dallas
Dallas, Texas, United States
OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Provides investment banking, securities, and wealth management services globally.
3+ YOEMid-level SRE with proficiency in Java, Python or Perl; experience with Linux, SDLC, observability tools (Prometheus, Grafana, ELK, OpenTelemetry) and cloud (AWS/Azure/GCP); strong communication and problem-solving; 3+ years preferred.
Java, Python, Perl, Prometheus, Grafana, ELK, OpenTelemetry, AWS, Azure, GCP, Linux, Hadoop
4d
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Westlake, Texas, United States
OnsiteFull Time
Fidelity Investments
Fidelity Investments: Provides investment management, retirement planning, and brokerage services.
3+ YOEBachelor's degree with 5 years or master's degree with 3 years of experience designing and automating container and cloud-based platforms; expertise with Kubernetes, infrastructure as code, distributed systems, monitoring, DevOps, and Python.
Kubernetes, Power BI, Grafana, Azure ARM, Terraform, Datadog, Splunk, Jenkins, Azure DevOps, Team Foundation Version Control, Cloud Formation Template, Amazon Web Services (AWS), Azure, LAMBDA, API Gateway, Fault Injection Service (FIS), Azure Chaos Studio, Python, Windows, Linux
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer II
Phoenix or Boca Raton or Dallas or Allen or Alpharetta or Atlanta or Buford or Gainesville
$105k-$175k/yr HybridFull Time
LexisNexis Risk Solutions
LexisNexis Risk SolutionsNYSE: RELX: Provides data and analytics for risk management and compliance.
5+ YOE5+ years in SRE/DevOps or Infra Engineering; strong AWS; Terraform; monitoring tools; CI/CD; GitHub workflows; ITSM with ServiceNow; Confluence docs; remote/hybrid capable.
AWS, Terraform, Grafana, Pingdom, Uptrends, Azure DevOps, GitHub, Confluence, Jira, ServiceNow
3d
Save
Mark Applied
Hide
Engineering - SRE Platforms - Site Reliability Engineer - Vice President - Dallas
Dallas, Texas, United States
OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
6+ YOERequires 6+ years in site reliability engineering, programming in Java, Python, or Go, cloud and container expertise, IaC and configuration management skills, Linux and distributed systems knowledge, and advanced monitoring experience.
Java, Python, Go, AWS, GCP, Docker, Kubernetes, Terraform, CloudFormation, Puppet, Chef, Ansible, Prometheus, Grafana, ELK, Datadog, PagerDuty, Jenkins, GitLab, Maven, Elastic Search, GCP Big Query, Kafka, Linux, Infrastructure as Code (IaC), Prompt Engineering, Retrieval-Augmented Generation (RAG), CI/CD
4d
Save
Mark Applied
Hide
IT Site Reliability Engineer — API Management Platforms
Dallas, Texas, United States
OnsiteFull Time
Texas Instruments
Texas InstrumentsNASDAQ: TXN: Designs and manufactures semiconductors and integrated circuits.
3+ YOEBachelor's degree or equivalent practical experience and 3+ years managing Apigee or similar API platforms. Requires platform administration, automation, troubleshooting, monitoring, and incident response expertise.
Apigee, Apigee Edge, Linux, Unix, RHEL, Rocky Linux, Python, Bash, PowerShell, REST API, OAuth 2.0, JWT, Elastic, Elasticsearch, Logstash, Kibana, Splunk, Prometheus, Grafana, TCP/IP, DNS, TLS/SSL, Terraform, Ansible, Chef, Puppet, Jenkins, GitLab CI, GitHub Actions, JFrog Artifactory, Sonatype Nexus, Docker, Kubernetes, AWS, GCP, Azure, Cassandra, ZooKeeper, Qpid, PostgreSQL
2w
Save
Mark Applied
Hide
Site Rel Engineer Sr. - Model Integration Platform
Cleveland or Phoenix or Pittsburgh or Denver or Birmingham or Dallas or Strongsville
$86k-$158k/yr OnsiteFull Time
PNC Financial Services
PNC Financial ServicesNYSE: PNC: Provides banking, lending, and investment services to customers.
3+ YOE3+ years experience with Linux/RHEL, site reliability, monitoring, incident response, storage and scheduler administration; proficiency with shell/bash and Python; bachelor's degree typically required.
Linux (RHEL), IBM GPFS / Spectrum Scale, Slurm Scheduler, Shell/Bash, Python, Spark, Jupyter, IBM Spectrum Conductor, AWS, Azure, Ansible, VMware, KVM, Docker, Kubernetes, Oracle, MySQL, Postgres, Tableau, Power BI, Apache Tomcat
2d
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineering – Paylo Platform
Alpharetta or Temple or Dallas or Houston
HybridFull Time
PDI Technologies
PDI Technologies: Software solutions for convenience retail and petroleum wholesale operations.
8+ YOE4+ MgmtRequires 8+ years in SRE, DevOps, or infrastructure engineering, 4+ years leading people, manager-management experience, and hands-on AWS, Azure, Kubernetes, Helm, Argo, Terraform/OpenTofu, Jenkins, and Datadog expertise.
AWS, Microsoft Azure, Kubernetes, Helm, Argo CD, Argo Workflows, Terraform, OpenTofu, Jenkins, Datadog, Kafka, SQS, SNS, PagerDuty
3w
Save
Mark Applied
Hide
Sr. Systems Engineer, Site Reliability
Dallas, Texas, United States
HybridFull Time
American Heart Association
American Heart Association: Nonprofit organization dedicated to fighting heart disease and stroke.
5+ YOEBachelor's degree or equivalent experience, minimum 5 years relevant experience, strong expertise in multi-cloud operations, IAM, security, automation, and infrastructure as code.
Azure, AWS, GCP, Oracle, Entra ID, Structured Query Language (SQL)
4d
Save
Mark Applied
Hide
Sr Software Engineer (Site Reliability) Austin or Dallas, TX
Austin or Dallas
HybridFull Time
HEB
HEB: A grocery retailer providing food, household products, and related services.
5+ YOERequires 5+ years designing or troubleshooting distributed systems, 3+ years in SRE and Terraform, CI pipeline experience, software architecture expertise, and proficiency with cloud, container, monitoring, and scripting technologies.
Google Kubernetes Engine, Kubernetes, AWS, Java, Spring, Terraform, GitLab Pipelines, GitHub Actions, GitLab, JIRA, Slack, Confluence, IntelliJ, PostgreSQL, Docker, Linux, Google Cloud Platform (GCP), REST, GraphQL, Datadog, Grafana, New Relic, Python, Ruby, Groovy, Bash
4d
Save
Mark Applied
Hide
Sr Software Engineer (Site Reliability) Austin or Dallas, TX
Austin or Dallas
HybridFull Time
H-E-B
H-E-B: Operates a major supermarket chain in Texas and Mexi
5+ YOERequires 5+ years with distributed systems, 3+ years in SRE, Terraform, and CI pipelines, plus experience with cloud, Kubernetes, Docker, Linux, databases, monitoring, APIs, and scripting languages.
Google Kubernetes Engine, K8s, AWS, Java, Spring, Terraform, Gitlab Pipelines, GitHub Actions, Gitlab, JIRA, Slack, Confluence, Intellij, microservices, PostgreSQL, Kubernetes, Docker, Linux, GCP, REST, GraphQL, Datadog, Grafana, New Relic, Python, Ruby, Groovy, Bash, Agile
4d
Save
Mark Applied
Hide
Full Stack Java/React Technical Lead - Vice President
Irving, Texas, United States
$126k-$189k/yr HybridFull Time
Citi
CitiNYSE: C: Providing global banking, investment, and wealth management services.
6+ YOE6+ years of full-stack software engineering experience; Java, Spring Boot, microservices, React or Angular, databases, generative AI, cloud, APIs, security, and site reliability expertise; bachelor's degree or equivalent experience.
Java, Spring Boot, Microservices, Angular, React, Oracle, PostgreSQL, Python, TensorFlow, PyTorch, JAX, Google ADK, Google AI/ML platforms, SDKs, RESTful APIs, Kafka, Elastic Search, NoSQL, AWS, DevOps, SRE