38 reliability automation engineer jobs at 24 companies in Addison, TX

1w
Save
Mark Applied
Hide
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$315k/yr HybridFull Time
Catalyst Brands
Catalyst Brands: Operates a portfolio of diverse retail clothing and apparel brands.
12+ YOE5+ Mgmt12+ years progressive engineering leadership with 5+ years leading reliability/performance/automation; deep SRE, performance, and automation experience; cloud and observability tool expertise; ability to lead distributed engineering teams.
AWS, Azure, Kubernetes, EKS, Jenkins, Kafka, Dynatrace, Splunk, Datadog, New Relic, Catchpoint, CloudWatch, ELK
1w
Save
Mark Applied
Hide
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$225k/yr HybridFull Time
Catalyst Brands
Catalyst Brands: Operates a portfolio of retail brands including JCPenney.
12+ YOE5+ Mgmt12+ years engineering leadership with 5+ years leading reliability, performance or automation functions; deep SRE, performance engineering, automation, cloud and observability experience; BA/BS preferred.
AWS, Azure, Kubernetes/EKS, Jenkins, Kafka, Dynatrace, Splunk, Datadog, New Relic, Catchpoint, CloudWatch, ELK
1w
Save
Mark Applied
Hide
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$225k/yr HybridFull Time
JCPenney
JCPenney: Retailer of apparel, home goods, and beauty products.
12+ YOE5+ Mgmt12+ years progressive engineering leadership with 5+ years leading reliability/performance/automation; deep SRE, performance engineering, automation, observability and cloud experience; strong architecture and incident management skills.
AWS, Azure, Kubernetes/EKS, Jenkins, Kafka, CDN, Dynatrace, Splunk, Datadog, New Relic, Catchpoint, CloudWatch, ELK
3w
Save
Mark Applied
Hide
Reliability Engineer
Westlake, Texas, United States
OnsiteFull Time
Fidelity Investments
Fidelity Investments: Financial services and investment management firm providing advisory solutions.
5+ YOEBachelor's or equivalent,5+ years deploying/supporting distributed systems,cloud and on-prem storage,Kubernetes (EKS/AKS/RKS),CI/CD automation,observability,backup/recovery,Python/NodeJS/Java and scripting.
Go, Angular, Python, JavaScript, AWS, RESTful services, Ruby, MVC, Jenkins CI/CD, Chef, Ansible, Bootstrap, HTML/CSS, Shell Scripting, MQ, OpenStack, PostgreSQL, PowerBI, Tableau, NodeJS, Java, Docker, Docker Compose, Git, Datadog, Splunk, Prometheus, Grafana, ELK/OpenSearch, OpenTelemetry, IAM, ARM, Terraform
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with GCP exposure, Terraform/IaC, CI/CD, observability, DevSecOps practices, incident response, and automation for reliability and resiliency.
GCP, Azure, Terraform, Terraform Enterprise, Log Analytics, Dynatrace, Resource Graph, CI/CD, IAM
1d
Save
Mark Applied
Hide
Engineer, Site Reliability
Atlanta or Frisco
$85k-$153k/yr OnsiteFull Time
T-Mobile
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
3+ YOEBachelor's degree and 3 years of related experience, or advanced degree and 1 year; automation, monitoring, CI/CD, cloud computing, scripting, incident management, and system reliability skills required.
CI/CD, Kubernetes, AWS
1d
Save
Mark Applied
Hide
Engineer, Site Reliability
Atlanta or Frisco
$85k-$153k/yr OnsiteFull Time
T-Mobile
T-MobileNASDAQ: TMUS: Provides wireless network services and mobile communication devices.
3+ YOEBachelor's degree and 3 years of related experience, or advanced degree and 1 year. Requires automation, monitoring, scripting, cloud computing, CI/CD, incident management, and system reliability skills.
CI/CD, Kubernetes, AWS
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides banking, investment, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with Terraform and GCP; strong IaC, observability, automation, incident response, and DevSecOps skills; ability to define SLIs/SLOs and mentor engineers.
Terraform, Terraform Enterprise, Google Cloud Platform (GCP), Azure, Log Analytics, Dynatrace, Resource Graph, CI/CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Idaho or Plano
$117k-$209k/yr RemoteFull Time
Autodesk
AutodeskNASDAQ: ADSK: Developing software for architecture, engineering, and entertainment industries.
7+ YOEU.S. citizen with 7+ years SRE/platform/cloud experience, strong reliability engineering skills (SLOs/SLIs, observability, incident management), cloud experience (AWS/Azure), automation using Python/Go/Java/PowerShell/Bash, and ability to operate production services in regulated GovCloud environments.
AWS, Azure, Python, Go, Java, PowerShell, Bash, Infrastructure as Code, CI/CD, Splunk, Dynatrace, Datadog, CloudWatch, Kubernetes
1w
Save
Mark Applied
Hide
Platform Reliability Engineer - Principal Engineer
Iselin or Irving or Charlotte
$159k-$305k/yr HybridFull Time
Wells Fargo
Wells FargoNYSE: WFC: Provides banking, investment, mortgage, and consumer finance products.
7+ YOE7+ years engineering experience, 5+ years supporting enterprise production environments, hands-on in one infrastructure domain, SRE practice experience, strong troubleshooting and automation skills.
Grafana, Splunk, Prometheus, AppDynamics, Cribl, ThousandEyes, Dynatrace, Python, Bash, PowerShell, Git, Ansible, Terraform, CICD
1mo
Save
Mark Applied
Hide
Senior Lead Site Reliability Engineer
Jersey City or Plano
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied SRE experience, formal SRE training/certification, experience with AI/ML platform reliability, SLO/SLI design, observability, automation, and mentoring peers.
Databricks, GPU clusters, Grafana, Dynatrace, Prometheus, Datadog, Splunk, AWS, Snowflake, Kubernetes (EKS), Apache Kafka, Apache Spark, Feature Stores, Vector Databases, LLM
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Copenhagen or Dallas or Dubai or London or Miami or Pune or Buenos Aires or Bogota or Mexico or Singapore
RemoteFull Time
CellPoint Digital
CellPoint Digital: Payment orchestration platform for the travel and hospitality industry.
6+ YOE6+ years SRE/DevOps experience, deep GCP and Kubernetes knowledge, Terraform expertise, incident leadership, automation, security collaboration, strong communication.
GCP, Kubernetes (GKE), CloudSQL, Spanner, Terraform, IAM
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
Denver or Dallas
HybridFull Time
Analytic Partners
Analytic Partners: Provides commercial analytics software and marketing measurement solutions.
4+ YOE4+ years in Platform Engineering/DevOps or related; strong Linux/Windows; automation with Python, Bash, or PowerShell; deep AWS and Azure experience; CI/CD; Infrastructure as Code; containers.
Linux, Windows, Python, Bash, PowerShell, AWS, Azure, Jenkins, GitHub Actions, Terraform, CloudFormation, Arm, Docker, Kubernetes, Nomad, Consul, Vault, Splunk, Sumo Logic
1d
Save
Mark Applied
Hide
Senior Site Reliability Engineer - 3 Month Contract
Dallas or Minot
RemoteContract
Orion Health
Orion HealthToronto Stock Exchange: AIDX: Developing population-scale health platforms and healthcare data interoperability software.
4+ YOERequires 4–6 years of site reliability or equivalent experience, systems/application support or development experience, scripting, cloud production support, infrastructure automation, networking, and a technical bachelor's degree or equivalent.
Amazon Web Services (AWS), Windows, Linux, Active Directory, Group Policy Object (GPO), DNS, PowerShell, Python, Bash, Puppet, Ansible, Kubernetes, CloudFormation, Terraform, Splunk, TCP/IP, DHCP, VLANs, VPNs, firewall, Load Balancers, Continuous Integration/Continuous Delivery (CI/CD), Red Hat, Oracle, SQL, HIPAA, HITRUST
1w
Save
Mark Applied
Hide
Principal Network Engineer - Reliability
Irving or Chandler or Charlotte
HybridFull Time
Wells Fargo
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
7+ YOE7+ years network engineering or reliability experience supporting enterprise network services; deep routing/switching and incident leadership; SRE/reliability practices; hands-on automation/scripting; strong communication and stakeholder influence.
Python, PowerShell, Bash, Ansible, Terraform
2mo
Save
Mark Applied
Hide
Site Reliability Engineer I
Arlington or Irving
HybridFull Time
GM Financial: Provides automotive financing and leasing services for dealers and consumers.
3+ YOE3-5 years cloud DevOps/SRE, Azure/AWS, Kubernetes, Terraform, CI/CD, Linux/Windows, automation, scripting (Python/PowerShell), OpenShift knowledge preferred.
Azure, AWS, GCP, Kubernetes, Terraform, Arm Templates, Powershell, Python, Git, Jenkins, Ansible, CI/CD, Azure CLI
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (Westlake, TX, US)
Westlake, Texas, United States
OnsiteFull Time
NTT DATA
NTT DATA: Global provider of IT and business consulting services.
5+ YOE5+ years SRE/DevOps experience with Terraform, hybrid cloud networking, load-balancing (F5/AVI), automation using Python/Ansible/Shell, CI/CD tooling, and production support including on-call rotations.
Terraform, F5 BIG-IP, Broadcom AVI, Jenkins, Azure DevOps, GitHub Actions, Python, Ansible, Shell
1mo
Save
Mark Applied
Hide
Senior Site Reliability Champion
Wayne or Charlotte or Dallas or Fort Worth
HybridFull Time
Vanguard
Vanguard: Provides mutual funds, ETFs, and investment management services.
Experience with observability tools, reliability metrics (SLIs/SLOs), monitoring/alerting, automation (Python, RPA), incident review leadership, and onboarding resiliency tooling.
Splunk, Honeycomb, CloudWatch, Dynatrace, AppDynamics, Python, Blue Prism, UiPath
4w
Save
Mark Applied
Hide
Director, Site Reliability Engineering
Frisco or Eagan
$159k-$295k/yr HybridFull Time
Thomson Reuters
Thomson ReutersNASDAQ: TRI: Provides professional software, data, and news services globally.
10+ YOE10+ years in SRE or related tech leadership with experience leading global teams, observability, incident management, automation, and resilience engineering.
1mo
Save
Mark Applied
Hide
Manager, AI Quality & Reliability Engineering
Edina or Irving or Chicago
$89k-$156k/yr OnsiteFull Time
Vizient: Provides performance improvement and supply chain services to hospitals
7+ YOE7+ years in quality engineering or testing, experience with AI/ML/LLM-enabled systems, test automation and validation frameworks, strong analytical and communication skills, US work authorization required.