14 reliability automation engineer jobs at 11 companies in Mabank, TX

2w
Save
Mark Applied
Hide
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$315k/yr HybridFull Time
Catalyst Brands
Catalyst Brands: Operates a portfolio of diverse retail clothing and apparel brands.
12+ YOE5+ Mgmt12+ years progressive engineering leadership with 5+ years leading reliability/performance/automation; deep SRE, performance, and automation experience; cloud and observability tool expertise; ability to lead distributed engineering teams.
AWS, Azure, Kubernetes, EKS, Jenkins, Kafka, Dynatrace, Splunk, Datadog, New Relic, Catchpoint, CloudWatch, ELK
2w
Save
Mark Applied
Hide
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$225k/yr HybridFull Time
Catalyst Brands
Catalyst Brands: Operates a portfolio of retail brands including JCPenney.
12+ YOE5+ Mgmt12+ years engineering leadership with 5+ years leading reliability, performance or automation functions; deep SRE, performance engineering, automation, cloud and observability experience; BA/BS preferred.
AWS, Azure, Kubernetes/EKS, Jenkins, Kafka, Dynatrace, Splunk, Datadog, New Relic, Catchpoint, CloudWatch, ELK
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Dallas or Charlotte or Wayne or Plano
HybridFull Time
Vanguard
Vanguard: Global investment management and financial services provider.
Experience with observability, monitoring, reliability metrics, alerting, automation, resilience engineering, incident response, and production troubleshooting; Python-based automation and chaos engineering experience are mentioned.
Splunk, Honeycomb, Amazon CloudWatch, Dynatrace, AppDynamics, Python, Blue Prism, UiPath
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Copenhagen or Dallas or Dubai or London or Miami or Pune or Buenos Aires or Bogota or Mexico or Singapore
RemoteFull Time
CellPoint Digital
CellPoint Digital: Payment orchestration platform for the travel and hospitality industry.
6+ YOE6+ years SRE/DevOps experience, deep GCP and Kubernetes knowledge, Terraform expertise, incident leadership, automation, security collaboration, strong communication.
GCP, Kubernetes (GKE), CloudSQL, Spanner, Terraform, IAM
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
Denver or Dallas
HybridFull Time
Analytic Partners
Analytic Partners: Provides commercial analytics software and marketing measurement solutions.
4+ YOE4+ years in Platform Engineering/DevOps or related; strong Linux/Windows; automation with Python, Bash, or PowerShell; deep AWS and Azure experience; CI/CD; Infrastructure as Code; containers.
Linux, Windows, Python, Bash, PowerShell, AWS, Azure, Jenkins, GitHub Actions, Terraform, CloudFormation, Arm, Docker, Kubernetes, Nomad, Consul, Vault, Splunk, Sumo Logic
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer - 3 Month Contract
Dallas or Minot
RemoteContract
Orion Health
Orion HealthToronto Stock Exchange: AIDX: Developing population-scale health platforms and healthcare data interoperability software.
4+ YOERequires 4–6 years of site reliability or equivalent experience, systems/application support or development experience, scripting, cloud production support, infrastructure automation, networking, and a technical bachelor's degree or equivalent.
Amazon Web Services (AWS), Windows, Linux, Active Directory, Group Policy Object (GPO), DNS, PowerShell, Python, Bash, Puppet, Ansible, Kubernetes, CloudFormation, Terraform, Splunk, TCP/IP, DHCP, VLANs, VPNs, firewall, Load Balancers, Continuous Integration/Continuous Delivery (CI/CD), Red Hat, Oracle, SQL, HIPAA, HITRUST
1mo
Save
Mark Applied
Hide
Senior Site Reliability Champion
Wayne or Charlotte or Dallas or Fort Worth
HybridFull Time
Vanguard
Vanguard: Provides mutual funds, ETFs, and investment management services.
Experience with observability tools, reliability metrics (SLIs/SLOs), monitoring/alerting, automation (Python, RPA), incident review leadership, and onboarding resiliency tooling.
Splunk, Honeycomb, CloudWatch, Dynatrace, AppDynamics, Python, Blue Prism, UiPath
3w
Save
Mark Applied
Hide
Sr. Systems Engineer, Site Reliability
Dallas, Texas, United States
HybridFull Time
American Heart Association
American Heart Association: Nonprofit organization dedicated to fighting heart disease and stroke.
5+ YOEBachelor's degree or equivalent experience, minimum 5 years relevant experience, strong expertise in multi-cloud operations, IAM, security, automation, and infrastructure as code.
Azure, AWS, GCP, Oracle, Entra ID, Structured Query Language (SQL)
1mo
Save
Mark Applied
Hide
Operations Engineering Manager, Fleet Reliability
Dallas or Bellevue
$143k-$191k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
7+ YOE2+ Mgmt7+ years in software or infrastructure engineering with 2+ years leadership; SRE fundamentals, incident management, observability, change management; strong automation and people development skills.
1mo
Save
Mark Applied
Hide
Software Engineering Manager - Site Reliability Center
Pittsburgh or Cleveland or Birmingham or Dallas or Denver or Phoenix
$100k-$204k/yr OnsiteFull Time
PNC Financial Services
PNC Financial ServicesNYSE: PNC: Provides banking, lending, and investment services to customers.
5+ YOE3+ MgmtLead SRE teams to ensure reliability, incident and change management, production support, automation, observability, and performance; 5+ years related experience with 3+ years management; hands-on with monitoring, cloud/infrastructure, databases and automation.
Dynatrace, BigPanda, Logscale, Linux, Windows, Oracle, SQL, MongoDB, Cassandra, Elasticsearch, Redis, MQ, Kafka, OCP, ShiftPlanning, ETL
1mo
Save
Mark Applied
Hide
Software Engineering Manager-Site Reliability Center-Twilight
Cleveland or Birmingham or Pittsburgh or Dallas or Denver or Phoenix
$100k-$204k/yr OnsiteFull Time
PNC Financial Services
PNC Financial ServicesNYSE: PNC: Provides banking, lending, and investment services to customers.
5+ YOE3+ Mgmt5+ years related experience and 3+ years management; SRE/production support/DevOps experience; incident/problem/change management; hands-on monitoring, cloud/infrastructure, automation; experience with Linux/Windows and databases.
Dynatrace, BigPanda, Logscale, OCP, Linux, Windows, Oracle, SQL, MongoDB, Cassandra, Elasticsearch, Redis, MQ, Kafka, ShiftPlanning
2mo
Save
Mark Applied
Hide
Senior Lead DevOps Engineer
McLean or Dallas or Memphis
HybridFull Time
Hilton
HiltonNYSE: HLT: Global hospitality providing hotel accommodation and lodging services.
7+ YOESeven years in technology; five years with CI tools (GitLab preferred); four years automating delivery; four years administering Kubernetes; four years on AWS; scripting in Bash; on-call reliability; API platforms in Kubernetes; travel up to 10%; hybrid role near US offices.
GitLab, Bamboo, Kubernetes, Docker, AWS, Terraform, Groovy, Shell, Python, Java, JavaScript, WSO2, Apigee, Kong, AWS API Gateway, MuleSoft, Kafka, MSK
1w
Save
Mark Applied
Hide
Associate Principal, Software Engineering: DevOps
Dallas, Texas, United States
$122k-$198k/yr HybridFull Time
Options Clearing Corporation
Options Clearing Corporation: Clearing and settlement for equity derivatives and options.
5+ YOEBachelor's degree or equivalent experience; 5+ years in DevOps, site reliability, or infrastructure engineering; automation, AWS, Kubernetes, Terraform, Linux, scripting, CI/CD, and Agile expertise.
Amazon Web Services (AWS), Kubernetes, Bash, Python, GitHub Actions, GitHub, Jenkins, Groovy, Helm, Terraform, Linux, Amazon EC2, Amazon S3, Amazon RDS, AWS Lambda, Amazon CloudWatch, AWS Identity and Access Management (IAM), Scrum, Kanban
3mo
Save
Mark Applied
Hide
Compliance Engineering, Dev Ops, Vice President, Dallas
Dallas, Texas, United States
OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Provides investment banking, securities, and wealth management services globally.
5+ YOESRE-focused engineer with experience in Java/Python, observability, containers, and cloud-native tech. Strong problem-solving and automation skills; able to design scalable, reliable systems.
Observability, Docker, Kubernetes, Prometheus, Grafana, ELK, OpenTelemetry, Terraform, Ansible, CloudFormation, Hadoop, Big Data