20 systems reliability engineer jobs at 17 companies in Aledo, TX

3w
Save
Mark Applied
Hide
Reliability Engineer
Westlake, Texas, United States
OnsiteFull Time
Fidelity Investments
Fidelity Investments: Financial services and investment management firm providing advisory solutions.
5+ YOEBachelor's or equivalent,5+ years deploying/supporting distributed systems,cloud and on-prem storage,Kubernetes (EKS/AKS/RKS),CI/CD automation,observability,backup/recovery,Python/NodeJS/Java and scripting.
Go, Angular, Python, JavaScript, AWS, RESTful services, Ruby, MVC, Jenkins CI/CD, Chef, Ansible, Bootstrap, HTML/CSS, Shell Scripting, MQ, OpenStack, PostgreSQL, PowerBI, Tableau, NodeJS, Java, Docker, Docker Compose, Git, Datadog, Splunk, Prometheus, Grafana, ELK/OpenSearch, OpenTelemetry, IAM, ARM, Terraform
2mo
Save
Mark Applied
Hide
Infrastructure Reliability Engineer
Manassas or Sterling or Portland or Chicago or Dallas Fort Worth
OnsiteFull Time
STACK Infrastructure
STACK Infrastructure: Developer and operator of sustainable wholesale data center infrastructure.
5+ YOE5–8 years in critical infrastructure; strong fluency in electrical systems; RCA/forensic troubleshooting; bachelor’s in engineering or equivalent.
Power distribution equipment, Waveform analysis, Fault analysis tools
3w
Save
Mark Applied
Hide
Reliability Engineer, Mechanical, NA (Design)
Shackelford County or Texas or Atlanta or Abilene or Dallas or Phoenix or Ashburn or Wisconsin
HybridFull Time
Vantage Data Centers
Vantage Data Centers: Provides hyperscale data center campuses for cloud and AI providers.
2+ YOEMechanical reliability engineer for data center cooling systems; 2–3 years critical facility experience preferred; bachelor’s degree preferred; experience with commissioning, maintenance program design, RCA, and technical support.
3w
Save
Mark Applied
Hide
Electrical Reliability Studies Engineer
New York or Los Angeles or Chicago or Houston or Tempe or Philadelphia or Dallas or North Miami Beach or Denver
$111k-$145k/yr HybridFull Time
Jacobs
JacobsNYSE: J: Global provider of professional engineering and technical services.
4+ YOEBachelor's in electrical engineering, 4+ years power-system and reliability analysis experience, familiarity with reliability indicators and NERC/FERC standards, strong analytical and communication skills.
ETAP, ReliaSoft BlockSim, Isograph Availability Workbench, Python
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin or Southlake
$129k-$175k/yr OnsiteFull Time
Charles Schwab
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
10+ YOEBachelor's in CS or related; 10+ years software development/SRE experience (8+ years DevOps/SRE), 8+ years CI/CD and observability, 5+ years leading reliability practices; strong automation, scripting (Python/shell), cloud and distributed systems experience; must be authorized to work in the U.S. without sponsorship.
Python, shell scripting, CI/CD, Splunk, Kubernetes, Terraform, AWS, GCP, Azure
2w
Save
Mark Applied
Hide
Sr. Systems Engineer, Site Reliability
Dallas, Texas, United States
HybridFull Time
American Heart Association
American Heart Association: Nonprofit organization dedicated to fighting heart disease and stroke.
5+ YOEBachelor's degree or equivalent experience, minimum 5 years relevant experience, strong expertise in multi-cloud operations, IAM, security, automation, and infrastructure as code.
Azure, AWS, GCP, Oracle, Entra ID, Structured Query Language (SQL)
6d
Save
Mark Applied
Hide
Director, Site Reliability Engineering
New York City or San Francisco or Dallas
$197k-$314k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
10+ YOE5+ MgmtBachelor's in a technical field,10+ years engineering experience with 5+ years leading SRE/Platform teams; experience with observability, incident management, distributed systems, and cloud architecture.
AWS, New Relic, Splunk, Datadog, Sentry, Honeycomb, Grafana, Prometheus, OpenTelemetry
1mo
Save
Mark Applied
Hide
Senior Lead Site Reliability Engineer - AI/ML and Data Platforms
Jersey City or Dallas
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied SRE experience, strong SLI/SLO/SLA and observability knowledge, experience with Grafana/Dynatrace/Prometheus/Datadog/Splunk, distributed systems expertise, mentoring and leadership experience, familiarity with safe AI usage in operations.
Grafana, Dynatrace, Prometheus, Datadog, Splunk, AWS, Databricks, Spark, Glue, MapReduce, Docker, Kubernetes, Terraform, Python
1mo
Save
Mark Applied
Hide
Manager, AI Quality & Reliability Engineering
Edina or Irving or Chicago
$89k-$156k/yr OnsiteFull Time
Vizient: Provides performance improvement and supply chain services to hospitals
7+ YOE7+ years in quality engineering or testing, experience with AI/ML/LLM-enabled systems, test automation and validation frameworks, strong analytical and communication skills, US work authorization required.
1w
Save
Mark Applied
Hide
Design for Safety & Reliability (DfSR) Practitioner
Fairfield or West Chester or Red Oak or Goodyear
$86k-$130k/yr HybridFull Time
Schneider Electric
Schneider ElectricEuronext Paris: SU: Provider of energy management and industrial automation solutions.
5+ YOEBachelor’s or Master’s in Mechanical/Electrical/Electronics/Mechatronics, 5+ years engineering or quality experience with DfSS/DfSR and data center systems, ReliaSoft proficiency, Agile experience; CSSGB/CSSBB or CRE desirable; Design FMEA experience preferred.
ReliaSoft, Agile
4d
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineering
Irving or San Leandro
OnsiteFull Time
Wells Fargo
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
7+ YOE3+ MgmtRequires 7+ years in systems engineering or technology architecture, 3+ years of management, 5+ years leading engineering or SRE teams, and experience with customer-facing platforms, incident management, SRE, DevOps, cloud, and production operations.
Splunk, Grafana, AppDynamics, Dynatrace, OpenTelemetry, Prometheus, Kubernetes, OpenShift, AWS, Microsoft Azure, Google Cloud Platform, Infrastructure as Code (IaC), CI/CD, AIOps, ITIL
5d
Save
Mark Applied
Hide
Senior Application Support Engineer / Site Reliability Engineer
Coppell, Texas, United States
OnsiteFull Time
DTCC
DTCC: Provides post-trade infrastructure for the global financial services industry
6+ YOE6+ years supporting enterprise applications; strong SRE knowledge, troubleshooting, observability, automation, messaging and distributed systems experience; familiarity with cloud and container platforms.
Java/J2EE, Oracle, SQL, DB2, IBM MQ, Kafka, Splunk, Grafana, AutoSys, ServiceNow, Linux/Unix, AWS, OpenShift Container Platform (OCP), IIB, Kubernetes, Docker, Python
1mo
Save
Mark Applied
Hide
Staff Software Engineer (Backend) - Everand Core
San Francisco or Ottawa or Phoenix or Toronto or Los Angeles or Denver or Salt Lake City or Atlanta or Chicago or Houston or Portland or New York City or Vancouver or San Diego or Sacramento or Jacksonville or Seattle or Mexico City or Austin or Miami or Boston or Dallas or Charlotte
$141k-$267k/yr HybridFull Time
Scribd
Scribd: Subscription-based digital library for e-books, audiobooks, and documents.
Significant backend engineering experience building and scaling distributed systems; strong coding in Ruby/Scala/Go/Python; expertise in reliability, observability, SLOs, and mentoring; cross-functional collaboration experience.
Ruby, Scala, Go, Python
3w
Save
Mark Applied
Hide
Operations Engineer - Data Center (SysOps)
Dallas, Texas, United States
$95k-$105k/yr HybridFull Time
Match Group
Match GroupNasdaq: MTCH: Provider of a global portfolio of online dating services.
5+ YOE5+ years onsite data center or systems engineering experience; Linux/Windows, AD/DNS/DHCP, HPE and Lenovo hardware, scripting (PowerShell,Bash,Python); able to lift 50 lbs; reliable for after-hours on-call; travel ~10%.
Claude, Glean, Linux, Windows, Active Directory, DNS, DHCP, HPE, Lenovo, PowerShell, Bash, Python, Proxmox, XCP-NG, HPE OneView, Lenovo XClarity
2w
Save
Mark Applied
Hide
Staff Machine Learning Engineer (Open to Remote)
Dallas or United States
$206k-$330k/yr RemoteFull Time
Triumph Financial: Provides financial and technology solutions for the transportation industry.
10+ YOE10+ years software engineering experience, 4+ years production ML, experience designing and operating distributed ML systems, model deployment and data pipeline expertise, strong reliability and production support skills.
Python, Clojure, Ruby, PySpark, AWS SageMaker Studio, PyTorch, HuggingFace, FastAPI, Zoom, Slack, MacBook
1mo
Save
Mark Applied
Hide
Staff Machine Learning Engineer - Leasing
Santa Barbara or San Diego or San Francisco or Denver or Dallas or Atlanta or Chicago or Washington, D.C.
HybridFull Time
AppFolio
AppFolioNASDAQ: APPF: Provides cloud-based property and investment management software.
Proven experience building production ML systems at scale, architectural leadership, training/fine-tuning LLMs, experience with LangChain/LangGraph and RAG patterns, AI safety/authorization, and production reliability discipline.
LangChain, LangGraph, vLLM, TensorRT, Triton
2w
Save
Mark Applied
Hide
Engineering Manager, Data Feeds
New York City or Boston or Chicago or Salt Lake City or Austin or Montreal or Canada or Washington or Dallas or United States or Vancouver or Toronto or Charlotte or Denver
$129k-$304k/yr RemoteFull Time
Chainlink Labs
Chainlink Labs: Building decentralized oracle networks for blockchain smart contracts.
Deep experience building and operating production blockchain infrastructure, strong engineering judgment, leadership and coaching experience, stakeholder management, and ability to deliver reliable, scalable systems.
6d
Save
Mark Applied
Hide
Sr SDE - AFX, AFX
Seattle or Dallas
$168k-$227k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
5+ YOE5+ years professional software development; expert in system design, reliability, and scaling; experience leading designs and mentoring engineers; experience across full SDLC; BS in CS preferred.
AWS, Salesforce, AWSentral
2mo
Save
Mark Applied
Hide
Manager III, Software Development - Content Data Platform
Austin or Bozeman or Denver or Minneapolis or Missoula or Portland or Salt Lake City or Seattle or Charlotte or Kalispell or Boise or Charleston or Dallas or Fort Worth or Phoenix or Richmond or Spokane or Vermont
$150k-$188k/yr RemoteFull Time
onXmaps
onXmaps: Digital mapping and navigation apps for outdoor recreation.
5+ Mgmt5+ years managing software engineers; experience building data platform/infrastructure and migration off legacy systems; hiring and team development skills; experience with data quality, SLOs, and operational reliability; US work authorization required; ability to travel.
GCP, Managed Spark, BigQuery, BigLake, Pub/Sub, Managed Airflow, Apache Iceberg, Spark, PySpark, DuckDB, dbt, Airflow, GDAL, PostGIS, Apache Sedona, DuckDB spatial extensions, Knowledge Graph, AI-assisted tools