25 systems reliability engineer jobs at 21 companies in Crowley, TX

2mo
Save
Mark Applied
Hide
Infrastructure Reliability Engineer
Manassas or Sterling or Portland or Chicago or Dallas Fort Worth
OnsiteFull Time
STACK Infrastructure
STACK Infrastructure: Developer and operator of sustainable wholesale data center infrastructure.
5+ YOE5–8 years in critical infrastructure; strong fluency in electrical systems; RCA/forensic troubleshooting; bachelor’s in engineering or equivalent.
Power distribution equipment, Waveform analysis, Fault analysis tools
3w
Save
Mark Applied
Hide
Reliability Engineer, Mechanical, NA (Design)
Shackelford County or Texas or Atlanta or Abilene or Dallas or Phoenix or Ashburn or Wisconsin
HybridFull Time
Vantage Data Centers
Vantage Data Centers: Provides hyperscale data center campuses for cloud and AI providers.
2+ YOEMechanical reliability engineer for data center cooling systems; 2–3 years critical facility experience preferred; bachelor’s degree preferred; experience with commissioning, maintenance program design, RCA, and technical support.
3w
Save
Mark Applied
Hide
Electrical Reliability Studies Engineer
New York or Los Angeles or Chicago or Houston or Tempe or Philadelphia or Dallas or North Miami Beach or Denver
$111k-$145k/yr HybridFull Time
Jacobs
JacobsNYSE: J: Global provider of professional engineering and technical services.
4+ YOEBachelor's in electrical engineering, 4+ years power-system and reliability analysis experience, familiarity with reliability indicators and NERC/FERC standards, strong analytical and communication skills.
ETAP, ReliaSoft BlockSim, Isograph Availability Workbench, Python
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin or Southlake
$129k-$175k/yr OnsiteFull Time
Charles Schwab
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
10+ YOEBachelor's in CS or related; 10+ years software development/SRE experience (8+ years DevOps/SRE), 8+ years CI/CD and observability, 5+ years leading reliability practices; strong automation, scripting (Python/shell), cloud and distributed systems experience; must be authorized to work in the U.S. without sponsorship.
Python, shell scripting, CI/CD, Splunk, Kubernetes, Terraform, AWS, GCP, Azure
2w
Save
Mark Applied
Hide
Site Reliability Engineer Lead
Plano or Chandler or Charlotte
OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
10+ YOE10+ years SRE/DevOps experience with expertise in distributed systems, observability, automation, IaC, cloud, incident response, capacity planning, and strong stakeholder skills.
Dynatrace, Grafana, Splunk, OpenTelemetry, Terraform, Ansible, Python, Kubernetes, ServiceNow
3w
Save
Mark Applied
Hide
Sr. Systems Engineer, Site Reliability
Dallas, Texas, United States
HybridFull Time
American Heart Association
American Heart Association: Nonprofit organization dedicated to fighting heart disease and stroke.
5+ YOEBachelor's degree or equivalent experience, minimum 5 years relevant experience, strong expertise in multi-cloud operations, IAM, security, automation, and infrastructure as code.
Azure, AWS, GCP, Oracle, Entra ID, Structured Query Language (SQL)
5d
Save
Mark Applied
Hide
Senior Site Reliability Engineer - 3 Month Contract
Dallas or Minot
RemoteContract
Orion Health
Orion HealthToronto Stock Exchange: AIDX: Developing population-scale health platforms and healthcare data interoperability software.
4+ YOERequires 4–6 years of site reliability or equivalent experience, systems/application support or development experience, scripting, cloud production support, infrastructure automation, networking, and a technical bachelor's degree or equivalent.
Amazon Web Services (AWS), Windows, Linux, Active Directory, Group Policy Object (GPO), DNS, PowerShell, Python, Bash, Puppet, Ansible, Kubernetes, CloudFormation, Terraform, Splunk, TCP/IP, DHCP, VLANs, VPNs, firewall, Load Balancers, Continuous Integration/Continuous Delivery (CI/CD), Red Hat, Oracle, SQL, HIPAA, HITRUST
1w
Save
Mark Applied
Hide
Director, Site Reliability Engineering
New York City or San Francisco or Dallas
$197k-$314k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
10+ YOE5+ MgmtBachelor's in a technical field,10+ years engineering experience with 5+ years leading SRE/Platform teams; experience with observability, incident management, distributed systems, and cloud architecture.
AWS, New Relic, Splunk, Datadog, Sentry, Honeycomb, Grafana, Prometheus, OpenTelemetry
2h
Save
Mark Applied
Hide
Lead Software Reliability Engineer
Irving or New York City or New Jersey or Tampa or Jacksonville
$145k-$175k/yr HybridFull Time, Contract
RE Partners
RE Partners: Providing technology consulting and digital transformation services for enterprise clients.
Several years of TDD experience, strong coding skills, JVM languages and/or TypeScript, secure coding, CI/CD, Docker, OpenShift, Linux, build automation, and critical-systems observability expertise.
Test-Driven Development (TDD), Java, NPM, Spring, Renovate, Liquibase, Ansible, Docker, OpenShift, TypeScript, Gradle, Maven, Microsoft?, Linux, JVM
1mo
Save
Mark Applied
Hide
Senior Lead Site Reliability Engineer - AI/ML and Data Platforms
Jersey City or Dallas
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied SRE experience, strong SLI/SLO/SLA and observability knowledge, experience with Grafana/Dynatrace/Prometheus/Datadog/Splunk, distributed systems expertise, mentoring and leadership experience, familiarity with safe AI usage in operations.
Grafana, Dynatrace, Prometheus, Datadog, Splunk, AWS, Databricks, Spark, Glue, MapReduce, Docker, Kubernetes, Terraform, Python
1mo
Save
Mark Applied
Hide
Manager, AI Quality & Reliability Engineering
Edina or Irving or Chicago
$89k-$156k/yr OnsiteFull Time
Vizient: Provides performance improvement and supply chain services to hospitals
7+ YOE7+ years in quality engineering or testing, experience with AI/ML/LLM-enabled systems, test automation and validation frameworks, strong analytical and communication skills, US work authorization required.
1w
Save
Mark Applied
Hide
Design for Safety & Reliability (DfSR) Practitioner
Fairfield or West Chester or Red Oak or Goodyear
$86k-$130k/yr HybridFull Time
Schneider Electric
Schneider ElectricEuronext Paris: SU: Provider of energy management and industrial automation solutions.
5+ YOEBachelor’s or Master’s in Mechanical/Electrical/Electronics/Mechatronics, 5+ years engineering or quality experience with DfSS/DfSR and data center systems, ReliaSoft proficiency, Agile experience; CSSGB/CSSBB or CRE desirable; Design FMEA experience preferred.
ReliaSoft, Agile
5d
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineering
Irving or San Leandro
OnsiteFull Time
Wells Fargo
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
7+ YOE3+ MgmtRequires 7+ years in systems engineering or technology architecture, 3+ years of management, 5+ years leading engineering or SRE teams, and experience with customer-facing platforms, incident management, SRE, DevOps, cloud, and production operations.
Splunk, Grafana, AppDynamics, Dynatrace, OpenTelemetry, Prometheus, Kubernetes, OpenShift, AWS, Microsoft Azure, Google Cloud Platform, Infrastructure as Code (IaC), CI/CD, AIOps, ITIL
6d
Save
Mark Applied
Hide
Senior Application Support Engineer / Site Reliability Engineer
Coppell, Texas, United States
OnsiteFull Time
DTCC
DTCC: Provides post-trade infrastructure for the global financial services industry
6+ YOE6+ years supporting enterprise applications; strong SRE knowledge, troubleshooting, observability, automation, messaging and distributed systems experience; familiarity with cloud and container platforms.
Java/J2EE, Oracle, SQL, DB2, IBM MQ, Kafka, Splunk, Grafana, AutoSys, ServiceNow, Linux/Unix, AWS, OpenShift Container Platform (OCP), IIB, Kubernetes, Docker, Python
6d
Save
Mark Applied
Hide
Hardware and Software High Availability Engineer – Aerospace & Mission-Critical Systems
Plano, Texas, United States
$150k-$220k/yr OnsiteFull Time
Nokia
NokiaNYSE: NOK: Sells telecommunications infrastructure and software for global network operators.
Bachelor's or Master's in Electrical/Computer/Aerospace Engineering; experience in high-reliability hardware and embedded systems; knowledge of radiation, thermal management, EMC/EMI, reliability engineering, and environmental qualification.
5d
Save
Mark Applied
Hide
Lead Software Engineer
United States or Louisville or Plano
$149k-$187k/yr RemoteFull Time
KFC Pan Europe
KFC Pan EuropeNYSE: YUM: Quick-service restaurant chain specializing in fried chicken products.
Advanced software engineering or architecture experience with TypeScript, Node.js, REST, OpenAPI, GraphQL, AWS, CI/CD, React, AI-assisted development, production reliability, and secure enterprise systems.
TypeScript, Node.js, REST API, OpenAPI, GraphQL, AWS, Lambda, API Gateway, CloudFront, Secrets Manager, CloudWatch, IAM, ACM, Route53, CDK, CI/CD, React, React Native, Apollo Client, Datadog, APM, Contentful, OIDC, JWKS, BFFClient, Yum Storefront GraphQL
1mo
Save
Mark Applied
Hide
Staff Software Engineer (Backend) - Everand Core
San Francisco or Ottawa or Phoenix or Toronto or Los Angeles or Denver or Salt Lake City or Atlanta or Chicago or Houston or Portland or New York City or Vancouver or San Diego or Sacramento or Jacksonville or Seattle or Mexico City or Austin or Miami or Boston or Dallas or Charlotte
$141k-$267k/yr HybridFull Time
Scribd
Scribd: Subscription-based digital library for e-books, audiobooks, and documents.
Significant backend engineering experience building and scaling distributed systems; strong coding in Ruby/Scala/Go/Python; expertise in reliability, observability, SLOs, and mentoring; cross-functional collaboration experience.
Ruby, Scala, Go, Python
3w
Save
Mark Applied
Hide
Operations Engineer - Data Center (SysOps)
Dallas, Texas, United States
$95k-$105k/yr HybridFull Time
Match Group
Match GroupNasdaq: MTCH: Provider of a global portfolio of online dating services.
5+ YOE5+ years onsite data center or systems engineering experience; Linux/Windows, AD/DNS/DHCP, HPE and Lenovo hardware, scripting (PowerShell,Bash,Python); able to lift 50 lbs; reliable for after-hours on-call; travel ~10%.
Claude, Glean, Linux, Windows, Active Directory, DNS, DHCP, HPE, Lenovo, PowerShell, Bash, Python, Proxmox, XCP-NG, HPE OneView, Lenovo XClarity
2w
Save
Mark Applied
Hide
Staff Machine Learning Engineer (Open to Remote)
Dallas or United States
$206k-$330k/yr RemoteFull Time
Triumph Financial: Provides financial and technology solutions for the transportation industry.
10+ YOE10+ years software engineering experience, 4+ years production ML, experience designing and operating distributed ML systems, model deployment and data pipeline expertise, strong reliability and production support skills.
Python, Clojure, Ruby, PySpark, AWS SageMaker Studio, PyTorch, HuggingFace, FastAPI, Zoom, Slack, MacBook
1mo
Save
Mark Applied
Hide
Staff Machine Learning Engineer - Leasing
Santa Barbara or San Diego or San Francisco or Denver or Dallas or Atlanta or Chicago or Washington, D.C.
HybridFull Time
AppFolio
AppFolioNASDAQ: APPF: Provides cloud-based property and investment management software.
Proven experience building production ML systems at scale, architectural leadership, training/fine-tuning LLMs, experience with LangChain/LangGraph and RAG patterns, AI safety/authorization, and production reliability discipline.
LangChain, LangGraph, vLLM, TensorRT, Triton