24 systems reliability engineer jobs at 19 companies in Allen, TX
1mo
Save
Mark Applied
Hide
1mo
Systems Engineer Senior Staff - Reliability and Maintainability
Fort Worth or Marietta or Palmdale
$128k-$226k/yrOnsiteFull Time
Lockheed MartinNYSE: LMT: Designs and manufactures global security and aerospace systems.
U.S. citizenship and active Secret clearance required; bachelor’s degree in a STEM discipline (or equivalent); experience leading reliability, maintainability, supportability or product support engineering activities; ability to develop technical deliverables.
Shackelford County or Texas or Atlanta or Abilene or Dallas or Phoenix or Ashburn or Wisconsin
HybridFull Time
Vantage Data Centers: Provides hyperscale data center campuses for cloud and AI providers.
2+ YOEMechanical reliability engineer for data center cooling systems; 2–3 years critical facility experience preferred; bachelor’s degree preferred; experience with commissioning, maintenance program design, RCA, and technical support.
Manassas or Sterling or Portland or Chicago or Dallas Fort Worth
OnsiteFull Time
STACK Infrastructure: Developer and operator of sustainable wholesale data center infrastructure.
5+ YOE5–8 years in critical infrastructure; strong fluency in electrical systems; RCA/forensic troubleshooting; bachelor’s in engineering or equivalent.
Power distribution equipment, Waveform analysis, Fault analysis tools
New York or Los Angeles or Chicago or Houston or Tempe or Philadelphia or Dallas or North Miami Beach or Denver
$111k-$145k/yrHybridFull Time
JacobsNYSE: J: Global provider of professional engineering and technical services.
4+ YOEBachelor's in electrical engineering, 4+ years power-system and reliability analysis experience, familiarity with reliability indicators and NERC/FERC standards, strong analytical and communication skills.
OptimumNYSE: OPTU: Provides broadband, television, and mobile connectivity services to customers.
2+ YOEBachelor's in telecommunications/computer engineering, 2+ years systems or mobile network operations experience; deep Unix/Linux administration, GCP experience, Terraform/Ansible, Python/Go scripting, SAN/NAS storage protocols, observability tooling.
Shield AI: Develops autonomous flight software and unmanned aircraft for defense.
1+ YOEBachelor’s in Materials/Mechanical/Reliability engineering; 1–3 years hardware/materials/reliability experience; knowledge of materials science, electronics, and mechanical systems; basic reliability concepts; strong problem-solving and teamwork; good communication.
AlconNYSE: ALC: Manufactures ophthalmic surgical equipment and vision care products.
2+ YOEDesign, program, install, modify, and maintain campus-wide automation and control systems; ensure validation, compliance, and reliability in regulated manufacturing environments.
Site Reliability Engineer [Multiple Positions Available]
Plano, Texas, United States
OnsiteFull Time
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
7+ YOEBachelor's in engineering or CS plus 7 years (or Master's plus 5); experience with system monitoring, log analysis, cloud platforms, containers, IaC, CI/CD, automation using Python/Java/.NET, and incident/RCA processes.
Grafana, Dynatrace, Prometheus, Datadog, Splunk, Python, Java, Spring Boot, .NET, Terraform, Ansible, CloudWatch, Jenkins, ECS, Kubernetes, Docker, TCP/IP, Google Compute Runtime, Amazon Web Services, Microsoft Azure
onsemi: Designs and manufactures semiconductor solutions for power and sensing.
3+ YOE3+ years semiconductor lab/reliability testing experience; Associate's in Electronics or related; hands-on with ESD simulators and Latchup systems; familiarity with JEDEC/ESDA standards; able to read schematics and use oscilloscopes, DMMs, curve tracers.
Vizient: Provides performance improvement and supply chain services to hospitals
7+ YOE7+ years in quality engineering or testing, experience with AI/ML/LLM-enabled systems, test automation and validation frameworks, strong analytical and communication skills, US work authorization required.
Lead Director, Site Reliability Engineering - Client Experience
Richardson, Texas, United States
$144k-$288k/yrHybridFull Time
CVS HealthNYSE: CVS: Provides retail pharmacy, health insurance, and pharmacy benefit management services.
10+ YOE5+ Mgmt10+ years in engineering/SRE; 5+ years managing engineers; 5+ years cloud experience (Azure/GCP); strong distributed systems and cloud-native expertise.
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
5+ YOE3+ Mgmt5+ years in SRE with 3+ years in architect/leadership; design scalable, fault-tolerant systems; strong observability; CI/CD; postmortems; SRE leadership.
BAE SystemsLondon Stock Exchange: BA: Designs and manufactures advanced defense and aerospace systems.
4+ YOEBachelor's in mechanical engineering, 4+ years experience (or 2+ with MS), hands-on lab and prototype/test experience, mechanical design and analysis for high-reliability systems, ability to obtain Secret clearance, familiarity with MIL-STD/ASTM/SAE standards and engineering tools.
San Francisco or Ottawa or Phoenix or Toronto or Los Angeles or Denver or Salt Lake City or Atlanta or Chicago or Houston or Portland or New York City or Vancouver or San Diego or Sacramento or Jacksonville or Seattle or Mexico City or Austin or Miami or Boston or Dallas or Charlotte
$141k-$267k/yrHybridFull Time
Scribd: Subscription-based digital library for e-books, audiobooks, and documents.
Significant backend engineering experience building and scaling distributed systems; strong coding in Ruby/Scala/Go/Python; expertise in reliability, observability, SLOs, and mentoring; cross-functional collaboration experience.
Match GroupNasdaq: MTCH: Provider of a global portfolio of online dating services.
5+ YOE5+ years onsite data center or systems engineering experience; Linux/Windows, AD/DNS/DHCP, HPE and Lenovo hardware, scripting (PowerShell,Bash,Python); able to lift 50 lbs; reliable for after-hours on-call; travel ~10%.
Santa Barbara or San Diego or San Francisco or Denver or Dallas or Atlanta or Chicago or Washington, D.C.
HybridFull Time
AppFolioNASDAQ: APPF: Provides cloud-based property and investment management software.
Proven experience building production ML systems at scale, architectural leadership, training/fine-tuning LLMs, experience with LangChain/LangGraph and RAG patterns, AI safety/authorization, and production reliability discipline.
Crunchyroll: Operates a global streaming platform for anime and manga.
15+ YOE7+ Mgmt15+ years in platform engineering/infrastructure/DevOps/SRE with 7+ years managing managers; experience with distributed systems, multi-cloud (AWS, GCP), platform strategy, reliability (SLIs/SLOs), automation, and executive stakeholder engagement.
5+ YOESRE-focused engineer with experience in Java/Python, observability, containers, and cloud-native tech. Strong problem-solving and automation skills; able to design scalable, reliable systems.
Observability, Docker, Kubernetes, Prometheus, Grafana, ELK, OpenTelemetry, Terraform, Ansible, CloudFormation, Hadoop, Big Data