35 reliability engineering manager jobs at 21 companies in Valley View, TX

1mo
Save
Mark Applied
Hide
Manager, AI Quality & Reliability Engineering
Edina or Irving or Chicago
$89k-$156k/yr OnsiteFull Time
Vizient: Provides performance improvement and supply chain services to hospitals
7+ YOE7+ years in quality engineering or testing, experience with AI/ML/LLM-enabled systems, test automation and validation frameworks, strong analytical and communication skills, US work authorization required.
1mo
Save
Mark Applied
Hide
Manager, AI Quality & Reliability Engineering
Edina or Irving or Chicago
$89k-$156k/yr OnsiteFull Time
Vizient
Vizient: Healthcare performance improvement and group purchasing organization.
7+ YOE7+ years in quality engineering or software testing, experience with AI/ML and LLM-enabled workflows, experience implementing validation and observability, strong analytical and communication skills, authorized to work in the U.S. without sponsorship.
LLMOps, AIOps
4w
Save
Mark Applied
Hide
Director, Site Reliability Engineering
Frisco or Eagan
$159k-$295k/yr HybridFull Time
Thomson Reuters
Thomson ReutersNASDAQ: TRI: Provides professional software, data, and news services globally.
10+ YOE10+ years in SRE or related tech leadership with experience leading global teams, observability, incident management, automation, and resilience engineering.
1d
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineering
Irving or San Leandro
OnsiteFull Time
Wells Fargo
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
7+ YOE3+ MgmtRequires 7+ years in systems engineering or technology architecture, 3+ years of management, 5+ years leading engineering or SRE teams, and experience with customer-facing platforms, incident management, SRE, DevOps, cloud, and production operations.
Splunk, Grafana, AppDynamics, Dynatrace, OpenTelemetry, Prometheus, Kubernetes, OpenShift, AWS, Microsoft Azure, Google Cloud Platform, Infrastructure as Code (IaC), CI/CD, AIOps, ITIL
3w
Save
Mark Applied
Hide
Site Reliability Engineering Senior Manager
Columbus or Frisco
$134k-$242k/yr HybridFull Time
Bread Financial
Bread FinancialNYSE: BFH: Provides credit cards, consumer loans, and personal savings accounts.
8+ YOE3+ Mgmt8+ years IT/SRE/DevOps experience, 3+ years leadership, AWS/Azure experience, SRE/DevOps expertise, AI/ML and AIOps familiarity, knowledge of regulatory compliance (PCI-DSS, SOX), ITILv4 preferred.
AWS, Azure, AIOps, AI/ML
4w
Save
Mark Applied
Hide
Associate Director, Maintenance and Reliability Engineering
Irving, Texas, United States
$123k-$148k/yr OnsiteFull Time
HelloFresh
HelloFreshFrankfurt Stock Exchange: HFG: Global meal kit delivery service and food solutions provider.
8+ YOE8+ years managing maintenance and reliability in food/perishables/distribution, high school diploma required, bachelor’s preferred, OSHA/GMP/LEAN knowledge, blueprint/schematic proficiency, refrigeration exposure, strong leadership and project management skills.
2mo
Save
Mark Applied
Hide
Principal Architect, Site Reliability Engineering
Southlake or Austin
$221k-$252k/yr OnsiteFull Time
Charles Schwab
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
5+ YOE3+ Mgmt5+ years in SRE with 3+ years in architect/leadership; design scalable, fault-tolerant systems; strong observability; CI/CD; postmortems; SRE leadership.
Prometheus, Grafana, Datadog, Splunk
1d
Save
Mark Applied
Hide
Engineer, Site Reliability
Atlanta or Frisco
$85k-$153k/yr OnsiteFull Time
T-Mobile
T-MobileNASDAQ: TMUS: Provides wireless network services and mobile communication devices.
3+ YOEBachelor's degree and 3 years of related experience, or advanced degree and 1 year. Requires automation, monitoring, scripting, cloud computing, CI/CD, incident management, and system reliability skills.
CI/CD, Kubernetes, AWS
1d
Save
Mark Applied
Hide
Engineer, Site Reliability
Atlanta or Frisco
$85k-$153k/yr OnsiteFull Time
T-Mobile
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
3+ YOEBachelor's degree and 3 years of related experience, or advanced degree and 1 year; automation, monitoring, CI/CD, cloud computing, scripting, incident management, and system reliability skills required.
CI/CD, Kubernetes, AWS
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Idaho or Plano
$117k-$209k/yr RemoteFull Time
Autodesk
AutodeskNASDAQ: ADSK: Developing software for architecture, engineering, and entertainment industries.
7+ YOEU.S. citizen with 7+ years SRE/platform/cloud experience, strong reliability engineering skills (SLOs/SLIs, observability, incident management), cloud experience (AWS/Azure), automation using Python/Go/Java/PowerShell/Bash, and ability to operate production services in regulated GovCloud environments.
AWS, Azure, Python, Go, Java, PowerShell, Bash, Infrastructure as Code, CI/CD, Splunk, Dynatrace, Datadog, CloudWatch, Kubernetes
1mo
Save
Mark Applied
Hide
Senior Lead Site Reliability Engineer
Jersey City or Plano
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied SRE experience, formal SRE training/certification, experience with AI/ML platform reliability, SLO/SLI design, observability, automation, and mentoring peers.
Databricks, GPU clusters, Grafana, Dynatrace, Prometheus, Datadog, Splunk, AWS, Snowflake, Kubernetes (EKS), Apache Kafka, Apache Spark, Feature Stores, Vector Databases, LLM
2w
Save
Mark Applied
Hide
Site Reliability Engineer Lead
Plano or Chandler or Charlotte
OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
10+ YOE10+ years SRE/DevOps experience with expertise in distributed systems, observability, automation, IaC, cloud, incident response, capacity planning, and strong stakeholder skills.
Dynatrace, Grafana, Splunk, OpenTelemetry, Terraform, Ansible, Python, Kubernetes, ServiceNow
2w
Save
Mark Applied
Hide
Site Reliability Engineer Lead
Plano or Chandler or Charlotte
OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides banking, investment, and financial risk management services.
10+ YOE10+ years SRE/DevOps experience with Linux/Unix and Windows, observability (Dynatrace,Grafana,Splunk,OpenTelemetry), IaC and automation (Terraform,Ansible,Python), Kubernetes, incident response and capacity planning.
Dynatrace, Grafana, Splunk, OpenTelemetry, Terraform, Ansible, Python, Kubernetes, Linux, Unix, Windows, ServiceNow
2w
Save
Mark Applied
Hide
Senior AI Ops & Incident/Site Reliability Engineer
Austin or Plano or Fort Mill or Boston or New York City or Tempe or San Diego
$94k-$171k/yr HybridFull Time
Perficient
Perficient: Provides digital transformation and AI consulting for global enterprises.
8+ YOEBachelor's degree or equivalent,8+ years in IT operations/SRE or production support,5+ years leading incident management,experience with Dynatrace,ServiceNow,cloud platforms and AIOps implementations.
Dynatrace, ServiceNow, AWS, Azure, GCP
2w
Save
Mark Applied
Hide
Software Engineer Lead - Site Reliability Engineering Center
Pittsburgh or Phoenix or Lakewood or Birmingham or Strongsville or Farmers Branch
$86k-$158k/yr OnsiteFull Time
PNC Financial Services
PNC Financial ServicesNYSE: PNC: Provides banking, lending, and investment services to customers.
3+ YOEBachelors degree and 3+ years experience in software engineering. Experience with Java ecosystem, cloud-native and SRE practices; strong problem solving and communication skills.
Java, Spring boot, J2EE, Kubernetes, Microservices, Mongo DB, Dynatrace, Elastic search, API, Postman, MQ, Kafka, ETL, SAFe, Mainframe, Oracle, SQL, Ansible
2mo
Save
Mark Applied
Hide
Package Architect
Dallas or Richardson
HybridFull Time
Nexperia
Nexperia: Manufactures discrete semiconductors, logic devices, and power MOSFET components.
8+ YOELead packaging strategy; bachelor's in engineering; 8+ years in semiconductors; 7+ years in packaging, DoE, data analysis, and reliability tests; 5+ years in project management; travel up to 5%.
Six Sigma, DoE, Data analysis, Semiconductor packaging, Reliability testing, Package design
6d
Save
Mark Applied
Hide
Lead Data Engineer – Physical AI Platform, Data Engineering
Irving, Texas, United States
$128k-$209k/yr OnsiteFull Time
Caterpillar
CaterpillarNYSE: CAT: Manufactures construction and mining equipment, engines, and gas turbines.
8+ YOE8+ years data engineering experience with Python/Java, AWS data services, SQL, CI/CD and microservices; leadership of data architecture, data quality, monitoring, and production reliability.
Python, Java, Kinesis, Amazon S3, DynamoDB, EventBridge, CloudWatch, SQL, Azure DevOps, Jira, Jenkins, Helios Data Platform
1d
Save
Mark Applied
Hide
Site Rel Eng III, GCP
Bethpage or Plano or Long Island or New York City
$134k-$220k/yr OnsiteFull Time
Optimum
OptimumNYSE: OPTU: Provides broadband, television, and mobile connectivity services to customers.
8+ YOERequires 8+ years in SRE, platform, cloud, DevOps, or network engineering and 5+ years operating GCP production workloads, with expertise in GKE, Kubernetes, Terraform, networking, IAM, observability, and automation.
Google Cloud Platform (GCP), GKE, Cloud Run, Compute Engine, Cloud SQL, Pub/Sub, Cloud Storage, Cloud Monitoring, Network Connectivity Center (NCC), Interconnect, Cloud VPN, Cloud Router, Shared VPC, Private Service Connect, Terraform, Kubernetes, IAM, Python, Bash, GitOps, FinOps, AI/ML
3w
Save
Mark Applied
Hide
Director of Packaging Technology
Plano, Texas, United States
OnsiteFull Time
Diodes Incorporated
Diodes IncorporatedNasdaq: DIOD: Global manufacturer of discrete, logic, and analog semiconductor components
15+ YOE10+ MgmtMSc/PhD in engineering or materials science,15+ years in semiconductor packaging,10+ years managing global teams,experience with package architectures, reliability testing, and supplier collaboration.
4w
Save
Mark Applied
Hide
GPUaaS-K8s Platform
Irving, Texas, United States
$95k-$110k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
6+ YOEOperate and maintain Kubernetes/OpenShift GPU platforms, enable GPU/accelerator workloads, implement CI/CD for GPU workloads, and manage scaling, reliability, and runbooks.
Kubernetes, OpenShift, CI/CD

Explore Jobs

Expand Your Job Search