19 platform reliability engineer jobs at 11 companies in Coppell, TX

5d
Save
Mark Applied
Hide
Platform Reliability Engineer - Principal Engineer
Iselin or Irving or Charlotte
$159k-$305k/yr HybridFull Time
Wells Fargo
Wells FargoNYSE: WFC: Provides banking, investment, mortgage, and consumer finance products.
7+ YOE7+ years engineering experience, 5+ years supporting enterprise production environments, hands-on in one infrastructure domain, SRE practice experience, strong troubleshooting and automation skills.
Grafana, Splunk, Prometheus, AppDynamics, Cribl, ThousandEyes, Dynatrace, Python, Bash, PowerShell, Git, Ansible, Terraform, CICD
1d
Save
Mark Applied
Hide
Senior Lead Platform Reliability Engineer
Charlotte or Irving or Chandler or West Des Moines or Iselin
$159k-$305k/yr HybridFull Time
Wells Fargo
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
7+ YOE7+ years systems engineering or architecture, 5+ years supporting enterprise production environments, deep expertise in one infrastructure domain, SRE practices, automation and troubleshooting across domains.
Grafana, Splunk, Prometheus, AppDynamics, Cribl, ThousandEyes, Dynatrace, Python, Bash, PowerShell, Git, Ansible, Terraform, CI/CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with GCP exposure, Terraform/IaC, CI/CD, observability, DevSecOps practices, incident response, and automation for reliability and resiliency.
GCP, Azure, Terraform, Terraform Enterprise, Log Analytics, Dynatrace, Resource Graph, CI/CD, IAM
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Idaho or Plano
$117k-$209k/yr RemoteFull Time
Autodesk
AutodeskNASDAQ: ADSK: Developing software for architecture, engineering, and entertainment industries.
7+ YOEU.S. citizen with 7+ years SRE/platform/cloud experience, strong reliability engineering skills (SLOs/SLIs, observability, incident management), cloud experience (AWS/Azure), automation using Python/Go/Java/PowerShell/Bash, and ability to operate production services in regulated GovCloud environments.
AWS, Azure, Python, Go, Java, PowerShell, Bash, Infrastructure as Code, CI/CD, Splunk, Dynatrace, Datadog, CloudWatch, Kubernetes
1mo
Save
Mark Applied
Hide
Senior Lead Site Reliability Engineer
Jersey City or Plano
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied SRE experience, formal SRE training/certification, experience with AI/ML platform reliability, SLO/SLI design, observability, automation, and mentoring peers.
Databricks, GPU clusters, Grafana, Dynatrace, Prometheus, Datadog, Splunk, AWS, Snowflake, Kubernetes (EKS), Apache Kafka, Apache Spark, Feature Stores, Vector Databases, LLM
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Denver or Dallas
HybridFull Time
Analytic Partners
Analytic Partners: Provides commercial analytics software and marketing measurement solutions.
4+ YOE4+ years in Platform Engineering/DevOps or related; strong Linux/Windows; automation with Python, Bash, or PowerShell; deep AWS and Azure experience; CI/CD; Infrastructure as Code; containers.
Linux, Windows, Python, Bash, PowerShell, AWS, Azure, Jenkins, GitHub Actions, Terraform, CloudFormation, Arm, Docker, Kubernetes, Nomad, Consul, Vault, Splunk, Sumo Logic
3w
Save
Mark Applied
Hide
GenAI Adoption-Platform Engineer
Irving, Texas, United States
$95k-$110k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
8+ YOE8+ years experience designing and operating GenAI platform environments, managing Kubernetes-based AI platforms, implementing CI/CD and IaC pipelines, enabling multi-tenant usage, and ensuring scalability and reliability.
Kubernetes, CI/CD, IaC
1mo
Save
Mark Applied
Hide
Site Reliability Engineer III
Jersey City or Dallas
$133k-$185k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
3+ YOE3+ years SRE experience, proficiency with SLI/SLO concepts, Python/Java/Spring Boot/.Net or PySpark, observability tools, CI/CD, cloud platforms (AWS), and experience implementing infrastructure-as-code.
Databricks, Snowflake, AWS, Kubernetes, Python, PySpark, Java, Spring Boot, .Net, Grafana, Dynatrace, Prometheus, Datadog, Splunk
3w
Save
Mark Applied
Hide
GPUaaS-K8s Platform
Irving, Texas, United States
$95k-$110k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
6+ YOEOperate and maintain Kubernetes/OpenShift GPU platforms, enable GPU/accelerator workloads, implement CI/CD for GPU workloads, and manage scaling, reliability, and runbooks.
Kubernetes, OpenShift, CI/CD
1d
Save
Mark Applied
Hide
Principal Security Engineer - Reliability
Irving or Chandler or Charlotte
HybridFull Time
Wells Fargo
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
7+ YOE7+ years network security experience, 5+ years with enterprise security platforms, SRE/reliability practices, automation (Python/Ansible/Terraform), and incident leadership; bachelor\u0002s in related field or equivalent experience.
Python, Ansible, Terraform, APIs
1mo
Save
Mark Applied
Hide
Senior Lead Site Reliability Engineer - AI/ML and Data Platforms
Jersey City or Dallas
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied SRE experience, strong SLI/SLO/SLA and observability knowledge, experience with Grafana/Dynatrace/Prometheus/Datadog/Splunk, distributed systems expertise, mentoring and leadership experience, familiarity with safe AI usage in operations.
Grafana, Dynatrace, Prometheus, Datadog, Splunk, AWS, Databricks, Spark, Glue, MapReduce, Docker, Kubernetes, Terraform, Python
1w
Save
Mark Applied
Hide
Senior AI Ops & Incident/Site Reliability Engineer
Austin or Plano or Fort Mill or Boston or New York City or Tempe or San Diego
$94k-$171k/yr HybridFull Time
Perficient
Perficient: Provides digital transformation and AI consulting for global enterprises.
8+ YOEBachelor's degree or equivalent,8+ years in IT operations/SRE or production support,5+ years leading incident management,experience with Dynatrace,ServiceNow,cloud platforms and AIOps implementations.
Dynatrace, ServiceNow, AWS, Azure, GCP
3w
Save
Mark Applied
Hide
Senior Engineer
Jersey City or Chandler or Plano
$122k-$200k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
7+ YOE7+ years software/platform/site reliability engineering with 5+ years AWS experience, cloud security and networking expertise, Terraform and scripting (Python/Go/Shell), CI/CD and GitOps experience.
AWS, Python, Go, Shell, Terraform, Docker, EKS, Kubernetes, Grafana, Prometheus, Splunk, Dynatrace, GitOps, KMS
2w
Save
Mark Applied
Hide
Sr. Lead Infrastructure Engineer - Storage SRA
Plano or Columbus or Houston
OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied site reliability/infrastructure engineering experience with enterprise storage platforms, SRE practices, observability, automation, and incident response.
Grafana, Dynatrace, Prometheus, Splunk, Netcool, Dell EMC PowerFlex, NetApp SolidFire, Pure Storage, PMAX
2mo
Save
Mark Applied
Hide
SRE Engineer
Dallas, Texas, United States
OnsiteFull Time
Mphasis
MphasisNational Stock Exchange of India: MPHASIS: Provides information technology and business process services.
7+ YOE7+ years experience building reliable cloud platforms; expertise in Linux, cloud architecture, Kubernetes, Terraform, Python/Go, observability, and SRE practices.
Linux, AWS, Azure, GCP, Kubernetes, service mesh, Terraform, Python, Go
2w
Save
Mark Applied
Hide
Senior Lead Software Engineer-AI Foundation Services
Plano or Jersey City or Wilmington or McLean
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years software engineering experience building cloud-native AI/ML platform services with Kubernetes, CI/CD, and infrastructure-as-code; proficiency in Python/Java/Go; strong production reliability and secure-by-design practices.
Kubernetes, CI/CD, infrastructure-as-code, Python, Java, Go, GPU
2mo
Save
Mark Applied
Hide
Senior AI Engineer, Architect
Plano, Texas, United States
OnsiteFull Time
PepsiCo
PepsiCoNASDAQ: PEP: Global food and beverage manufacturer and distributor.
10+ YOEBachelor's in CS/AI/ML or equivalent; Master's preferred; 10+ years in ML/AI; experience designing and operating enterprise platforms with reliability and governance requirements.
Python, Machine Learning, Temporal, CI/CD, RBAC, OIDC/SAML, OTel
1mo
Save
Mark Applied
Hide
Manager III, Software Development - Content Data Platform
Austin or Bozeman or Denver or Minneapolis or Missoula or Portland or Salt Lake City or Seattle or Charlotte or Kalispell or Boise or Charleston or Dallas or Fort Worth or Phoenix or Richmond or Spokane or Vermont
$150k-$188k/yr RemoteFull Time
onXmaps
onXmaps: Digital mapping and navigation apps for outdoor recreation.
5+ Mgmt5+ years managing software engineers; experience building data platform/infrastructure and migration off legacy systems; hiring and team development skills; experience with data quality, SLOs, and operational reliability; US work authorization required; ability to travel.
GCP, Managed Spark, BigQuery, BigLake, Pub/Sub, Managed Airflow, Apache Iceberg, Spark, PySpark, DuckDB, dbt, Airflow, GDAL, PostGIS, Apache Sedona, DuckDB spatial extensions, Knowledge Graph, AI-assisted tools
2mo
Save
Mark Applied
Hide
Senior Lead DevOps Engineer
McLean or Dallas or Memphis
HybridFull Time
Hilton
HiltonNYSE: HLT: Global hospitality providing hotel accommodation and lodging services.
7+ YOESeven years in technology; five years with CI tools (GitLab preferred); four years automating delivery; four years administering Kubernetes; four years on AWS; scripting in Bash; on-call reliability; API platforms in Kubernetes; travel up to 10%; hybrid role near US offices.
GitLab, Bamboo, Kubernetes, Docker, AWS, Terraform, Groovy, Shell, Python, Java, JavaScript, WSO2, Apigee, Kong, AWS API Gateway, MuleSoft, Kafka, MSK