19 infrastructure reliability engineer jobs at 15 companies in Addison, TX

2mo
Save
Mark Applied
Hide
Infrastructure Reliability Engineer
Manassas or Sterling or Portland or Chicago or Dallas Fort Worth
OnsiteFull Time
STACK Infrastructure
STACK Infrastructure: Developer and operator of sustainable wholesale data center infrastructure.
5+ YOE5–8 years in critical infrastructure; strong fluency in electrical systems; RCA/forensic troubleshooting; bachelor’s in engineering or equivalent.
Power distribution equipment, Waveform analysis, Fault analysis tools
1w
Save
Mark Applied
Hide
Sr. Lead Infrastructure Engineer - Storage SRA
Plano or Columbus or Houston
OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied site reliability/infrastructure engineering experience with enterprise storage platforms, SRE practices, observability, automation, and incident response.
Grafana, Dynatrace, Prometheus, Splunk, Netcool, Dell EMC PowerFlex, NetApp SolidFire, Pure Storage, PMAX
1d
Save
Mark Applied
Hide
Platform Reliability Engineer - Principal Engineer
Iselin or Irving or Charlotte
$159k-$305k/yr HybridFull Time
Wells Fargo
Wells FargoNYSE: WFC: Provides banking, investment, mortgage, and consumer finance products.
7+ YOE7+ years engineering experience, 5+ years supporting enterprise production environments, hands-on in one infrastructure domain, SRE practice experience, strong troubleshooting and automation skills.
Grafana, Splunk, Prometheus, AppDynamics, Cribl, ThousandEyes, Dynatrace, Python, Bash, PowerShell, Git, Ansible, Terraform, CICD
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Denver or Dallas
HybridFull Time
Analytic Partners
Analytic Partners: Provides commercial analytics software and marketing measurement solutions.
4+ YOE4+ years in Platform Engineering/DevOps or related; strong Linux/Windows; automation with Python, Bash, or PowerShell; deep AWS and Azure experience; CI/CD; Infrastructure as Code; containers.
Linux, Windows, Python, Bash, PowerShell, AWS, Azure, Jenkins, GitHub Actions, Terraform, CloudFormation, Arm, Docker, Kubernetes, Nomad, Consul, Vault, Splunk, Sumo Logic
1w
Save
Mark Applied
Hide
TE080 Senior Infrastructure Engineer
Plano, Texas, United States
OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
Lead complex technical designs, develop automation and deployment playbooks, translate business requirements into technical blueprints, mentor team members, and improve reliability and efficiency of services.
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (Onsite Hybrid) (Plano, TX, US)
Plano, Texas, United States
$97k-$145k/yr HybridFull Time
NTT DATA
NTT DATA: Global provider of IT and business consulting services.
5+ YOE5+ years SRE experience with New Relic, GitHub Enterprise/Actions, CI/CD for Java/.NET, SLI/SLO/alerting, incident response/RCA, production troubleshooting, and supporting application+infrastructure workloads.
New Relic, GitHub Enterprise, GitHub Actions, Java, .NET, GitHub Copilot, JFrog Artifactory, JFrog Xray, SonarQube, GitHub Advanced Security, Databricks, ServiceNow ITOM, CodeQL, Dependabot, SQL, Angular
1mo
Save
Mark Applied
Hide
Site Reliability Engineer III
Jersey City or Dallas
$133k-$185k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
3+ YOE3+ years SRE experience, proficiency with SLI/SLO concepts, Python/Java/Spring Boot/.Net or PySpark, observability tools, CI/CD, cloud platforms (AWS), and experience implementing infrastructure-as-code.
Databricks, Snowflake, AWS, Kubernetes, Python, PySpark, Java, Spring Boot, .Net, Grafana, Dynatrace, Prometheus, Datadog, Splunk
1w
Save
Mark Applied
Hide
Sr. Systems Engineer, Site Reliability
Dallas, Texas, United States
OnsiteFull Time
American Heart Association
American Heart Association: Non-profit organization funding cardiovascular research and public health education.
5+ YOEBachelor's degree or equivalent, 5+ years relevant experience, expertise in multi-cloud operations, IAM, security, automation, DevOps, and infrastructure reliability.
Entra ID, Azure, AWS, GCP, Oracle, SQL
2w
Save
Mark Applied
Hide
Infrastructure/Cloud Architect
Dallas or Frisco
$115k-$219k/yr RemoteFull Time
Kyndryl
KyndrylNYSE: KD: Manages and modernizes mission-critical IT infrastructure systems.
Design and guide cloud, Linux, container, Kubernetes, and edge infrastructure; collaborate with engineering and delivery teams; ensure scalability, reliability, automation, and operational readiness.
Microsoft Project, Kubernetes, Linux
1w
Save
Mark Applied
Hide
Sr. Systems Engineer, Site Reliability
Dallas, Texas, United States
HybridFull Time
American Heart Association
American Heart Association: Nonprofit organization dedicated to fighting heart disease and stroke.
5+ YOEBachelor's degree or equivalent experience, minimum 5 years relevant experience, strong expertise in multi-cloud operations, IAM, security, automation, and infrastructure as code.
Azure, AWS, GCP, Oracle, Entra ID, Structured Query Language (SQL)
1mo
Save
Mark Applied
Hide
Senior DevOps & Infrastructure Lead
Dallas, Texas, United States
$100k-$130k/yr RemoteFull Time
RV LIFE
RV LIFE: Software and tools for planning and navigating RV trips.
Hands-on senior DevOps experience administering Linux servers and managing DigitalOcean/AWS/Cloudflare environments; strong IaC, DB migration, observability, CI/CD, scripting, runbooks, and on-call reliability skills.
DigitalOcean, AWS, Cloudflare, DigitalOcean App Platform, Laravel Cloud, Cloudflare Pages, Cloudflare Workers, Datadog, Terraform, Pulumi, AWS CDK, Serverless Framework, CloudFormation, Jira, GitHub, ChatGPT, Claude, Cursor, GitHub Copilot, Linux, MySQL, MongoDB, OpenSearch, Redis, Node.js, React, React Native, PHP, Laravel, IAM, CloudWatch, S3, SSM, Secrets Manager, Lambda, VPC, WAF
1mo
Save
Mark Applied
Hide
Operations Engineering Manager, Fleet Reliability
Dallas or Bellevue
$143k-$191k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
7+ YOE2+ Mgmt7+ years in software or infrastructure engineering with 2+ years leadership; SRE fundamentals, incident management, observability, change management; strong automation and people development skills.
1mo
Save
Mark Applied
Hide
Software Engineering Manager - Site Reliability Center
Pittsburgh or Cleveland or Birmingham or Dallas or Denver or Phoenix
$100k-$204k/yr OnsiteFull Time
PNC Financial Services
PNC Financial ServicesNYSE: PNC: Provides banking, lending, and investment services to customers.
5+ YOE3+ MgmtLead SRE teams to ensure reliability, incident and change management, production support, automation, observability, and performance; 5+ years related experience with 3+ years management; hands-on with monitoring, cloud/infrastructure, databases and automation.
Dynatrace, BigPanda, Logscale, Linux, Windows, Oracle, SQL, MongoDB, Cassandra, Elasticsearch, Redis, MQ, Kafka, OCP, ShiftPlanning, ETL
1mo
Save
Mark Applied
Hide
Software Engineering Manager-Site Reliability Center-Twilight
Cleveland or Birmingham or Pittsburgh or Dallas or Denver or Phoenix
$100k-$204k/yr OnsiteFull Time
PNC Financial Services
PNC Financial ServicesNYSE: PNC: Provides banking, lending, and investment services to customers.
5+ YOE3+ Mgmt5+ years related experience and 3+ years management; SRE/production support/DevOps experience; incident/problem/change management; hands-on monitoring, cloud/infrastructure, automation; experience with Linux/Windows and databases.
Dynatrace, BigPanda, Logscale, OCP, Linux, Windows, Oracle, SQL, MongoDB, Cassandra, Elasticsearch, Redis, MQ, Kafka, ShiftPlanning
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Robinhood Command Center
New York or Menlo Park or Bellevue or Washington or Denver or Westlake or Chicago or Lake Mary or Clearwater or Gainesville
$196k-$230k/yr HybridFull Time
Robinhood
RobinhoodNASDAQ: HOOD: Provides a commission-free platform for investing and financial services.
5+ YOE5+ years software engineering, 2+ years reliability/infrastructure or production operations, incident leadership experience, deep reliability and observability knowledge, familiarity with OpenTelemetry/Prometheus/Grafana, strong communication and mentoring skills.
OpenTelemetry, Prometheus, Grafana
2w
Save
Mark Applied
Hide
Senior Lead Software Engineer-AI Foundation Services
Plano or Jersey City or Wilmington or McLean
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years software engineering experience building cloud-native AI/ML platform services with Kubernetes, CI/CD, and infrastructure-as-code; proficiency in Python/Java/Go; strong production reliability and secure-by-design practices.
Kubernetes, CI/CD, infrastructure-as-code, Python, Java, Go, GPU
6d
Save
Mark Applied
Hide
Engineering Manager, Data Feeds
New York City or Boston or Chicago or Salt Lake City or Austin or Montreal or Canada or Washington or Dallas or United States or Vancouver or Toronto or Charlotte or Denver
$129k-$304k/yr RemoteFull Time
Chainlink Labs
Chainlink Labs: Building decentralized oracle networks for blockchain smart contracts.
Deep experience building and operating production blockchain infrastructure, strong engineering judgment, leadership and coaching experience, stakeholder management, and ability to deliver reliable, scalable systems.
1mo
Save
Mark Applied
Hide
Senior Director, Platform Engineering - Enterprise Technology
Dallas, Texas, United States
HybridFull Time
Crunchyroll
Crunchyroll: Operates a global streaming platform for anime and manga.
15+ YOE7+ Mgmt15+ years in platform engineering/infrastructure/DevOps/SRE with 7+ years managing managers; experience with distributed systems, multi-cloud (AWS, GCP), platform strategy, reliability (SLIs/SLOs), automation, and executive stakeholder engagement.
AWS, GCP
1mo
Save
Mark Applied
Hide
Manager III, Software Development - Content Data Platform
Austin or Bozeman or Denver or Minneapolis or Missoula or Portland or Salt Lake City or Seattle or Charlotte or Kalispell or Boise or Charleston or Dallas or Fort Worth or Phoenix or Richmond or Spokane or Vermont
$150k-$188k/yr RemoteFull Time
onXmaps
onXmaps: Digital mapping and navigation apps for outdoor recreation.
5+ Mgmt5+ years managing software engineers; experience building data platform/infrastructure and migration off legacy systems; hiring and team development skills; experience with data quality, SLOs, and operational reliability; US work authorization required; ability to travel.
GCP, Managed Spark, BigQuery, BigLake, Pub/Sub, Managed Airflow, Apache Iceberg, Spark, PySpark, DuckDB, dbt, Airflow, GDAL, PostGIS, Apache Sedona, DuckDB spatial extensions, Knowledge Graph, AI-assisted tools