40 cloud reliability engineer jobs at 19 companies in Krum, TX
2d
Save
Mark Applied
Hide
2d
Systems Reliability Engineer
Overland Park or Atlanta or Frisco or Bellevue
$84k-$151k/yrOnsiteFull Time
T-MobileNASDAQ: TMUS: Provides wireless network services and mobile communication devices.
2+ YOEBachelor's degree required; 2–4+ years preferred. Requires DevOps, cloud, automation, monitoring, scripting, APIs, cybersecurity, and reliability engineering experience, plus U.S. work authorization.
C, C#, Java, Perl, Python, Go, Jenkins, CloudBees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Microsoft Graph API, REST API, Microsoft Power Apps, Microsoft Power Automate, Microsoft Entra, SailPoint, ServiceNow, Azure, Microsoft Azure DevOps Pipelines, Linux, Shell
5+ YOEBachelor's or equivalent,5+ years deploying/supporting distributed systems,cloud and on-prem storage,Kubernetes (EKS/AKS/RKS),CI/CD automation,observability,backup/recovery,Python/NodeJS/Java and scripting.
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with GCP exposure, Terraform/IaC, CI/CD, observability, DevSecOps practices, incident response, and automation for reliability and resiliency.
AutodeskNASDAQ: ADSK: Developing software for architecture, engineering, and entertainment industries.
7+ YOEU.S. citizen with 7+ years SRE/platform/cloud experience, strong reliability engineering skills (SLOs/SLIs, observability, incident management), cloud experience (AWS/Azure), automation using Python/Go/Java/PowerShell/Bash, and ability to operate production services in regulated GovCloud environments.
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
3+ YOEBachelor's degree and 3 years of related experience, or advanced degree and 1 year; automation, monitoring, CI/CD, cloud computing, scripting, incident management, and system reliability skills required.
Bank of AmericaNYSE: BAC: Provides banking, investment, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with Terraform and GCP; strong IaC, observability, automation, incident response, and DevSecOps skills; ability to define SLIs/SLOs and mentor engineers.
Charles SchwabNYSE: SCHW: Financial services, brokerage, and investment management provider.
10+ YOEBachelor's in CS or related; 10+ years software development/SRE experience (8+ years DevOps/SRE), 8+ years CI/CD and observability, 5+ years leading reliability practices; strong automation, scripting (Python/shell), cloud and distributed systems experience; must be authorized to work in the U.S. without sponsorship.
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
3+ YOEFormal SRE training/certification and 3+ years applied SRE experience; experience with SLO/SLI, observability, CI/CD, cloud, containers, Python/Ansible/Terraform, and networking.
NTT DATATokyo Stock Exchange: 9613: Global provider of business and technology consulting and IT services.
5+ YOE5+ years SRE/DevOps experience with Terraform, cloud and hybrid infrastructure, load balancing (F5/AVI), automation (Python/Ansible/Shell), CI/CD, and production support.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
10+ YOE5+ MgmtBachelor's in a technical field,10+ years engineering experience with 5+ years leading SRE/Platform teams; experience with observability, incident management, distributed systems, and cloud architecture.
AWS, New Relic, Splunk, Datadog, Sentry, Honeycomb, Grafana, Prometheus, OpenTelemetry
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$315k/yrHybridFull Time
Catalyst Brands: Operates a portfolio of diverse retail clothing and apparel brands.
12+ YOE5+ Mgmt12+ years progressive engineering leadership with 5+ years leading reliability/performance/automation; deep SRE, performance, and automation experience; cloud and observability tool expertise; ability to lead distributed engineering teams.
3+ YOEMid-level SRE with proficiency in Java, Python or Perl; experience with Linux, SDLC, observability tools (Prometheus, Grafana, ELK, OpenTelemetry) and cloud (AWS/Azure/GCP); strong communication and problem-solving; 3+ years preferred.
NTT DATA: Global provider of IT and business consulting services.
5+ YOE5+ years SRE/DevOps experience with Terraform, hybrid cloud networking, load-balancing (F5/AVI), automation using Python/Ansible/Shell, CI/CD tooling, and production support including on-call rotations.
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$225k/yrHybridFull Time
Catalyst Brands: Operates a portfolio of retail brands including JCPenney.
12+ YOE5+ Mgmt12+ years engineering leadership with 5+ years leading reliability, performance or automation functions; deep SRE, performance engineering, automation, cloud and observability experience; BA/BS preferred.
Senior AI Ops & Incident/Site Reliability Engineer
Austin or Plano or Fort Mill or Boston or New York City or Tempe or San Diego
$94k-$171k/yrHybridFull Time
Perficient: Provides digital transformation and AI consulting for global enterprises.
8+ YOEBachelor's degree or equivalent,8+ years in IT operations/SRE or production support,5+ years leading incident management,experience with Dynatrace,ServiceNow,cloud platforms and AIOps implementations.
American Heart Association: Nonprofit organization dedicated to fighting heart disease and stroke.
5+ YOEBachelor's degree or equivalent experience, minimum 5 years relevant experience, strong expertise in multi-cloud operations, IAM, security, automation, and infrastructure as code.
Azure, AWS, GCP, Oracle, Entra ID, Structured Query Language (SQL)
Software Engineer Lead - Site Reliability Engineering Center
Pittsburgh or Phoenix or Lakewood or Birmingham or Strongsville or Farmers Branch
$86k-$158k/yrOnsiteFull Time
PNC Financial ServicesNYSE: PNC: Provides banking, lending, and investment services to customers.
3+ YOEBachelors degree and 3+ years experience in software engineering. Experience with Java ecosystem, cloud-native and SRE practices; strong problem solving and communication skills.
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
7+ YOE3+ MgmtRequires 7+ years in systems engineering or technology architecture, 3+ years of management, 5+ years leading engineering or SRE teams, and experience with customer-facing platforms, incident management, SRE, DevOps, cloud, and production operations.
Splunk, Grafana, AppDynamics, Dynatrace, OpenTelemetry, Prometheus, Kubernetes, OpenShift, AWS, Microsoft Azure, Google Cloud Platform, Infrastructure as Code (IaC), CI/CD, AIOps, ITIL