28 cloud reliability engineer jobs at 17 companies in Wills Point, TX

3d
Save
Mark Applied
Hide
Cloud Site Reliability Engineer
Dallas, Texas, United States
$85-$90/hr RemoteContract
Stefanini
Stefanini: Global provider of IT consulting and digital business solutions.
7+ YOEBachelor's degree or equivalent experience; 7+ years software development, 5+ Python, 3+ AWS and SRE experience. Requires Terraform, CI/CD, testing, observability, Agile, and ITSM expertise.
Terraform, AWS, Python, CI/CD, Grafana, AWS CloudWatch, AWS Canary, GoLang, EC2, VPC, S3, Lambda, IAM, CloudFormation, EventBridge, Step Functions
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with GCP exposure, Terraform/IaC, CI/CD, observability, DevSecOps practices, incident response, and automation for reliability and resiliency.
GCP, Azure, Terraform, Terraform Enterprise, Log Analytics, Dynatrace, Resource Graph, CI/CD, IAM
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Idaho or Plano
$117k-$209k/yr RemoteFull Time
Autodesk
AutodeskNASDAQ: ADSK: Developing software for architecture, engineering, and entertainment industries.
7+ YOEU.S. citizen with 7+ years SRE/platform/cloud experience, strong reliability engineering skills (SLOs/SLIs, observability, incident management), cloud experience (AWS/Azure), automation using Python/Go/Java/PowerShell/Bash, and ability to operate production services in regulated GovCloud environments.
AWS, Azure, Python, Go, Java, PowerShell, Bash, Infrastructure as Code, CI/CD, Splunk, Dynatrace, Datadog, CloudWatch, Kubernetes
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides banking, investment, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with Terraform and GCP; strong IaC, observability, automation, incident response, and DevSecOps skills; ability to define SLIs/SLOs and mentor engineers.
Terraform, Terraform Enterprise, Google Cloud Platform (GCP), Azure, Log Analytics, Dynatrace, Resource Graph, CI/CD
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Research Triangle Park or San Jose or Milpitas or Richardson or Santa Clara
$127k-$182k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE5+ years SRE/Cloud Ops experience, Docker and Kubernetes proficiency, scripting in Python/Go/Bash, monitoring and incident response experience, Linux and networking knowledge, CI/CD and IaC familiarity, bachelor’s degree or equivalent.
Docker, Kubernetes, Python, Go, Bash, Git
3w
Save
Mark Applied
Hide
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$315k/yr HybridFull Time
Catalyst Brands
Catalyst Brands: Operates a portfolio of diverse retail clothing and apparel brands.
12+ YOE5+ Mgmt12+ years progressive engineering leadership with 5+ years leading reliability/performance/automation; deep SRE, performance, and automation experience; cloud and observability tool expertise; ability to lead distributed engineering teams.
AWS, Azure, Kubernetes, EKS, Jenkins, Kafka, Dynatrace, Splunk, Datadog, New Relic, Catchpoint, CloudWatch, ELK
2mo
Save
Mark Applied
Hide
Compliance Engineering, Site Reliability Engineer SRE, Associate, Dallas
Dallas, Texas, United States
OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Provides investment banking, securities, and wealth management services globally.
3+ YOEMid-level SRE with proficiency in Java, Python or Perl; experience with Linux, SDLC, observability tools (Prometheus, Grafana, ELK, OpenTelemetry) and cloud (AWS/Azure/GCP); strong communication and problem-solving; 3+ years preferred.
Java, Python, Perl, Prometheus, Grafana, ELK, OpenTelemetry, AWS, Azure, GCP, Linux, Hadoop
2d
Save
Mark Applied
Hide
Site Reliability Engineer III
Plano, Texas, United States
OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
3+ YOERequires SRE training or certification and 3+ years of experience, with programming or infrastructure-as-code skills, observability, cloud, containers, CI/CD, networking, and SLO/SLI expertise.
Python, Ansible, Terraform, Grafana, Dynatrace, Prometheus, Datadog, Splunk, Kubernetes, ECS, Docker, Jenkins, GitLab, Linux, Windows
4d
Save
Mark Applied
Hide
Site Reliability Engineer Intern
Dallas, Texas, United States
OnsiteFull Time, Internship
Copart
CopartNASDAQ: CPRT: Provides global online vehicle auction and remarketing services.
Experience with production incident management, Linux, Windows, scripting, automation, monitoring, troubleshooting, and observability tools. Strong communication and analytical skills required; programming, virtualization, and cloud experience preferred.
Python, Ansible, Datadog, Kubernetes, Linux, Windows, VMware vSphere, Unix, AWS, GCP
1w
Save
Mark Applied
Hide
Engineering - SRE Platforms - Site Reliability Engineer - Vice President - Dallas
Dallas, Texas, United States
OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
6+ YOERequires 6+ years in site reliability engineering, programming in Java, Python, or Go, cloud and container expertise, IaC and configuration management skills, Linux and distributed systems knowledge, and advanced monitoring experience.
Java, Python, Go, AWS, GCP, Docker, Kubernetes, Terraform, CloudFormation, Puppet, Chef, Ansible, Prometheus, Grafana, ELK, Datadog, PagerDuty, Jenkins, GitLab, Maven, Elastic Search, GCP Big Query, Kafka, Linux, Infrastructure as Code (IaC), Prompt Engineering, Retrieval-Augmented Generation (RAG), CI/CD
3w
Save
Mark Applied
Hide
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$225k/yr HybridFull Time
Catalyst Brands
Catalyst Brands: Operates a portfolio of retail brands including JCPenney.
12+ YOE5+ Mgmt12+ years engineering leadership with 5+ years leading reliability, performance or automation functions; deep SRE, performance engineering, automation, cloud and observability experience; BA/BS preferred.
AWS, Azure, Kubernetes/EKS, Jenkins, Kafka, Dynatrace, Splunk, Datadog, New Relic, Catchpoint, CloudWatch, ELK
3w
Save
Mark Applied
Hide
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$225k/yr HybridFull Time
JCPenney
JCPenney: Retailer of apparel, home goods, and beauty products.
12+ YOE5+ Mgmt12+ years progressive engineering leadership with 5+ years leading reliability/performance/automation; deep SRE, performance engineering, automation, observability and cloud experience; strong architecture and incident management skills.
AWS, Azure, Kubernetes/EKS, Jenkins, Kafka, CDN, Dynatrace, Splunk, Datadog, New Relic, Catchpoint, CloudWatch, ELK
1mo
Save
Mark Applied
Hide
Sr. Systems Engineer, Site Reliability
Dallas, Texas, United States
HybridFull Time
American Heart Association
American Heart Association: Nonprofit organization dedicated to fighting heart disease and stroke.
5+ YOEBachelor's degree or equivalent experience, minimum 5 years relevant experience, strong expertise in multi-cloud operations, IAM, security, automation, and infrastructure as code.
Azure, AWS, GCP, Oracle, Entra ID, Structured Query Language (SQL)
1w
Save
Mark Applied
Hide
Distinguished Engineer - Global Payment Network
McLean or Richmond or Riverwoods or Dallas
$245k-$307k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Provides credit card, banking, and auto loan services.
7+ YOEBachelor's degree, 7+ years of software engineering, 5+ years designing distributed systems, and 5+ years with public cloud technologies. Strong expertise in performance, reliability, scalability, and operability required.
agentic AI, Java, Python, Go, JavaScript, TypeScript, Swift, PCI-DSS
1w
Save
Mark Applied
Hide
Sr. Distinguished Engineer - Global Payment Network (Remote Eligible)
Chicago or Dallas or United States or McLean or Richmond or Riverwoods or Chicago or Dallas or McLean or Richmond or Riverwoods
$286k-$359k/yr RemoteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
9+ YOEBachelor's degree and 9+ years in software engineering, including 7+ years designing distributed systems and using public cloud technologies; strong systems design, coding, reliability, and technical leadership expertise.
Java, Python, Go, JavaScript, TypeScript, Swift, Microsoft AI, PCI-DSS
1w
Save
Mark Applied
Hide
Sr Software Engineer (Site Reliability) Austin or Dallas, TX
Austin or Dallas
HybridFull Time
H-E-B
H-E-B: Operates a major supermarket chain in Texas and Mexi
5+ YOERequires 5+ years with distributed systems, 3+ years in SRE, Terraform, and CI pipelines, plus experience with cloud, Kubernetes, Docker, Linux, databases, monitoring, APIs, and scripting languages.
Google Kubernetes Engine, K8s, AWS, Java, Spring, Terraform, Gitlab Pipelines, GitHub Actions, Gitlab, JIRA, Slack, Confluence, Intellij, microservices, PostgreSQL, Kubernetes, Docker, Linux, GCP, REST, GraphQL, Datadog, Grafana, New Relic, Python, Ruby, Groovy, Bash, Agile
1w
Save
Mark Applied
Hide
Sr Software Engineer (Site Reliability) Austin or Dallas, TX
Austin or Dallas
HybridFull Time
HEB
HEB: A grocery retailer providing food, household products, and related services.
5+ YOERequires 5+ years designing or troubleshooting distributed systems, 3+ years in SRE and Terraform, CI pipeline experience, software architecture expertise, and proficiency with cloud, container, monitoring, and scripting technologies.
Google Kubernetes Engine, Kubernetes, AWS, Java, Spring, Terraform, GitLab Pipelines, GitHub Actions, GitLab, JIRA, Slack, Confluence, IntelliJ, PostgreSQL, Docker, Linux, Google Cloud Platform (GCP), REST, GraphQL, Datadog, Grafana, New Relic, Python, Ruby, Groovy, Bash
1mo
Save
Mark Applied
Hide
Software Engineering Manager - Site Reliability Center
Pittsburgh or Cleveland or Birmingham or Dallas or Denver or Phoenix
$100k-$204k/yr OnsiteFull Time
PNC Financial Services
PNC Financial ServicesNYSE: PNC: Provides banking, lending, and investment services to customers.
5+ YOE3+ MgmtLead SRE teams to ensure reliability, incident and change management, production support, automation, observability, and performance; 5+ years related experience with 3+ years management; hands-on with monitoring, cloud/infrastructure, databases and automation.
Dynatrace, BigPanda, Logscale, Linux, Windows, Oracle, SQL, MongoDB, Cassandra, Elasticsearch, Redis, MQ, Kafka, OCP, ShiftPlanning, ETL
2w
Save
Mark Applied
Hide
Site Rel Eng III, GCP
Bethpage or Plano or Long Island or New York City
$134k-$220k/yr OnsiteFull Time
Optimum
OptimumNYSE: OPTU: Provides broadband, television, and mobile connectivity services to customers.
8+ YOERequires 8+ years in SRE, platform, cloud, DevOps, or network engineering and 5+ years operating GCP production workloads, with expertise in GKE, Kubernetes, Terraform, networking, IAM, observability, and automation.
Google Cloud Platform (GCP), GKE, Cloud Run, Compute Engine, Cloud SQL, Pub/Sub, Cloud Storage, Cloud Monitoring, Network Connectivity Center (NCC), Interconnect, Cloud VPN, Cloud Router, Shared VPC, Private Service Connect, Terraform, Kubernetes, IAM, Python, Bash, GitOps, FinOps, AI/ML
1w
Save
Mark Applied
Hide
Tech Lead - Java/ Scala/ Akka
Chicago or Miami or Atlanta or Dallas or New Jersey or Nashville
$82k-$193k/yr OnsiteFull Time
Capgemini
CapgeminiEuronext Paris: CAP: Provides global IT consulting and digital transformation services.
10+ YOERequires 10+ years of backend engineering, Java/JVM expertise, scalable microservices, Spring, Kafka, distributed systems, technical leadership, production reliability, and cross-team influence.
Java, Scala, Go, Spring Framework, Spring Boot, Play Framework, Akka, Apache Kafka, Docker, Kubernetes, CI/CD, Cloud Platforms, REST API, Electronic Medical Records (EMR)