59 observability engineer jobs at 37 companies in Mabank, TX
1w
Save
Mark Applied
Hide
1w
Senior Systems Operations Engineer
Charlotte or Des Moines or Minneapolis or Dallas
$42-$74/hrHybridFull Time
Wells FargoNYSE: WFC: Global provider of banking, investment, and mortgage financial services.
4+ YOE4+ years systems engineering experience, 3+ years production application support, observability tooling, Java/.NET and RDBMS expertise, incident/problem/change management, automation and CI/CD experience.
AppDynamics, Splunk, Elastic, BigPanda, Grafana, Microsoft Application Insights, Java, .NET, Oracle, MSSQL, MongoDB, Jenkins, Artifactory, Harness, Terraform, Ansible, F5, AVI
Dallas or Austin or Houston or San Antonio or Naperville or Nashville or Salt Lake City or Denver or Winchester
OnsiteFull Time
AG&E Associates: Provides integrated structural engineering services for complex building projects.
5+ YOEPE license required with 5+ years construction experience; SE preferred; site observations; coordination with contractors; knowledge of steel and concrete; seismic design.
Gravitate: AI-powered software for fuel supply and logistics optimization.
6+ YOE6+ years platform/DevOps experience with Kubernetes, Terraform, GCP, CI/CD (GitHub Actions, ArgoCD/Flux), Python, observability stacks, and cloud security/IAM; strong communication and automation skills.
3+ YOEMid-level SRE with proficiency in Java, Python or Perl; experience with Linux, SDLC, observability tools (Prometheus, Grafana, ELK, OpenTelemetry) and cloud (AWS/Azure/GCP); strong communication and problem-solving; 3+ years preferred.
DTCC: Provides post-trade infrastructure for the global financial services industry
6+ YOE6+ years systems engineering experience with endpoint/patch management, observability, telemetry analysis, PowerShell/KQL/Power BI skills; bachelor\u0002s preferred.
SCCM/MECM, Microsoft Intune, Windows Update for Business, PowerShell, KQL, Log Analytics, Power BI, SysTrack, Lakeside, ZDX, Azure, AVD, Windows 365, VDI
Vanguard: Global investment management and financial services provider.
Experience with observability, monitoring, reliability metrics, alerting, automation, resilience engineering, incident response, and production troubleshooting; Python-based automation and chaos engineering experience are mentioned.
Splunk, Honeycomb, Amazon CloudWatch, Dynatrace, AppDynamics, Python, Blue Prism, UiPath
Senior Lead Site Reliability Engineer - AI/ML and Data Platforms
Jersey City or Dallas
$171k-$260k/yrOnsiteFull Time
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied SRE experience, strong SLI/SLO/SLA and observability knowledge, experience with Grafana/Dynatrace/Prometheus/Datadog/Splunk, distributed systems expertise, mentoring and leadership experience, familiarity with safe AI usage in operations.
Pure StorageNYSE: PSTG: Provides all-flash enterprise data storage and management solutions.
Production Python engineering experience, enterprise DDI (DNS/DHCP/IPAM) and datacenter networking expertise, automation with Ansible/Puppet, CMDB integration, observability and API development.
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE3+ MgmtRequires 5+ years in solution or sales engineering and 3+ years managing technical presales teams. Bachelor's degree or equivalent experience, enterprise sales expertise, and up to 30% travel required.
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
12+ YOE12+ years in API development and enterprise integration with deep MuleSoft expertise, API security, event-driven architectures, observability, databases, Agile/DevSecOps and strong stakeholder communication.
San Francisco or Ottawa or Phoenix or Toronto or Los Angeles or Denver or Salt Lake City or Atlanta or Chicago or Houston or Portland or New York City or Vancouver or San Diego or Sacramento or Jacksonville or Seattle or Mexico City or Austin or Miami or Boston or Dallas or Charlotte
$141k-$267k/yrHybridFull Time
Scribd: Subscription-based digital library for e-books, audiobooks, and documents.
Significant backend engineering experience building and scaling distributed systems; strong coding in Ruby/Scala/Go/Python; expertise in reliability, observability, SLOs, and mentoring; cross-functional collaboration experience.
KyndrylNYSE: KD: Manages and modernizes mission-critical IT infrastructure systems.
Production-grade Python and advanced SQL, ETL/ELT pipeline design with Airflow/dbt/Kafka, cloud deployments (AWS/Azure/GCP), vector DB experience, data observability, and strong communication skills.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
10+ YOE5+ MgmtBachelor's in a technical field,10+ years engineering experience with 5+ years leading SRE/Platform teams; experience with observability, incident management, distributed systems, and cloud architecture.
AWS, New Relic, Splunk, Datadog, Sentry, Honeycomb, Grafana, Prometheus, OpenTelemetry
St. Petersburg or Detroit or St. Louis or Richmond or Raleigh or Dallas or Indianapolis or Minneapolis
RemoteFull Time
Kobie: Designs and manages customer loyalty programs for global brands.
3+ YOE3+ years professional Python, 1+ year LLMs in production, experience with LangChain/LangGraph or similar, observability tools (CloudWatch, LangSmith, Langfuse, MLflow, OpenTelemetry), Git and Docker, SQL and Snowflake experience preferred.
Director Reliability, Automation & Performance Engineering
Dallas, Texas, United States
$135k-$315k/yrHybridFull Time
Catalyst Brands: Operates a portfolio of diverse retail clothing and apparel brands.
12+ YOE5+ Mgmt12+ years progressive engineering leadership with 5+ years leading reliability/performance/automation; deep SRE, performance, and automation experience; cloud and observability tool expertise; ability to lead distributed engineering teams.
GapNYSE: GAP: Retailer of apparel and accessories across multiple global brands.
Senior leader overseeing enterprise network engineering; expertise in WAN/LAN/Wi‑Fi, cloud networking, SD‑WAN, IaC, automation, observability; experience in large multi‑site or retail environments; bachelor's in CS or Engineering; advanced degree preferred.
Infrastructure as Code, Automation, DevOps, SRE, Networking automation, Observability
Coralogix: AI-powered observability and security data platform.
5+ YOE5+ years in customer‑facing technical roles; hands‑on deployment, configuration, integration, and troubleshooting; deep observability and cloud‑native experience; coding/scripting (Python, Go, Java, JavaScript, Bash); strong communication and commercial judgment.