127 application reliability engineer jobs at 92 companies in United States

1w
Save
Mark Applied
Hide
Application Reliability Engineer
United States
$60k-$100k/yr RemoteFull Time
Innodata
InnodataThe Nasdaq Stock Market LLC: INOD: Public data engineering providing AI training, evaluation, safety, and deployment services to technology companies and enterprises.
3+ YOERequires 3–7 years in application, software, cloud engineering, or support; hands-on GCP and Google App Engine experience; microservices, APIs, production operations, programming, SQL, CI/CD, and incident troubleshooting skills.
Google Cloud Platform (GCP), Google App Engine (GAE), Cloud Monitoring, Cloud Logging, Error Reporting, Cloud Trace, REST, gRPC, Protocol Buffers (protobuf), Pub/Sub, Cloud Tasks, Google Cloud SQL, Firestore, Datastore, Cloud Spanner, BigQuery, Cloud Build, Jenkins, GitHub Actions, GitLab CI, Cloud Run, GKE, Cloud Functions, Apigee, API Gateway, Cloud Scheduler, Terraform, Docker, Kubernetes, Looker, Tableau, Vertex AI, Python, Java, Node.js, JavaScript, Go, SQL
1mo
Save
Mark Applied
Hide
Application & System Reliability Engineer (68306)
United States
$143k-$210k/yr RemoteFull Time
Eaton
EatonNYSE: ETN: Intelligent power management providing energy-efficient solutions.
7+ YOEBachelor's degree, minimum seven years electrical sales/application engineering experience, valid driver\u0002s license, US work authorization without sponsorship, ability to travel up to 50%.
DFMEA, 8D
1w
Save
Mark Applied
Hide
Reliability Engineer Sr
Stratford, Connecticut, United States
$96k-$178k/yr OnsiteFull Time
Sikorsky Aircraft Corporation
Sikorsky Aircraft CorporationNYSE: LMT: Lockheed Martin-owned aerospace manufacturer designing, building, and supporting military and commercial helicopters.
3+ YOEBachelor’s degree in engineering or related field with 3–5 years of related experience or equivalent. Requires reliability analysis, statistics, quality assurance, data analytics, or application development experience.
Weibull analysis, Design of Experiments (DOE), JMP, Python, R, MATLAB, Failure Reporting, Analysis, and Corrective Action System (FRACAS), Failure Modes, Effects, and Criticality Analysis (FMECA), Fault Tree Analysis (FTA), EVMS, 1LMX
3mo
Save
Mark Applied
Hide
Reliability Engineer
McLean, Virginia, United States
$100k-$125k/yr OnsiteFull Time
Amentum
AmentumNYSE: AMTM: Global provider of engineering, technical, and mission-critical services.
5+ YOEBachelor's degree in engineering or a related scientific field plus 5 years of job-related experience or equivalent; strong communication, analytical, computer, and integrated software application skills; ability to obtain TS/SCI clearance with polygraph.
computer systems
2mo
Save
Mark Applied
Hide
Site Reliability Engineer - Application Support (Director)
New York, New York, United States
$120k-$165k/yr OnsiteFull Time
Morgan Stanley Wealth Management
Morgan Stanley Wealth ManagementNYSE: MS: Private wealth-management business serving individuals, families, businesses, institutions and foundations with advice, brokerage and financial planning.
5+ YOE3+ Mgmt5+ years supporting or developing enterprise applications, 3+ years leading teams, DevOps/SRE experience, observability tools, AWS services, scripting in Python, knowledge of Java and database engineering, on-call support experience.
Grafana, Prometheus, Splunk, Kibana, EC2, ECS, S3, Fargate, Aurora, Lambda, Python, Java, Perl, PowerShell, Unix scripting, C#, .NET, JavaScript, React, GraphQL, Django, Celery, PostgreSQL, Golang, ElasticSearch, RabbitMQ, Kafka, Adobe Experience Cloud, ITIL
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (Application Software)
Hawthorne, California, United States
$125k-$195k/yr OnsiteFull Time
SpaceX
SpaceXNasdaq: SPCX: Designing, manufacturing, and launching advanced rockets and spacecraft.
3+ YOE3+ years SRE/DevOps experience or bachelor's in CS/IS/engineering; 1+ year Python; Linux experience; familiarity with build systems, containers, databases, IaC, and effective communication.
Python, Linux, Bazel, Buck, Make, Docker, Kubernetes, vSphere, QEMU, KVM, Postgres, MySQL, ClickHouse, Terraform, Ansible, Puppet, JavaScript, C#, C++
3mo
Save
Mark Applied
Hide
Application Reliability & Support Engineer
Fort Meade, Maryland, United States
$87k-$182k/yr OnsiteFull Time
CACI
CACINYSE: CACI: Provider of specialized IT and mission-critical government services.
2+ YOETS/SCI with Polygraph; 10+ years with HS; 8 with an Associate; 6 with a Bachelor; 4 with a Master; 2 with a Doctorate; Linux admin; AWS; monitoring tools; 24/7 support.
Grafana, Kibana, Splunk, CloudWatch, AWS Console, AWS CLI, NiFi, Cribl, Kafka, Logstash, Terraform, Ansible, GitLab Pipelines, Docker
2w
Save
Mark Applied
Hide
Electrical Reliability Engineer
Hampton, Virginia, United States
$105k-$140k/yr OnsiteFull Time
Amentum
AmentumNYSE: AMTM: Global provider of engineering, technical, and mission-critical services.
4+ YOEBachelor's degree in electrical engineering or related discipline with 4–5 years' experience, or advanced degree with applicable experience; U.S. citizenship, clearance eligibility, driver's license, and electrical reliability expertise required.
Microsoft Office, Microsoft Word, Microsoft Excel, Microsoft Outlook, IBM Maximo, CMMS, FMEA, RCM, RCA, RCFA
1w
Save
Mark Applied
Hide
Senior Maintenance Reliability Engineer
Bridgewater, Virginia, United States
$101k-$150k/yr OnsiteFull Time
Perdue Farms
Perdue Farms: Fourth-generation, family-owned food and agricultural.
5+ YOEBachelor's degree in reliability, mechanical, or electrical engineering; 5+ years of industrial manufacturing reliability experience; CMMS, RCA, RCM, precision maintenance, and Microsoft applications experience required.
Microsoft Word, Microsoft Excel, Microsoft Outlook, IBM Maximo, Computerized Maintenance Management System (CMMS), Root Cause Analysis (RCA), Reliability Centered Maintenance (RCM), Root Cause Failure Analysis (RCFA), Failure Mode Effects Analysis (FMEA), Lockout/Tagout
3w
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Buffalo, New York, United States
$140k-$233k/yr OnsiteFull Time
M&T Bank
M&T BankNYSE: MTB: A diversified financial services providing banking and wealth management.
7+ YOEExpert in reliability engineering, SLO/SLI frameworks, incident and problem management, observability, automation, cloud platforms, and production operations; 7+ years systems analysis/application development or equivalent.
AWS, Azure, CI/CD, SDLC, SLO/SLI
1mo
Save
Mark Applied
Hide
Reliability Engineer
York, Pennsylvania, United States
OnsiteFull Time
RHI Magnesita
RHI MagnesitaVienna Stock Exchange: RHIM: Global leader in high-grade refractory products and solutions.
5+ YOEB.S. in Engineering, 5+ years maintenance experience, TPM and RCM experience, CMMS (SAP) and Microsoft applications proficiency, mechanical aptitude, RCA/FMEA experience.
Microsoft applications, CMMS, SAP
2w
Save
Mark Applied
Hide
Platform / Site Reliability Engineer
New York City, New York, United States
OnsiteFull Time
Sunset
Sunset: Private startup wind-down service helping founders close companies through legal, tax, and operational work.
Production cloud infrastructure and reliability experience across multiple services, strong software engineering skills, infrastructure and application coding, incident leadership, recovery expertise, and AI engineering tool proficiency.
AWS, Terraform, CI/CD, SOC 2
1w
Save
Mark Applied
Hide
Reliability Engineer 4 (Observability Specialist )
Chicago or Atlanta or Cupertino or Gresham or Denver or Charlotte or Brookfield or Irving or Hopkins or Earth City
$124k-$146k/yr HybridFull Time
U.S. Bank
U.S. BankNew York Stock Exchange: USB: Diversified financial services and banking institution.
6+ YOEBachelor's degree or equivalent experience and 6–8 years in reliability, SRE, IT service management, production support, application development, or related work; expertise in observability and stakeholder leadership.
Datadog, Dynatrace, Splunk, Grafana, Prometheus, New Relic, Elastic, OpenTelemetry, Kubernetes
3mo
Save
Mark Applied
Hide
System Reliability Engineer, III - VI
Tucker, Georgia, United States
OnsiteFull Time
Georgia Transmission Corporation
Georgia Transmission Corporation: Nonprofit electric cooperative that plans, builds, and maintains high-voltage transmission infrastructure for Georgia member cooperatives.
4+ YOEBS Electrical Engineering; multiple E-level experience in power utility; PE/EIT applicable but not required; proficient in Excel, PowerBI, data tools; strong communication and planning skills.
Microsoft Excel, Power BI, Microsoft Access, Microsoft Word, PowerPoint
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Reston or Austin
$81k-$187k/yr OnsiteFull Time
Oracle Corporation
Oracle CorporationNYSE: ORCL: Cloud infrastructure and enterprise software solutions provider.
3+ YOERequires 3+ years SRE/operations experience, deep Linux knowledge, scripting and automation (Python, Bash, Perl, Java, Go), and TS/SCI clearance for US applicants.
Linux, Unix, Docker, Kubernetes, Terraform, Python, Bash, Perl, Java, Go, Ruby, JavaScript, Chef, Puppet
1mo
Save
Mark Applied
Hide
Senior AI Site Reliability Engineer, AI.x
Austin, Texas, United States
$170k-$220k/yr OnsiteFull Time
Charles Schwab
Charles SchwabNYSE: SCHW: Financial services firm providing brokerage, banking, and advisory services.
8+ YOE8+ years software engineering experience with 4+ years SRE, experience with cloud-native containers, IaC, CI/CD, observability, on-call rotations, and deploying/operating LLM-powered applications.
Terraform, Google Cloud Platform, Gemini, Claude, OpenAI
6d
Save
Mark Applied
Hide
Site Reliability Engineer
Vancouver or Toronto or Tel Aviv or Plano or London or Amsterdam or Tbilisi or Medellín or Foster City
$100k-$125k/yr HybridFull Time
Tipalti
Tipalti: Private fintech providing accounts-payable, payments, procurement, expense, and treasury automation for mid-market businesses.
4+ YOERequires 4+ years of software engineering experience, production deployment and application lifecycle ownership, architecture knowledge, troubleshooting skills, and English communication. SRE, cloud, monitoring, and scripting experience preferred.
.NET, TypeScript, OpenTelemetry, Prometheus, Bash, PowerShell, Python, AWS, GCP, Azure
2w
Save
Mark Applied
Hide
Database Reliability Engineer
Livermore, California, United States
$146k-$223k/yr HybridFull Time, Contract
Lawrence Livermore National Laboratory
Lawrence Livermore National Laboratory: Federally funded research and development center for national security.
Bachelor's degree or equivalent experience; broad database reliability, automation, administration, monitoring, application server, operating system, storage, backup, and recovery experience; U.S. citizenship and ability to obtain DOE Q clearance.
Oracle, MySQL, Microsoft SQL Server, MongoDB, Cassandra, PostgreSQL, DynamoDB, WebLogic, Tomcat, SQL, Oracle Forms/Reports, PL/SQL, Java, Python, Windows, Linux, Oracle Enterprise Manager, Datadog, Grafana, SQL Server Management Studio, Icinga, Ansible, SCCM, PowerShell
3mo
Save
Mark Applied
Hide
Sr. Reliability Engineer, Power Modules and AI Applications
San Jose or Kirkland
$115k-$170k/yr OnsiteFull Time
Monolithic Power Systems
Monolithic Power SystemsNASDAQ: MPWR: Leading provider of high-performance power semiconductor solutions.
5+ YOE5+ years in reliability testing/qualification of IC devices and packages; MSc/PhD in materials/EE/physics/mechanical; JMP/Weibull; SEM/CSAM/X-ray/OBIRCH/TIVA; AI data center power architectures a plus.
JMP, Reliasoft, SEM, CSAM, X-ray, OBIRCH/TIVA, Weibull
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer (Arlington, VA) - Secret Clearance Required - Relocation Provided
Arlington, Virginia, United States
$180k-$220k/yr OnsiteFull Time
Onebrief
Onebrief: Private defense-software providing AI-powered planning, collaboration, simulation, and decision-support tools to military staffs.
5+ YOEActive Secret clearance, 5+ years SRE or software engineering with strong TypeScript and application-level reliability experience; CI/CD, incident response, Kubernetes, observability expertise.
TypeScript, Node, Prometheus, Loki, Alloy, Grafana, GitHub Actions, GitLab CI/CD, Jenkins, Python, Go, Bash, kubectl, Kubernetes, Terraform, Ansible, AWS, AWS GovCloud, Grafana stack, ELK, Datadog, Istio, Linkerd, VMware, Proxmox, Nutanix, Hyper-V