316 service reliability engineer jobs at 204 companies in United States

1mo
Save
Mark Applied
Hide
Service Reliability Engineer (SRE)
Seattle, Washington, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designing and manufacturing consumer electronics, software, and digital services.
Design and build secure end-to-end server-side solutions and APIs for large-scale media and service platforms; collaborate across teams to deliver reliable services globally.
2mo
Save
Mark Applied
Hide
Reliability Engineer, Truck Service
United States or Ohio
RemoteFull Time
BP
BPLondon Stock Exchange: BP: Integrated energy producing low-carbon solutions and fuels.
Bachelor's in engineering or equivalent experience; experience in reliability engineering or fleet maintenance; knowledge of RCM, FMEA, RCA; proficiency with data analysis/visualization and advanced Excel; EAM/CMMS experience preferred; up to 75% travel.
Microsoft Office, Microsoft Excel, EAM/CMMS
4d
Save
Mark Applied
Hide
Service Reliability Engineer
Liberty Lake, Washington, United States
$80k-$110k/yr HybridFull Time
OpenEye
OpenEye: Private commercial cloud video surveillance serving businesses with AI-driven analytics, business intelligence, and loss-prevention tools.
1+ YOERequires 1–5 years of related experience, cloud, CI/CD, infrastructure automation, monitoring, scripting or development experience, TCP/IP knowledge, Agile familiarity, and strong problem-solving and communication skills.
AWS, Coralogix, TypeScript, MySQL, CrateDB, Git, Java, JavaScript, C#, C++, Datadog, Prometheus, Grafana, Jira
3w
Save
Mark Applied
Hide
Service Reliability Engineer
London or Manchester or New York City
HybridFull Time
Fitch Group
Fitch Group: Global provider of financial information, credit ratings, and analytics.
Deep SRE, DevOps, or platform engineering experience with AWS, Azure, Docker, Kubernetes, Linux, Windows, CI/CD, cloud security, networking, and Python, PowerShell, or Bash.
AWS, Azure, Docker, Kubernetes, Linux, Windows, IIS, .NET, Java Spring Boot, GitHub Actions, Bamboo, Python, PowerShell, Bash, Datadog, Microsoft Teams, AWS Bedrock, SageMaker, Model Context Protocol (MCP), IAM, OPA, AWS Config, AWS CloudTrail, AWS Security Hub, Wiz, DNS, CIS, NIST, ISO 27001
2mo
Save
Mark Applied
Hide
Reliability Engineer - Southwestern PA
Jeannette or Charleroi or Rices Landing
OnsiteFull Time
FirstEnergy Pennsylvania Electric Company
FirstEnergy Pennsylvania Electric Company: Investor-owned electric distribution utility serving Pennsylvania customers.
0+ YOEABET-accredited BS in Engineering/Engineering Technology (or PE/alternate degree), 0+ years engineering experience (0-5), FE preferred, knowledge of NESC/NEC, valid driver's license, able to travel across service territory and work irregular hours.
2mo
Save
Mark Applied
Hide
Service Management Reliability Engineer
O'Fallon, Missouri, United States
$96k-$163k/yr OnsiteFull Time
Mastercard
MastercardNYSE: MA: Global payments technology powering digital economies.
Knowledge of service management, ITSM and ITIL practices; experience with monitoring tools (Splunk, Dynatrace) preferred; strong critical thinking, risk awareness, automation mindset, and cross-functional collaboration skills. Bachelor's in IT/CS/Engineering preferred.
Splunk, Dynatrace, ITIL
2w
Save
Mark Applied
Hide
Platform / Site Reliability Engineer
New York City, New York, United States
OnsiteFull Time
Sunset
Sunset: Private startup wind-down service helping founders close companies through legal, tax, and operational work.
Production cloud infrastructure and reliability experience across multiple services, strong software engineering skills, infrastructure and application coding, incident leadership, recovery expertise, and AI engineering tool proficiency.
AWS, Terraform, CI/CD, SOC 2
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Idaho or Plano
$117k-$209k/yr RemoteFull Time
Autodesk
AutodeskNASDAQ: ADSK: Global provider of software for design, engineering, and manufacturing.
7+ YOEU.S. citizen with 7+ years SRE/platform/cloud experience, strong reliability engineering skills (SLOs/SLIs, observability, incident management), cloud experience (AWS/Azure), automation using Python/Go/Java/PowerShell/Bash, and ability to operate production services in regulated GovCloud environments.
AWS, Azure, Python, Go, Java, PowerShell, Bash, Infrastructure as Code, CI/CD, Splunk, Dynatrace, Datadog, CloudWatch, Kubernetes
2w
Save
Mark Applied
Hide
Reliability Engineer 4 (Observability Specialist )
Chicago or Atlanta or Cupertino or Gresham or Denver or Charlotte or Brookfield or Irving or Hopkins or Earth City
$124k-$146k/yr HybridFull Time
U.S. Bank
U.S. BankNew York Stock Exchange: USB: Diversified financial services and banking institution.
6+ YOEBachelor's degree or equivalent experience and 6–8 years in reliability, SRE, IT service management, production support, application development, or related work; expertise in observability and stakeholder leadership.
Datadog, Dynatrace, Splunk, Grafana, Prometheus, New Relic, Elastic, OpenTelemetry, Kubernetes
1mo
Save
Mark Applied
Hide
Site Reliability Engineer - System Service Global
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Global technology specializing in AI-powered content platforms.
Bachelor's in related field and strong experience with large-scale Linux host management, core data-center services (DNS, NTP, DHCP, NAT, APT, Kerberos), DevOps tooling, SRE practices, and troubleshooting.
BIND, PowerDNS, NTP, DHCP, NAT, APT, Kerberos, Ansible, Salt, Puppet, CI/CD, Python, Go, Bash, Linux
2mo
Save
Mark Applied
Hide
Customer Reliability Engineer
Chicago, Illinois, United States
$103k-$159k/yr HybridFull Time
iManage
iManage: AI knowledge-work software helping legal, accounting, and financial-services organizations manage documents, email, and governed knowledge.
3+ YOE3+ years in technical escalation/Support/CRE/SRE roles; incident response for P1/P2; experience with distributed cloud services; SQL, Python, Bash/Shell, PowerShell, REST APIs; AKS/Azure services; Splunk, Grafana, Kibana, Prometheus; strong communication and troubleshooting.
SQL, Python, Bash/Shell, PowerShell, REST APIs, Azure Kubernetes Service (AKS), Azure services, Splunk, Grafana, Kibana, Prometheus
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
San Francisco, California, United States
$149k-$224k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: The #1 AI CRM driving customer success together.
5+ YOE5+ years systems and software engineering experience for large-scale internet services; expertise in SRE principles, containers, observability, incident management, Python and Go, and applying AI/ML to operations.
Temporal, Airflow, Argo Workflows, Docker, Kubernetes, DNS, HTTP, Grafana, Prometheus, ELK, Splunk, Datadog, Python, Go, Linux, Claude Code, GitHub Copilot, Codex, Cursor, AWS, GCP, MCP
3w
Save
Mark Applied
Hide
Site Reliability Engineer
Hyderabad or New York City or Chicago or London or Singapore or Tokyo or Hong Kong or Europe or United States or Asia-Pacific
HybridFull Time
Pico
Pico: Private financial-markets technology providing trading infrastructure, connectivity, market data, software, and analytics to institutions.
Bachelor's degree or relevant experience; financial markets technology experience; Linux, networking, computer architecture, programming or scripting, customer service, communication, and collaborative teamwork skills.
Linux, Python, C, C++, Java
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (DevOps)
Annapolis Junction or Reston
$112k-$222k/yr OnsiteFull Time
Accenture Federal Services
Accenture Federal ServicesNew York Stock Exchange: ACN: Technology and management consulting for U.S. federal agencies.
5+ YOEBS in CS/CE, 5+ years relevant experience, strong programming and data engineering (SQL, Python, Spark), Azure data services, PKI knowledge, valid US passport, TS/SCI with Polygraph required.
SQL, Python, Azure Data Factory, Azure Synapse, Apache Spark, Azure Data Explorer, Azure Blob Storage, Kusto Query Language (KQL), Git, JSON, Powershell, C#, LLMs
2w
Save
Mark Applied
Hide
Senior Staff Service Reliability and Operational Intelligence Engineer
Santa Clara, California, United States
$188k-$270k/yr OnsiteFull Time
IonQ
IonQNYSE: IONQ: Public quantum technology delivering trapped-ion computing, networking, sensing, and security systems to enterprise and government customers.
12+ YOERequires 12+ years in production, site reliability, platform engineering, or cloud operations; AWS or GCP; distributed systems, Kubernetes, observability, incident response, automation, and cross-team technical leadership.
AWS, GCP, Kubernetes, Python, Go, Jira, Confluence, Amazon Bedrock AgentCore, CI/CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Raleigh, North Carolina, United States
$119k-$196k/yr HybridFull Time
Red Hat
Red Hat: Provider of enterprise open source software solutions.
5+ YOE5+ years operating production services on Kubernetes/OpenShift, 3+ years programming in Python/Go, 2+ years with cloud providers, Linux and networking knowledge, SRE principles and on-call experience.
Python, Go, Splunk, Splunk IM, Prometheus, Grafana, Catchpoint, DataDog, OpenShift, Kubernetes, OpenShift Pipelines, Tekton, GitLab, ArgoCD, GitOps, Operator SDK, Linux, AWS, Google, Azure
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Austin, Texas, United States
$111k-$172k/yr HybridFull Time
Visa
VisaNYSE: V: Global leader in digital payments and transaction technology.
2+ YOE2+ years with a Bachelor's or 5+ years experience; hands-on Azure, Kubernetes, Terraform, IaC/GitOps, CI/CD, observability, service mesh (Istio preferred); strong SRE, troubleshooting, documentation, and English (B2+).
Azure, AWS, Kubernetes, Terraform, GitOps, CI/CD, Istio, App Mesh, Linkerd, Service Mesh
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Newton, Massachusetts, United States
$160k-$205k/yr OnsiteFull Time
Manifold
Manifold: The Enterprise Agent Platform for life sciences that helps biopharma and research teams analyze governed biomedical data.
7+ YOE7+ years in infrastructure/DevOps/SRE with deep cloud (AWS/GCP/Azure), Terraform, CI/CD (Github Action), container tooling, identity systems, data platform services, and experience operating secure multi-account environments.
AWS, GCP, Azure, Terraform, Github Action, Okta, Auth0, Docker, ECS, Packer, Tailscale, WireGuard, Snowflake, Airflow, dbt, PostgreSQL, LLM, CI/CD
4w
Save
Mark Applied
Hide
Sr. Database Reliability Engineer
San Jose, California, United States
$139k-$258k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Empowering everyone to create through innovative digital experiences.
7+ YOE7+ years operating highly available database platforms; strong experience with MongoDB/Cassandra/MySQL/PostgreSQL, cloud (AWS/Azure), managed DB services, IaC (Terraform/Chef/Ansible), Kubernetes/Docker, Python; bachelor's or equivalent experience.
MongoDB, Cassandra, MySQL, PostgreSQL, Percona XtraDB Cluster, MariaDB Galera Cluster, AWS, Azure, Amazon RDS, Keyspaces, DynamoDB, Azure SQL, Cosmos DB, MongoDB Atlas, Terraform, Chef, Ansible, Kubernetes, Docker, Python
2mo
Save
Mark Applied
Hide
Sr. Quality & Reliability Engineer, Hardware Engineering Services
Seattle, Washington, United States
$159k-$215k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Multinational technology focused on e-commerce and cloud computing.
5+ YOEBachelor's in a related field, 5+ years quality/reliability or mechanical design engineering, experience with DFMEA/ALT/HALT, root cause analysis, test plan development, supplier quality, and statistical methods (Weibull).
DFMEA, HALT, ALT, Weibull analysis