163 software reliability engineer jobs at 114 companies in Elmhurst, NY

3w
Save
Mark Applied
Hide
Sr Software Engineer - Reliability Engineering
North Hills, New York, United States
$122k-$203k/yr HybridFull Time
Cox Enterprises
Cox Enterprises: Providing global communications, automotive services, and media solutions.
5+ YOE5+ years in software/platform/infrastructure engineering, strong coding (Python/Go/Java), AWS and Terraform experience, SRE and observability knowledge, system design and incident response skills.
Python, Go, Java, AWS EC2, AWS RDS, AWS DynamoDB, AWS S3, AWS Aurora, AWS Lambda, AWS VPCs, AWS Athena, Terraform, Docker, Kubernetes, Linux, Windows, New Relic, Splunk, Prometheus, CI/CD pipelines
1w
Save
Mark Applied
Hide
Lead Software Reliability Engineer
Irving or New York City or New Jersey or Tampa or Jacksonville
$145k-$175k/yr HybridFull Time, Contract
RE Partners
RE Partners: Providing technology consulting and digital transformation services for enterprise clients.
Several years of TDD experience, strong coding skills, JVM languages and/or TypeScript, secure coding, CI/CD, Docker, OpenShift, Linux, build automation, and critical-systems observability expertise.
Test-Driven Development (TDD), Java, NPM, Spring, Renovate, Liquibase, Ansible, Docker, OpenShift, TypeScript, Gradle, Maven, Microsoft?, Linux, JVM
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1mo
Save
Mark Applied
Hide
Staff Software Engineer, Reliability Engineering & SDLC Governance
Carlsbad or Germantown or San Jose or San Francisco or New York City
$165k-$261k/yr OnsiteFull Time
Viasat
ViasatNASDAQ: VSAT: Provides global satellite broadband and secure networking communication services.
8+ YOE8+ years software engineering experience with SDLC governance, SRE/DevOps knowledge, observability, incident analysis, and cross-team influence to improve reliability and customer experience.
Prometheus, Grafana, OpenTelemetry, Datadog, AS9115
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
1w
Save
Mark Applied
Hide
Platform / Site Reliability Engineer
New York City, New York, United States
OnsiteFull Time
Sunset
Sunset: Handles legal and operational tasks for winding down startups.
Production cloud infrastructure and reliability experience across multiple services, strong software engineering skills, infrastructure and application coding, incident leadership, recovery expertise, and AI engineering tool proficiency.
AWS, Terraform, CI/CD, SOC 2
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer
New York City or Toronto
$184k-$240k/yr OnsiteFull Time
Movable Ink
Movable Ink: Provides AI-powered content personalization for digital marketing campaigns.
6+ YOE6+ years SRE/Software Engineering experience designing and operating scalable, multi-cloud distributed systems; expertise with observability, IaC, Kubernetes, and multiple programming languages.
Apache Pulsar, Apache Kafka, Grafana Loki, ScyllaDB, Cassandra, Prometheus, Thanos, Grafana Alloy, Tempo, Terraform, Chef, EKS, GKE, NodeJS, Golang, Ruby, Python, shell
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
New York or San Ramon or Reno
$153k-$210k/yr HybridFull Time
Ridgeline
Ridgeline: Cloud-native platform for investment management operations.
3+ YOE3–6 years SRE/DevOps experience, 2+ years on AWS, proficiency with Terraform, observability, CI/CD, Python/Go/Bash, incident response, and strong communication and troubleshooting skills.
Claude Code, Cursor, Terraform, AWS, EC2, ECS, EKS, RDS, S3, IAM, CloudWatch, GitHub Actions, CircleCI, Buildkite, Python, Go, Bash, Kubernetes, Helm, Kotlin, Node.js, TypeScript
1w
Save
Mark Applied
Hide
Site Reliability Engineer
New York or Hong Kong or London or Singapore
$105k-$300k/yr OnsiteFull Time
Citadel
Citadel: Global alternative investment management firm
Bachelor's degree in computer science, related STEM discipline, or equivalent experience; proficiency in a modern structured programming language, software development practices, distributed systems, and strong communication skills.
Python, SQL, JavaScript, CSS, React, CI/CD
1d
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
HybridFull Time
Chariot
Chariot: Payment infrastructure for charitable giving and Donor Advised Funds.
4+ YOERequires 4+ years software development, 2+ years backend application development, bachelor's degree preferred, and proficiency with Go, Node, Docker, Terraform, Kubernetes, Postgres, REST APIs, gRPC, and AWS.
Go, Node, Docker, Terraform, Kubernetes, Postgres, REST APIs, gRPC, AWS
2mo
Save
Mark Applied
Hide
Senior Software Engineer - Observability and Reliability
New York City or San Francisco or London or Sydney
$170k-$240k/yr OnsiteFull Time
Sigma Computing
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
5+ YOE5+ years building high-quality software, strong CS fundamentals, experience building observability tools, proficiency with Go, OpenTelemetry, Kubernetes, participation in on-call rotations, cloud service administration (GCP/AWS/Azure) preferred.
Go, OpenTelemetry, Kubernetes, GCP, AWS, Azure, SQL, Python
1d
Save
Mark Applied
Hide
Sr. Staff Engineer Software, Infrastructure Reliability (Chronosphere)
San Francisco or Denver or Austin or Jacksonville or Bridgeport or Seattle or Boston or New York City
$126k-$205k/yr RemoteFull Time
Palo Alto Networks
Palo Alto NetworksNASDAQ: PANW: Provides enterprise-grade network, cloud, and endpoint security software.
8+ YOERequires 8+ years of relevant experience, backend programming proficiency, cloud-native and distributed systems expertise, Linux and networking knowledge, debugging skills, and experience with AWS or GCP and Kubernetes.
Go, Java, Python, Rust, AWS, GCP, Kubernetes, Linux, Terraform, Cursor, Claude
1w
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineer - Remote
Basking Ridge, New Jersey, United States
$113k-$193k/yr RemoteFull Time
UnitedHealth Group
UnitedHealth GroupNYSE: UNH: Provides health insurance and technology-enabled health care services.
10+ YOE5+ MgmtBachelor’s degree in a relevant field, 10+ years in software, SRE, platform, DevOps, infrastructure, or technology operations, and 5+ years leading engineering or operational teams. Requires production, cloud, ITSM, and reliability experience.
Azure, AWS, Infrastructure-as-Code, Kubernetes, OpenShift, Datadog, Splunk, Grafana, Prometheus, OpenTelemetry, AIOps, ChatOps, LLM
1w
Save
Mark Applied
Hide
Site Reliability Engineer, Global Banking & Markets, Vice President
New York City, New York, United States
$150k-$250k/yr OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
8+ YOE8+ years of software or reliability engineering experience; proficiency in a major programming language; cloud, distributed systems, SRE, automation, observability, incident response, and risk management expertise.
Java, GCP, AWS, Kubernetes, Docker, Claude Code, GitHub Copilot Agent Mode, Devin, Gemini Code Assist, Apache Kafka, Spring Boot, gRPC, Protocol Buffers, Apache Camel, Spring Integration, Terraform, Helm, Prometheus, Grafana, OpenTelemetry, SQL, NoSQL, Vert.x, Netty, Microsoft?
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
$100k-$125k/yr OnsiteFull Time
Dirac
Dirac: AI-driven platform automating manufacturing work instructions from CAD files.
Builder mindset with ability to explain projects; eager to learn, re-build, and improve; strong debugging and troubleshooting skills; programming or scripting knowledge; familiarity with container tech and networking.
Kubernetes, Docker, Networking, DNS, Routing, ITAR
2w
Save
Mark Applied
Hide
Site Reliability Engineer, Global Banking & Markets, Vice President
New York City, New York, United States
$150k-$250k/yr OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Provides investment banking, securities, and wealth management services globally.
8+ YOE8+ years reliability/software engineering experience, strong Java (Java 17+) skills, cloud (GCP/AWS), Kubernetes/Docker, SLO/SLI incident experience, AI-assisted tooling familiarity, excellent communication and risk acumen.
Java 17, Claude Code, GitHub Copilot Agent Mode, Devin, Gemini Code Assist, Apache Kafka, GCP, AWS, Kubernetes, Docker, Terraform, Helm, Prometheus, Grafana, OpenTelemetry, Spring Boot, gRPC, Protocol Buffers, Apache Camel, Spring Integration, Vert.x, Netty
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
New York or Mountain View
$200k-$220k/yr HybridFull Time
ASAPP
ASAPP: Develops AI-powered software to automate enterprise contact center interactions.
10+ YOE10+ years production software experience; distributed systems; automation; AWS; Terraform; Python/Go; Kubernetes; on-call; strong communication.
Terraform, Python, Go, AWS, Kubernetes
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability
New York or Dublin or Singapore
$160k-$200k/yr OnsiteFull Time
Ripple
Ripple: Provides blockchain solutions for global payments and liquidity.
5+ YOE5+ years in SRE/DevOps/Platform Engineering; strong observability and CI/CD experience; security integration into pipelines; IaC with Terraform; knowledge of SLOs/SLIs and secrets management.
Azure DevOps, GitHub Actions, Octopus Deploy, New Relic, Datadog, Terraform, HashiCorp Vault, Azure Key Vault, SAST, DAST, SCA
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer - Paze
San Francisco or Chicago or Scottsdale or New York City or Phoenix or Washington
$116k-$174k/yr HybridFull Time
Early Warning Services
Early Warning Services: Operates payment and risk solutions for the financial industry.
8+ YOE8+ years SRE/engineering experience, bachelor’s degree preferred, hands-on with Python/Go/Java, Docker, CI/CD, Linux, cloud and messaging/datastore technologies; on-call rotation and incident leadership experience.
Python, Go, Java, Docker, Kafka, SQS, JMS, Oracle, Dynamo DB, Aurora, Redis, memcached, Linux, GIT, Chef, Maven, Jenkins, Kubernetes, Swarm, AWS