175 software reliability engineer jobs at 128 companies in Secaucus, NJ

1mo
Save
Mark Applied
Hide
Sr Software Engineer - Reliability Engineering
North Hills, New York, United States
$122k-$203k/yr HybridFull Time
Cox Automotive
Cox Automotive: Privately held automotive services and software serving dealers, fleets, lenders, automakers, and car shoppers.
5+ YOE5+ years in software/platform/infrastructure engineering, strong coding (Python/Go/Java), AWS and Terraform experience, SRE and observability knowledge, system design and incident response skills.
Python, Go, Java, AWS EC2, AWS RDS, AWS DynamoDB, AWS S3, AWS Aurora, AWS Lambda, AWS VPCs, AWS Athena, Terraform, Docker, Kubernetes, Linux, Windows, New Relic, Splunk, Prometheus, CI/CD pipelines
1w
Save
Mark Applied
Hide
Software Engineer, Site Reliability
United States or Columbus or Austin or San Francisco or New York City
$142k-$197k/yr RemoteFull Time
Upstart
UpstartNasdaq: UPST: Public AI lending marketplace connecting consumers with banks and credit unions for personal, auto, and home-equity loans.
3+ YOE3+ years in software or site reliability engineering; programming in Python, Go, JavaScript, or TypeScript; production software or infrastructure experience; cloud, distributed systems, observability, and incident response experience.
Python, Go, JavaScript, TypeScript, Kubernetes, AWS, Datadog, Sumo Logic, CloudWatch
2w
Save
Mark Applied
Hide
Lead Software Reliability Engineer
Irving or New York City or New Jersey or Tampa or Jacksonville
$145k-$175k/yr HybridFull Time, Contract
RE Partners
RE Partners: Woman-owned business technology consulting firm delivering application development, enterprise modernization, and engineering services to global brands.
Several years of TDD experience, strong coding skills, JVM languages and/or TypeScript, secure coding, CI/CD, Docker, OpenShift, Linux, build automation, and critical-systems observability expertise.
Test-Driven Development (TDD), Java, NPM, Spring, Renovate, Liquibase, Ansible, Docker, OpenShift, TypeScript, Gradle, Maven, Microsoft?, Linux, JVM
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: Private supply-chain AI software serving corporations, government agencies, and banks with risk and compliance technology.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1mo
Save
Mark Applied
Hide
Staff Software Engineer, Reliability Engineering & SDLC Governance
Carlsbad or Germantown or San Jose or San Francisco or New York City
$165k-$261k/yr OnsiteFull Time
Viasat
ViasatNASDAQ: VSAT: Global communications providing satellite connectivity and defense solutions.
8+ YOE8+ years software engineering experience with SDLC governance, SRE/DevOps knowledge, observability, incident analysis, and cross-team influence to improve reliability and customer experience.
Prometheus, Grafana, OpenTelemetry, Datadog, AS9115
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Customer engagement platform for cross-channel marketing and analytics.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
2w
Save
Mark Applied
Hide
Platform / Site Reliability Engineer
New York City, New York, United States
OnsiteFull Time
Sunset
Sunset: Private startup wind-down service helping founders close companies through legal, tax, and operational work.
Production cloud infrastructure and reliability experience across multiple services, strong software engineering skills, infrastructure and application coding, incident leadership, recovery expertise, and AI engineering tool proficiency.
AWS, Terraform, CI/CD, SOC 2
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer
New York City or Toronto
$184k-$240k/yr OnsiteFull Time
Movable Ink
Movable Ink: AI-powered marketing software helping marketers personalize customer experiences across email, mobile, and web.
6+ YOE6+ years SRE/Software Engineering experience designing and operating scalable, multi-cloud distributed systems; expertise with observability, IaC, Kubernetes, and multiple programming languages.
Apache Pulsar, Apache Kafka, Grafana Loki, ScyllaDB, Cassandra, Prometheus, Thanos, Grafana Alloy, Tempo, Terraform, Chef, EKS, GKE, NodeJS, Golang, Ruby, Python, shell
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
New York or San Ramon or Reno
$153k-$210k/yr HybridFull Time
Ridgeline
Ridgeline: Private investment management software serving asset and wealth managers through an AI-native cloud platform.
3+ YOE3–6 years SRE/DevOps experience, 2+ years on AWS, proficiency with Terraform, observability, CI/CD, Python/Go/Bash, incident response, and strong communication and troubleshooting skills.
Claude Code, Cursor, Terraform, AWS, EC2, ECS, EKS, RDS, S3, IAM, CloudWatch, GitHub Actions, CircleCI, Buildkite, Python, Go, Bash, Kubernetes, Helm, Kotlin, Node.js, TypeScript
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Home services marketplace helping homeowners find and hire local professionals for repairs, maintenance, and improvements.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
2w
Save
Mark Applied
Hide
Site Reliability Engineer
New York or Hong Kong or London or Singapore
$105k-$300k/yr OnsiteFull Time
Citadel
Citadel: Private multi-strategy alternative investment manager serving public and private institutions through global market strategies.
Bachelor's degree in computer science, related STEM discipline, or equivalent experience; proficiency in a modern structured programming language, software development practices, distributed systems, and strong communication skills.
Python, SQL, JavaScript, CSS, React, CI/CD
2mo
Save
Mark Applied
Hide
Senior Software Engineer - Observability and Reliability
New York City or San Francisco or London or Sydney
$170k-$240k/yr OnsiteFull Time
Sigma Computing
Sigma Computing: Private cloud analytics platform helping business and technical teams analyze live warehouse data and build AI applications.
5+ YOE5+ years building high-quality software, strong CS fundamentals, experience building observability tools, proficiency with Go, OpenTelemetry, Kubernetes, participation in on-call rotations, cloud service administration (GCP/AWS/Azure) preferred.
Go, OpenTelemetry, Kubernetes, GCP, AWS, Azure, SQL, Python
1w
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
HybridFull Time
Chariot
Chariot: US fintech helping nonprofits receive and process donor-advised fund gifts and grant payments.
4+ YOERequires 4+ years software development, 2+ years backend application development, bachelor's degree preferred, and proficiency with Go, Node, Docker, Terraform, Kubernetes, Postgres, REST APIs, gRPC, and AWS.
Go, Node, Docker, Terraform, Kubernetes, Postgres, REST APIs, gRPC, AWS
1w
Save
Mark Applied
Hide
Sr. Staff Engineer Software, Infrastructure Reliability (Chronosphere)
San Francisco or Denver or Austin or Jacksonville or Bridgeport or Seattle or Boston or New York City
$126k-$205k/yr RemoteFull Time
Palo Alto Networks
Palo Alto NetworksNASDAQ: PANW: Global cybersecurity platform providing network, cloud, and AI-driven security solutions.
8+ YOERequires 8+ years of relevant experience, backend programming proficiency, cloud-native and distributed systems expertise, Linux and networking knowledge, debugging skills, and experience with AWS or GCP and Kubernetes.
Go, Java, Python, Rust, AWS, GCP, Kubernetes, Linux, Terraform, Cursor, Claude
2w
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineer - Remote
Basking Ridge, New Jersey, United States
$113k-$193k/yr RemoteFull Time
UnitedHealth Group
UnitedHealth GroupNYSE: UNH: Diversified health care helping people live healthier lives.
10+ YOE5+ MgmtBachelor’s degree in a relevant field, 10+ years in software, SRE, platform, DevOps, infrastructure, or technology operations, and 5+ years leading engineering or operational teams. Requires production, cloud, ITSM, and reliability experience.
Azure, AWS, Infrastructure-as-Code, Kubernetes, OpenShift, Datadog, Splunk, Grafana, Prometheus, OpenTelemetry, AIOps, ChatOps, LLM
2w
Save
Mark Applied
Hide
Site Reliability Engineer, Global Banking & Markets, Vice President
New York City, New York, United States
$150k-$250k/yr OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities and investment management firm.
8+ YOE8+ years of software or reliability engineering experience; proficiency in a major programming language; cloud, distributed systems, SRE, automation, observability, incident response, and risk management expertise.
Java, GCP, AWS, Kubernetes, Docker, Claude Code, GitHub Copilot Agent Mode, Devin, Gemini Code Assist, Apache Kafka, Spring Boot, gRPC, Protocol Buffers, Apache Camel, Spring Integration, Terraform, Helm, Prometheus, Grafana, OpenTelemetry, SQL, NoSQL, Vert.x, Netty, Microsoft?
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
$100k-$125k/yr OnsiteFull Time
Dirac
Dirac: Private U.S. manufacturing-technology providing BuildOS, an AI-driven production-planning platform, to industrial manufacturers.
Builder mindset with ability to explain projects; eager to learn, re-build, and improve; strong debugging and troubleshooting skills; programming or scripting knowledge; familiarity with container tech and networking.
Kubernetes, Docker, Networking, DNS, Routing, ITAR
1d
Save
Mark Applied
Hide
Site Reliability Engineer
Ontario or Canada or New York City or Ireland
$88k-$127k/yr RemoteFull Time
Greenhouse
Greenhouse: Private hiring software helping organizations source, interview, and onboard candidates with structured, AI-powered recruiting tools.
3+ YOE3+ years in site reliability or infrastructure, production AWS and Kubernetes experience, software delivery experience, and strong fluency with AI coding tools such as Claude. Must be eligible to work in Canada.
AWS, Kubernetes, RDS, Redis, OpenSearch, Claude
3w
Save
Mark Applied
Hide
Site Reliability Engineer, Global Banking & Markets, Vice President
New York City, New York, United States
$150k-$250k/yr OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities and investment management firm.
8+ YOE8+ years reliability/software engineering experience, strong Java (Java 17+) skills, cloud (GCP/AWS), Kubernetes/Docker, SLO/SLI incident experience, AI-assisted tooling familiarity, excellent communication and risk acumen.
Java 17, Claude Code, GitHub Copilot Agent Mode, Devin, Gemini Code Assist, Apache Kafka, GCP, AWS, Kubernetes, Docker, Terraform, Helm, Prometheus, Grafana, OpenTelemetry, Spring Boot, gRPC, Protocol Buffers, Apache Camel, Spring Integration, Vert.x, Netty
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
New York or Mountain View
$200k-$220k/yr HybridFull Time
ASAPP
ASAPP: ASAPP is a private enterprise AI software providing agentic customer-service platforms for contact centers.
10+ YOE10+ years production software experience; distributed systems; automation; AWS; Terraform; Python/Go; Kubernetes; on-call; strong communication.
Terraform, Python, Go, AWS, Kubernetes