151 software reliability engineer jobs at 105 companies in Lakewood, NJ

3d
Save
Mark Applied
Hide
Lead Software Reliability Engineer
Irving or New York City or New Jersey or Tampa or Jacksonville
$145k-$175k/yr HybridFull Time, Contract
RE Partners
RE Partners: Providing technology consulting and digital transformation services for enterprise clients.
Several years of TDD experience, strong coding skills, JVM languages and/or TypeScript, secure coding, CI/CD, Docker, OpenShift, Linux, build automation, and critical-systems observability expertise.
Test-Driven Development (TDD), Java, NPM, Spring, Renovate, Liquibase, Ansible, Docker, OpenShift, TypeScript, Gradle, Maven, Microsoft?, Linux, JVM
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1mo
Save
Mark Applied
Hide
Staff Software Engineer, Reliability Engineering & SDLC Governance
Carlsbad or Germantown or San Jose or San Francisco or New York City
$165k-$261k/yr OnsiteFull Time
Viasat
ViasatNASDAQ: VSAT: Provides global satellite broadband and secure networking communication services.
8+ YOE8+ years software engineering experience with SDLC governance, SRE/DevOps knowledge, observability, incident analysis, and cross-team influence to improve reliability and customer experience.
Prometheus, Grafana, OpenTelemetry, Datadog, AS9115
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer
New York City or Toronto
$184k-$240k/yr OnsiteFull Time
Movable Ink
Movable Ink: Provides AI-powered content personalization for digital marketing campaigns.
6+ YOE6+ years SRE/Software Engineering experience designing and operating scalable, multi-cloud distributed systems; expertise with observability, IaC, Kubernetes, and multiple programming languages.
Apache Pulsar, Apache Kafka, Grafana Loki, ScyllaDB, Cassandra, Prometheus, Thanos, Grafana Alloy, Tempo, Terraform, Chef, EKS, GKE, NodeJS, Golang, Ruby, Python, shell
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
New York or San Ramon or Reno
$153k-$210k/yr HybridFull Time
Ridgeline
Ridgeline: Cloud-native platform for investment management operations.
3+ YOE3–6 years SRE/DevOps experience, 2+ years on AWS, proficiency with Terraform, observability, CI/CD, Python/Go/Bash, incident response, and strong communication and troubleshooting skills.
Claude Code, Cursor, Terraform, AWS, EC2, ECS, EKS, RDS, S3, IAM, CloudWatch, GitHub Actions, CircleCI, Buildkite, Python, Go, Bash, Kubernetes, Helm, Kotlin, Node.js, TypeScript
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
5d
Save
Mark Applied
Hide
Site Reliability Engineer
New York or Hong Kong or London or Singapore
$105k-$300k/yr OnsiteFull Time
Citadel
Citadel: Global alternative investment management firm
Bachelor's degree in computer science, related STEM discipline, or equivalent experience; proficiency in a modern structured programming language, software development practices, distributed systems, and strong communication skills.
Python, SQL, JavaScript, CSS, React, CI/CD
2mo
Save
Mark Applied
Hide
Senior Software Engineer - Observability and Reliability
New York City or San Francisco or London or Sydney
$170k-$240k/yr OnsiteFull Time
Sigma Computing
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
5+ YOE5+ years building high-quality software, strong CS fundamentals, experience building observability tools, proficiency with Go, OpenTelemetry, Kubernetes, participation in on-call rotations, cloud service administration (GCP/AWS/Azure) preferred.
Go, OpenTelemetry, Kubernetes, GCP, AWS, Azure, SQL, Python
5d
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineer - Remote
Basking Ridge, New Jersey, United States
$113k-$193k/yr RemoteFull Time
UnitedHealth Group
UnitedHealth GroupNYSE: UNH: Provides health insurance and technology-enabled health care services.
10+ YOE5+ MgmtBachelor’s degree in a relevant field, 10+ years in software, SRE, platform, DevOps, infrastructure, or technology operations, and 5+ years leading engineering or operational teams. Requires production, cloud, ITSM, and reliability experience.
Azure, AWS, Infrastructure-as-Code, Kubernetes, OpenShift, Datadog, Splunk, Grafana, Prometheus, OpenTelemetry, AIOps, ChatOps, LLM
1d
Save
Mark Applied
Hide
Site Reliability Engineer, Global Banking & Markets, Vice President
New York City, New York, United States
$150k-$250k/yr OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
8+ YOE8+ years of software or reliability engineering experience; proficiency in a major programming language; cloud, distributed systems, SRE, automation, observability, incident response, and risk management expertise.
Java, GCP, AWS, Kubernetes, Docker, Claude Code, GitHub Copilot Agent Mode, Devin, Gemini Code Assist, Apache Kafka, Spring Boot, gRPC, Protocol Buffers, Apache Camel, Spring Integration, Terraform, Helm, Prometheus, Grafana, OpenTelemetry, SQL, NoSQL, Vert.x, Netty, Microsoft?
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
$100k-$125k/yr OnsiteFull Time
Dirac
Dirac: AI-driven platform automating manufacturing work instructions from CAD files.
Builder mindset with ability to explain projects; eager to learn, re-build, and improve; strong debugging and troubleshooting skills; programming or scripting knowledge; familiarity with container tech and networking.
Kubernetes, Docker, Networking, DNS, Routing, ITAR
1w
Save
Mark Applied
Hide
Site Reliability Engineer, Global Banking & Markets, Vice President
New York City, New York, United States
$150k-$250k/yr OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Provides investment banking, securities, and wealth management services globally.
8+ YOE8+ years reliability/software engineering experience, strong Java (Java 17+) skills, cloud (GCP/AWS), Kubernetes/Docker, SLO/SLI incident experience, AI-assisted tooling familiarity, excellent communication and risk acumen.
Java 17, Claude Code, GitHub Copilot Agent Mode, Devin, Gemini Code Assist, Apache Kafka, GCP, AWS, Kubernetes, Docker, Terraform, Helm, Prometheus, Grafana, OpenTelemetry, Spring Boot, gRPC, Protocol Buffers, Apache Camel, Spring Integration, Vert.x, Netty
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
New York or Mountain View
$200k-$220k/yr HybridFull Time
ASAPP
ASAPP: Develops AI-powered software to automate enterprise contact center interactions.
10+ YOE10+ years production software experience; distributed systems; automation; AWS; Terraform; Python/Go; Kubernetes; on-call; strong communication.
Terraform, Python, Go, AWS, Kubernetes
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability
New York or Dublin or Singapore
$160k-$200k/yr OnsiteFull Time
Ripple
Ripple: Provides blockchain solutions for global payments and liquidity.
5+ YOE5+ years in SRE/DevOps/Platform Engineering; strong observability and CI/CD experience; security integration into pipelines; IaC with Terraform; knowledge of SLOs/SLIs and secrets management.
Azure DevOps, GitHub Actions, Octopus Deploy, New Relic, Datadog, Terraform, HashiCorp Vault, Azure Key Vault, SAST, DAST, SCA
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer - Paze
San Francisco or Chicago or Scottsdale or New York City or Phoenix or Washington
$116k-$174k/yr HybridFull Time
Early Warning Services
Early Warning Services: Operates payment and risk solutions for the financial industry.
8+ YOE8+ years SRE/engineering experience, bachelor’s degree preferred, hands-on with Python/Go/Java, Docker, CI/CD, Linux, cloud and messaging/datastore technologies; on-call rotation and incident leadership experience.
Python, Go, Java, Docker, Kafka, SQS, JMS, Oracle, Dynamo DB, Aurora, Redis, memcached, Linux, GIT, Chef, Maven, Jenkins, Kubernetes, Swarm, AWS
1mo
Save
Mark Applied
Hide
Staff Software Engineer, Banking
San Francisco or New York City or Seattle or Vancouver
$240k-$285k/yr HybridFull Time
Brex
Brex: Corporate cards and spend management software for businesses.
8+ YOE8+ years software engineering experience with multi-team technical leadership, distributed systems architecture, reliability, cross-functional collaboration, strong communication, and mentoring senior engineers.
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Toronto or New York City or North America
$144k-$200k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development and distributed systems experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud platform expertise (AWS, GCP, Azure); Linux and networking knowledge; participate in 24/7 on-call.
Python, Go, Argo Workflows, ArgoCD, Kubernetes, AWS, Google Cloud Platform (GCP), Azure, Linux, CI/CD
2w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
New York, New York, United States
$152k-$215k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOEFormal SRE training or certification, 5+ years applied SRE/production management experience, leadership in production support, experience with Front Office sales platforms, observability and incident management expertise.
Dynatrace, Splunk, Geneos, Grafana, AWS, Microsoft Azure, GCP, Python, Shell, PowerShell, Ansible, Terraform, Kubernetes, OpenShift
1d
Save
Mark Applied
Hide
Staff Software Engineer, Education
San Francisco or New York City
$320k-$405k/yr HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
7+ YOE7+ years of software engineering experience; strong full-stack and backend skills in databases, APIs, authentication, reliability, and infrastructure; experience scaling platforms and using AI tools.
SAML, OAuth, Open Badges, LTI, LLM

Explore Jobs

Expand Your Job Search