128 software reliability engineer jobs at 96 companies in Montclair, NJ

2d
Save
Mark Applied
Hide
Sr Software Engineer - Reliability Engineering
North Hills, New York, United States
$122k-$203k/yr HybridFull Time
Cox Enterprises
Cox Enterprises: Providing global communications, automotive services, and media solutions.
5+ YOE5+ years in software/platform/infrastructure engineering, strong coding (Python/Go/Java), AWS and Terraform experience, SRE and observability knowledge, system design and incident response skills.
Python, Go, Java, AWS EC2, AWS RDS, AWS DynamoDB, AWS S3, AWS Aurora, AWS Lambda, AWS VPCs, AWS Athena, Terraform, Docker, Kubernetes, Linux, Windows, New Relic, Splunk, Prometheus, CI/CD pipelines
3w
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1mo
Save
Mark Applied
Hide
Software Engineer, Reliability Platforms
San Francisco or Sunnyvale or New York City
$160k-$235k/yr OnsiteFull Time
DoorDash
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
5+ YOE5+ years in infrastructure/platform/backend engineering; fluent in Go or similar; AWS, containerization, and IaC experience (Terraform or Pulumi); SRE concepts (SLOs, error budgets); platform engineering mindset and familiarity with AI tools.
Go, AWS, Terraform, Pulumi
1mo
Save
Mark Applied
Hide
Staff Quality & Reliability Engineer
San Francisco or New York City
HybridFull Time
Beast Industries
Beast Industries: Produces digital media and consumer goods for MrBeast brands.
Expert in software quality engineering and site reliability for consumer-scale distributed systems; owns test strategy, SLOs/error budgets, CI/CD test gates, observability, incident response, and reliability tooling.
CI/CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
2mo
Save
Mark Applied
Hide
Product Reliability Engineer
New York City or London or Copenhagen
$185k-$250k/yr HybridFull Time
Normal Computing
Normal Computing: Building probabilistic AI and thermodynamic computing for semiconductor design.
QA automation, reliability engineering, scripting, CI/CD, testing frameworks, and ability to automate validation; experience with LLMs and agents; strong debugging and issue reproduction.
Scripting, Automation, CI/CD, QA automation, LLMs/Agents, Testing frameworks
2w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
New York City or Toronto
$184k-$240k/yr OnsiteFull Time
Movable Ink
Movable Ink: Provides AI-powered content personalization for digital marketing campaigns.
6+ YOE6+ years SRE/Software Engineering experience designing and operating scalable, multi-cloud distributed systems; expertise with observability, IaC, Kubernetes, and multiple programming languages.
Apache Pulsar, Apache Kafka, Grafana Loki, ScyllaDB, Cassandra, Prometheus, Thanos, Grafana Alloy, Tempo, Terraform, Chef, EKS, GKE, NodeJS, Golang, Ruby, Python, shell
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
4w
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
New York or San Ramon or Reno
$153k-$210k/yr HybridFull Time
Ridgeline
Ridgeline: Cloud-native platform for investment management operations.
3+ YOE3–6 years SRE/DevOps experience, 2+ years on AWS, proficiency with Terraform, observability, CI/CD, Python/Go/Bash, incident response, and strong communication and troubleshooting skills.
Claude Code, Cursor, Terraform, AWS, EC2, ECS, EKS, RDS, S3, IAM, CloudWatch, GitHub Actions, CircleCI, Buildkite, Python, Go, Bash, Kubernetes, Helm, Kotlin, Node.js, TypeScript
1mo
Save
Mark Applied
Hide
Senior Software Engineer - Observability and Reliability
New York City or San Francisco or London or Sydney
$170k-$240k/yr OnsiteFull Time
Sigma Computing
Sigma Computing: Cloud-native analytics platform featuring a spreadsheet-style interface.
5+ YOE5+ years building high-quality software, strong CS fundamentals, experience building observability tools, proficiency with Go, OpenTelemetry, Kubernetes, participation in on-call rotations, cloud service administration (GCP/AWS/Azure) preferred.
Go, OpenTelemetry, Kubernetes, GCP, AWS, Azure, SQL, Python
1mo
Save
Mark Applied
Hide
GOV Site Reliability Engineer
United States or Kansas or Washington or California or Texas or Illinois or North Carolina or Colorado or Massachusetts or Pennsylvania or Virginia or Oregon or Nevada or Hawaii or New York or Georgia or Ohio or Arizona or Seattle or San Francisco or New York City
$110k-$183k/yr RemoteFull Time
Veeam
Veeam: Data resilience and security for hybrid cloud environments
3+ YOE3+ years in software engineering with 1+ year in SRE/Platform/DevOps, cloud experience (Azure or comparable), observability (Prometheus, Grafana, OpenTelemetry, ELK), IaC (Terraform/Terragrunt/Pulumi), Kubernetes, CI/CD tooling, programming in TypeScript/JS, Go, Java, or C#, and experience in compliance-oriented environments.
VDC, Prometheus, Grafana, OpenTelemetry, ELK stack, Terraform, Terragrunt, Pulumi, Kubernetes, GitHub Actions, Azure DevOps, GitLab CI, ArgoCD, TypeScript, JS, Go, Java, C#, Azure Government, AWS GovCloud
2w
Save
Mark Applied
Hide
Senior Lead Site Reliability Engineer-Core Engineering Solutions
Jersey City, New Jersey, United States
$171k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years software engineering experience with site reliability practices, public cloud modernization, observability (SLOs/alerts/telemetry), incident management, and using enterprise AI for reliability workflows.
Grafana, Dynatrace, Prometheus, Datadog, Splunk
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
$111k-$160k/yr HybridFull Time
Mizuho Financial Group
Mizuho Financial GroupTokyo Stock Exchange: 8411: Global financial group providing banking and investment services.
3+ YOEBachelor’s degree or equivalent; 3+ years SRE/automation experience; strong Grafana, cloud (AWS/Azure/GCP), containers (Docker/Kubernetes); CI/CD; scripting (Python/Bash/Go); on-call experience.
Grafana, Ansible, Terraform, Jenkins, AWS, Azure, Google Cloud, Docker, Kubernetes, CI/CD, Python, Bash, Go
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
$100k-$125k/yr OnsiteFull Time
Dirac
Dirac: AI-driven platform automating manufacturing work instructions from CAD files.
Builder mindset with ability to explain projects; eager to learn, re-build, and improve; strong debugging and troubleshooting skills; programming or scripting knowledge; familiarity with container tech and networking.
Kubernetes, Docker, Networking, DNS, Routing, ITAR
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
New York or Mountain View
$200k-$220k/yr HybridFull Time
ASAPP
ASAPP: Develops AI-powered software to automate enterprise contact center interactions.
10+ YOE10+ years production software experience; distributed systems; automation; AWS; Terraform; Python/Go; Kubernetes; on-call; strong communication.
Terraform, Python, Go, AWS, Kubernetes
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability
New York or Dublin or Singapore
$160k-$200k/yr OnsiteFull Time
Ripple
Ripple: Provides blockchain solutions for global payments and liquidity.
5+ YOE5+ years in SRE/DevOps/Platform Engineering; strong observability and CI/CD experience; security integration into pipelines; IaC with Terraform; knowledge of SLOs/SLIs and secrets management.
Azure DevOps, GitHub Actions, Octopus Deploy, New Relic, Datadog, Terraform, HashiCorp Vault, Azure Key Vault, SAST, DAST, SCA
1mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer - Paze
Scottsdale or San Francisco or Chicago or New York City or Phoenix
$106k-$156k/yr HybridFull Time
Early Warning Services
Early Warning Services: Operates payment and risk solutions for the financial industry.
3+ YOEBachelor's degree preferred, 3+ years experience in technical/software environments, Linux administration, scripting, observability, incident/problem management, Git, security, on-call rotation; experience with CI/CD, AWS, Docker, Kubernetes preferred.
Git, AWS, Docker, Kubernetes, Swarm, CI/CD, Linux, Java, Ruby, Python, JavaScript, Go, TCP/UDP/IP
4w
Save
Mark Applied
Hide
Staff Software Engineer, Banking
San Francisco or New York City or Seattle or Vancouver
$240k-$285k/yr HybridFull Time
Brex
Brex: Corporate cards and spend management software for businesses.
8+ YOE8+ years software engineering experience with multi-team technical leadership, distributed systems architecture, reliability, cross-functional collaboration, strong communication, and mentoring senior engineers.
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Toronto or New York City or North America
$144k-$200k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development and distributed systems experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud platform expertise (AWS, GCP, Azure); Linux and networking knowledge; participate in 24/7 on-call.
Python, Go, Argo Workflows, ArgoCD, Kubernetes, AWS, Google Cloud Platform (GCP), Azure, Linux, CI/CD
1w
Save
Mark Applied
Hide
Principal Software Engineer
Redmond or Mountain View or New York
$143k-$275k/yr HybridFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
6+ YOEBachelor's in CS or related and 6+ years engineering experience (or equivalent), coding in .Net/Java/JavaScript/Rust/Python; experience with distributed systems, AI evaluation, and reliability engineering.
.Net, Java, JavaScript, Rust, Python