139 platform reliability engineer jobs at 78 companies in Secaucus, NJ
2w
Save
Mark Applied
Hide
2w
Platform / Site Reliability Engineer
New York City, New York, United States
OnsiteFull Time
Sunset: Private startup wind-down service helping founders close companies through legal, tax, and operational work.
Production cloud infrastructure and reliability experience across multiple services, strong software engineering skills, infrastructure and application coding, incident leadership, recovery expertise, and AI engineering tool proficiency.
Grow Therapy: Mental health technology platform connecting patients, independent providers, and insurers with affordable, insurance-covered therapy and medication care.
6+ YOE6+ years operating production systems; hands-on AWS, Kubernetes (EKS), Terraform; experience defining SLOs/SLAs and observability (DataDog); strong communication and systems-thinking skills; PostgreSQL experience a plus.
7+ YOE7+ years systems engineering or architecture, 5+ years supporting enterprise production environments, deep expertise in one infrastructure domain, SRE practices, automation and troubleshooting across domains.
Konica MinoltaTokyo Stock Exchange: 4902: Provider of office technology, IT services, and digital solutions.
8+ YOERequires 8+ years administering enterprise databases and middleware, including Microsoft SQL Server, Linux/RHEL, WebSphere, OpenLDAP, SSL/TLS, DNS, scripting, and incident resolution. Bachelor's preferred.
Microsoft SQL Server, SSIS, SQL Agent, T-SQL, Linux, RHEL, Apache, PHP, SSL/TLS, Java keystores, DNS, IBM WebSphere Application Server (WAS), OpenLDAP, Windows Server, RDP, PowerShell, bash, IBM DB2, SSRS, Azure Data Factory, Power BI, Tidal, Microsoft Entra ID, Uptrends, MuleSoft, SAP, Salesforce, Ansible, Terraform, PowerShell DSC, Microsoft Azure, Microsoft 365, GitHub Copilot, Datadog, Dynatrace, Docker, OpenLiberty, Ollama, vLLM, REST, API, Tenable, Bitsight
London or Toronto or New York City or Montreal or Kitchener or San Francisco
HybridFull Time
Index Exchange: Independent ad-tech supply-side platform helping media owners monetize digital content and enabling brands to buy programmatic advertising.
8+ YOERequires 8+ years in platform engineering, SRE, infrastructure engineering, or DevOps; deep Linux and Kubernetes expertise; IaC at scale; Go or Python; networking; and cross-team technical strategy.
Staff , Site Reliability Engineer - Cloud Platform
New York City, New York, United States
$190k-$210k/yrHybridFull Time
Butterfly NetworkNew York Stock Exchange: BFLY: Public U.S. medical technology making handheld point-of-care ultrasound hardware and AI-powered clinical software for healthcare professionals.
8+ YOE8+ years managing production systems; deep hands-on AWS and Kubernetes experience; strong programming/scripting and automation skills; observability platform ownership; incident response leadership and mentoring ability.
Forge GlobalNYSE: FRGE: Financial technology operating a private-market marketplace and data, custody, and investment solutions for companies and investors.
8+ YOE5+ Mgmt8+ years software engineering experience with infrastructure/platform focus, 5+ years people leadership, deep cloud/observability/incident response experience, strong distributed systems judgment, and ability to set platform strategy.
Exiger: Private supply-chain AI software serving corporations, government agencies, and banks with risk and compliance technology.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
Bank of AmericaNYSE: BAC: Global financial services and banking institution.
4+ YOE4+ years cloud/platform engineering experience with GCP exposure, Terraform/IaC, CI/CD, observability, DevSecOps practices, incident response, and automation for reliability and resiliency.
Bank of AmericaNYSE: BAC: Global financial services and banking institution.
4+ YOE4+ years cloud/platform engineering experience with Terraform and GCP; strong IaC, observability, automation, incident response, and DevSecOps skills; ability to define SLIs/SLOs and mentor engineers.
Brooklyn or New York City or Los Angeles or Santa Monica or United States
$230k-$260k/yrHybridFull Time
Radix Health: Healthcare technology helping providers achieve fair reimbursement through integrated IDR software, data, and AI.
8+ YOE8+ years in SRE, infrastructure, platform engineering, or large-scale production systems; expertise in cloud infrastructure, distributed systems, networking, containers, orchestration, infrastructure as code, observability, automation, and incident management.
Arca Wealth: AI-native wealth management providing personalized, advisor-led financial services through expert advisors and purpose-built AI.
Experienced platform engineer to own infrastructure, backend systems, developer experience, sandboxing for agents, reliability, performance, and M&A onboarding at scale.
Senior VoIP Operations & Reliability Engineer (Carrier-Class Voice Platform)
Newton, New Jersey, United States
OnsiteFull Time
Planet Networks: Connecting People, Places, and Things Since 1994
Senior hands-on experience operating carrier-scale VoIP systems (SIP, Kamailio/OpenSIPS, Asterisk), reliability engineering, incident response, SLOs/SLIs, observability, and Linux automation.
Ripple: Enables institutions to move, manage, and tokenize value.
7+ YOE7+ years SRE/Platform experience focused on observability, New Relic, Terraform, PowerShell, Azure/AWS, incident management (Incident.IO/PagerDuty/OpsGenie), and coaching engineering teams.
New Relic, NRQL, Terraform, PowerShell, Azure, AWS, Azure DevOps, Octopus Deploy, Incident.IO, PagerDuty, OpsGenie, Slack, Python, Bash, Jira, SQL Server
Claryo: AI-powered warehouse intelligence serving logistics operators with AI agents for monitoring, forecasting, and automation.
3+ YOE3+ years SRE/infrastructure experience, strong Linux and networking fundamentals, experience with Kubernetes, cloud platforms, observability tooling, and debugging distributed systems in production.
deCircle: Boutique recruitment partner for high-tech and Web3 organizations.
3+ YOE3+ years in platform/infra or reliability engineering; experience building test infrastructure for payments/ledgers; familiarity with formal verification, property-based or chaos testing preferred; strong ownership of internal tooling and CI/CD.
Cox Automotive: Privately held automotive services and software serving dealers, fleets, lenders, automakers, and car shoppers.
5+ YOE5+ years in software/platform/infrastructure engineering, strong coding (Python/Go/Java), AWS and Terraform experience, SRE and observability knowledge, system design and incident response skills.
Fitch Group: Global provider of financial information, credit ratings, and analytics.
Deep SRE, DevOps, or platform engineering experience with AWS, Azure, Docker, Kubernetes, Linux, Windows, CI/CD, cloud security, networking, and Python, PowerShell, or Bash.
San Francisco or San Jose or Seattle or Los Angeles or San Diego or Portland or Reno or Ontario or Bakersfield or Phoenix or Riverside or Sacramento or Dallas or Houston or Irving or Chicago or Cleveland or Columbus or Cincinnati or Detroit or Newark or Dulles or Indianapolis or Austin or Brooklyn or Las Vegas or Kansas City or Philadelphia or Pensacola or South San Francisco or St. Louis
$120k-$147k/yrHybridFull Time
Jitsu: Privately held U.S. last-mile delivery provider serving e-commerce brands and high-volume shippers.
5+ YOE5+ years as a DevOps or Site Reliability Engineer, with deep CI/CD, cloud infrastructure, Terraform, GitOps, Kubernetes, Helm, databases, monitoring, reliability, and infrastructure security experience.