110 site reliability manager jobs at 57 companies in Hackensack, NJ
4w
Save
Mark Applied
Hide
4w
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger: AI-powered supply chain risk and compliance management software.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years applied SRE experience, proficiency in reliability/scalability/security, experience with CI/CD, containers, observability, programming in Python/Java/.Net, and using enterprise AI for SRE workflows.
8+ YOE3+ MgmtBachelor's in CS or equivalent, 8+ years software development, 3+ years people management, experience with distributed systems, on-call and postmortem practices, strong communication and critical thinking.
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
10+ YOE5+ MgmtBachelor's in a technical field,10+ years engineering experience with 5+ years leading SRE/Platform teams; experience with observability, incident management, distributed systems, and cloud architecture.
AWS, New Relic, Splunk, Datadog, Sentry, Honeycomb, Grafana, Prometheus, OpenTelemetry
Ripple: Provides blockchain solutions for global payments and liquidity.
7+ YOE7+ years SRE/Platform experience focused on observability, New Relic, Terraform, PowerShell, Azure/AWS, incident management (Incident.IO/PagerDuty/OpsGenie), and coaching engineering teams.
New Relic, NRQL, Terraform, PowerShell, Azure, AWS, Azure DevOps, Octopus Deploy, Incident.IO, PagerDuty, OpsGenie, Slack, Python, Bash, Jira, SQL Server
Optimal Market Technologies: Operates a broker-dealer platform for wholesale options execution.
Hands-on Linux and network administration, strong scripting/automation, Infrastructure-as-Code experience, production support and incident response ownership, staff management/mentoring, familiarity with C++, Python, SQL, Azure, and real-time trading systems.
C++, Python, SQL, Linux, CentOS7, RHEL9, Microsoft Azure, Microsoft Azure Virtual Desktop (AVD), Claude Code, FIX Protocol, PostgreSQL, AERON
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
8+ YOE8+ years managing production systems; deep hands-on AWS and Kubernetes experience; strong programming/scripting and automation skills; observability platform ownership; incident response leadership and mentoring ability.
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yrRemoteFull Time
Thumbtack: Online marketplace connecting homeowners with local service professionals.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
Site Reliability Engineer , Engineering Enablement (Remote)
Boston or Atlanta or Chicago or Washington or New York City or Herndon
$138k-$198k/yrRemoteFull Time
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
3+ YOEBachelor's+5 or Master's+3 experience; 3+ years writing production Python or Ruby; 3+ years managing Terraform/Ansible IaC; Unix/Linux experience; containerization and CI/CD experience; ability to operate at scale and participate in on-call rotations.
Python, Ruby, Terraform, Ansible, Unix/Linux, Docker, Kubernetes, DORA, SPACE
BTIG: Provides global institutional trading and investment banking services.
2+ YOE2+ years IT support experience, strong customer-facing skills, Windows 10/11 and Microsoft 365 proficiency, experience with SCCM/Endpoint Manager and ticketing systems (ServiceNow); willing to obtain MS900.
ServiceNow, Windows 10, Windows 11, Microsoft Office 365, Microsoft OneDrive, System Center Configuration Manager, Endpoint Manager, Zoom, Bloomberg, Thomson Reuters, ICE, Fidessa, Redi+, Global Relay, Cisco, Active Directory, VPNWIFI
Executive Director – Site Reliability Engineering – WM Technology
New York, New York, United States
$195k-$215k/yrOnsiteFull Time
Morgan StanleyNYSE: MS: Global financial services firm providing investment and wealth management.
12+ YOEBachelor's degree and 12+ years in technology production management; leadership in incident management and SRE practices; deep understanding of asset-management workflows, monitoring/observability, cloud/DevOps, and strong communication skills.
Senior Site Reliability Engineer (In-Office Required)
New York City, New York, United States
$156k-$262k/yrOnsiteFull Time
NebiusNasdaq: NBIS: Builds cloud infrastructure and software for artificial intelligence development.
5+ YOE5+ years DevOps/SRE experience; strong Kubernetes and managed cloud expertise; infrastructure as code (Terraform), GitOps, CI/CD, observability, production incident management, and large-scale distributed systems experience.
Senior Site Reliability Engineer (In-Office Required)
New York City, New York, United States
$156k-$262k/yrOnsiteFull Time
TavilyNASDAQ: NBIS: Real-time web search engine for AI agents.
5+ YOE5+ years in DevOps/SRE, strong Kubernetes and infrastructure-as-code (Terraform) experience, familiarity with GitOps and CI/CD, observability and incident management experience, ability to operate large-scale distributed systems.
Senior AI Ops & Incident/Site Reliability Engineer
Austin or Plano or Fort Mill or Boston or New York City or Tempe or San Diego
$94k-$171k/yrHybridFull Time
Perficient: Provides digital transformation and AI consulting for global enterprises.
8+ YOEBachelor's degree or equivalent,8+ years in IT operations/SRE or production support,5+ years leading incident management,experience with Dynatrace,ServiceNow,cloud platforms and AIOps implementations.
Senior Site Reliability Engineer - AVP - Credit Trade Floor
New York, New York, United States
$115k-$160k/yrOnsiteFull Time
BarclaysLondon Stock Exchange: BARC: Global bank providing retail, corporate, and investment financial services.
Extensive systems engineering and SRE experience with Windows/Linux, Kubernetes, AWS/Azure, Python, monitoring/observability (ITRS Geneos), and SQL; strong incident management and stakeholder communication skills.