184 site reliability manager jobs at 72 companies in Harlem, NY

1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Jersey City or McLean or Richmond
HybridFull Time
Exiger
Exiger: Private supply-chain AI software serving corporations, government agencies, and banks with risk and compliance technology.
6+ YOEBachelor's or Master's (or equivalent), 6+ years software/systems engineering with >=4 years in SRE or production/platform reliability, strong Linux/Unix and networking knowledge, experience with SLIs/SLOs, observability, automation, chaos engineering, incident management, and familiarity with AWS and secure/gov environments.
AWS, Codex, Claude, Chaos Monkey, Gremlin, LitmusChaos, Snowflake, Redshift, Apache Iceberg, Go, C, Java, Linux/Unix
1d
Save
Mark Applied
Hide
Site Reliability Engineer - Enterprise Technology
New York City, New York, United States
$200k-$250k/yr OnsiteFull Time
Hudson River Trading
Hudson River Trading: Private quantitative trading firm providing liquidity across global markets and directly to financial-market clients.
5+ YOERequires 5+ years in site reliability or related disciplines, Python, Linux, Kubernetes, observability, containerized infrastructure, CI/CD, IaC, configuration management, and cloud platform experience.
Linux, Kubernetes, Python, Jenkins, GitHub Actions, ArgoCD, Terraform, SaltStack, Chef, Puppet, Ansible, AWS, Azure, GCP
1d
Save
Mark Applied
Hide
Site Reliability Engineer - Enterprise Technology
New York City, New York, United States
$200k-$250k/yr OnsiteFull Time
Hudson River Trading
Hudson River Trading: Private quantitative trading firm providing liquidity across global markets and directly to financial-market clients.
5+ YOERequires 5+ years in site reliability or related disciplines, Python, containerized infrastructure, CI/CD, IaC, configuration management, cloud platforms, and Linux, Kubernetes, and observability expertise.
Linux, Kubernetes, Python, Jenkins, GitHub Actions, ArgoCD, Terraform, SaltStack, Chef, Puppet, Ansible, AWS, Azure, GCP
1w
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
HybridFull Time
Chariot
Chariot: US fintech helping nonprofits receive and process donor-advised fund gifts and grant payments.
4+ YOERequires 4+ years software development, 2+ years backend application development, bachelor's degree preferred, and proficiency with Go, Node, Docker, Terraform, Kubernetes, Postgres, REST APIs, gRPC, and AWS.
Go, Node, Docker, Terraform, Kubernetes, Postgres, REST APIs, gRPC, AWS
4w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
Plano or Jersey City
$152k-$215k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services and investment banking firm.
5+ YOE5+ years applied SRE experience, proficiency in reliability/scalability/security, experience with CI/CD, containers, observability, programming in Python/Java/.Net, and using enterprise AI for SRE workflows.
Python, Java, Spring Boot, .Net, CI/CD
2d
Save
Mark Applied
Hide
Manager, Site Reliability Engineering (Auth0)
New York City or Washington or California or Colorado or Illinois or Washington
$182k-$251k/yr HybridFull Time
Auth0
Auth0Nasdaq: OKTA: Public American identity-security providing cloud-based authentication, authorization, and access-management services to organizations.
8+ YOE3+ MgmtRequires 8+ years industry experience, 3+ years SRE or software engineering team leadership, AWS/Azure, Terraform, containers, Kubernetes, microservices, databases, Go or Python, and U.S. Person status.
AWS, Azure, Terraform, Kubernetes, Go, Python
2w
Save
Mark Applied
Hide
Senior Manager, Site Reliability Engineer - Remote
Basking Ridge, New Jersey, United States
$113k-$193k/yr RemoteFull Time
UnitedHealth Group
UnitedHealth GroupNYSE: UNH: Diversified health care helping people live healthier lives.
10+ YOE5+ MgmtBachelor’s degree in a relevant field, 10+ years in software, SRE, platform, DevOps, infrastructure, or technology operations, and 5+ years leading engineering or operational teams. Requires production, cloud, ITSM, and reliability experience.
Azure, AWS, Infrastructure-as-Code, Kubernetes, OpenShift, Datadog, Splunk, Grafana, Prometheus, OpenTelemetry, AIOps, ChatOps, LLM
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York, New York, United States
$178k-$258k/yr RemoteFull Time
Adobe
AdobeNASDAQ: ADBE: Empowering everyone to create through innovative digital experiences.
Requires a computer science bachelor's degree or equivalent experience, Python, production ML inference, AWS cloud infrastructure, Kubernetes, vulnerability management, distributed-systems debugging, and on-call participation.
Python, PHP, Node.js, Ruby, SageMaker, OpenAI, Bedrock, EC2 Auto Scaling Groups, Kubernetes, AWS, Azure, GCP, AMI, LangGraph, LLM gateway, MCP, Aurora PostgreSQL, Memcached, Terraform, Terragrunt, Atlantis, Chef, Ansible, SSM, Docker, bash, Jenkins, Argo CD, New Relic, Splunk, Grafana, Prometheus, Fastly, Datadome, WAF, vLLM, LangSmith, Lambda
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Brooklyn or New York City or Los Angeles or Santa Monica or United States
$230k-$260k/yr HybridFull Time
Radix Health
Radix Health: Healthcare technology helping providers achieve fair reimbursement through integrated IDR software, data, and AI.
8+ YOE8+ years in SRE, infrastructure, platform engineering, or large-scale production systems; expertise in cloud infrastructure, distributed systems, networking, containers, orchestration, infrastructure as code, observability, automation, and incident management.
AI, HIPAA, PHI, SOC 2, 401(k)
1mo
Save
Mark Applied
Hide
Staff Engineer, Site Reliability
New York City or Los Angeles or United States or Canada
RemoteFull Time
Babylist
Babylist: Private baby-registry and e-commerce platform helping expecting parents plan, shop, and prepare.
Hands-on Terraform expertise, proven AWS experience (EKS, RDS, networking, CDNs), production Kubernetes operation, CI/CD design, observability and alerting, on-call/incident management, cross-functional developer support, familiarity with AI tooling.
Ruby on Rails, AWS, Sidekiq, MySQL, Redis, Terraform, EKS, RDS, Kubernetes, CircleCI, GitHub Actions, Datadog, Sentry, PagerDuty, Cronitor, Claude, ChatGPT
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Observability
Chicago or New York City
$160k-$200k/yr HybridFull Time
Ripple
Ripple: Enables institutions to move, manage, and tokenize value.
7+ YOE7+ years SRE/Platform experience focused on observability, New Relic, Terraform, PowerShell, Azure/AWS, incident management (Incident.IO/PagerDuty/OpsGenie), and coaching engineering teams.
New Relic, NRQL, Terraform, PowerShell, Azure, AWS, Azure DevOps, Octopus Deploy, Incident.IO, PagerDuty, OpsGenie, Slack, Python, Bash, Jira, SQL Server
2mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer (SRE)
Chicago or New York City
$175k-$220k/yr HybridFull Time
Optimal Market Technologies
Optimal Market Technologies: Private FINRA-registered broker-dealer providing options execution, ATS, routing, and algorithms to retail brokers and institutional trading firms.
Hands-on Linux and network administration, strong scripting/automation, Infrastructure-as-Code experience, production support and incident response ownership, staff management/mentoring, familiarity with C++, Python, SQL, Azure, and real-time trading systems.
C++, Python, SQL, Linux, CentOS7, RHEL9, Microsoft Azure, Microsoft Azure Virtual Desktop (AVD), Claude Code, FIX Protocol, PostgreSQL, AERON
2w
Save
Mark Applied
Hide
Sr. Site Reliability Engineer - Paze
Scottsdale or Chicago or San Francisco or New York City or Phoenix or California or Illinois
$106k-$156k/yr HybridFull Time
Early Warning Services
Early Warning Services: U.S. bank-owned fintech and consumer reporting agency providing identity, fraud-risk, and real-time payment solutions to financial institutions.
3+ YOEBachelor's degree in business, computer science, or related field; 3+ years of related technical or software development experience; Linux administration, Git, scripting, observability, incident management, and enterprise-scale experience required.
Linux, Git, Java, Ruby, Python, JavaScript, Go, AWS, Docker, Kubernetes, Swarm, CI/CD, TCP/UDP/IP
1w
Save
Mark Applied
Hide
Site Reliability Engineer- Team Lead
New York City or Europe or United States or Asia-Pacific
$170k-$190k/yr HybridFull Time
Pico
Pico: Private financial-markets technology providing trading infrastructure, connectivity, market data, software, and analytics to institutions.
Bachelor's degree or higher in engineering or related discipline, operational team leadership experience, Linux performance expertise, networking knowledge, financial technology experience, and programming or scripting skills.
Linux, Python, C, C++, Java
2w
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Playout
Stamford, Connecticut, United States
$145k-$175k/yr HybridFull Time
NBCUniversal
NBCUniversalNASDAQ: CMCSA: A leading global media and entertainment.
8+ YOEBachelor's degree or equivalent experience, 8 years of engineering experience in broadcast playout, Linux administration, cloud and networking expertise, monitoring, containerization, and 24/7 on-call availability.
Linux, Splunk, Grafana, ServiceNow, Docker, Kubernetes, AWS, Snell, Harris, Imagine, Amagi, Evertz, GrassValley, Harmonic, CoralBay, Veset, TS, HEVC, H.264, HLS, CMAF, SCTE-35, SCTE-224, ESAM, SRT, RIST, Slack
1w
Save
Mark Applied
Hide
Site Reliability Team Leader
Tel Aviv-Yafo or Israel or New York City or London or Edinburgh or Brazil or Estonia or Ukraine
OnsiteFull Time
Optimove
Optimove: Private SaaS providing AI-powered customer engagement and personalized marketing software to consumer brands.
5+ YOE5+ years in SRE, platform, DevOps, or infrastructure engineering; Kubernetes, GCP/AWS, programming, automation, CI/CD, observability, Linux, networking, distributed systems, and strong communication skills.
Kubernetes, GCP, AWS, Python, Go, Bash, CI/CD, Datadog, Prometheus, Grafana, Linux, Terraform, Ansible, Kafka, Pub/Sub, Redis, OpenTelemetry, Canary, Blue/Green, Feature Flags
2mo
Save
Mark Applied
Hide
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yr OnsiteFull Time
TikTok USDS Joint Venture LLC
TikTok USDS Joint Venture LLC: Ensuring U.S. data security and content integrity for TikTok.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
Python, Go, Java, Shell, Linux, Docker, Kubernetes, Prometheus, Grafana
4w
Save
Mark Applied
Hide
Staff , Site Reliability Engineer - Cloud Platform
New York City, New York, United States
$190k-$210k/yr HybridFull Time
Butterfly Network
Butterfly NetworkNew York Stock Exchange: BFLY: Public U.S. medical technology making handheld point-of-care ultrasound hardware and AI-powered clinical software for healthcare professionals.
8+ YOE8+ years managing production systems; deep hands-on AWS and Kubernetes experience; strong programming/scripting and automation skills; observability platform ownership; incident response leadership and mentoring ability.
Kubernetes, EKS, AWS, NewRelic, Datadog, DICOM, HL7, FHIR, PACS, VNA, EMR
2mo
Save
Mark Applied
Hide
Senior Software Engineer, Site Reliability Engineering
San Francisco or San Jose or New York City or Seattle or Austin or Washington or California or Massachusetts or New Jersey or Washington or United States
$179k-$273k/yr RemoteFull Time
Thumbtack
Thumbtack: Home services marketplace helping homeowners find and hire local professionals for repairs, maintenance, and improvements.
5+ YOE5+ years managing infrastructure and systems; extensive AWS and Linux fluency; proficiency in Python, Go, PHP, and JavaScript; experience with distributed systems, observability, and on-call rotations; strong communication and troubleshooting skills.
AWS, Linux, Python, Go, PHP, JavaScript, DNS, TLS, HTTP/S, TCP/IP
3w
Save
Mark Applied
Hide
Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid)
New York City or Austin or Sunnyvale or Redmond
$140k-$215k/yr HybridFull Time
CrowdStrike
CrowdStrikeNASDAQ: CRWD: AI-native platform for cybersecurity and threat protection.
10+ YOE10+ years building distributed systems, 5+ years developing SaaS microservices, expert programming skills, distributed-systems expertise, architectural leadership, and a Computer Science degree or equivalent experience.
Go, Java, Scala, Kotlin, Python, Node.js, Kubernetes, AWS, Cassandra, Kafka, Elasticsearch, OpenSearch, Google Cloud Platform (GCP), Oracle Cloud Infrastructure (OCI), GitHub, Stack Overflow

Explore Jobs

Expand Your Job Search