15 staff site reliability engineer jobs at 13 companies in New York

1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Brooklyn or New York City or Los Angeles or Santa Monica or United States
$230k-$260k/yr HybridFull Time
Radix Health
Radix Health: Healthcare technology helping providers achieve fair reimbursement through integrated IDR software, data, and AI.
8+ YOE8+ years in SRE, infrastructure, platform engineering, or large-scale production systems; expertise in cloud infrastructure, distributed systems, networking, containers, orchestration, infrastructure as code, observability, automation, and incident management.
AI, HIPAA, PHI, SOC 2, 401(k)
2mo
Save
Mark Applied
Hide
Staff Engineer, Site Reliability
New York City or Los Angeles or United States or Canada
RemoteFull Time
Babylist
Babylist: Private baby-registry and e-commerce platform helping expecting parents plan, shop, and prepare.
Hands-on Terraform expertise, proven AWS experience (EKS, RDS, networking, CDNs), production Kubernetes operation, CI/CD design, observability and alerting, on-call/incident management, cross-functional developer support, familiarity with AI tooling.
Ruby on Rails, AWS, Sidekiq, MySQL, Redis, Terraform, EKS, RDS, Kubernetes, CircleCI, GitHub Actions, Datadog, Sentry, PagerDuty, Cronitor, Claude, ChatGPT
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer (FedRAMP)
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Identity management and access control software provider.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
AWS, Terraform, Chef, Ansible, Puppet, Apache httpd, nginx, Apache Tomcat, Bash, Python, Golang, git, gdb, strace, ltrace, tcpdump, Wireshark, Docker, Kubernetes
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
New York City, New York, United States
$241k-$270k/yr RemoteFull Time
Garner Health
Garner Health: Private healthtech helping employers and members find high-quality in-network doctors and reduce healthcare costs.
7+ YOE7+ years operating production cloud infrastructure at scale; deep Kubernetes and Terraform expertise; Python or Go skills; reliability practice design, mentoring, cost optimization, and regulated-environment experience preferred.
Amazon Web Services (AWS), Kubernetes, Terraform, Istio, Python, Go, TypeScript, Postgres, NATS, Datadog, GitLab, Claude
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Release Engineering
New York, New York, United States
$208k-$274k/yr HybridFull Time
Plaid
Plaid: Fintech data network connecting consumers’ financial accounts to apps and services.
8+ YOE8+ years in backend/SRE/platform engineering; experience designing SLO/SLI programs, progressive delivery, canary rollouts, metric-gated analysis, and automated rollback; proficiency in Go or similar; familiarity with Kubernetes, Prometheus, ArgoCD; strong leadership and incident response skills.
Go, Kubernetes, Prometheus, ArgoCD
5d
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Santa Barbara or California or Illinois or Texas or Minnesota or Colorado or Georgia or New York or Massachusetts or Connecticut
$175k-$230k/yr HybridFull Time
PayJunction
PayJunction: Privately held U.S. payment processor serving businesses with in-store, online, and mobile payments.
10+ years relevant experience; 5+ years Linux administration; 3+ years AWS, containers, infrastructure as code, configuration management, and automation scripting; physical server and data center experience required.
AWS, Terraform, Puppet, OpenVox, Ansible, Linux, containers, scripting, BI
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Toronto or New York City or North America
$144k-$200k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Unified data platform for building modern applications.
6+ YOE6+ years software development and distributed systems experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud platform expertise (AWS, GCP, Azure); Linux and networking knowledge; participate in 24/7 on-call.
Python, Go, Argo Workflows, ArgoCD, Kubernetes, AWS, Google Cloud Platform (GCP), Azure, Linux, CI/CD
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer - Paze
San Francisco or Chicago or Scottsdale or New York City or Phoenix or Washington
$116k-$174k/yr HybridFull Time
Early Warning Services
Early Warning Services: U.S. bank-owned fintech and consumer reporting agency providing identity, fraud-risk, and real-time payment solutions to financial institutions.
8+ YOE8+ years SRE/engineering experience, bachelor’s degree preferred, hands-on with Python/Go/Java, Docker, CI/CD, Linux, cloud and messaging/datastore technologies; on-call rotation and incident leadership experience.
Python, Go, Java, Docker, Kafka, SQS, JMS, Oracle, Dynamo DB, Aurora, Redis, memcached, Linux, GIT, Chef, Maven, Jenkins, Kubernetes, Swarm, AWS
1mo
Save
Mark Applied
Hide
Sr/Staff Site Reliability Engineer, Consumer Apps
Chicago or New York City
HybridFull Time
Attain
Attain: Private consumer data and advertising technology helping brands measure, target, and optimize media using permissioned purchase data.
6+ YOEExperience building cloud-native infrastructure, strong automation and observability skills, fluency directing AI coding agents, Terraform/Helm/Kubernetes knowledge, and database/streaming experience.
Claude Code, Cursor, Terraform, Helm, Kubernetes, Istio, GCP, Google BigQuery, Spanner, CloudSQL, Prometheus, Grafana, GitLab, Docker, Kafka, Amazon Kinesis, AWS SNS, Google Pub/Sub, AWS Lambda, Google Cloud Functions, Google Cloud Run, Datadog, AWS
3w
Save
Mark Applied
Hide
Staff Platform Site Reliability Engineer
London or Toronto or New York City or Montreal or Kitchener or San Francisco
HybridFull Time
Index Exchange
Index Exchange: Independent ad-tech supply-side platform helping media owners monetize digital content and enabling brands to buy programmatic advertising.
8+ YOERequires 8+ years in platform engineering, SRE, infrastructure engineering, or DevOps; deep Linux and Kubernetes expertise; IaC at scale; Go or Python; networking; and cross-team technical strategy.
Kubernetes, Terraform, Ansible, GitOps, ArgoCD, Go, Python, Linux, EKS, GKE, Ceph, Hadoop, Spark, HBase, Kafka, Prometheus, Grafana, ELK, Mimir, Loki, Tempo, Vault, AWS, GCP
5d
Save
Mark Applied
Hide
Staff Site Reliability Engineer - Kubernetes
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Identity management and access control software provider.
Requires 4+ years with Kubernetes, Helm, and Terraform; 5+ years with AWS; cloud-native architecture, automation, scripting, CI/CD, monitoring, and multi-region environments experience. U.S. Person status required.
Kubernetes, Helm, Karpenter, Istio, AWS, Amazon EKS, Amazon ECS, Amazon S3, Amazon VPC, Amazon RDS, AWS IAM, Terraform, AWS CloudFormation, Jenkins, GitLab, CircleCI, Ansible, Spinnaker, Python, Bash, Go, Prometheus, Grafana, Amazon CloudWatch, ELK Stack, Docker, CI/CD, RBAC
2mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer (SRE)
Chicago or New York City
$175k-$220k/yr HybridFull Time
Optimal Market Technologies
Optimal Market Technologies: Private FINRA-registered broker-dealer providing options execution, ATS, routing, and algorithms to retail brokers and institutional trading firms.
Hands-on Linux and network administration, strong scripting/automation, Infrastructure-as-Code experience, production support and incident response ownership, staff management/mentoring, familiarity with C++, Python, SQL, Azure, and real-time trading systems.
C++, Python, SQL, Linux, CentOS7, RHEL9, Microsoft Azure, Microsoft Azure Virtual Desktop (AVD), Claude Code, FIX Protocol, PostgreSQL, AERON
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Boston or Miami or New Jersey or New York City or Princeton or Raleigh or Washington or Toronto or North America
$127k-$249k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Unified data platform for building modern applications.
6+ YOE6+ years software development experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud (AWS, Google Cloud Platform, Microsoft Azure) expertise; Linux and networking knowledge.
Argo Workflows, ArgoCD, Kubernetes, Python, Go, AWS, Google Cloud Platform (GCP), Microsoft Azure, Linux
3w
Save
Mark Applied
Hide
Staff+ Site Reliability Engineer, Safeguards ML Infra
San Francisco or Seattle or New York City
$405k-$485k/yr HybridFull Time
Anthropic
Anthropic: AI research developing safe and steerable AI systems.
8+ YOEProduction change-management experience, high-stakes release and on-call experience, AWS/GCP operations, Python proficiency, and a bachelor's degree or equivalent experience.
Python, Rust, AWS, GCP, AWS Bedrock, GCP Vertex, Claude
1mo
Save
Mark Applied
Hide
Member of Technical Staff - SRE
New York, New York, United States
$100k-$300k/yr OnsiteFull Time
Basis
Basis: Private AI software building autonomous accounting agents for accounting firms.
5+ YOE5+ years building and operating production infrastructure; strong software engineering; cloud, networking, databases, security; IaC, CI/CD, containerization; observability and incident management; on-call and incident leadership.
Terraform, CloudFormation, OpenTelemetry, Prometheus, Grafana, BetterStack, PagerDuty, Neon, Modal