18 staff site reliability engineer jobs at 16 companies in Queens, NY

1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Brooklyn or New York City or Los Angeles or Santa Monica or United States
$230k-$260k/yr HybridFull Time
Pivotal Health
Pivotal Health: AI platform automating healthcare insurance claim disputes for providers.
8+ YOE8+ years in SRE, infrastructure, platform engineering, or large-scale production systems; expertise in cloud infrastructure, distributed systems, networking, containers, orchestration, infrastructure as code, observability, automation, and incident management.
AI, HIPAA, PHI, SOC 2, 401(k)
1mo
Save
Mark Applied
Hide
Staff Engineer, Site Reliability
New York City or Los Angeles or United States or Canada
RemoteFull Time
Babylist
Babylist: Operates an e-commerce platform and universal registry for baby products.
Hands-on Terraform expertise, proven AWS experience (EKS, RDS, networking, CDNs), production Kubernetes operation, CI/CD design, observability and alerting, on-call/incident management, cross-functional developer support, familiarity with AI tooling.
Ruby on Rails, AWS, Sidekiq, MySQL, Redis, Terraform, EKS, RDS, Kubernetes, CircleCI, GitHub Actions, Datadog, Sentry, PagerDuty, Cronitor, Claude, ChatGPT
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
New York or Mountain View
$200k-$220k/yr HybridFull Time
ASAPP
ASAPP: Develops AI-powered software to automate enterprise contact center interactions.
10+ YOE10+ years production software experience; distributed systems; automation; AWS; Terraform; Python/Go; Kubernetes; on-call; strong communication.
Terraform, Python, Go, AWS, Kubernetes
3mo
Save
Mark Applied
Hide
Senior Staff Site Reliability Engineer
San Francisco or New York City or Chicago
$245k-$270k/yr HybridFull Time
Ironclad
Ironclad: AI platform for digital contract lifecycle management
8+ YOE8+ years DevOps/SRE; 5+ years coding; Kubernetes and GCP expertise; build resilient infra; GitOps with Terraform/Pulumi, CircleCI, ArgoCD; AI tooling experience; strong communication; cross-functional collaboration.
Kubernetes, Google Cloud Platform, Terraform, Pulumi, CircleCI, ArgoCD, Claude Code, Cursor, Zed
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer (FedRAMP)
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Provide secure identity management and authentication for enterprises.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
AWS, Terraform, Chef, Ansible, Puppet, Apache httpd, nginx, Apache Tomcat, Bash, Python, Golang, git, gdb, strace, ltrace, tcpdump, Wireshark, Docker, Kubernetes
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
New York City, New York, United States
$241k-$270k/yr RemoteFull Time
Garner Health
Garner Health: Identifies high-quality doctors to lower employer healthcare costs.
7+ YOE7+ years operating production cloud infrastructure at scale; deep Kubernetes and Terraform expertise; Python or Go skills; reliability practice design, mentoring, cost optimization, and regulated-environment experience preferred.
Amazon Web Services (AWS), Kubernetes, Terraform, Istio, Python, Go, TypeScript, Postgres, NATS, Datadog, GitLab, Claude
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Release Engineering
New York, New York, United States
$208k-$274k/yr HybridFull Time
Plaid
Plaid: Provides financial data connectivity and payment infrastructure via APIs.
8+ YOE8+ years in backend/SRE/platform engineering; experience designing SLO/SLI programs, progressive delivery, canary rollouts, metric-gated analysis, and automated rollback; proficiency in Go or similar; familiarity with Kubernetes, Prometheus, ArgoCD; strong leadership and incident response skills.
Go, Kubernetes, Prometheus, ArgoCD
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Toronto or New York City or North America
$144k-$200k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development and distributed systems experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud platform expertise (AWS, GCP, Azure); Linux and networking knowledge; participate in 24/7 on-call.
Python, Go, Argo Workflows, ArgoCD, Kubernetes, AWS, Google Cloud Platform (GCP), Azure, Linux, CI/CD
4w
Save
Mark Applied
Hide
Staff , Site Reliability Engineer - Cloud Platform
New York City, New York, United States
$190k-$210k/yr HybridFull Time
Butterfly Network
Butterfly NetworkNYSE: BFLY: Handheld whole-body ultrasound scanners powered by semiconductor technology.
8+ YOE8+ years managing production systems; deep hands-on AWS and Kubernetes experience; strong programming/scripting and automation skills; observability platform ownership; incident response leadership and mentoring ability.
Kubernetes, EKS, AWS, NewRelic, Datadog, DICOM, HL7, FHIR, PACS, VNA, EMR
2w
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Playout
Stamford, Connecticut, United States
$145k-$175k/yr HybridFull Time
NBCUniversal
NBCUniversalNASDAQ: CMCSA: Produces and distributes entertainment content and theme park experiences.
8+ YOEBachelor's degree or equivalent experience, 8 years of engineering experience in broadcast playout, Linux administration, cloud and networking expertise, monitoring, containerization, and 24/7 on-call availability.
Linux, Splunk, Grafana, ServiceNow, Docker, Kubernetes, AWS, Snell, Harris, Imagine, Amagi, Evertz, GrassValley, Harmonic, CoralBay, Veset, TS, HEVC, H.264, HLS, CMAF, SCTE-35, SCTE-224, ESAM, SRT, RIST, Slack
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer - Paze
San Francisco or Chicago or Scottsdale or New York City or Phoenix or Washington
$116k-$174k/yr HybridFull Time
Early Warning Services
Early Warning Services: Operates payment and risk solutions for the financial industry.
8+ YOE8+ years SRE/engineering experience, bachelor’s degree preferred, hands-on with Python/Go/Java, Docker, CI/CD, Linux, cloud and messaging/datastore technologies; on-call rotation and incident leadership experience.
Python, Go, Java, Docker, Kafka, SQS, JMS, Oracle, Dynamo DB, Aurora, Redis, memcached, Linux, GIT, Chef, Maven, Jenkins, Kubernetes, Swarm, AWS
1mo
Save
Mark Applied
Hide
Sr/Staff Site Reliability Engineer, Consumer Apps
Chicago or New York City
HybridFull Time
Attain
Attain: Provides real-time purchase data and measurement solutions for marketers.
6+ YOEExperience building cloud-native infrastructure, strong automation and observability skills, fluency directing AI coding agents, Terraform/Helm/Kubernetes knowledge, and database/streaming experience.
Claude Code, Cursor, Terraform, Helm, Kubernetes, Istio, GCP, Google BigQuery, Spanner, CloudSQL, Prometheus, Grafana, GitLab, Docker, Kafka, Amazon Kinesis, AWS SNS, Google Pub/Sub, AWS Lambda, Google Cloud Functions, Google Cloud Run, Datadog, AWS
3w
Save
Mark Applied
Hide
Staff Platform Site Reliability Engineer
London or Toronto or New York City or Montreal or Kitchener or San Francisco
HybridFull Time
Index Exchange
Index Exchange: Independent ad-tech supply-side platform helping media owners monetize digital content and enabling brands to buy programmatic advertising.
8+ YOERequires 8+ years in platform engineering, SRE, infrastructure engineering, or DevOps; deep Linux and Kubernetes expertise; IaC at scale; Go or Python; networking; and cross-team technical strategy.
Kubernetes, Terraform, Ansible, GitOps, ArgoCD, Go, Python, Linux, EKS, GKE, Ceph, Hadoop, Spark, HBase, Kafka, Prometheus, Grafana, ELK, Mimir, Loki, Tempo, Vault, AWS, GCP
9h
Save
Mark Applied
Hide
Staff Site Reliability Engineer - Kubernetes
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Identity management and access control software provider.
Requires 4+ years with Kubernetes, Helm, and Terraform; 5+ years with AWS; cloud-native architecture, automation, scripting, CI/CD, monitoring, and multi-region environments experience. U.S. Person status required.
Kubernetes, Helm, Karpenter, Istio, AWS, Amazon EKS, Amazon ECS, Amazon S3, Amazon VPC, Amazon RDS, AWS IAM, Terraform, AWS CloudFormation, Jenkins, GitLab, CircleCI, Ansible, Spinnaker, Python, Bash, Go, Prometheus, Grafana, Amazon CloudWatch, ELK Stack, Docker, CI/CD, RBAC
1mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer (SRE)
Chicago or New York City
$175k-$220k/yr HybridFull Time
Optimal Market Technologies
Optimal Market Technologies: Operates a broker-dealer platform for wholesale options execution.
Hands-on Linux and network administration, strong scripting/automation, Infrastructure-as-Code experience, production support and incident response ownership, staff management/mentoring, familiarity with C++, Python, SQL, Azure, and real-time trading systems.
C++, Python, SQL, Linux, CentOS7, RHEL9, Microsoft Azure, Microsoft Azure Virtual Desktop (AVD), Claude Code, FIX Protocol, PostgreSQL, AERON
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Boston or Miami or New Jersey or New York City or Princeton or Raleigh or Washington or Toronto or North America
$127k-$249k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud (AWS, Google Cloud Platform, Microsoft Azure) expertise; Linux and networking knowledge.
Argo Workflows, ArgoCD, Kubernetes, Python, Go, AWS, Google Cloud Platform (GCP), Microsoft Azure, Linux
2w
Save
Mark Applied
Hide
Staff+ Site Reliability Engineer, Safeguards ML Infra
San Francisco or Seattle or New York City
$405k-$485k/yr HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
8+ YOEProduction change-management experience, high-stakes release and on-call experience, AWS/GCP operations, Python proficiency, and a bachelor's degree or equivalent experience.
Python, Rust, AWS, GCP, AWS Bedrock, GCP Vertex, Claude
1mo
Save
Mark Applied
Hide
Member of Technical Staff - SRE
New York, New York, United States
$100k-$300k/yr OnsiteFull Time
Basis
Basis: AI agents that automate complex accounting and tax workflows.
5+ YOE5+ years building and operating production infrastructure; strong software engineering; cloud, networking, databases, security; IaC, CI/CD, containerization; observability and incident management; on-call and incident leadership.
Terraform, CloudFormation, OpenTelemetry, Prometheus, Grafana, BetterStack, PagerDuty, Neon, Modal