18 staff site reliability engineer jobs at 16 companies in Newark, NJ

1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Brooklyn or New York City or Los Angeles or Santa Monica or United States
$230k-$260k/yr HybridFull Time
Radix Health
Radix Health: Healthcare technology helping providers achieve fair reimbursement through integrated IDR software, data, and AI.
8+ YOE8+ years in SRE, infrastructure, platform engineering, or large-scale production systems; expertise in cloud infrastructure, distributed systems, networking, containers, orchestration, infrastructure as code, observability, automation, and incident management.
AI, HIPAA, PHI, SOC 2, 401(k)
2mo
Save
Mark Applied
Hide
Staff Engineer, Site Reliability
New York City or Los Angeles or United States or Canada
RemoteFull Time
Babylist
Babylist: Private baby-registry and e-commerce platform helping expecting parents plan, shop, and prepare.
Hands-on Terraform expertise, proven AWS experience (EKS, RDS, networking, CDNs), production Kubernetes operation, CI/CD design, observability and alerting, on-call/incident management, cross-functional developer support, familiarity with AI tooling.
Ruby on Rails, AWS, Sidekiq, MySQL, Redis, Terraform, EKS, RDS, Kubernetes, CircleCI, GitHub Actions, Datadog, Sentry, PagerDuty, Cronitor, Claude, ChatGPT
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
New York or Mountain View
$200k-$220k/yr HybridFull Time
ASAPP
ASAPP: ASAPP is a private enterprise AI software providing agentic customer-service platforms for contact centers.
10+ YOE10+ years production software experience; distributed systems; automation; AWS; Terraform; Python/Go; Kubernetes; on-call; strong communication.
Terraform, Python, Go, AWS, Kubernetes
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer (FedRAMP)
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Identity management and access control software provider.
8+ YOE8+ years operations experience in cloud and Linux, strong networking and web server knowledge, proficiency with Terraform/Chef and scripting (Bash, Python, Go), experience with automation tools and on-call duty.
AWS, Terraform, Chef, Ansible, Puppet, Apache httpd, nginx, Apache Tomcat, Bash, Python, Golang, git, gdb, strace, ltrace, tcpdump, Wireshark, Docker, Kubernetes
3mo
Save
Mark Applied
Hide
Senior Staff Site Reliability Engineer
San Francisco or New York City or Chicago
$245k-$270k/yr HybridFull Time
Ironclad
Ironclad: Private software providing AI contract lifecycle management tools for legal and business teams.
8+ YOE8+ years DevOps/SRE; 5+ years coding; Kubernetes and GCP expertise; build resilient infra; GitOps with Terraform/Pulumi, CircleCI, ArgoCD; AI tooling experience; strong communication; cross-functional collaboration.
Kubernetes, Google Cloud Platform, Terraform, Pulumi, CircleCI, ArgoCD, Claude Code, Cursor, Zed
1w
Save
Mark Applied
Hide
Staff Site Reliability Engineer
New York City, New York, United States
$241k-$270k/yr RemoteFull Time
Garner Health
Garner Health: Private healthtech helping employers and members find high-quality in-network doctors and reduce healthcare costs.
7+ YOE7+ years operating production cloud infrastructure at scale; deep Kubernetes and Terraform expertise; Python or Go skills; reliability practice design, mentoring, cost optimization, and regulated-environment experience preferred.
Amazon Web Services (AWS), Kubernetes, Terraform, Istio, Python, Go, TypeScript, Postgres, NATS, Datadog, GitLab, Claude
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Release Engineering
New York, New York, United States
$208k-$274k/yr HybridFull Time
Plaid
Plaid: Fintech data network connecting consumers’ financial accounts to apps and services.
8+ YOE8+ years in backend/SRE/platform engineering; experience designing SLO/SLI programs, progressive delivery, canary rollouts, metric-gated analysis, and automated rollback; proficiency in Go or similar; familiarity with Kubernetes, Prometheus, ArgoCD; strong leadership and incident response skills.
Go, Kubernetes, Prometheus, ArgoCD
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Toronto or New York City or North America
$144k-$200k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Unified data platform for building modern applications.
6+ YOE6+ years software development and distributed systems experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud platform expertise (AWS, GCP, Azure); Linux and networking knowledge; participate in 24/7 on-call.
Python, Go, Argo Workflows, ArgoCD, Kubernetes, AWS, Google Cloud Platform (GCP), Azure, Linux, CI/CD
1mo
Save
Mark Applied
Hide
Staff , Site Reliability Engineer - Cloud Platform
New York City, New York, United States
$190k-$210k/yr HybridFull Time
Butterfly Network
Butterfly NetworkNew York Stock Exchange: BFLY: Public U.S. medical technology making handheld point-of-care ultrasound hardware and AI-powered clinical software for healthcare professionals.
8+ YOE8+ years managing production systems; deep hands-on AWS and Kubernetes experience; strong programming/scripting and automation skills; observability platform ownership; incident response leadership and mentoring ability.
Kubernetes, EKS, AWS, NewRelic, Datadog, DICOM, HL7, FHIR, PACS, VNA, EMR
2w
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Playout
Stamford, Connecticut, United States
$145k-$175k/yr HybridFull Time
NBCUniversal
NBCUniversalNASDAQ: CMCSA: A leading global media and entertainment.
8+ YOEBachelor's degree or equivalent experience, 8 years of engineering experience in broadcast playout, Linux administration, cloud and networking expertise, monitoring, containerization, and 24/7 on-call availability.
Linux, Splunk, Grafana, ServiceNow, Docker, Kubernetes, AWS, Snell, Harris, Imagine, Amagi, Evertz, GrassValley, Harmonic, CoralBay, Veset, TS, HEVC, H.264, HLS, CMAF, SCTE-35, SCTE-224, ESAM, SRT, RIST, Slack
2mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer - Paze
San Francisco or Chicago or Scottsdale or New York City or Phoenix or Washington
$116k-$174k/yr HybridFull Time
Early Warning Services
Early Warning Services: U.S. bank-owned fintech and consumer reporting agency providing identity, fraud-risk, and real-time payment solutions to financial institutions.
8+ YOE8+ years SRE/engineering experience, bachelor’s degree preferred, hands-on with Python/Go/Java, Docker, CI/CD, Linux, cloud and messaging/datastore technologies; on-call rotation and incident leadership experience.
Python, Go, Java, Docker, Kafka, SQS, JMS, Oracle, Dynamo DB, Aurora, Redis, memcached, Linux, GIT, Chef, Maven, Jenkins, Kubernetes, Swarm, AWS
1mo
Save
Mark Applied
Hide
Sr/Staff Site Reliability Engineer, Consumer Apps
Chicago or New York City
HybridFull Time
Attain
Attain: Private consumer data and advertising technology helping brands measure, target, and optimize media using permissioned purchase data.
6+ YOEExperience building cloud-native infrastructure, strong automation and observability skills, fluency directing AI coding agents, Terraform/Helm/Kubernetes knowledge, and database/streaming experience.
Claude Code, Cursor, Terraform, Helm, Kubernetes, Istio, GCP, Google BigQuery, Spanner, CloudSQL, Prometheus, Grafana, GitLab, Docker, Kafka, Amazon Kinesis, AWS SNS, Google Pub/Sub, AWS Lambda, Google Cloud Functions, Google Cloud Run, Datadog, AWS
3w
Save
Mark Applied
Hide
Staff Platform Site Reliability Engineer
London or Toronto or New York City or Montreal or Kitchener or San Francisco
HybridFull Time
Index Exchange
Index Exchange: Independent ad-tech supply-side platform helping media owners monetize digital content and enabling brands to buy programmatic advertising.
8+ YOERequires 8+ years in platform engineering, SRE, infrastructure engineering, or DevOps; deep Linux and Kubernetes expertise; IaC at scale; Go or Python; networking; and cross-team technical strategy.
Kubernetes, Terraform, Ansible, GitOps, ArgoCD, Go, Python, Linux, EKS, GKE, Ceph, Hadoop, Spark, HBase, Kafka, Prometheus, Grafana, ELK, Mimir, Loki, Tempo, Vault, AWS, GCP
3d
Save
Mark Applied
Hide
Staff Site Reliability Engineer - Kubernetes
Bellevue or Chicago or New York City or San Francisco or Washington
$174k-$267k/yr HybridFull Time
Okta
OktaNASDAQ: OKTA: Identity management and access control software provider.
Requires 4+ years with Kubernetes, Helm, and Terraform; 5+ years with AWS; cloud-native architecture, automation, scripting, CI/CD, monitoring, and multi-region environments experience. U.S. Person status required.
Kubernetes, Helm, Karpenter, Istio, AWS, Amazon EKS, Amazon ECS, Amazon S3, Amazon VPC, Amazon RDS, AWS IAM, Terraform, AWS CloudFormation, Jenkins, GitLab, CircleCI, Ansible, Spinnaker, Python, Bash, Go, Prometheus, Grafana, Amazon CloudWatch, ELK Stack, Docker, CI/CD, RBAC
2mo
Save
Mark Applied
Hide
Lead Site Reliability Engineer (SRE)
Chicago or New York City
$175k-$220k/yr HybridFull Time
Optimal Market Technologies
Optimal Market Technologies: Private FINRA-registered broker-dealer providing options execution, ATS, routing, and algorithms to retail brokers and institutional trading firms.
Hands-on Linux and network administration, strong scripting/automation, Infrastructure-as-Code experience, production support and incident response ownership, staff management/mentoring, familiarity with C++, Python, SQL, Azure, and real-time trading systems.
C++, Python, SQL, Linux, CentOS7, RHEL9, Microsoft Azure, Microsoft Azure Virtual Desktop (AVD), Claude Code, FIX Protocol, PostgreSQL, AERON
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Boston or Miami or New Jersey or New York City or Princeton or Raleigh or Washington or Toronto or North America
$127k-$249k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Unified data platform for building modern applications.
6+ YOE6+ years software development experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud (AWS, Google Cloud Platform, Microsoft Azure) expertise; Linux and networking knowledge.
Argo Workflows, ArgoCD, Kubernetes, Python, Go, AWS, Google Cloud Platform (GCP), Microsoft Azure, Linux
2w
Save
Mark Applied
Hide
Staff+ Site Reliability Engineer, Safeguards ML Infra
San Francisco or Seattle or New York City
$405k-$485k/yr HybridFull Time
Anthropic
Anthropic: AI research developing safe and steerable AI systems.
8+ YOEProduction change-management experience, high-stakes release and on-call experience, AWS/GCP operations, Python proficiency, and a bachelor's degree or equivalent experience.
Python, Rust, AWS, GCP, AWS Bedrock, GCP Vertex, Claude
1mo
Save
Mark Applied
Hide
Member of Technical Staff - SRE
New York, New York, United States
$100k-$300k/yr OnsiteFull Time
Basis
Basis: Private AI software building autonomous accounting agents for accounting firms.
5+ YOE5+ years building and operating production infrastructure; strong software engineering; cloud, networking, databases, security; IaC, CI/CD, containerization; observability and incident management; on-call and incident leadership.
Terraform, CloudFormation, OpenTelemetry, Prometheus, Grafana, BetterStack, PagerDuty, Neon, Modal