95 cloud reliability engineer jobs at 75 companies in New York

2w
Save
Mark Applied
Hide
Cloud Reliability Engineer
Englewood Cliffs or New York City
$135k-$165k/yr HybridFull Time
Versant
VersantNasdaq: VSNT: Operates cable television networks and digital media entertainment platforms.
3+ YOEBachelor's degree or equivalent experience; 3–7 years in SRE/Cloud/DevOps roles; strong AWS experience (enterprise scale), Terraform and CloudFormation, scripting (Python/PowerShell/Bash), CI/CD, monitoring/observability, incident management.
AWS, AWS Organizations, Control Tower, Identity Center, Terraform, CloudFormation, Python, PowerShell, Bash, CI/CD
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York, New York, United States
$177k-$209k/yr HybridFull Time
Peloton
PelotonNASDAQ: PTON: Sells connected fitness equipment and streaming exercise classes.
Kubernetes expertise; observability/monitoring mindset; CI/CD experience; IaC (Terraform/Pulumi); cloud (AWS); security and reliability focus; programming (Python/Go/Java/C).
Kubernetes, Observability, Jenkins, ArgoCD, Harness, Tekton, Terraform, Pulumi, AWS, Python, Golang, Java, C, Nginx, Ubuntu, Chef
1w
Save
Mark Applied
Hide
Site Reliability Engineer
New York, New York, United States
HybridFull Time
Mistral AI
Mistral AI: Developing frontier artificial intelligence models and enterprise AI solutions.
7+ YOE7+ years SRE/DevOps experience, Master’s in CS/Engineering or related, strong cloud and distributed systems skills, Kubernetes/CI-CD/infra-as-code proficiency, scripting experience, observability and on-call experience.
Kubernetes, Flux, Terraform, Docker, Prometheus, Grafana, ELK Stack, Datadog, CloudFormation, Python, Go, Bash, Slurm, Fluidstack, Coreweave, Vast
3mo
Save
Mark Applied
Hide
Senior / Staff Site Reliability Engineer
New York, New York, United States
$175k-$230k/yr HybridFull Time
Sage
Sage: Modern software and sensors for senior living communities.
7+ YOE7-12+ years in software/infrastructure engineering; expert in cloud, networks, databases, and automation; strong SRE practices; capable of leading incident response and reliability initiatives.
Datadog, Prometheus, Grafana, Terraform, Pulumi, Kubernetes, AWS, Amazon Web Services, Google Cloud Platform, PostgreSQL, MySQL, Go, Python, Java
1w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
New York City or Toronto
$184k-$240k/yr OnsiteFull Time
Movable Ink
Movable Ink: Provides AI-powered content personalization for digital marketing campaigns.
6+ YOE6+ years SRE/Software Engineering experience designing and operating scalable, multi-cloud distributed systems; expertise with observability, IaC, Kubernetes, and multiple programming languages.
Apache Pulsar, Apache Kafka, Grafana Loki, ScyllaDB, Cassandra, Prometheus, Thanos, Grafana Alloy, Tempo, Terraform, Chef, EKS, GKE, NodeJS, Golang, Ruby, Python, shell
2w
Save
Mark Applied
Hide
Senior Database Reliability Engineer
Southfield or San Francisco or Seattle or Boston or New York City or Los Angeles or San Diego or United States
$104k-$153k/yr RemoteFull Time
Credit Acceptance
Credit AcceptanceNASDAQ: CACC: Provides vehicle financing programs for subprime credit consumers.
7+ YOE7+ years in DB engineering/DBA/DBRE roles; strong SQL, performance tuning, HA, backup/recovery; bachelor's or equivalent experience; experience with cloud DBs, automation, observability, and on-call rotations.
Terraform, Ansible, AWS DMS, AWS RDS / Aurora (PostgreSQL preferred), DynamoDB, Oracle, SQL Server, MongoDB, MySQL, Datadog, Grafana, Prometheus, OpenTelemetry, Jenkins, GitHub Actions, Python, CyberArk, Liquibase, Flyway, Hibernate, JPA, SQLAlchemy, REST APIs, Linux/Unix, SQL
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Robotics & Cloud Infrastructure
Brooklyn or New York City or Richmond or Europe
$164k-$220k/yr RemoteFull Time
Bedrock Ocean Exploration
Bedrock Ocean Exploration: Maps the ocean floor using autonomous underwater robotic vehicles.
5+ YOE5+ years SRE/DevOps experience with on-call ownership; strong automation using Python/Go/Bash; Terraform and AWS hands-on; containerization (Docker, Kubernetes); observability (Prometheus, Grafana); Linux and networking expertise; East Coast location and US work authorization required.
Python, Go, Bash, Terraform, AWS, Docker, Kubernetes, Prometheus, Grafana, ROS 2, ROS, Jetson, Linux, IAM
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States or San Francisco or New York City
$101k-$199k/yr OnsiteFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
1+ YOEMaster's or Bachelor's in CS/IT (or equivalent experience), 1+ years managing physical infrastructure, on-call experience, experience with large-scale cloud/distributed systems preferred, and ability to pass Microsoft security screening.
Azure, InfiniBand, GPUs
4w
Save
Mark Applied
Hide
Customer Reliability Engineer, Airflow
United States or San Francisco or Boston or Atlanta or Austin or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus
$125k-$130k/yr RemoteFull Time
Astronomer
Astronomer: Managed data orchestration platform powered by Apache Airflow.
4+ YOEData engineering background, 4 years Python, 1 year Airflow administration/DAG creation, Kubernetes/Docker experience, cloud provider (AWS/GCP/Azure) experience, troubleshooting, strong communication, and mentoring experience.
Apache Airflow, Python, Kubernetes, Docker, AWS, GCP, Azure, SQL, PostgreSQL, Databricks, Snowflake, Redshift, dbt, Zoom
4w
Save
Mark Applied
Hide
Production Reliability Engineer
Garden City, New York, United States
$50k-$70k/yr HybridFull Time
Friedman Vartolo
Friedman Vartolo: Specializes in real estate law and default legal services.
5+ YOE5+ years SRE/DevOps experience, incident-response leadership, SLO/error-budget experience, strong Azure (or AWS/GCP) knowledge, observability tooling, ability to read TypeScript/Python/C#, and cloud security fundamentals.
Azure, AWS, GCP, Azure Monitor, Application Insights, Datadog, New Relic, TypeScript, Python, C#, Huntress, Microsoft Defender, Vanta, Databricks, SOC2, ISO 27001
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City, New York, United States
$240k-$300k/yr OnsiteFull Time
Legora: AI workspace for legal document research and drafting.
Extensive experience operating and improving production systems; strong automation, observability, and incident management; proficiency with cloud and Kubernetes.
Kubernetes, Cloud, Observability, Automation, Monitoring
2d
Save
Mark Applied
Hide
Site Reliability Engineer - VP
New York City, New York, United States
$170k-$230k/yr OnsiteFull Time
Barclays
BarclaysLondon Stock Exchange: BARC: Global bank providing retail, corporate, and investment financial services.
Proven SRE experience implementing reliability practices, defining SLOs/error budgets, incident response, cloud platforms, IaC, observability, capacity planning, and mentoring teams.
AWS, Azure, GCP, Prometheus, CloudWatch, Terraform, CloudFormation, Kubernetes, REST API
1mo
Save
Mark Applied
Hide
GOV Site Reliability Engineer
United States or Kansas or Washington or California or Texas or Illinois or North Carolina or Colorado or Massachusetts or Pennsylvania or Virginia or Oregon or Nevada or Hawaii or New York or Georgia or Ohio or Arizona or Seattle or San Francisco or New York City
$110k-$183k/yr RemoteFull Time
Veeam
Veeam: Data resilience and security for hybrid cloud environments
3+ YOE3+ years in software engineering with 1+ year in SRE/Platform/DevOps, cloud experience (Azure or comparable), observability (Prometheus, Grafana, OpenTelemetry, ELK), IaC (Terraform/Terragrunt/Pulumi), Kubernetes, CI/CD tooling, programming in TypeScript/JS, Go, Java, or C#, and experience in compliance-oriented environments.
VDC, Prometheus, Grafana, OpenTelemetry, ELK stack, Terraform, Terragrunt, Pulumi, Kubernetes, GitHub Actions, Azure DevOps, GitLab CI, ArgoCD, TypeScript, JS, Go, Java, C#, Azure Government, AWS GovCloud
3mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Bellevue or New York City
$164k-$213k/yr HybridFull Time
iSpot.tv
iSpot.tv: Provides real-time measurement and analytics for TV advertising.
10+ YOE3+ Mgmt10+ years in software engineering, cloud architecture, and/or SRE; strong AWS, Kubernetes, Terraform; Spark; CI/CD; leadership; excellent communication.
AWS, EKS, ECR, RDS, SQS, SNS, VPC, MWAA, S3, Terraform, CloudFormation, Kubernetes, kubectl, Helm, ArgoCD, CircleCI, Spark, EMR, Databricks, Glue, OTel, DataDog, GenAI tools
2w
Save
Mark Applied
Hide
Site Reliability Engineer IV
Buffalo, New York, United States
$140k-$233k/yr HybridFull Time
M&T Bank
M&T BankNYSE: MTB: Provides retail, commercial, and wealth management banking services.
7+ YOEExpert-level SRE experience, 7+ years systems analysis/application development (or 9+ with associate), advanced scripting, observability, IaC (Terraform), and cloud (Azure) experience.
OpenTelemetry (OTel), Dynatrace, Terraform, Microsoft Azure, Azure Monitor, Application Insights, PowerShell, Python, Bash
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Fabric
New York City or Toronto or United States
$127k-$249k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
10+ YOE10+ years in distributed systems with strong networking, cloud (AWS/Azure/GCP), service mesh, multi-cloud design, and on-call experience.
TCP/IP, DNS, TLS/mTLS, BGP, VPN, VPC, Subnets, CDNs, Service Mesh, Kubernetes, AWS, Azure, GCP
6d
Save
Mark Applied
Hide
Site Reliability Engineer II
Scottsdale or San Francisco or Chicago or New York City
$86k-$126k/yr HybridFull Time
Early Warning Services
Early Warning Services: Operates payment and risk solutions for the financial industry.
2+ YOEBachelor's or equivalent, minimum 2 years DevOps/Dev/SRE experience, Linux/Unix experience, infrastructure automation (Chef/Ansible/Puppet, Terraform), containerization (Docker,Kubernetes), cloud (AWS/GCP/Azure), on-call rotation.
Linux, Unix, Chef, Ansible, Puppet, Terraform, Docker, Kubernetes, AWS, GCP, Azure, Java, Ruby, Python, JavaScript, Go
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
Python, Go, Java, Shell, Linux, Docker, Kubernetes, Prometheus, Grafana
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
$111k-$160k/yr HybridFull Time
Mizuho Financial Group
Mizuho Financial GroupTokyo Stock Exchange: 8411: Global financial group providing banking and investment services.
3+ YOEBachelor’s degree or equivalent; 3+ years SRE/automation experience; strong Grafana, cloud (AWS/Azure/GCP), containers (Docker/Kubernetes); CI/CD; scripting (Python/Bash/Go); on-call experience.
Grafana, Ansible, Terraform, Jenkins, AWS, Azure, Google Cloud, Docker, Kubernetes, CI/CD, Python, Bash, Go
3mo
Save
Mark Applied
Hide
Site Reliability and Infrastructure Engineer
New York City, New York, United States
$160k-$215k/yr HybridFull Time
Treeswift
Treeswift: Provides AI-augmented vegetation management and asset monitoring for utilities.
7+ YOEExperienced SRE/infrastructure engineer with 7-10 years in observability, cloud, and DevOps; strong Linux, IaC, Kubernetes, and Python.
Terraform, Kubernetes, AWS, Airflow, Astronomer, ECR, S3, SQS, Lambda, Step Functions, ECS, Docker

Explore Jobs

Expand Your Job Search