106 cloud reliability engineer jobs at 78 companies in Suffern, NY

4w
Save
Mark Applied
Hide
Cloud Reliability Engineer
Englewood Cliffs or New York City
$135k-$165k/yr HybridFull Time
Versant
VersantNasdaq: VSNT: Operates cable television networks and digital media entertainment platforms.
3+ YOEBachelor's degree or equivalent experience; 3–7 years in SRE/Cloud/DevOps roles; strong AWS experience (enterprise scale), Terraform and CloudFormation, scripting (Python/PowerShell/Bash), CI/CD, monitoring/observability, incident management.
AWS, AWS Organizations, Control Tower, Identity Center, Terraform, CloudFormation, Python, PowerShell, Bash, CI/CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with GCP exposure, Terraform/IaC, CI/CD, observability, DevSecOps practices, incident response, and automation for reliability and resiliency.
GCP, Azure, Terraform, Terraform Enterprise, Log Analytics, Dynatrace, Resource Graph, CI/CD, IAM
2mo
Save
Mark Applied
Hide
Senior Associate - Platform Reliability Engineer - Cloud Technology
Lebanon, New Jersey, United States
$145k-$212k/yr OnsiteFull Time
New York Life Insurance
New York Life Insurance: Providing life insurance, annuities, and long-term investment solutions.
4+ YOE4+ years (with Master's) or 6+ years (with Bachelor's) experience delivering automated, scalable cloud solutions; experience with AWS/Azure, Terraform, CI/CD (Jenkins/Azure DevOps), monitoring, security remediation, and related enterprise tools.
Terraform, GitHub, Jenkins, AWS, Microsoft Azure, AWS CloudFormation, Wiz, Microsoft Azure DevOps, New Relic, Artifactory, Nexus, SonarQube, Fisheye
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York, New York, United States
$177k-$209k/yr HybridFull Time
Peloton
PelotonNASDAQ: PTON: Sells connected fitness equipment and streaming exercise classes.
Kubernetes expertise; observability/monitoring mindset; CI/CD experience; IaC (Terraform/Pulumi); cloud (AWS); security and reliability focus; programming (Python/Go/Java/C).
Kubernetes, Observability, Jenkins, ArgoCD, Harness, Tekton, Terraform, Pulumi, AWS, Python, Golang, Java, C, Nginx, Ubuntu, Chef
2w
Save
Mark Applied
Hide
Site Reliability Engineer
New York, New York, United States
HybridFull Time
Mistral AI
Mistral AI: Developing frontier artificial intelligence models and enterprise AI solutions.
7+ YOE7+ years SRE/DevOps experience, Master’s in CS/Engineering or related, strong cloud and distributed systems skills, Kubernetes/CI-CD/infra-as-code proficiency, scripting experience, observability and on-call experience.
Kubernetes, Flux, Terraform, Docker, Prometheus, Grafana, ELK Stack, Datadog, CloudFormation, Python, Go, Bash, Slurm, Fluidstack, Coreweave, Vast
3h
Save
Mark Applied
Hide
Staff , Site Reliability Engineer - Cloud Platform
New York City, New York, United States
$190k-$210k/yr HybridFull Time
Butterfly Network
Butterfly NetworkNYSE: BFLY: Handheld whole-body ultrasound scanners powered by semiconductor technology.
8+ YOE8+ years managing production systems; deep hands-on AWS and Kubernetes experience; strong programming/scripting and automation skills; observability platform ownership; incident response leadership and mentoring ability.
Kubernetes, EKS, AWS, NewRelic, Datadog, DICOM, HL7, FHIR, PACS, VNA, EMR
2w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
New York City or Toronto
$184k-$240k/yr OnsiteFull Time
Movable Ink
Movable Ink: Provides AI-powered content personalization for digital marketing campaigns.
6+ YOE6+ years SRE/Software Engineering experience designing and operating scalable, multi-cloud distributed systems; expertise with observability, IaC, Kubernetes, and multiple programming languages.
Apache Pulsar, Apache Kafka, Grafana Loki, ScyllaDB, Cassandra, Prometheus, Thanos, Grafana Alloy, Tempo, Terraform, Chef, EKS, GKE, NodeJS, Golang, Ruby, Python, shell
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Robotics & Cloud Infrastructure
Brooklyn or New York City or Richmond or Europe
$164k-$220k/yr RemoteFull Time
Bedrock Ocean Exploration
Bedrock Ocean Exploration: Maps the ocean floor using autonomous underwater robotic vehicles.
5+ YOE5+ years SRE/DevOps experience with on-call ownership; strong automation using Python/Go/Bash; Terraform and AWS hands-on; containerization (Docker, Kubernetes); observability (Prometheus, Grafana); Linux and networking expertise; East Coast location and US work authorization required.
Python, Go, Bash, Terraform, AWS, Docker, Kubernetes, Prometheus, Grafana, ROS 2, ROS, Jetson, Linux, IAM
1mo
Save
Mark Applied
Hide
Customer Reliability Engineer, Airflow
United States or San Francisco or Boston or Atlanta or Austin or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus
$125k-$130k/yr RemoteFull Time
Astronomer
Astronomer: Managed data orchestration platform powered by Apache Airflow.
4+ YOEData engineering background, 4 years Python, 1 year Airflow administration/DAG creation, Kubernetes/Docker experience, cloud provider (AWS/GCP/Azure) experience, troubleshooting, strong communication, and mentoring experience.
Apache Airflow, Python, Kubernetes, Docker, AWS, GCP, Azure, SQL, PostgreSQL, Databricks, Snowflake, Redshift, dbt, Zoom
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Greenwich, Connecticut, United States
HybridFull Time
Interactive Brokers
Interactive BrokersNASDAQ: IBKR: Automated global electronic brokerage and trading services provider.
5+ YOE5+ years experience in Linux/Unix, networking and coding; experience with cloud (AWS or Azure), Terraform or CloudFormation, Docker and Kubernetes; bachelor's or master's in CS/STEM; CI/CD, on-call rotation, mentoring skills.
CI/CD, Terraform, CloudFormation, AWS, Azure, Docker, Kubernetes, Linux, Unix
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City, New York, United States
$240k-$300k/yr OnsiteFull Time
Legora: AI workspace for legal document research and drafting.
Extensive experience operating and improving production systems; strong automation, observability, and incident management; proficiency with cloud and Kubernetes.
Kubernetes, Cloud, Observability, Automation, Monitoring
1mo
Save
Mark Applied
Hide
GOV Site Reliability Engineer
United States or Kansas or Washington or California or Texas or Illinois or North Carolina or Colorado or Massachusetts or Pennsylvania or Virginia or Oregon or Nevada or Hawaii or New York or Georgia or Ohio or Arizona or Seattle or San Francisco or New York City
$110k-$183k/yr RemoteFull Time
Veeam
Veeam: Data resilience and security for hybrid cloud environments
3+ YOE3+ years in software engineering with 1+ year in SRE/Platform/DevOps, cloud experience (Azure or comparable), observability (Prometheus, Grafana, OpenTelemetry, ELK), IaC (Terraform/Terragrunt/Pulumi), Kubernetes, CI/CD tooling, programming in TypeScript/JS, Go, Java, or C#, and experience in compliance-oriented environments.
VDC, Prometheus, Grafana, OpenTelemetry, ELK stack, Terraform, Terragrunt, Pulumi, Kubernetes, GitHub Actions, Azure DevOps, GitLab CI, ArgoCD, TypeScript, JS, Go, Java, C#, Azure Government, AWS GovCloud
3mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Bellevue or New York City
$164k-$213k/yr HybridFull Time
iSpot.tv
iSpot.tv: Provides real-time measurement and analytics for TV advertising.
10+ YOE3+ Mgmt10+ years in software engineering, cloud architecture, and/or SRE; strong AWS, Kubernetes, Terraform; Spark; CI/CD; leadership; excellent communication.
AWS, EKS, ECR, RDS, SQS, SNS, VPC, MWAA, S3, Terraform, CloudFormation, Kubernetes, kubectl, Helm, ArgoCD, CircleCI, Spark, EMR, Databricks, Glue, OTel, DataDog, GenAI tools
6d
Save
Mark Applied
Hide
Systems Reliability Engineer (SRE)
San Francisco or New York City
$150k-$170k/yr OnsiteFull Time
Claryo
Claryo: AI-powered spatial software for optimizing warehouse operations
3+ YOE3+ years SRE/infrastructure experience, strong Linux and networking fundamentals, experience with Kubernetes, cloud platforms, observability tooling, and debugging distributed systems in production.
Linux, Kubernetes, GCP, AWS, Azure, Prometheus, Grafana, OpenTelemetry, Kafka, RTSP, WebRTC
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE)
Jersey City or United States
$190k-$240k/yr RemoteFull Time
Tradeweb Markets
Tradeweb MarketsNasdaq: TW: Operates electronic marketplaces for fixed income and derivatives trading.
6+ YOE6+ years technology operations/engineering experience, 4+ years SRE or comparable, AWS and cloud-native experience, ArgoCD/Kustomize/Pulumi/K8s experience, Python scripting, and a Bachelor’s in Computer Science or related field.
ArgoCD, Kustomize, Pulumi, K8s, LGTM, Python, AWS, Lambda, EKS, SNS, SMS, GitSecOps, Linux, Unix
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Fabric
New York City or Toronto or United States
$127k-$249k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
10+ YOE10+ years in distributed systems with strong networking, cloud (AWS/Azure/GCP), service mesh, multi-cloud design, and on-call experience.
TCP/IP, DNS, TLS/mTLS, BGP, VPN, VPC, Subnets, CDNs, Service Mesh, Kubernetes, AWS, Azure, GCP
2w
Save
Mark Applied
Hide
Site Reliability Engineer II
Scottsdale or San Francisco or Chicago or New York City
$86k-$126k/yr HybridFull Time
Early Warning Services
Early Warning Services: Operates payment and risk solutions for the financial industry.
2+ YOEBachelor's or equivalent, minimum 2 years DevOps/Dev/SRE experience, Linux/Unix experience, infrastructure automation (Chef/Ansible/Puppet, Terraform), containerization (Docker,Kubernetes), cloud (AWS/GCP/Azure), on-call rotation.
Linux, Unix, Chef, Ansible, Puppet, Terraform, Docker, Kubernetes, AWS, GCP, Azure, Java, Ruby, Python, JavaScript, Go
1mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer
Berkeley Heights or Alpharetta or Sunnyvale
$128k-$216k/yr OnsiteFull Time
Fiserv
FiservNew York Stock Exchange: FI: Provides financial technology and payment processing services to institutions.
5+ YOE5+ years production experience with AWS, Kubernetes, and Linux; strong Terraform, CI/CD (GitHub Actions), Docker, GitHub, RDBMS/Document storage, and scripting (Python/Bash/Node/Ruby); experience designing scalable cloud systems.
Amazon Web Services, Kubernetes, GitHub Actions, Terraform, New Relic, Dynatrace, Datadog, Docker, GitHub, Python, Bash, Node, Ruby on Rails
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
Python, Go, Java, Shell, Linux, Docker, Kubernetes, Prometheus, Grafana
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
$111k-$160k/yr HybridFull Time
Mizuho Financial Group
Mizuho Financial GroupTokyo Stock Exchange: 8411: Global financial group providing banking and investment services.
3+ YOEBachelor’s degree or equivalent; 3+ years SRE/automation experience; strong Grafana, cloud (AWS/Azure/GCP), containers (Docker/Kubernetes); CI/CD; scripting (Python/Bash/Go); on-call experience.
Grafana, Ansible, Terraform, Jenkins, AWS, Azure, Google Cloud, Docker, Kubernetes, CI/CD, Python, Bash, Go