101 cloud reliability engineer jobs at 76 companies in Stamford, CT

4w
Save
Mark Applied
Hide
Cloud Reliability Engineer
Englewood Cliffs or New York City
$135k-$165k/yr HybridFull Time
Versant
VersantNasdaq: VSNT: Operates cable television networks and digital media entertainment platforms.
3+ YOEBachelor's degree or equivalent experience; 3–7 years in SRE/Cloud/DevOps roles; strong AWS experience (enterprise scale), Terraform and CloudFormation, scripting (Python/PowerShell/Bash), CI/CD, monitoring/observability, incident management.
AWS, AWS Organizations, Control Tower, Identity Center, Terraform, CloudFormation, Python, PowerShell, Bash, CI/CD
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Jersey City or Charlotte or Plano
$153k-$192k/yr OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Provides global banking, investing, and financial risk management services.
4+ YOE4+ years cloud/platform engineering experience with GCP exposure, Terraform/IaC, CI/CD, observability, DevSecOps practices, incident response, and automation for reliability and resiliency.
GCP, Azure, Terraform, Terraform Enterprise, Log Analytics, Dynatrace, Resource Graph, CI/CD, IAM
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York, New York, United States
$177k-$209k/yr HybridFull Time
Peloton
PelotonNASDAQ: PTON: Sells connected fitness equipment and streaming exercise classes.
Kubernetes expertise; observability/monitoring mindset; CI/CD experience; IaC (Terraform/Pulumi); cloud (AWS); security and reliability focus; programming (Python/Go/Java/C).
Kubernetes, Observability, Jenkins, ArgoCD, Harness, Tekton, Terraform, Pulumi, AWS, Python, Golang, Java, C, Nginx, Ubuntu, Chef
3w
Save
Mark Applied
Hide
Site Reliability Engineer
New York, New York, United States
HybridFull Time
Mistral AI
Mistral AI: Developing frontier artificial intelligence models and enterprise AI solutions.
7+ YOE7+ years SRE/DevOps experience, Master’s in CS/Engineering or related, strong cloud and distributed systems skills, Kubernetes/CI-CD/infra-as-code proficiency, scripting experience, observability and on-call experience.
Kubernetes, Flux, Terraform, Docker, Prometheus, Grafana, ELK Stack, Datadog, CloudFormation, Python, Go, Bash, Slurm, Fluidstack, Coreweave, Vast
1d
Save
Mark Applied
Hide
Staff , Site Reliability Engineer - Cloud Platform
New York City, New York, United States
$190k-$210k/yr HybridFull Time
Butterfly Network
Butterfly NetworkNYSE: BFLY: Handheld whole-body ultrasound scanners powered by semiconductor technology.
8+ YOE8+ years managing production systems; deep hands-on AWS and Kubernetes experience; strong programming/scripting and automation skills; observability platform ownership; incident response leadership and mentoring ability.
Kubernetes, EKS, AWS, NewRelic, Datadog, DICOM, HL7, FHIR, PACS, VNA, EMR
2w
Save
Mark Applied
Hide
Lead Site Reliability Engineer
New York City or Toronto
$184k-$240k/yr OnsiteFull Time
Movable Ink
Movable Ink: Provides AI-powered content personalization for digital marketing campaigns.
6+ YOE6+ years SRE/Software Engineering experience designing and operating scalable, multi-cloud distributed systems; expertise with observability, IaC, Kubernetes, and multiple programming languages.
Apache Pulsar, Apache Kafka, Grafana Loki, ScyllaDB, Cassandra, Prometheus, Thanos, Grafana Alloy, Tempo, Terraform, Chef, EKS, GKE, NodeJS, Golang, Ruby, Python, shell
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Robotics & Cloud Infrastructure
Brooklyn or New York City or Richmond or Europe
$164k-$220k/yr RemoteFull Time
Bedrock Ocean Exploration
Bedrock Ocean Exploration: Maps the ocean floor using autonomous underwater robotic vehicles.
5+ YOE5+ years SRE/DevOps experience with on-call ownership; strong automation using Python/Go/Bash; Terraform and AWS hands-on; containerization (Docker, Kubernetes); observability (Prometheus, Grafana); Linux and networking expertise; East Coast location and US work authorization required.
Python, Go, Bash, Terraform, AWS, Docker, Kubernetes, Prometheus, Grafana, ROS 2, ROS, Jetson, Linux, IAM
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Greenwich, Connecticut, United States
HybridFull Time
Interactive Brokers
Interactive BrokersNASDAQ: IBKR: Automated global electronic brokerage and trading services provider.
5+ YOE5+ years experience in Linux/Unix, networking and coding; experience with cloud (AWS or Azure), Terraform or CloudFormation, Docker and Kubernetes; bachelor's or master's in CS/STEM; CI/CD, on-call rotation, mentoring skills.
CI/CD, Terraform, CloudFormation, AWS, Azure, Docker, Kubernetes, Linux, Unix
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City, New York, United States
$240k-$300k/yr OnsiteFull Time
Legora: AI workspace for legal document research and drafting.
Extensive experience operating and improving production systems; strong automation, observability, and incident management; proficiency with cloud and Kubernetes.
Kubernetes, Cloud, Observability, Automation, Monitoring
1mo
Save
Mark Applied
Hide
GOV Site Reliability Engineer
United States or Kansas or Washington or California or Texas or Illinois or North Carolina or Colorado or Massachusetts or Pennsylvania or Virginia or Oregon or Nevada or Hawaii or New York or Georgia or Ohio or Arizona or Seattle or San Francisco or New York City
$110k-$183k/yr RemoteFull Time
Veeam
Veeam: Data resilience and security for hybrid cloud environments
3+ YOE3+ years in software engineering with 1+ year in SRE/Platform/DevOps, cloud experience (Azure or comparable), observability (Prometheus, Grafana, OpenTelemetry, ELK), IaC (Terraform/Terragrunt/Pulumi), Kubernetes, CI/CD tooling, programming in TypeScript/JS, Go, Java, or C#, and experience in compliance-oriented environments.
VDC, Prometheus, Grafana, OpenTelemetry, ELK stack, Terraform, Terragrunt, Pulumi, Kubernetes, GitHub Actions, Azure DevOps, GitLab CI, ArgoCD, TypeScript, JS, Go, Java, C#, Azure Government, AWS GovCloud
3mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Bellevue or New York City
$164k-$213k/yr HybridFull Time
iSpot.tv
iSpot.tv: Provides real-time measurement and analytics for TV advertising.
10+ YOE3+ Mgmt10+ years in software engineering, cloud architecture, and/or SRE; strong AWS, Kubernetes, Terraform; Spark; CI/CD; leadership; excellent communication.
AWS, EKS, ECR, RDS, SQS, SNS, VPC, MWAA, S3, Terraform, CloudFormation, Kubernetes, kubectl, Helm, ArgoCD, CircleCI, Spark, EMR, Databricks, Glue, OTel, DataDog, GenAI tools
1w
Save
Mark Applied
Hide
Systems Reliability Engineer (SRE)
San Francisco or New York City
$150k-$170k/yr OnsiteFull Time
Claryo
Claryo: AI-powered spatial software for optimizing warehouse operations
3+ YOE3+ years SRE/infrastructure experience, strong Linux and networking fundamentals, experience with Kubernetes, cloud platforms, observability tooling, and debugging distributed systems in production.
Linux, Kubernetes, GCP, AWS, Azure, Prometheus, Grafana, OpenTelemetry, Kafka, RTSP, WebRTC
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE)
Jersey City or United States
$190k-$240k/yr RemoteFull Time
Tradeweb Markets
Tradeweb MarketsNasdaq: TW: Operates electronic marketplaces for fixed income and derivatives trading.
6+ YOE6+ years technology operations/engineering experience, 4+ years SRE or comparable, AWS and cloud-native experience, ArgoCD/Kustomize/Pulumi/K8s experience, Python scripting, and a Bachelor’s in Computer Science or related field.
ArgoCD, Kustomize, Pulumi, K8s, LGTM, Python, AWS, Lambda, EKS, SNS, SMS, GitSecOps, Linux, Unix
23h
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Production Engineer - ThousandEyes
San Francisco or Seattle or Austin or New York City
$165k-$241k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOE5+ years experience; proficiency in Python or Go; expertise with Kubernetes, cloud (AWS), Unix/Linux; strong SRE principles, incident response, and security-minded engineering.
Python, Go, Kubernetes, Service Mesh, Prometheus, OpenTelemetry, ArgoCD, CNCF, AWS, Unix, Linux
3mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Fabric
New York City or Toronto or United States
$127k-$249k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
10+ YOE10+ years in distributed systems with strong networking, cloud (AWS/Azure/GCP), service mesh, multi-cloud design, and on-call experience.
TCP/IP, DNS, TLS/mTLS, BGP, VPN, VPC, Subnets, CDNs, Service Mesh, Kubernetes, AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
Customer Reliability Engineer - Infrastructure
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yr RemoteFull Time
Astronomer
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
Apache Airflow, AWS, Azure, CI/CD, GCP, Infrastructure as Code (IaC), Kubernetes, Linux, Python
2w
Save
Mark Applied
Hide
Site Reliability Engineer II
Scottsdale or San Francisco or Chicago or New York City
$86k-$126k/yr HybridFull Time
Early Warning Services
Early Warning Services: Operates payment and risk solutions for the financial industry.
2+ YOEBachelor's or equivalent, minimum 2 years DevOps/Dev/SRE experience, Linux/Unix experience, infrastructure automation (Chef/Ansible/Puppet, Terraform), containerization (Docker,Kubernetes), cloud (AWS/GCP/Azure), on-call rotation.
Linux, Unix, Chef, Ansible, Puppet, Terraform, Docker, Kubernetes, AWS, GCP, Azure, Java, Ruby, Python, JavaScript, Go
1mo
Save
Mark Applied
Hide
Site Reliability Engineer, Tech Infrastructure - USDS
New York, New York, United States
$137k-$259k/yr OnsiteFull Time
TikTok USDS Joint Venture
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
3+ YOE3+ years SRE/systems engineering experience, bachelor\u0002s in CS or related, proficiency in Python/Go/Java/Shell, Linux, cloud and distributed systems, monitoring and incident management.
Python, Go, Java, Shell, Linux, Docker, Kubernetes, Prometheus, Grafana
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
New York City, New York, United States
$111k-$160k/yr HybridFull Time
Mizuho Financial Group
Mizuho Financial GroupTokyo Stock Exchange: 8411: Global financial group providing banking and investment services.
3+ YOEBachelor’s degree or equivalent; 3+ years SRE/automation experience; strong Grafana, cloud (AWS/Azure/GCP), containers (Docker/Kubernetes); CI/CD; scripting (Python/Bash/Go); on-call experience.
Grafana, Ansible, Terraform, Jenkins, AWS, Azure, Google Cloud, Docker, Kubernetes, CI/CD, Python, Bash, Go
4mo
Save
Mark Applied
Hide
Site Reliability and Infrastructure Engineer
New York City, New York, United States
$160k-$215k/yr HybridFull Time
Treeswift
Treeswift: Provides AI-augmented vegetation management and asset monitoring for utilities.
7+ YOEExperienced SRE/infrastructure engineer with 7-10 years in observability, cloud, and DevOps; strong Linux, IaC, Kubernetes, and Python.
Terraform, Kubernetes, AWS, Airflow, Astronomer, ECR, S3, SQS, Lambda, Step Functions, ECS, Docker

Explore Jobs

Expand Your Job Search