37 cloud reliability engineer jobs at 26 companies in South Peabody, MA
3w
Save
Mark Applied
Hide
3w
Principal Site Reliability Engineer (Hybrid)
Merrimack or San Diego
$118k-$201k/yrHybridFull Time
BAE SystemsLSE: BA.: Global defense, aerospace, and security technology.
4+ YOERequires 4–6+ years of site reliability engineering, Juniper networking, cloud technologies, automation, storage, virtualization, and security clearance eligibility; Security+ required or obtainable within 90 days.
Manifold: The Enterprise Agent Platform for life sciences that helps biopharma and research teams analyze governed biomedical data.
7+ YOE7+ years in infrastructure/DevOps/SRE with deep cloud (AWS/GCP/Azure), Terraform, CI/CD (Github Action), container tooling, identity systems, data platform services, and experience operating secure multi-account environments.
DraftKingsNASDAQ: DKNG: Provide online sports betting, fantasy sports, and casino gaming.
8+ YOEBachelor's degree in CS or related; 8+ years designing and operating distributed cloud/on-prem infrastructure (3+ years at principal/staff level); deep Kubernetes, AWS/GCP, IaC (Terraform/Pulumi), Go/Python, GitOps experience; leadership and communication skills.
Kubernetes, Google Kubernetes Engine, Amazon Elastic Kubernetes Service, RKE2, Terraform, Pulumi, Go, Python, GitOps, AWS, Google Cloud Platform, Linux
Planet FitnessNYSE: PLNT: Leading fitness center franchisor and operator.
7+ YOE7+ years leading SRE/DevOps with cloud (AWS/Azure/GCP), incident management, SLO/SLI experience, CI/CD and observability expertise; bachelor\u0002s degree or equivalent experience.
AWS, Azure, GCP, CI/CD, Infrastructure as Code (IaC)
United States or New York City or Boston or San Francisco
$125k-$130k/yrRemoteFull Time
Astronomer: Private software providing managed Apache Airflow data orchestration for enterprise data teams.
4+ YOERequires data engineering background, 4 years with Python, 1 year administering Airflow and creating DAGs, Kubernetes, Docker, containers, cloud provider experience, troubleshooting, communication, autonomy, and mentoring experience.
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Actively enrolled in a bachelor's program, available for a 16-week full-time co-op, and knowledgeable in Linux, monitoring, troubleshooting, automation, scripting, cloud platforms, and production support.
Linux, Python, Go, Bash, IBM Cloud, AWS, Microsoft Azure, Google Cloud Platform, Kubernetes, OpenShift, Ansible, Terraform, Jenkins, IBM Continuous Delivery, ArgoCD, Instana, New Relic, Grafana, Prometheus, PostgreSQL, CouchDB, Redis, Kafka, Spark, SQL, NoSQL, CI/CD
Veson Nautical: Private maritime software serving shipowners, charterers, traders, and operators with commercial freight management solutions.
5+ YOEBachelor's degree or equivalent experience; 5+ years of GCP experience, production Kubernetes/GKE, Terraform, cloud networking, and Python, Go, or TypeScript programming skills.
Google Cloud Platform, Bigtable, Cloud SQL, Dataflow, Datastore, Google Kubernetes Engine (GKE), Google Cloud Storage (GCS), Google Cloud Key Management Service (KMS), Pub/Sub, Amazon Web Services, Kubernetes, Amazon Elastic Kubernetes Service (EKS), Terraform, Terragrunt, Atlantis, GitLab Pipelines, ArgoCD, Octopus Deploy, ElasticSearch, Kubernetes Operator, PostgreSQL, SQL Server, BigQuery, Splunk, Grafana, Grafana Tempo, OpenTelemetry, Cloud Armor Enterprise, OpsGenie, Renovate, Sentry, Claude, Amazon Bedrock, Gemini, Vertex AI, Python, Go, TypeScript, GitLab CI
8+ YOERequires 8+ years of relevant experience, backend programming proficiency, cloud-native and distributed systems expertise, Linux and networking knowledge, debugging skills, and experience with AWS or GCP and Kubernetes.
Boston or Miami or New Jersey or New York City or Princeton or Raleigh or Washington or Toronto or North America
$127k-$249k/yrHybridFull Time
MongoDBNASDAQ: MDB: Unified data platform for building modern applications.
6+ YOE6+ years software development experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud (AWS, Google Cloud Platform, Microsoft Azure) expertise; Linux and networking knowledge.
Argo Workflows, ArgoCD, Kubernetes, Python, Go, AWS, Google Cloud Platform (GCP), Microsoft Azure, Linux
TSD Mobility Solutions: Provider of automotive dealership software and business solutions.
5+ YOE5+ years in DevOps/SRE with hands-on AWS, CI/CD (Jenkins), automated deployments for on-prem and cloud, Windows IIS and Linux administration, scripting (Bash, Python, PowerShell), and infrastructure-as-code (Terraform/Ansible).
Site Reliability Engineer - Disaster Recovery & Business Continuity
Boston or Chicago
$130k-$150k/yrHybridFull Time
Charles River AssociatesNasdaq Global Select Market: CRAI: Public global economic, financial, and management consulting firm serving law firms, corporations, accounting firms, and governments.
Experience with IT service continuity, disaster recovery, cross-functional coordination, and documentation.
Windows, Microsoft 365, Cloud, SaaS, Backups, Virtualization, Identity, Networking
Senior Manager, Site Reliability & Operational Resilience
Morristown or Boston or St. Petersburg or St. Louis or Atlanta or Hyderabad
$139k-$177k/yrHybridFull Time
Zelis: Providing healthcare payment and claims cost management solutions.
8+ YOE3+ MgmtRequires 8+ years in SRE, production, platform, DevOps, cloud, or infrastructure engineering; 3+ years leading people; enterprise resilience experience; bachelor's degree or equivalent; no visa sponsorship.
LogicMonitor, New Relic, Splunk, Datadog, Python, PowerShell, Go, Terraform, Azure, AWS, Kubernetes, OpenTelemetry, Jira Service Management
Senior System Architect, Infrastructure Reliability
Santa Clara or Austin or Westford or Durham or Redmond
$184k-$357k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
6+ YOEBS, MS, or PhD in computer science or electrical engineering, or equivalent experience; 6+ years in systems programming; expertise in distributed systems, C++ and Python, HPC or cloud RCA pipelines, CPU metrics, and cluster managers.
Member of Technical Staff – Senior Engineer, Data Infrastructure & Data Operations
San Francisco or Cambridge
$255k-$340k/yrOnsiteFull Time
Walden Robotics: Private full-stack physical AI building and deploying general-purpose robots for manufacturing and logistics.
Experience building production data infrastructure and high-throughput pipelines, cloud-based data platform development, platform reliability and cost ownership, and collaboration with ML teams.
Senior System Architect, Infrastructure Reliability
Santa Clara or Austin or Westford or Durham or Redmond
$184k-$357k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
6+ YOE6+ years systems programming experience; BS/MS/PhD in CS or EE (or equivalent); experience building RCA pipelines for HPC/cloud; deep CPU/GPU architecture knowledge; strong C++ and Python; familiarity with Slurm/LSF/Kubernetes.