25 platform reliability engineer jobs at 19 companies in Woburn, MA
2w
Save
Mark Applied
Hide
2w
Reliability Engineer
Boston or Fort Meade
RemoteFull Time
Global Enterprise Services: Provides IT and cybersecurity services for federal government programs.
8+ YOE8+ years experience in reliability engineering for cloud platforms; IAT-2 and cloud certification(s); Bachelor\u0002s degree required; secret clearance.
8+ YOE8+ years SRE or platform engineering experience; defining SLO/SLI, building observability, incident management, Kubernetes at scale, automation in Python or Go, CI/CD and IaC familiarity.
WHOOP: Wearable technology for personalized health and performance tracking.
5+ YOE5+ years in DevOps/Platform/SRE or backend roles; deep Kubernetes experience; hands-on AWS and Terraform; knowledge of AI runtimes; strong system design, reliability, and mentoring skills.
Manifold: AI platform for life sciences data and research collaboration.
7+ YOE7+ years in infrastructure/DevOps/SRE with deep cloud (AWS/GCP/Azure), Terraform, CI/CD (Github Action), container tooling, identity systems, data platform services, and experience operating secure multi-account environments.
Ahold Delhaize USAEuronext Amsterdam: AD: Operates a portfolio of omnichannel grocery brands.
8+ YOE8+ years in platform or infrastructure engineering, automation, CI/CD, container/runtime services, incident response, reliability engineering, and effective communication. Bachelor's degree or equivalent experience required.
Boston or Miami or New Jersey or New York City or Princeton or Raleigh or Washington or Toronto or North America
$127k-$249k/yrHybridFull Time
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud (AWS, Google Cloud Platform, Microsoft Azure) expertise; Linux and networking knowledge.
Argo Workflows, ArgoCD, Kubernetes, Python, Go, AWS, Google Cloud Platform (GCP), Microsoft Azure, Linux
Senior AI Ops & Incident/Site Reliability Engineer
Austin or Plano or Fort Mill or Boston or New York City or Tempe or San Diego
$94k-$171k/yrHybridFull Time
Perficient: Provides digital transformation and AI consulting for global enterprises.
8+ YOEBachelor's degree or equivalent,8+ years in IT operations/SRE or production support,5+ years leading incident management,experience with Dynatrace,ServiceNow,cloud platforms and AIOps implementations.
Bengaluru or San Francisco or Boston or New York City or Austin or Tokyo or London
HybridFull Time
Postman: Platform for building, testing, and managing software APIs.
6+ YOE6+ years backend experience building distributed systems; strong proficiency in one or more backend languages (JavaScript, Python, Java, Go, PHP, C++); experience with APIs, event-driven architectures, reliability, and observability; mentoring experience.
OnRamp: SaaS platform automating B2B customer onboarding and implementation processes.
Deep hands-on AWS experience, infrastructure-as-code (Terraform or CDK), containers and CI/CD, building observability/reliability, security/compliance (SOC 2/HIPAA), and using AI/LLM tooling and coding agents to automate operations.
CVS HealthNYSE: CVS: Provides retail pharmacy, health insurance, and pharmacy benefit management services.
8+ YOERequires 8+ years in SRE, platform, or production systems engineering; major program ownership; LLM operations, fleet-scale reliability, change management, TIC expertise, observability, Kafka, Kubernetes, policy automation, and a bachelor's degree or equivalent.
Member of Technical Staff – Senior Engineer, Data Infrastructure & Data Operations
San Francisco or Cambridge
$255k-$340k/yrOnsiteFull Time
Walden Robotics: Builds general-purpose robots and develops the teams and infrastructure to scale robot applications and improve quality of life.
Experience building production data infrastructure and high-throughput pipelines, cloud-based data platform development, platform reliability and cost ownership, and collaboration with ML teams.
WEXNYSE: WEX: Provides global payment processing and business information management services.
Staff/lead level engineer with deep experience architecting autonomous/agentic AI systems, cross-platform architecture, security-by-design, CI/CD, cloud reliability, and mentoring senior engineers.
Verily: Developing data-driven technologies for clinical research and precision health.
8+ YOEBA/BS or equivalent; 8+ years building reliable, scalable cloud-native services in Go, Python, C++; experience with cloud platforms, Kubernetes, ArgoCD, Backstage, GitHub Actions; strong system design and communication skills.
Go, Python, C++, Google Cloud Platform, Amazon Web Services, Azure, GitHub Actions, Kubernetes, ArgoCD, Backstage, Terraform
HumanaNYSE: HUM: Provides health insurance plans and clinical healthcare services.
8+ YOE2+ MgmtBachelor's in CS or related, 6+ years full-stack experience, 8+ years progressive experience, 2+ years project leadership, platform reliability/SRE experience, testing and observability expertise, polyglot stack familiarity.
Columbus or Boston or New York City or Chicago or Austin or Los Angeles
$150k-$190k/yrRemoteFull Time
Loop Returns: Software platform automating e-commerce returns and post-purchase experiences.
3+ YOE3+ years engineering management experience, experience with platform reliability/system health, AI agent adoption, technical depth in high-risk domains, ability to manage and grow engineers.