36 site reliability engineer jobs at 23 companies in Needham, MA

1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Cambridge or United States
$76k-$136k/yr HybridFull Time
Akamai
AkamaiNASDAQ: AKAM: Provides content delivery, cybersecurity, and cloud computing services globally.
Bachelor's in Computer Science/Engineering or equivalent; experience in SRE/Software Engineering for large-scale distributed systems; Terraform and IAC experience; familiarity with SaltStack/Ansible/Chef/Puppet; Linux, CI/CD, observability, and participation in on-call rotation.
Terraform, SaltStack, Ansible, Chef, Puppet, Linux, CI/CD, Infrastructure as Code (IAC)
3d
Save
Mark Applied
Hide
Site Reliability Engineer
Waltham, Massachusetts, United States
$166k-$220k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology building autonomous military hardware and software.
3+ YOE3+ years SRE/DevOps/field or production support experience; strong Linux and networking fundamentals; on-call rotation experience; ability to diagnose cross-stack issues; eligibility for U.S. Secret clearance.
Linux, Python, Bash, Nix, NixOS, systemd, PagerDuty, VPN
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer I
Boston or Seattle or Atlanta
$134k-$215k/yr HybridFull Time
Axon
AxonNASDAQ: AXON: Develops weapons and software for law enforcement and safety.
7+ YOEBachelor's in CS/Engineering, 7+ years software engineering experience, expertise in distributed systems, Kubernetes, cloud (Azure/AWS/GCP), observability, Kafka, Terraform/Pulumi, and experience with agentic AI/LLM tooling preferred.
Kubernetes, Terraform, Pulumi, Kafka, Grafana, Datadog, New Relic, MySQL, Cassandra, PostgreSQL, Azure, AWS, GCP
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer
Newton, Massachusetts, United States
$160k-$205k/yr OnsiteFull Time
Manifold
Manifold: AI platform for life sciences data and research collaboration.
7+ YOE7+ years in infrastructure/DevOps/SRE with deep cloud (AWS/GCP/Azure), Terraform, CI/CD (Github Action), container tooling, identity systems, data platform services, and experience operating secure multi-account environments.
AWS, GCP, Azure, Terraform, Github Action, Okta, Auth0, Docker, ECS, Packer, Tailscale, WireGuard, Snowflake, Airflow, dbt, PostgreSQL, LLM, CI/CD
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Somerville, Massachusetts, United States
$160k-$200k/yr HybridFull Time
Tulip
Tulip: Provides a no-code platform for industrial frontline operations.
5+ YOE5+ years experience with observability tools, OpenTelemetry instrumentation, Prometheus metrics, experience with time-series data and producing SLIs/SLOs, strong systems reasoning and communication.
Grafana, Loki, Tempo, Mimir, OpenTelemetry, Prometheus, promQL, TypeScript, Go, Kubernetes, MongoDB, PostGres, Alloy, Claude Skills, Gemini Gems
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer- Eng
Lowell, Massachusetts, United States
$130k-$186k/yr HybridFull Time
UKG
UKG: Provider of workforce management and human capital management software.
5+ YOE5+ years software/systems/cloud engineering; public cloud experience (GCP/AWS/Azure); observability, SLOs, incident response, Linux, coding in Python/Java/C++; GitHub Actions and dashboarding (Splunk/Grafana).
GCP, AWS, Azure, Python, Java, C++, Linux, GitHub Actions, Splunk, Grafana, Kubernetes, Terraform, Ansible
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE)
Woonsocket or Scottsdale
$93k-$204k/yr HybridFull Time
CVS Health
CVS HealthNYSE: CVS: Provides retail pharmacy, health insurance, and pharmacy benefit management services.
5+ YOE5+ years SRE/DevOps experience, observability and monitoring expertise, cloud and containerization knowledge, 2+ years Java/Python and AI/AIOps experience, CI/CD and source control experience.
Splunk, Dynatrace, Datadog, Prometheus, Grafana, Java, Python, AWS, Microsoft Azure, Google Cloud, Rancher, Docker, Kubernetes, OpenShift, GitHub, BitBucket, Jenkins, Apigee, Data power
1d
Save
Mark Applied
Hide
Site Reliability Engineer III, DevEx
United States or Atlanta or Boston
$140k-$165k/yr RemoteFull Time
Flock Safety
Flock Safety: Sells AI-powered cameras and software for public safety surveillance.
Experience writing production Go or TypeScript, proficiency with Kubernetes, Helm, Terraform, GitHub Actions, and AWS; observability and CI/CD expertise; participates in on-call rotations.
Go, TypeScript, Kubernetes, Helm, Terraform, GitHub Actions, AWS
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Boston or Miami or Pittsburgh or Raleigh
$127k-$249k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years in software development and distributed systems; proficiency in Python or Go; experience with stateful storage systems; containerization (Kubernetes); cloud platforms (AWS, GCP, Azure); Linux networking; customer-focused and automation-minded.
Python, Go, Kubernetes, AWS, Google Cloud Platform, Azure, Linux, TCP/IP, DNS, TLS, Networking
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - Government Cloud
Boston or Dublin or United States
$210k-$220k/yr RemoteFull Time
Tines
Tines: No-code workflow automation for security and IT teams.
5+ YOE5+ years in infrastructure/DevOps/cloud engineering with strong AWS experience; hands-on IaC (CDK or Terraform), container image pipelines and hardening, observability, FedRAMP/CMMC/FISMA familiarity, documentation and assessment experience; U.S. citizenship required.
AWS GovCloud, AWS, CDK, Terraform, FedRAMP, FISMA, CMMC, FIPS, CI/CD, VPC, OpenSearch, Ruby, Rails, React, TypeScript, Postgres, Redis, Kubernetes
1w
Save
Mark Applied
Hide
Principal Site Reliability Engineer, Machine Learning
Cambridge, Massachusetts, United States
$142k-$178k/yr OnsiteFull Time
Cambridge Mobile Telematics
Cambridge Mobile Telematics: Providing telematics and behavioral analytics for safer driving and insurance.
7+ YOEBachelor's or equivalent, 7+ years SRE/IT experience, AWS (EC2,EKS, S3,RDS), Databricks, Ray, Terraform, Python, Linux, Datadog/CloudWatch, strong incident response and system design skills.
Ray, AWS EKS, Databricks, CloudWatch, Datadog, EC2, S3, RDS/Aurora, Dynamo, SQS, Lambda, IAM, Terraform, Python, Docker, Kubernetes, Unity Catalog, CI/CD
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Boston, Massachusetts, United States
$128k-$160k/yr OnsiteFull Time
DraftKings
DraftKingsNASDAQ: DKNG: Provide online sports betting, fantasy sports, and casino gaming.
4+ YOE4+ years managing distributed cloud and on‑prem environments, strong AWS and Kubernetes experience, proficiency in Go or Python, networking and Linux knowledge, bachelor's in CS or equivalent experience.
Rancher Fleet, Flux, Helm, Karpenter, HPA, KEDA, Datadog, Go, Python, Docker, containerd, vSphere, Nutanix, AWS, GCP, Linux
2d
Save
Mark Applied
Hide
Sr. Site Reliability Engineer
Waltham, Massachusetts, United States
$130k-$140k/yr HybridFull Time
SS&C Technologies
SS&C TechnologiesNasdaq: SSNC: Provides software and technology for financial and healthcare sectors.
7+ YOEExperience in Unix/Linux, Java, Python or Bash, Kubernetes, Docker, cloud (AWS/Azure), IaC (Terraform/Ansible), monitoring tools, networking protocols, and on-call incident response.
Unix, Linux, Java, Python, Bash, Kubernetes, Docker, Azure, AWS, CloudWatch, EKS, EFS, S3, RedShift, Terraform, Ansible, HTTP(s), JMS, TCP, UDP, Splunk, Datadog, Dynatrace, Zabbix, Prometheus, Akamai, Cloudflare, DNS, CDN, DataStream, WAF, Oracle, PostgreSQL, MongoDB, RabbitMQ, Interconnect, AMQ
3w
Save
Mark Applied
Hide
Lead Site Reliability Engineer, Engineering Enablement (Remote)
Boston, Massachusetts, United States
$164k-$235k/yr RemoteFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
8+ YOE8+ years experience (Bachelor+8 or Master+6), software development background with 5+ years coding (Ruby or Python), systems-at-scale experience, automation/config-as-code, Unix/Linux proficiency, mentoring experience.
Ruby, Python, Unix, Linux, container orchestration
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Boston, Massachusetts, United States
$140k-$211k/yr OnsiteFull Time
Federal Reserve System
Federal Reserve System: The central bank of the United States.
Senior SRE with AWS, Terraform, Docker, Linux, CI/CD, IaC, observability, and automation experience; able to operate large-scale distributed systems in a production environment.
AWS, EC2, EKS, RDS, Aurora, S3, Route 53, ELB, IAM, Terraform, Consul, Vault, Ansible, Python, Java, Go, Docker, ECR, OpenSearch, Dynatrace, Grafana, Prometheus, CloudWatch, ChaosToolkit, Gremlin, Chaos Monkey
1mo
Save
Mark Applied
Hide
Senior DevOps Site Reliability Engineer
North Andover, Massachusetts, United States
HybridFull Time
Reynolds and Reynolds
Reynolds and Reynolds: Provides software and services for automotive retailers.
5+ YOE5+ years in DevOps/SRE with hands-on AWS, CI/CD (Jenkins), automated deployments for on-prem and cloud, Windows IIS and Linux administration, scripting (Bash, Python, PowerShell), and infrastructure-as-code (Terraform/Ansible).
Jenkins, AWS, EC2, S3, RDS, Lambda, VPC, IAM, Windows IIS, Linux, Azure DevOps, Bash, Python, PowerShell, Terraform, Ansible, Web Deploy (MSDeploy), PowerShell DSC, Docker, Kubernetes, Datadog, Grafana, CloudWatch
1mo
Save
Mark Applied
Hide
Sr. Control System Engineer/Site Reliability Engineer (SRE)
Boston, Massachusetts, United States
$160k-$225k/yr OnsiteFull Time
QuEra Computing
QuEra Computing: Develops and operates neutral-atom quantum computing systems.
10+ YOEDesign, implement, and maintain hardware and software control systems for quantum computers; strong Linux/Windows administration, networking (LAN/WAN/VLAN/DNS/DHCP/TCP/IP), scripting (Python/Bash/Go), containerization, CI/CD, infrastructure-as-code, observability, and rack server experience; 10+ years experience.
Hardware-in-the-loop (HIL), Kubernetes, Docker, Git, Python, Bash, Go, GitLab CI, Jenkins, Ansible, Terraform, Grafana, Prometheus, ELK stack, CI/CD, Ubuntu, Debian, Redhat, Linux, Windows, VLAN, DNS, DHCP, TCP/IP
2mo
Save
Mark Applied
Hide
Site Reliability Engineer - Disaster Recovery & Business Continuity
Boston or Chicago
$130k-$150k/yr HybridFull Time
Charles River Associates
Charles River AssociatesNasdaq: CRAI: Provides global economic, financial, and management consulting services.
Experience with IT service continuity, disaster recovery, cross-functional coordination, and documentation.
Windows, Microsoft 365, Cloud, SaaS, Backups, Virtualization, Identity, Networking
1w
Save
Mark Applied
Hide
Senior AI Ops & Incident/Site Reliability Engineer
Austin or Plano or Fort Mill or Boston or New York City or Tempe or San Diego
$94k-$171k/yr HybridFull Time
Perficient
Perficient: Provides digital transformation and AI consulting for global enterprises.
8+ YOEBachelor's degree or equivalent,8+ years in IT operations/SRE or production support,5+ years leading incident management,experience with Dynatrace,ServiceNow,cloud platforms and AIOps implementations.
Dynatrace, ServiceNow, AWS, Azure, GCP
3w
Save
Mark Applied
Hide
Senior Application Support Engineer / Site Reliability Engineer (SRE)
Boston, Massachusetts, United States
$75k-$150k/yr HybridFull Time
DTCC
DTCC: Provides post-trade infrastructure for the global financial services industry
6+ YOE6–8 years supporting enterprise applications; strong SRE knowledge; experience with distributed systems, cloud platforms, middleware, monitoring, automation, and incident management.
Java, J2EE, Oracle, SQL, DB2, IBM MQ, Kafka, Splunk, Grafana, AutoSys, ServiceNow, ITIL, Linux, Unix, AWS, OpenShift Container Platform, OpenShift, Kubernetes, Docker, IIB, JCL, Python