45 staff infrastructure engineer jobs at 30 companies in Needham, MA
2mo
Save
Mark Applied
Hide
2mo
Staff Software Engineer, ML Infrastructure
Boston, Massachusetts, United States
$147k-$215k/yrHybridFull Time
SimpliSafe: Provides wireless home security systems and professional monitoring services.
8+ YOE8+ years in software engineering; strong distributed systems; Kubernetes and AWS; Python; ML infra experience preferred; strong communication; staff-level technical leadership.
Member of Technical Staff – Senior Engineer, Data & Training Infrastructure
San Francisco or Cambridge
$255k-$340k/yrOnsiteFull Time
Walden Robotics: Builds general-purpose robots and the data and training infrastructure to improve quality of life.
Strong experience in systems/ML infrastructure and distributed training, performance optimization, observability, and technical leadership; ability to mentor engineers and remove bottlenecks.
PfizerNYSE: PFE: Develops and manufactures vaccines and medicines for global healthcare
6+ YOEBachelor's degree and 6+ years in cloud infrastructure engineering; expert AWS or GCP experience, HPC frameworks, CI/CD, observability, distributed systems, networking, identity, and security.
AWS, GCP, Terraform, Microsoft CloudFormation, CloudWatch, Prometheus, Grafana, EKS, GKE, Kubernetes, NVIDIA GPU, AWS ParallelCluster, Parallel Computing Services, Google Cloud Cluster Toolkit, Linux, CI/CD, Infrastructure as Code (IaC)
Lightfield: AI-powered CRM automating customer interaction capture and insights.
Experience owning production data systems, strong software engineering fundamentals, debugging across stack, query performance and schema evolution, product-oriented judgment.
Lila Sciences: Develops an AI platform for autonomous scientific research and discovery.
8+ YOE8+ years in software or data engineering; strong Python/SQL; experience building data infrastructure, ingestion, storage, orchestration; cloud (AWS, Kubernetes); relational/NoSQL databases; cross-functional collaboration.
Boston or Kyiv or Seattle or Lviv or San Francisco
HybridFull Time
DataRobot: Provides an enterprise platform for building and governing AI.
7+ YOE7-10+ years in engineering with 5+ years in infrastructure/platform/backend; Kubernetes internals; Python/Go; multi-cloud or hybrid deployments; IaC; GitOps; strong mentoring and leadership.
Fresenius Medical CareNew York Stock Exchange: FMS: Provides dialysis services and manufactures renal care medical devices
8+ YOEBachelor’s degree in computer science, information technology, or related field; 8+ years related experience, including 4+ years in system and network administration; Windows/Linux, networking, virtualization, and cloud infrastructure expertise.
Terraform, Windows, Linux, VMware, Hyper-V, routers, switches, firewalls, Microsoft Certified: Windows Server, CompTIA Network+, CompTIA Security+, Cisco CCNA
New York City or Italy or Poland or Stockholm or London or Berlin or Amsterdam or Paris or Toronto or Seattle or Florida or Frankfurt or United Kingdom or Colorado or Vancouver or Boston or Munich or Atlanta or South San Francisco or Texas or United States
RemoteFull Time
AeroVect: Develops autonomous driving software for airport logistics vehicles.
7+ YOE7+ years building cloud infrastructure and data systems; strong Python; experience with Kafka,Kubernetes,gRPC,AWS,Terraform/CloudFormation,CI/CD; distributed systems and observability skills.
8+ YOEBachelor's degree or equivalent,8+ years testing and launching software,5+ years building large-scale infrastructure or distributed systems,4+ years designing file and block storage,3+ years software design; EMphasis on storage performance and leadership.
Member of Technical Staff - AI Cloud Infrastructure
Oakland or Boston or Washington or California or Massachusetts or District of Columbia
HybridFull Time
Emerald AI: Managing data center power with AI-driven workload orchestration.
7+ YOERequires 7+ years in infrastructure or platform engineering, managed cloud or AI platform architecture, Kubernetes, Slurm, parallel filesystems, Linux, Terraform, Ansible, Python or Go, and GPU networking.
Suno: AI platform for creating full songs from text prompts.
5+ YOE5+ years full-stack engineering experience; built and shipped internal tooling; hands-on experience with LLMs, AI APIs, and agentic workflows; strong cross-functional communication; experience owning production infrastructure.
Inari: Designs gene-edited seeds to increase crop yield and sustainability.
7+ YOE2+ Mgmt7+ years engineering experience, 2+ years technical leadership, hands-on platform architecture in data/AI, proficiency with Python, AWS, Kubernetes/EKS, and building CI/CD and infrastructure for AI systems.
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
5+ YOE5+ years in high-performance computing or AI infrastructure, deep GPU and inference optimization expertise, experience with CUDA/Triton/ROCm, attention-layer and kernel-level optimization, strong system design and leadership through influence.
Manifold: AI platform for life sciences data and research collaboration.
7+ YOE7+ years in infrastructure/DevOps/SRE with deep cloud (AWS/GCP/Azure), Terraform, CI/CD (Github Action), container tooling, identity systems, data platform services, and experience operating secure multi-account environments.
ToastNYSE: TOST: Cloud-based technology platform for the restaurant and hospitality industry.
Proven experience delivering and operating mission-critical, high-throughput distributed services; strong ownership, system design, and cross-team leadership skills; experience with Kotlin/Java and cloud infrastructure.
Senior/Staff Site Reliability Engineer - Data Center
Boston, Massachusetts, United States
$166k-$224k/yrHybridFull Time
PathAI: AI-powered platform for pathology research and clinical diagnostics.
8+ YOERequires 8+ years of relevant experience, infrastructure operations expertise, a bachelor's degree in Computer Science or equivalent experience, and knowledge of datacenter, virtualization, storage, automation, monitoring, and incident response.
Blitzy: Autonomous AI platform for enterprise software development.
8+ YOE4+ Mgmt8+ years software engineering experience with 4+ years in engineering leadership; expertise in distributed systems and cloud infrastructure; experience scaling teams, shipping complex products, and presenting to executives.