5 gpu inference performance engineer jobs at 5 companies in Boston, MA
1mo
Save
Mark Applied
Hide
1mo
Staff Engineer, Inference Optimizations
Boston, Massachusetts, United States
$191k-$239k/yrRemoteFull Time
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
5+ YOE5+ years in high-performance computing or AI infrastructure, deep GPU and inference optimization expertise, experience with CUDA/Triton/ROCm, attention-layer and kernel-level optimization, strong system design and leadership through influence.
Physical Superintelligence: Building AI systems to discover new physics at scale.
5+ YOERequires 5+ years with GPU and large-scale multi-node compute workloads, distributed training performance, networking, parallel file systems, AI training and inference optimization, and capacity decisions.
GPU, H100, B200, InfiniBand, RDMA, parallel file systems
Senior Deep Learning Framework Communications Engineer
Santa Clara or Austin or Westford or Durham or United States
$152k-$288k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOE5+ years software engineering experience in HPC/AI, experience with PyTorch/JAX and inference engines, Python/C++/CUDA development, performance benchmarking and profilers, understanding of multi-GPU communication and compilers.