81 gpu jobs at 55 companies in Taunton, MA

1w
Save
Mark Applied
Hide
Senior Software Engineer - GPU Local AI Platforms
Santa Clara or Westford or Austin or Durham or Seattle
$224k-$431k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOE12+ years software engineering experience in GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
Python, C++, CUDA, Triton, Docker, OCI, NCCL, RCCL, NVIDIA Container Toolkit
2w
Save
Mark Applied
Hide
Senior Product Manager (AI Infrastructure & GPU)
Cambridge or United States
$139k-$251k/yr HybridFull Time
Akamai
AkamaiNASDAQ: AKAM: Provides content delivery, cybersecurity, and cloud computing services globally.
12+ YOE12 years experience and Bachelor's in CS/Engineering (or equivalent); deep expertise in GPU architectures/CUDA, cloud networking, product strategy, financial modeling, and AI/HPC workload patterns.
CUDA
2mo
Save
Mark Applied
Hide
Silicon Validation Software Engineer- GPU IP Validation and Integration
Waltham, Massachusetts, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Develop graphics validation software and integrate it into system-level test environments; background in graphics, video encoding/processing, file systems, CPU/cache, kernel programming, or embedded systems.
2mo
Save
Mark Applied
Hide
Software Solutions Architect
Austin or Boxborough or Markham
$212k-$318k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Senior-level architect with strong software engineering, enterprise delivery, and customer-facing leadership; deep AI/ML and GPU/CPU experience.
C/C++, Python, Docker, Kubernetes, ONNX, PyTorch, ROCm, CUDA, TensorRT
1w
Save
Mark Applied
Hide
Systems ML Engineer (Member of the Technical Staff)
Cambridge, Massachusetts, United States
OnsiteFull Time
Transfyr
Transfyr: Building physical AI infrastructure for scientific research and automation.
Experience optimizing and deploying large-scale ML models for training and inference, profiling and custom GPU kernel development, distributed training, cloud and edge deployment, and infrastructure automation.
Nsight, PyTorch Profiler, PyTorch Distributed, PyTorch, JAX, Triton, CUDA, Kubernetes, Terraform, AWS, NCCL
5d
Save
Mark Applied
Hide
Senior Software Developer - Amazon Robotics , Autonomous Mobility
Denver or North Reading
$168k-$227k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
5+ years professional software development, strong C++ and Python skills, experience with perception/SLAM/sensor fusion, ROS, CUDA/GPU programming, AWS, system design and mentoring.
C++, Python, Robot Operating System (ROS), Simultaneous Localization and Mapping (SLAM), CUDA, AWS
4d
Save
Mark Applied
Hide
Staff/Principal DevOps Engineer, AI Inference
Cambridge, Massachusetts, United States
$192k-$272k/yr OnsiteFull Time
Lila Sciences
Lila Sciences: Develops an AI platform for autonomous scientific research and discovery.
Expertise operating GPU/accelerator infrastructure for ML, Kubernetes and AWS deployment experience, infrastructure-as-code (Terraform, Helm), Python proficiency, networking and performance optimization for low-latency inference.
Kubernetes, vLLM, Triton Inference Server, TGI, Terraform, Helm, EKS, EC2, S3, EFA, IAM, NCCL, Python, Rust, Go, CUDA
1mo
Save
Mark Applied
Hide
AI Systems Engineer
Boston, Massachusetts, United States
OnsiteFull Time
Glia AI
Glia AI: Automated AI systems engineering for high-performance infrastructure.
Experience with AI inference and distributed serving, strong systems and GPU optimization skills, proficiency in Python and PyTorch, familiarity with inference frameworks (vLLM, Triton, Ray Serve); PhD or equivalent experience preferred.
vLLM, PyTorch, Triton inference server, Ray Serve, CUDA, Python
3d
Save
Mark Applied
Hide
Hardware Systems Engineer
Maynard, Massachusetts, United States
$92k-$112k/yr OnsiteFull Time
Penguin Solutions
Penguin SolutionsNASDAQ: PENG: Provides high-performance AI and edge computing infrastructure solutions.
3+ YOEBachelor's in electrical engineering or related, 3+ years experience, strong computer architecture and server integration knowledge, storage/network familiarity, IPMI/Redfish, Python scripting, GPU toolkits, and virtualization/container experience.
SAN, NAS, CSIs, CNIs, IPMI, Redfish, Python, CUDA, ROCm, OpenCL, Kubernetes, VMware, Microsoft Hyper-V
3d
Save
Mark Applied
Hide
AI HPC Infrastructure Engineer
Boston, Massachusetts, United States
$150k-$170k/yr HybridFull Time
Analysis Group
Analysis Group: Provides economic, financial, and strategy consulting to businesses and governments.
5+ YOEBachelor's degree and 5+ years Linux/HPC systems administration experience; expertise with NVIDIA GPU stack, SLURM, GPFS, Kubernetes, Python/R, LLM tuning, Posit Workbench; must be authorized to work in the US.
Posit Workbench (RStudio Server Pro), MPI, OpenMP, PyTorch Distributed, Horovod, DeepSpeed, SLURM, Kubernetes, CUDA, cuDNN, NCCL, MLflow, Kubeflow, Docker, Singularity/Apptainer, GPFS (IBM Spectrum Scale), Bright Cluster Manager, Ansible, RDP, SSH, PyTorch, TensorFlow, Weights & Biases, NVIDIA GPU Operator, LiteLLM, Kong AI Gateway, Portkey
1mo
Save
Mark Applied
Hide
Sr. Principal Software Engineer
Burlington or United States or Europe or Asia or North America
$141k-$226k/yr RemoteFull Time
Cerence
CerenceNASDAQ: CRNC: Develops AI-powered voice assistants and software for automotive vehicles.
Proven experience optimizing ML inference in production, deep GPU architecture knowledge, hands-on CUDA kernel development, quantization techniques (INT8/INT4/FP4/FP8/AWQ/GPTQ), and edge/embedded deployment expertise.
vLLM, TensorRT‑LLM, llama.cpp, QAIRT, CUDA, AWQ, GPTQ
1w
Save
Mark Applied
Hide
Staff Engineer, Inference Optimizations
Boston, Massachusetts, United States
$191k-$239k/yr RemoteFull Time
DigitalOcean
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
5+ YOE5+ years in high-performance computing or AI infrastructure, deep GPU and inference optimization expertise, experience with CUDA/Triton/ROCm, attention-layer and kernel-level optimization, strong system design and leadership through influence.
AITER, CUDA, ROCm, TensorRT, Triton, FlashAttention
1mo
Save
Mark Applied
Hide
Staff Research Scientist, Localization and Mapping
Cambridge or Suzhou or Singapore
OnsiteFull Time
Venti Technologies
Venti Technologies: Develops autonomous robotic vehicles for industrial and logistics transport.
7+ YOEPhD or Master’s in CS/Robotics/Engineering, 7+ years in robotics/localization/mapping, strong C++ and GPU programming skills, experience with real-time systems and ROS/ROS2, expertise in SLAM, sensor fusion, state estimation and HD mapping.
C++, GPU programming, ROS/ROS2, SLAM algorithms, Kalman Filters, LiDAR, GNSS, IMU, Cameras, HD mapping systems
2mo
Save
Mark Applied
Hide
Forward Deployed Engineer
Boston or San Francisco or United States
FieldFull Time
Code Metal
Code Metal: Verifiable AI-powered code translation for mission-critical industries.
7+ YOE7+ years engineering experience preferred; strong programming in Python/C++/CUDA/Matlab/VHDL/Verilog/Rust; experience with GPU/FPGA/embedded systems; ability to obtain U.S. Top Secret clearance; willing to travel 30-75%.
Python, C++, CUDA, Matlab, VHDL, Verilog, Rust, GPU, FPGA
1mo
Save
Mark Applied
Hide
Senior Engineering Manager, ML Platform
Waltham, Massachusetts, United States
$198k-$300k/yr OnsiteFull Time
Boston Dynamics
Boston Dynamics: Building advanced mobile robots for industrial and warehouse automation.
7+ YOE2+ Mgmt7+ years engineering experience with 2+ years management; experience building/scaling ML/platform/infrastructure teams; expertise in GPU/distributed compute, large-scale data storage, and pipeline frameworks; hands-on coding and strong cross-functional communication.
Kubernetes, Slurm, Ray, GPU, TPU
1mo
Save
Mark Applied
Hide
Senior Imaging Software Engineer
Marlborough, Massachusetts, United States
OnsiteFull Time
IPG Photonics
IPG PhotonicsNASDAQ: IPGP: Develops and manufactures high-performance fiber lasers and systems.
5+ YOE5+ years R&D in medical imaging and control systems; expertise in computer vision, image processing, and ML; experience with embedded and real-time systems; GPU programming (CUDA/OpenCL) preferred; proficiency in C++, Python, Qt, OpenCV, TensorFlow/PyTorch; familiarity with medical device design controls.
C++, Python, Qt, CUDA, OpenCL, OpenCV, TensorFlow, PyTorch
2mo
Save
Mark Applied
Hide
Senior Distinguished Engineer, AI Compute (Remote Eligible)
San Francisco or Cambridge or Richmond or McLean or New York or San Jose or San Francisco
$286k-$392k/yr RemoteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
10+ YOE10+ years AI/ML engineering; Python/Go/Scala/Java; distributed compute; CPU/GPU infrastructure; ML/AI workloads.
Python, Go, Scala, Java, Spark, Dask, Ray, Flink, Kubernetes, AWS Lambda
2mo
Save
Mark Applied
Hide
Machine Learning Architect
Boston, Massachusetts, United States
OnsiteFull Time
Ayo Semiconductor
Ayo Semiconductor: Building photonic processors to accelerate AI computing workloads.
PhD in ML or equivalent industry experience; distributed training on GPUs; built and trained neural networks; Python proficiency; PyTorch, TensorFlow, or JAX experience.
Python, PyTorch, TensorFlow, JAX
2w
Save
Mark Applied
Hide
Member of Technical Staff – Senior Engineer, Reinforcement Learning for Wholebody Control
Cambridge, Massachusetts, United States
$180k-$240k/yr OnsiteFull Time
Walden Robotics
Walden Robotics: A developing general-purpose robots and related control and simulation systems to improve quality of life.
Proven RL experience for continuous control, sim-to-real transfer on real robots, control theory knowledge, large-scale GPU simulation, strong Python and C++ skills, and software engineering for training/deployment pipelines.
Python, C++
3w
Save
Mark Applied
Hide
Lead Quantitative Software Engineer / Front-Office Quant Developer, VP
Boston, Massachusetts, United States
OnsiteFull Time
State Street
State StreetNYSE: STT: Provides investment servicing and management to institutional investors.
Expertise in modern C++ (C++20/23) and core Java; experience with Python, kdb+/q, SQL, Linux, Boost, QuantLib, and GPU programming; strong quantitative background in stochastic calculus, Monte Carlo, and calibration methods.
C++, Java, Python, kdb+/q, SQL, Linux, Boost, QuantLib, Nvidia CUDA, OpenCL, NumPy, Pandas, Git, Jira, CI/CD