9 gpu kernel development engineer jobs at 6 companies in Massachusetts
2w
Save
Mark Applied
Hide
2w
Senior Software Engineer - GPU Local AI Platforms
Santa Clara or Westford or Austin or Durham or Seattle
$224k-$431k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOE12+ years software engineering experience in GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
Santa Clara or Austin or Westford or Durham or Seattle
$224k-$431k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
12+ YOE12+ years software engineering experience with GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
Senior Software Engineer, AI and DL Kernel Libraries
Santa Clara or Georgia or Texas or Colorado or Washington or California or Oregon or Massachusetts
$184k-$288k/yrRemoteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
6+ YOEMasters (or equivalent experience) in CS/EE, 6+ years ML/DL systems experience, strong Python and C/C++ skills, GPU kernel development experience (CUDA, Triton, cuTile), familiarity with deep learning frameworks and inference runtimes.
Silicon Validation Software Engineer- GPU IP Validation and Integration
Waltham, Massachusetts, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Develop graphics validation software and integrate it into system-level test environments; background in graphics, video encoding/processing, file systems, CPU/cache, kernel programming, or embedded systems.
Systems ML Engineer (Member of the Technical Staff)
Cambridge, Massachusetts, United States
OnsiteFull Time
Transfyr: Building physical AI infrastructure for scientific research and automation.
Experience optimizing and deploying large-scale ML models for training and inference, profiling and custom GPU kernel development, distributed training, cloud and edge deployment, and infrastructure automation.
Burlington or United States or Europe or Asia or North America
$141k-$226k/yrRemoteFull Time
CerenceNASDAQ: CRNC: Develops AI-powered voice assistants and software for automotive vehicles.
Proven experience optimizing ML inference in production, deep GPU architecture knowledge, hands-on CUDA kernel development, quantization techniques (INT8/INT4/FP4/FP8/AWQ/GPTQ), and edge/embedded deployment expertise.
Systems & Technology Research: Develops advanced technology for national security and defense applications.
5+ YOEAbility to obtain Top Secret clearance, 5+ years experience, proficiency with GNU/Linux toolchains, C/C++, Python or MATLAB, VHDL/Verilog, embedded real-time software, and strong communication and leadership skills.
MatrixSpace: AI-powered radar systems for airspace safety and drone detection.
Bachelor's degree or equivalent experience; professional experience building, deploying, and maintaining production embedded software on constrained edge devices; strong C/C++ skills; working knowledge of Golang and Python 3.8+; Yocto/embedded Linux experience; debugging and performance optimization; cross-team collaboration.