Santa Clara or Westford or Austin or Durham or Seattle
$224k-$431k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOE12+ years software engineering experience in GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
12+ YOE12 years experience and Bachelor's in CS/Engineering (or equivalent); deep expertise in GPU architectures/CUDA, cloud networking, product strategy, financial modeling, and AI/HPC workload patterns.
Silicon Validation Software Engineer- GPU IP Validation and Integration
Waltham, Massachusetts, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Develop graphics validation software and integrate it into system-level test environments; background in graphics, video encoding/processing, file systems, CPU/cache, kernel programming, or embedded systems.
Systems ML Engineer (Member of the Technical Staff)
Cambridge, Massachusetts, United States
OnsiteFull Time
Transfyr: Building physical AI infrastructure for scientific research and automation.
Experience optimizing and deploying large-scale ML models for training and inference, profiling and custom GPU kernel development, distributed training, cloud and edge deployment, and infrastructure automation.
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
5+ years professional software development, strong C++ and Python skills, experience with perception/SLAM/sensor fusion, ROS, CUDA/GPU programming, AWS, system design and mentoring.
C++, Python, Robot Operating System (ROS), Simultaneous Localization and Mapping (SLAM), CUDA, AWS
Glia AI: Automated AI systems engineering for high-performance infrastructure.
Experience with AI inference and distributed serving, strong systems and GPU optimization skills, proficiency in Python and PyTorch, familiarity with inference frameworks (vLLM, Triton, Ray Serve); PhD or equivalent experience preferred.
vLLM, PyTorch, Triton inference server, Ray Serve, CUDA, Python
Penguin SolutionsNASDAQ: PENG: Provides high-performance AI and edge computing infrastructure solutions.
3+ YOEBachelor's in electrical engineering or related, 3+ years experience, strong computer architecture and server integration knowledge, storage/network familiarity, IPMI/Redfish, Python scripting, GPU toolkits, and virtualization/container experience.
Analysis Group: Provides economic, financial, and strategy consulting to businesses and governments.
5+ YOEBachelor's degree and 5+ years Linux/HPC systems administration experience; expertise with NVIDIA GPU stack, SLURM, GPFS, Kubernetes, Python/R, LLM tuning, Posit Workbench; must be authorized to work in the US.
Burlington or United States or Europe or Asia or North America
$141k-$226k/yrRemoteFull Time
CerenceNASDAQ: CRNC: Develops AI-powered voice assistants and software for automotive vehicles.
Proven experience optimizing ML inference in production, deep GPU architecture knowledge, hands-on CUDA kernel development, quantization techniques (INT8/INT4/FP4/FP8/AWQ/GPTQ), and edge/embedded deployment expertise.
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
5+ YOE5+ years in high-performance computing or AI infrastructure, deep GPU and inference optimization expertise, experience with CUDA/Triton/ROCm, attention-layer and kernel-level optimization, strong system design and leadership through influence.
Staff Research Scientist, Localization and Mapping
Cambridge or Suzhou or Singapore
OnsiteFull Time
Venti Technologies: Develops autonomous robotic vehicles for industrial and logistics transport.
7+ YOEPhD or Master’s in CS/Robotics/Engineering, 7+ years in robotics/localization/mapping, strong C++ and GPU programming skills, experience with real-time systems and ROS/ROS2, expertise in SLAM, sensor fusion, state estimation and HD mapping.
C++, GPU programming, ROS/ROS2, SLAM algorithms, Kalman Filters, LiDAR, GNSS, IMU, Cameras, HD mapping systems
Code Metal: Verifiable AI-powered code translation for mission-critical industries.
7+ YOE7+ years engineering experience preferred; strong programming in Python/C++/CUDA/Matlab/VHDL/Verilog/Rust; experience with GPU/FPGA/embedded systems; ability to obtain U.S. Top Secret clearance; willing to travel 30-75%.
Boston Dynamics: Building advanced mobile robots for industrial and warehouse automation.
7+ YOE2+ Mgmt7+ years engineering experience with 2+ years management; experience building/scaling ML/platform/infrastructure teams; expertise in GPU/distributed compute, large-scale data storage, and pipeline frameworks; hands-on coding and strong cross-functional communication.
IPG PhotonicsNASDAQ: IPGP: Develops and manufactures high-performance fiber lasers and systems.
5+ YOE5+ years R&D in medical imaging and control systems; expertise in computer vision, image processing, and ML; experience with embedded and real-time systems; GPU programming (CUDA/OpenCL) preferred; proficiency in C++, Python, Qt, OpenCV, TensorFlow/PyTorch; familiarity with medical device design controls.
Ayo Semiconductor: Building photonic processors to accelerate AI computing workloads.
PhD in ML or equivalent industry experience; distributed training on GPUs; built and trained neural networks; Python proficiency; PyTorch, TensorFlow, or JAX experience.
Member of Technical Staff – Senior Engineer, Reinforcement Learning for Wholebody Control
Cambridge, Massachusetts, United States
$180k-$240k/yrOnsiteFull Time
Walden Robotics: A developing general-purpose robots and related control and simulation systems to improve quality of life.
Proven RL experience for continuous control, sim-to-real transfer on real robots, control theory knowledge, large-scale GPU simulation, strong Python and C++ skills, and software engineering for training/deployment pipelines.
Lead Quantitative Software Engineer / Front-Office Quant Developer, VP
Boston, Massachusetts, United States
OnsiteFull Time
State StreetNYSE: STT: Provides investment servicing and management to institutional investors.
Expertise in modern C++ (C++20/23) and core Java; experience with Python, kdb+/q, SQL, Linux, Boost, QuantLib, and GPU programming; strong quantitative background in stochastic calculus, Monte Carlo, and calibration methods.