QualcommNASDAQ: QCOM: Designs and manufactures semiconductors and wireless telecommunications products.
Bachelor's in a related field, knowledge of graphics-suitable languages (C, C++), developing and verifying GPU drivers/features; Master's and 1+ year GPU experience preferred.
Santa Clara or Westford or Austin or Durham or Seattle
$224k-$431k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOE12+ years software engineering experience in GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
12+ YOE12 years experience and Bachelor's in CS/Engineering (or equivalent); deep expertise in GPU architectures/CUDA, cloud networking, product strategy, financial modeling, and AI/HPC workload patterns.
Analog DevicesNASDAQ: ADI: Designs and manufactures semiconductors for signal processing and power management.
Expert in AI infrastructure with on-prem, hybrid, and cloud-native architectures; lead cross-team initiatives; strong Kubernetes, GPU, and IaC; multi-region experience.
Systems ML Engineer (Member of the Technical Staff)
Cambridge, Massachusetts, United States
OnsiteFull Time
Transfyr: Building physical AI infrastructure for scientific research and automation.
Experience optimizing and deploying large-scale ML models for training and inference, profiling and custom GPU kernel development, distributed training, cloud and edge deployment, and infrastructure automation.
5+ YOEBachelor's in CS or equivalent, 5+ years technical experience, 2+ years running AI/ML workloads on GPUs/TPUs with Kubernetes/GKE, 2+ years producing developer content; Python preferred.
Glia AI: Automated AI systems engineering for high-performance infrastructure.
Experience with AI inference and distributed serving, strong systems and GPU optimization skills, proficiency in Python and PyTorch, familiarity with inference frameworks (vLLM, Triton, Ray Serve); PhD or equivalent experience preferred.
vLLM, PyTorch, Triton inference server, Ray Serve, CUDA, Python
Burlington or United States or Europe or Asia or North America
$141k-$226k/yrRemoteFull Time
CerenceNASDAQ: CRNC: Develops AI-powered voice assistants and software for automotive vehicles.
Proven experience optimizing ML inference in production, deep GPU architecture knowledge, hands-on CUDA kernel development, quantization techniques (INT8/INT4/FP4/FP8/AWQ/GPTQ), and edge/embedded deployment expertise.
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
5+ YOE5+ years in high-performance computing or AI infrastructure, deep GPU and inference optimization expertise, experience with CUDA/Triton/ROCm, attention-layer and kernel-level optimization, strong system design and leadership through influence.
Staff Research Scientist, Localization and Mapping
Cambridge or Suzhou or Singapore
OnsiteFull Time
Venti Technologies: Develops autonomous robotic vehicles for industrial and logistics transport.
7+ YOEPhD or Master’s in CS/Robotics/Engineering, 7+ years in robotics/localization/mapping, strong C++ and GPU programming skills, experience with real-time systems and ROS/ROS2, expertise in SLAM, sensor fusion, state estimation and HD mapping.
C++, GPU programming, ROS/ROS2, SLAM algorithms, Kalman Filters, LiDAR, GNSS, IMU, Cameras, HD mapping systems
Code Metal: Verifiable AI-powered code translation for mission-critical industries.
7+ YOE7+ years engineering experience preferred; strong programming in Python/C++/CUDA/Matlab/VHDL/Verilog/Rust; experience with GPU/FPGA/embedded systems; ability to obtain U.S. Top Secret clearance; willing to travel 30-75%.
Boston Dynamics: Building advanced mobile robots for industrial and warehouse automation.
7+ YOE2+ Mgmt7+ years engineering experience with 2+ years management; experience building/scaling ML/platform/infrastructure teams; expertise in GPU/distributed compute, large-scale data storage, and pipeline frameworks; hands-on coding and strong cross-functional communication.
Merlin Labs: Develops autonomous pilot systems for commercial and military aircraft.
10+ YOE4+ Mgmt10+ years engineering experience with 4+ years technical leadership; production-scale data pipeline, dataset management, labeling, versioning, and integration with large GPU/TPU training infrastructure; simulation and safety-critical domain experience; strong cross-functional communication.
MLOps, GPU, TPU, DO-178C, DO-254, ISO 26262, SOTIF
IPG PhotonicsNASDAQ: IPGP: Develops and manufactures high-performance fiber lasers and systems.
5+ YOE5+ years R&D in medical imaging and control systems; expertise in computer vision, image processing, and ML; experience with embedded and real-time systems; GPU programming (CUDA/OpenCL) preferred; proficiency in C++, Python, Qt, OpenCV, TensorFlow/PyTorch; familiarity with medical device design controls.
Ayo Semiconductor: Building photonic processors to accelerate AI computing workloads.
PhD in ML or equivalent industry experience; distributed training on GPUs; built and trained neural networks; Python proficiency; PyTorch, TensorFlow, or JAX experience.