2 gpu inference performance engineer jobs at 1 company in Seaside, CA

2w
Save
Mark Applied
Hide
Software Development Engineer, AI/ML, AWS Neuron, Model Inference
Cupertino, California, United States
$165k-$224k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
3+ YOEBachelor's degree or equivalent; 3+ years professional software development and systems design experience; C++ or Python; machine learning, LLM, performance, memory, parallel computing, debugging, and profiling expertise.
AWS Neuron, Inferentia, Trainium, PyTorch, JAX, Python, C++, CUDA, CUTLASS, FlashInfer, Triton, vLLM, SGLang, TensorRT
6d
Save
Mark Applied
Hide
Senior Software Development Engineer, AI/ML, AWS Neuron, Model Inference
Cupertino, California, United States
$193k-$262k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
5+ YOEBachelor's degree and 5+ years of professional software development and systems design experience. Requires C++ or Python, machine learning and LLM knowledge, system performance, memory management, parallel computing, debugging, and profiling.
Amazon Neuron, PyTorch, JAX, Python, C++, CUDA, CUTLASS, FlashInfer, Triton, vLLM, SGLang, TensorRT