5 performance modeling engineer jobs at 2 companies in Seaside, CA

3w
Save
Mark Applied
Hide
Fellow Software Engineer — AI Performance & Reliability
San Jose or Bellevue
$235k-$402k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
PhD or equivalent in AI/ML/CS, strong software engineering, experience profiling and optimizing ML models and AI workloads, proficiency in Python/C++, ML frameworks, customer-facing troubleshooting and performance analysis.
Python, C++, PyTorch, TensorFlow, JAX, ROCm, HIP, CUDA, Triton, XLA, MLIR, NCCL
1mo
Save
Mark Applied
Hide
Sr. Staff Software Development Engineer - Collectives and Network optimization
San Jose, California, United States
$179k-$306k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Senior engineer with deep knowledge of network, NIC and GPU architecture, performance optimization and modeling, experience with AI frameworks (PyTorch, JAX, vLLM, SGLang) and ROCm; PhD or master's in CS/EE or related preferred; strong communication and leadership.
PyTorch, JAX, vLLM, SGLang, ROCm
2w
Save
Mark Applied
Hide
Software Development Engineer, AI/ML, AWS Neuron, Model Inference
Cupertino, California, United States
$165k-$224k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
3+ YOEBachelor's degree or equivalent; 3+ years professional software development and systems design experience; C++ or Python; machine learning, LLM, performance, memory, parallel computing, debugging, and profiling expertise.
AWS Neuron, Inferentia, Trainium, PyTorch, JAX, Python, C++, CUDA, CUTLASS, FlashInfer, Triton, vLLM, SGLang, TensorRT
3d
Save
Mark Applied
Hide
Senior Software Development Engineer, AI/ML, AWS Neuron, Model Inference
Cupertino, California, United States
$193k-$262k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
5+ YOEBachelor's degree and 5+ years of professional software development and systems design experience. Requires C++ or Python, machine learning and LLM knowledge, system performance, memory management, parallel computing, debugging, and profiling.
Amazon Neuron, PyTorch, JAX, Python, C++, CUDA, CUTLASS, FlashInfer, Triton, vLLM, SGLang, TensorRT
1mo
Save
Mark Applied
Hide
Software Development Manager, LLM Inference Model Enablement, Neuron SDK
Cupertino, California, United States
$213k-$288k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
7+ YOE3+ MgmtManage engineering team to onboard and optimize LLMs for inference on Trainium; strong background in LLM architectures, model performance optimization, and inference techniques; experience with PyTorch and Neuron stack.
PyTorch, AWS Neuron, Neuron compiler, Neuron runtime