1 model deployment engineer job at 1 company in Roanoke, VA

PromotedHiringCafe
ML Engineer - Inference & Model Deployment
Cupertino, CA, US
$250k-$310k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Turn powerful AI and ML models into fast, reliable production systems. Own inference latency, throughput, model-serving architecture, multi-GPU systems, and production deployment for millions of users.
Python, PyTorch, vLLM, SGLang, TensorRT, LLMs
4w
Save
Mark Applied
Hide
Manager, Engineering - App Engine (CUDA)
Ann Arbor or Blacksburg or Fort Worth or United States
$161k-$193k/yr HybridFull Time
Torc Robotics
Torc Robotics: Software for autonomous heavy-duty trucking.
3+ MgmtBachelor's or Master's in a technical field, 3+ years people management, deep expertise in CUDA/GPU parallel computing, C++ and Linux, experience with PyTorch/TensorRT/ONNX and model deployment, safety-critical systems knowledge.
CUDA, PyTorch, TensorRT, ONNX, C++, Linux