1 model deployment engineer job at 1 company in Wooster, OH

PromotedHiringCafe
ML Engineer - Inference & Model Deployment
Cupertino, CA, US
$250k-$310k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Turn powerful AI and ML models into fast, reliable production systems. Own inference latency, throughput, model-serving architecture, multi-GPU systems, and production deployment for millions of users.
Python, PyTorch, vLLM, SGLang, TensorRT, LLMs
2w
Save
Mark Applied
Hide
Director of AI Engineering
Cleveland, Ohio, United States
OnsiteFull Time
Flexjet
Flexjet: Private jet provider offering fractional ownership and charter services.
10+ YOE5+ Mgmt10+ years software/ML engineering experience, 5+ years leadership, MLOps, cloud infrastructure, model lifecycle, Generative AI and LLM deployment experience.
AWS, Azure, Google Cloud Platform (GCP), Kubernetes, Docker, Terraform, Python, SQL, Bash, Git, CI/CD