2 model deployment engineer jobs at 2 companies in Lodi, CA
🚀PromotedHiringCafe
ML Engineer - Inference & Model Deployment
Cupertino, CA, US
$250k-$310k/yrOn-SiteFull Time
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Turn powerful AI and ML models into fast, reliable production systems. Own inference latency, throughput, model-serving architecture, multi-GPU systems, and production deployment for millions of users.
Vagaro: Cloud-based management platform for beauty and wellness businesses.
2+ YOE2+ years building LLM/LLM-based or generative AI applications; proficiency in Python and SQL; experience with cloud platforms (AWS/GCP/Azure), Docker, Kubernetes, ML model deployment, and collaboration with product teams.
New York or Rochester or Albany or Melville or San Francisco or Los Angeles or Sacramento or Dallas or Miami or Philadelphia or Washington or Atlanta
$74k-$212k/yrHybridFull Time
PwC: Providing audit, tax, and management consulting services to businesses.
5+ YOE2+ MgmtBachelor's degree; 5+ years AI engineering or related field; Master's preferred; cloud-native microservices; Docker/Kubernetes; CI/CD; AI/ML model deployment; experience leading teams.