1 junior mlops engineer job at 1 company in Gilroy, CA
🚀PromotedHiringCafe
ML Engineer - Inference & Model Deployment
Cupertino, CA, US
$250k-$310k/yrOn-SiteFull Time
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Turn powerful AI and ML models into fast, reliable production systems. Own inference latency, throughput, model-serving architecture, multi-GPU systems, and production deployment for millions of users.
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Build the ML and AI search behind HiringCafe — ranking, recommenders, retrieval, and LLM agents that surface jobs people would never find on their own.
2+ YOEBachelor's or equivalent,2+ years software development (Python,C++),1+ year GenAI experience,experience with ADK and Google Cloud,EMR not mentioned,excellent communication and critical thinking.
Python, C++, Agent Development Kit (ADK), Google Cloud Platform (GCP), MLOps