10 model deployment engineer jobs at 5 companies in Del Mar, CA
🚀PromotedHiringCafe
ML Engineer - Inference & Model Deployment
Cupertino, CA, US
$250k-$310k/yrOn-SiteFull Time
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Turn powerful AI and ML models into fast, reliable production systems. Own inference latency, throughput, model-serving architecture, multi-GPU systems, and production deployment for millions of users.
QualcommNASDAQ: QCOM: Designs and manufactures semiconductors and wireless telecommunications products.
6+ YOE6+ years experience (PhD+6 / MS+7 / BS+8) in engineering; expertise in computer vision (optical flow, depth, tracking, SLAM), ML model development and on-device deployment; strong Python/C++ skills, ML frameworks, embedded system optimization, and technical leadership.
QualcommNASDAQ: QCOM: Designs and manufactures semiconductors and wireless telecommunications products.
2+ YOEBachelor's in CS/Engineering/Information Systems (or higher) with 2+ years relevant experience; strong Python and systems-language skills; experience with ML frameworks, model optimization, deployment, and distributed/GPU environments.
QualcommNASDAQ: QCOM: Designs and manufactures semiconductors and wireless telecommunications products.
4+ YOEDegree in CS/Engineering/Information Systems (BS/MS/PhD) with 4+–6+ years relevant engineering experience; hands-on ML model optimization (quantization, pruning), PyTorch/ONNX, Python, edge deployment, and software engineering best practices.
Efficient AI Systems, Principal Engineer (On-Device/Edge)
San Diego, California, United States
$207k-$310k/yrOnsiteFull Time
QualcommNASDAQ: QCOM: Designs and manufactures semiconductors and wireless telecommunications products.
6+ YOEAdvanced degree in engineering/computer science and 6–8+ years systems engineering experience; strong ML, deep learning, vision/ LiDAR, model quantization, and deployment experience on resource-constrained devices; leadership and cross-functional collaboration skills.
ICW Group: Provides workers' compensation and specialty property and casualty insurance.
3+ YOEBachelor's in a technical discipline, 3+ years AI/ML engineering (1–2 years in generative AI/LLMs), cloud model deployment experience (preferably AWS), Snowflake feature engineering, Python and ML framework proficiency, MLOps and containerization experience.
10+ YOEBachelor's degree or equivalent experience,10+ years cloud-native architecture experience,customer-facing engagement,cloud/on-prem engineering,programming and ML model deployment expertise.
5+ YOEPhD (5+ yrs) or MS (8+ yrs) in CS/Stats/Economics (or related), 6+ yrs ML experience, 4+ yrs GenAI experience, expertise in LLMs, RL, recommendations, causal inference, model deployment, and strong problem-solving and collaboration skills.
AI Lead – Autonomous Driving/Reasoning/Vision Language Action Models
San Diego, California, United States
$212k-$318k/yrOnsiteFull Time
QualcommNASDAQ: QCOM: Designs and manufactures semiconductors and wireless telecommunications products.
6+ YOEAdvanced degree in CS/EE/ME or related with 6+ years (PhD) / 7+ (MS) / 8+ (BS) systems engineering experience; deep learning, autonomous driving stacks, Python, PyTorch/TensorFlow, embedded/edge deployment experience, mentoring and system design skills.
Principal Engineer, Agentic AI Solutions Lead – Industrial & Embedded IoT, Edge AI On‑Prem Appliance
San Diego or Santa Clara
$193k-$289k/yrOnsiteFull Time
QualcommNASDAQ: QCOM: Designs and manufactures semiconductors and wireless telecommunications products.
6+ YOEBachelor's (8+ yrs) or Master's (7+ yrs) or PhD (6+ yrs) in CS/EE/IS; deep AI/ML experience (LLMs, RAG, agentic systems), hands-on software skills (Python/C++), model optimization and hardware-aware deployment, and customer-facing solution delivery.