67 model deployment engineer jobs at 38 companies in Prunedale, CA
2mo
Save
Mark Applied
Hide
2mo
ML Engineer - Inference & Model Deployment
Cupertino, California, United States
$250k-$310k/yrOnsiteFull Time
Hiring.Cafe: An AI-powered job search engine and aggregator.
Experience deploying and optimizing deep learning models in production, multi-GPU inference, profiling/benchmarking model performance, inference optimization techniques, and cloud/distributed systems familiarity.
John DeereNYSE: DE: Manufactures agricultural, construction, and forestry machinery and equipment.
3+ YOE3+ years software engineering with modern C++, applied ML for perception, experience with sensor data pipelines, model training/deployment, and system-level debugging.
Applied MaterialsNASDAQ: AMAT: Produces equipment and services for chip and display manufacturing.
5+ YOEBachelor's in engineering/CS/data science; 5+ years (or 2+ with a Master's) in application engineering, algorithm development, or similar; experience in industrial/manufacturing environments; model lifecycle and production deployment; Python/R/C# and AI/ML familiarity.
Applied MaterialsNASDAQ: AMAT: Manufacturers of equipment for semiconductor and display production.
2+ YOEBachelor's in engineering/computer science/data science required; 5+ years relevant experience (or 2+ with a Master’s). Strong analytics, AI/ML familiarity, Python/R/C# programming, model lifecycle and production deployment experience.
Sonatus: Develops software platforms for AI-enabled software-defined vehicles.
10+ YOE10+ years ML engineering with 3+ years in Edge AI/embedded systems, Bachelor’s in CS/EE/Software Engineering, expert Python, C++14/17, PyTorch/TensorFlow, edge deployment and model optimization experience.
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
10+ YOEExtensive AI/ML leadership with deep learning, NLP, LLMs, large-scale model development and production deployment; strong communication and mentorship skills.
UnityNYSE: U: Provides software for creating real-time 3D interactive content.
8+ YOE4+ Mgmt8+ years software/ML engineering with 4+ years on-device/edge inference; production deployment of transformer/diffusion models; WebGPU/WGSL and GPU API performance tuning; proficiency with TypeScript/JavaScript and Python; leadership experience.
Staff Software Engineer/ Tech Lead - Onboard Model Consolidation
Mountain View, California, United States
$251k-$310k/yrOnsiteFull Time
Waymo: Autonomous driving technology for ride-hailing and logistics.
8+ YOE8+ years professional software development; BS/MS in CS/EE/Robotics/related or equivalent experience; extensive C++ experience building large-scale, high-performance systems; leadership on cross-functional projects; ML deployment and inference expertise.
Senior Principal Engineer- MLOps & AI Machinery, ADAS/AV
Sunnyvale, California, United States
$240k-$320k/yrHybridFull Time
Bosch: Global manufacturer of automotive and industrial engineering technology.
10+ YOEMaster's or PhD in CS/Robotics/EE/AI, 10+ years in software/system engineering for autonomous driving or ADAS, experience releasing L2+ AI systems, knowledge of training pipelines, model optimization, and embedded deployment.
TensorFlow, PyTorch, Python, C++, MLOps, CICD, SIL, HIL
E-Space: Builds sustainable LEO satellite networks for global IoT connectivity
Plan and execute vibration, shock, acoustic, and thermal-vacuum tests for large deployable antenna structures; instrumentation, data acquisition, test procedures, anomaly disposition, and model-test correlation.
Lead AI Engineer -- Advanced AI (applied ML, LLMs, agentic AI, ML Ops)
Brooklyn Park or Sunnyvale
$132k-$286k/yrHybridFull Time
TargetNYSE: TGT: General merchandise retailer operating physical stores and e-commerce.
5+ YOEDegree in quantitative field or equivalent experience,5+ years applied ML/AI experience,experience with LLMs,agentic systems,model deployment,software engineering practices and strong communication.
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
15+ YOE3+ Mgmt15+ years software engineering experience, deep Windows internals and security, LLM inference and GPU acceleration experience, proficiency in C++ and Python, experience with agent frameworks and local model deployment.
Senior Machine Learning Engineer, Model Serving Infrastructure (Multiple Positions)
San Jose, California, United States
$265k-$388k/yrOnsiteFull Time
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOEMaster's (plus 2 years) or Bachelor's (plus 5 years) in a quantitative field; 2+ years coding in Python or C++; Linux development experience; ML, system design, and production deployment experience.
Seattle or United States or Redwood City or Santa Clara or Austin
$126k-$264k/yrOnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
6+ YOE6+ years experience building and productionizing ML models; strong software engineering, model deployment, monitoring, data quality and debugging skills; stakeholder collaboration and mentoring experience.
Reevo: AI-native revenue operating system for go-to-market teams.
5+ YOE5+ years of software engineering experience focused on AI/ML systems, with expertise in LLMs, RAG, AI frameworks, model deployment, monitoring, prompt engineering, evaluation, vector databases, and semantic search.
RAG, large language models, AI/ML, ML operations, vector databases, embedding systems, semantic search
CoStar GroupNASDAQ: CSGP: Provides global real estate information, analytics, and online marketplaces.
3+ YOEBachelor's in CS/Data Science/Engineering or equivalent, 3+ years ML engineering experience with model optimization and deployment, Python, TensorFlow/PyTorch, cloud (AWS/Azure/GCP), Git, strong communication and problem-solving skills.
Principal Perception Engineer, Obstacle Foundation Models - Autonomous Vehicles
Santa Clara, California, United States
$272k-$431k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
15+ YOE15+ years developing deep-learning perception systems, technical leadership experience, proficiency in PyTorch, Python and/or C++, EM/deployment experience, strong communication and collaboration, BS/MS/PhD in CS/EE or equivalent.