84 model deployment engineer jobs at 66 companies in Ross, CA
2mo
Save
Mark Applied
Hide
2mo
AI Deployment Strategist
San Francisco, California, United States
HybridFull Time
Pigment: AI-powered business planning and performance management platform.
2+ YOEEngineering or computer science degree, 2+ years in technical client-facing/implementation roles, experience in data modeling and AI/ML, proficiency with formulas and structured modeling, programming (Python/SQL), and APIs/data pipelines.
Applied Research Scientist / Engineer - Deployment
Palo Alto, California, United States
OnsiteFull Time
Rhoda AI: Developing generalist robotic intelligence for real-world industrial automation.
Strong ML research and engineering skills with hands-on experience fine-tuning or adapting large models; translate customer requirements into model adaptations; customer-facing applied research or solutions engineering experience; staff-level/ senior execution expectations.
OpenAI: Develops artificial intelligence models and generative AI software services.
Proven experience leading engineering teams, shipping production systems at scale, and owning model deployment, experimentation, and measurement tooling; strong communication and cross-functional collaboration skills.
Black Forest Labs: Developing frontier generative AI models for visual intelligence.
Proven robotics/forward-deployed engineering experience, customer-facing integration skills, experience with action/VLA models, model deployment and latency optimization, strong communication and collaboration.
Black Forest Labs: Develops generative AI models for image and video creation.
Proven robotics engineering experience; hands-on with action/VLA models (π0/π0.5, LeRobot); customer-facing deployments, model hosting, inference/edge constraints, and strong communication skills.
Paradigm: Venture capital firm focused on crypto and frontier technologies.
Experienced engineer with infra, security, and AI model experience; familiarity with durable control planes, runtimes, secrets, observability, and self-hosted deployment.
Verse: Software to plan, procure, and manage clean energy portfolios.
5+ YOE5+ years building production optimization models, strong Python software engineering, deployment of scalable optimization services, energy industry experience (electricity markets, battery storage), Master's degree in quantitative field.
P-1 AI: Developing AI agents for industrial engineering and physical design.
Experience shipping data-driven or AI systems to production (Python preferred); physical engineering background; building integrations, fine-tuning models, customer-facing deployment and troubleshooting experience.
Pika: AI-powered platform for generating and editing professional videos
5+ YOE5+ years engineering experience in inference acceleration, GPU programming (CUDA, NCCL), model deployment, quantization, attention optimization, and parallelism for production-scale AI systems.
Sonatus: Develops software platforms for AI-enabled software-defined vehicles.
10+ YOE10+ years ML engineering with 3+ years in Edge AI/embedded systems, Bachelor’s in CS/EE/Software Engineering, expert Python, C++14/17, PyTorch/TensorFlow, edge deployment and model optimization experience.
Artos: AI-powered document authoring platform for life sciences R&D.
2+ YOE2+ years building and deploying AI/ML applications, hands-on LLM experience, backend engineering with Python, API development (FastAPI/Django), cloud container deployment, R&D on model capabilities, and evaluation tooling experience.
Vantaca: AI software for community association and HOA management.
8+ YOE8+ years in infrastructure/DevOps/SRE; strong cloud expertise; experience with CI/CD, PostgreSQL, Redis, APM, model serving, vector databases, GPU optimization, and LLM deployment.
PostgreSQL, Redis, APM, CI/CD, vector databases, model serving frameworks, LLM
UnityNYSE: U: Provides software for creating real-time 3D interactive content.
8+ YOE4+ Mgmt8+ years software/ML engineering with 4+ years on-device/edge inference; production deployment of transformer/diffusion models; WebGPU/WGSL and GPU API performance tuning; proficiency with TypeScript/JavaScript and Python; leadership experience.
MetaNASDAQ: META: Develops social networking platforms and virtual reality technologies.
Bachelor's in CS/CE or equivalent experience; experience shipping ML models to production; background in computer vision/ML with focus on gesture recognition/pose estimation; model optimization for on-device deployment; cross-functional collaboration.
Fireworks AI: Provides high-performance generative AI model inference and deployment infrastructure.
5+ YOE5+ years in customer-facing technical engineering roles, strong Python and Kubernetes skills, experience with LLM inference, model serving and fine-tuning, cloud GPU deployment across major clouds, and exceptional communication.
Python, Kubernetes, vLLM, SGLang, TensorRT-LLM, AWS, Microsoft Azure, GCP, Azure AI Foundry, AWS Bedrock, SageMaker, GCP Vertex
PinterestNYSE: PINS: Visual discovery engine for finding inspiration and creative ideas.
4+ YOE4+ years industry experience, strong production Python, statistics and ML fundamentals, experience with adtech/CTV/RTB preferred, familiarity with LLMs and model deployment/monitoring, Bachelor's in CS/Math/Engineering or equivalent.
San Francisco or New York City or Portland or United States or Canada
$167k-$208k/yrHybridFull Time
Mercury: Banking services and financial software designed for startup companies.
5+ YOE5+ years in ML engineering/MLOps or backend engineering; production ML service experience; strong Python and API framework skills (FastAPI/Flask); model deployment, CI/CD, registries, observability, SQL, low-latency stores, and streaming pipelines.
Staff Software Engineer/ Tech Lead - Onboard Model Consolidation
Mountain View, California, United States
$251k-$310k/yrOnsiteFull Time
Waymo: Autonomous driving technology for ride-hailing and logistics.
8+ YOE8+ years professional software development; BS/MS in CS/EE/Robotics/related or equivalent experience; extensive C++ experience building large-scale, high-performance systems; leadership on cross-functional projects; ML deployment and inference expertise.
Lead AI Engineer -- Advanced AI (applied ML, LLMs, agentic AI, ML Ops)
Brooklyn Park or Sunnyvale
$132k-$286k/yrHybridFull Time
TargetNYSE: TGT: General merchandise retailer operating physical stores and e-commerce.
5+ YOEDegree in quantitative field or equivalent experience,5+ years applied ML/AI experience,experience with LLMs,agentic systems,model deployment,software engineering practices and strong communication.