169 ai deployment engineer jobs at 76 companies in Live Oak, CA
2mo
Save
Mark Applied
Hide
2mo
Forward Deployment Engineer
Menlo Park, California, United States
OnsiteFull Time
Hippocratic AI: Develops safety-focused generative AI for autonomous clinical healthcare conversations.
0+ YOEBachelor's in engineering/CS from a top university, 0–2+ years software engineering experience with production AI deployments, experience with LLMs/RAG, Python, cloud (AWS/GCP/Azure), EHR/FHIR/HL7 integrations, strong communication skills.
LLMs, RAG, Python, APIs, AWS, GCP, Azure, FHIR, HL7, EHR
Uniphore: Develops an AI cloud platform for enterprise customer engagement.
6+ YOEDegree in CS/Data Science,6+ years engineering experience (2+ in customer-facing roles),proven AI/ML production deployments,full-stack and DevOps skills,knowledge of LLMs,RAG,vector DBs and agent orchestration.
ThinkingAI: Enterprise platform for deploying autonomous AI agent teams.
3+ YOERequires 3+ years in data analysis, AI product delivery, or technical customer-facing work; SQL, Python, LLMs, vector databases, agent frameworks, API/SDK integration, and a bachelor's degree or higher.
Sonatus: Develops software platforms for AI-enabled software-defined vehicles.
10+ YOE10+ years ML engineering with 3+ years in Edge AI/embedded systems, Bachelor’s in CS/EE/Software Engineering, expert Python, C++14/17, PyTorch/TensorFlow, edge deployment and model optimization experience.
Capital OneNYSE: COF: Provides credit card, banking, and auto loan services.
4+ YOEBachelor's plus 4 years or master's plus 2 years developing AI/ML technologies; 4 years programming in Python, Go, Scala, or Java. Cloud AI deployment experience preferred.
McLean or San Francisco or Cambridge or San Jose or New York
$230k-$286k/yrOnsiteFull Time
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's in CS/AI/EE/CE+6yrs or Master's+4yrs experience in AI/ML, 6+ years programming in Python/Go/Scala/Java, cloud AI deployment experience, leadership and systems engineering skills.
Lead AI Engineer -- Advanced AI (applied ML, LLMs, agentic AI, ML Ops)
Brooklyn Park or Sunnyvale
$132k-$286k/yrHybridFull Time
TargetNYSE: TGT: General merchandise retailer operating physical stores and e-commerce.
5+ YOEDegree in quantitative field or equivalent experience,5+ years applied ML/AI experience,experience with LLMs,agentic systems,model deployment,software engineering practices and strong communication.
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
10+ YOEExtensive AI/ML leadership with deep learning, NLP, LLMs, large-scale model development and production deployment; strong communication and mentorship skills.
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
8+ YOE8+ years applied ML/AI engineering; 3+ years delivering production AI systems; MS in CS/Robotics/EE or related; PhD preferred; strong multimodal/3D vision and deployable inference pipelines; strong Python and ML library skills.
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
5+ YOE5+ years AI/ML experience, 1+ year building agentic AI, strong Python, LangGraph/LangChain and LLM experience, prompt engineering, cloud AI platforms, APIs, CI/CD, enterprise deployment and leadership skills.
Python, LangGraph, LangChain, LLMs, GPT, Claude, Llama, Azure AI Foundry, AWS Bedrock, Google Gemini Enterprise, APIs, CI/CD, Claude Code, Codex, Databricks, MLOps, LLMOps
UST: Global provider of digital transformation and IT services.
5+ YOERequires 5+ years in AI/ML, 1+ year building Agentic AI solutions, Python, LangGraph, LangChain, LLMs, prompt engineering, enterprise AI architecture, cloud deployment, APIs, CI/CD, and technical leadership.
Python, LangGraph, LangChain, GPT, Claude, Llama, Azure AI Foundry, AWS Bedrock, Google Gemini Enterprise, CI/CD, Databricks, Claude Code, Codex, APIs, MLOps, LLMOps, HSA, FSA
Fireworks AI: Provides high-performance generative AI model inference and deployment infrastructure.
5+ YOE5+ years in customer-facing technical engineering roles, strong Python and Kubernetes skills, experience with LLM inference, model serving and fine-tuning, cloud GPU deployment across major clouds, and exceptional communication.
Python, Kubernetes, vLLM, SGLang, TensorRT-LLM, AWS, Microsoft Azure, GCP, Azure AI Foundry, AWS Bedrock, SageMaker, GCP Vertex
Agilent TechnologiesNew York Stock Exchange: A: Providing instruments, software, and consumables for laboratory scientific discovery.
8+ YOEFull‑stack engineering experience building LLM applications, agents, RAG, evaluation, observability, and production deployment; familiarity with MCP/A2A patterns; ability to operate under regulated constraints; typically 8+ years experience.
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
5+ YOE5+ years in engineering/data roles with 2+ years deploying GenAI/LLMs, strong Python skills, experience with LangChain/LlamaIndex/AutoGen, cloud AI services and vector databases, and production deployment experience.
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
5+ YOEMaster's+3yrs or Bachelor's+5yrs in engineering fields; 5+ years networking for accelerator systems, Ethernet and RDMA fabrics; experience with AI/HPC/cloud deployments; security screening required.
Conviva: Real-time analytics and operational intelligence for digital businesses.
3+ YOE3+ years software engineering experience; strong Python backend skills; experience building and operating LLM-powered agents, MCP/agent integrations, cloud deployments (GCP/AWS), and observability for production systems.
Deployment Manager – Global Data Center Build and Deploy
Sunnyvale or United States
FieldFull Time
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
10+ YOE10+ years data center deployment experience, hands-on hyperscale AI/HPC infrastructure deployment, structured cabling expertise, ability to direct technicians, DCIM familiarity, and experience with high-power/high-density environments.
3+ YOEMaster's degree or higher in a technical field; 3+ years in HPC, AI infrastructure, model optimization, or embedded deployment; C++, Python, CUDA, OpenMP, inference engines, GPU architectures, and system profiling expertise.
Albany or Arlington or Atlanta or Austin or Beaverton or Bentonville or Boston or Carmel or Charlotte or Chicago or Cincinnati or Cleveland or Columbus or Culver City or Denver or Des Moines or Detroit or Hartford or Houston or Irving or Kirkland or Miami or Milwaukee or Minneapolis or Morristown or Mountain View or Nashville or New York City or Oklahoma City or Overland Park or Philadelphia or Pittsburgh or Raleigh or Redmond or Sacramento or San Diego or San Francisco or Scottsdale or Seattle or St. Louis or St. Petersburg or Walnut Creek or United States
$54k-$206k/yrHybridFull Time
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
3+ YOERequires a bachelor's degree or equivalent experience, 3+ years in machine learning, 2+ years with Snowflake Cortex AI, cloud AI/ML deployment, distributed systems, and Python-based frameworks.
Sr. Staff AI Engineer - On-Prem AI Infrastructure & Agentic Systems
San Jose, California, United States
$140k-$165k/yrOnsiteFull Time
SK hynix Memory Solutions AmericaKorea Exchange: 000660: Develops semiconductor controllers and firmware for enterprise data storage.
2+ YOE2+ years in AI/ML engineering with on-prem/private-cloud deployment, experience building agentic AI and RAG pipelines, model fine-tuning (LoRA/QLoRA), Python/Linux/Docker/Kubernetes proficiency, and familiarity with vector DBs and AI serving frameworks.