293 ai modeling engineer jobs at 127 companies in Soquel, CA

2mo
Save
Mark Applied
Hide
Principal AI Performance Modeling Architect
Santa Clara or Austin
$203k-$348k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Extensive, senior experience optimizing large-scale ML systems and GPU architectures; CUDA programming; memory hierarchies; distributed training; transformer models.
PyTorch, CUDA, TensorRT, OpenAI Triton, Ray, Megatron-LM, NSight Compute, nvprof, PyTorch Profiler, KV cache optimization, Flash Attention, InfiniBand, RDMA, NVLink
2w
Save
Mark Applied
Hide
Senior Staff AI Engineer, Edge AI
Sunnyvale, California, United States
$227k-$300k/yr HybridFull Time
Sonatus
Sonatus: Develops software platforms for AI-enabled software-defined vehicles.
10+ YOE10+ years ML engineering with 3+ years in Edge AI/embedded systems, Bachelor’s in CS/EE/Software Engineering, expert Python, C++14/17, PyTorch/TensorFlow, edge deployment and model optimization experience.
Python, C++14, C++17, PyTorch, TensorFlow, ONNX, TFLite, TVM, scikit-learn, tslearn, statsmodels, NVIDIA TensorRT, Qualcomm SNPE, Gemini, OpenAI, Claude, Linux, QNX, CAN, DBC, UDS, SOME/IP, MQTT, ARM
1mo
Save
Mark Applied
Hide
Distinguished Engineer - AI
San Jose, California, United States
$266k-$396k/yr OnsiteFull Time
NetApp
NetAppNasdaq: NTAP: Provides intelligent data infrastructure for hybrid cloud environments.
15+ YOEExpert in AI inferencing and distributed systems at scale with 15+ years experience; hands-on with inference engines, model optimization, GPU/TPU orchestration, Kubernetes, RDMA/DPDK; strong architecture, communication, and mentorship skills.
TensorRT, vLLM, ONNX Runtime, Triton, RDMA, DPDK, Kubernetes, GPU, TPU
2w
Save
Mark Applied
Hide
Performance Modeling Architect – AI Systems
Durham or Santa Clara or Boston or Austin
$200k-$500k/yr OnsiteFull Time
Velaura AI
Velaura AI: Developing ultra-low-power silicon and IP for AI accelerators.
Experienced in computer/system architecture and performance modeling; building simulation or analytical models; strong programming skills (Python, C++); knowledge of CPUs/GPUs/accelerators; ability to analyze system bottlenecks.
Python, C++
1mo
Save
Mark Applied
Hide
Staff AI Software Engineer, Siri Core Modeling
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Work on Siri core modeling to build a privacy-first, context-aware AI assistant across Apple platforms with strong user-centered design focus.
iOS, iPadOS, macOS, watchOS, visionOS
4d
Save
Mark Applied
Hide
Traffic Modeling System Engineer for Missile Tracking (Space Force)
Scottsdale or Chantilly or Huntsville or San Jose or Grand Forks
$150k-$166k/yr HybridFull Time
General Dynamics Mission Systems
General Dynamics Mission SystemsNYSE: GD: Engineers secure communication and mission-critical defense technology systems.
8+ YOEBachelor's in systems engineering or related field, 8+ years experience (or Master's+6), DoD Top Secret clearance with polygraph eligibility, U.S. citizenship, systems engineering and traffic modeling experience with satellite/ground networks, Agile preferred, cloud/AI/image processing knowledge preferred.
2mo
Save
Mark Applied
Hide
AI Engineer
Palo Alto, California, United States
$100k-$200k/yr OnsiteFull Time
OPPO
OPPO: Develops and manufactures smartphones and connected consumer electronics.
Master’s in CS/AI/ML/CV/NLP or Bachelor’s with 2+ years; hands-on with LLM APIs and foundation models; strong prompt/context engineering and AI agent skills; production experience with AI systems.
LLM APIs, Foundation models, LangChain, OpenAI API, Anthropic API, Gemini API, Vector databases, RAG pipelines, Agent orchestration frameworks
2mo
Save
Mark Applied
Hide
Senior Staff AI Engineer | Senior Technical Lead - AI Modeling
Sunnyvale, California, United States
$241k-$326k/yr HybridFull Time
LinkedInNASDAQ: MSFT: Professional social network for career development and job recruitment.
5+ YOE5+ years in AI/ML engineering; 2+ years as a technical lead; strong Python and PyTorch skills; experience with LLMs, RL, and recommender systems.
Python, PyTorch, Large Language Models (LLMs), Reinforcement Learning
1mo
Save
Mark Applied
Hide
AI Engineer
Mountain View, California, United States
$150k-$190k/yr OnsiteFull Time
AI Fund
AI Fund: Builds and invests in artificial intelligence startups.
3+ YOE3+ years software engineering experience with end-to-end ownership of production AI applications, experience with large language or multimodal models, SQL/NoSQL, frontend/backend/cloud, retrieval and agentic systems, and strong communication.
large language models, multimodal models, SQL, NoSQL, APIs, cloud infrastructure
1mo
Save
Mark Applied
Hide
AI Field Engineer - AI Natives
San Mateo, California, United States
FieldFull Time
Fireworks AI
Fireworks AI: Provides high-performance generative AI model inference and deployment infrastructure.
5+ YOE5+ years in customer-facing technical engineering roles, strong Python and Kubernetes skills, experience with LLM inference, model serving and fine-tuning, cloud GPU deployment across major clouds, and exceptional communication.
Python, Kubernetes, vLLM, SGLang, TensorRT-LLM, AWS, Microsoft Azure, GCP, Azure AI Foundry, AWS Bedrock, SageMaker, GCP Vertex
1mo
Save
Mark Applied
Hide
Distinguished Engineer - AI (San Jose, CA, US, 95128)
San Jose, California, United States
$266k-$396k/yr OnsiteFull Time
NetApp
NetAppNASDAQ: NTAP: Sells enterprise data storage and cloud management software.
15+ YOE15+ years building low-latency, fault-tolerant distributed systems and AI/ML inference platforms; expertise with inference engines, model optimization, storage for AI, RDMA/DPDK, and Kubernetes-based orchestration.
TensorRT, vLLM, ONNX Runtime, Triton, RDMA, DPDK, Kubernetes
1mo
Save
Mark Applied
Hide
Principal Supply Chain Modeling Engineer
Santa Clara, California, United States
$240k-$380k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
15+ YOE15+ years quantitative modeling experience; expert in Python (Pandas, NumPy, SciPy), MATLAB, and Microsoft Excel; 2+ years hands-on AI/ML deployment experience; bachelor’s in quantitative field or equivalent; extreme attention to data integrity and VP-level communication skills.
Python, Pandas, NumPy, SciPy, MATLAB, Microsoft Excel, AWS, Azure, GCP
1mo
Save
Mark Applied
Hide
Principal AI Engineer - AI Builder
Mountain View or New York
OnsiteFull Time
Intuit
IntuitNASDAQ: INTU: Provides financial software for accounting, tax, and personal finance.
8+ YOE8+ years building and deploying production-scale AI/ML solutions; advanced degree in CS or related field preferred; deep expertise training LLMs and multimodal models; experience with cloud platforms and AI tooling.
Claude Code, Cursor, Codex, AWS, GCP, Azure
1w
Save
Mark Applied
Hide
Staff AI Engineer, Enterprise Applied AI
Palo Alto or Irvine or Normal or Plymouth or Atlanta
$173k-$258k/yr OnsiteFull Time
Rivian
RivianNASDAQ: RIVN: Designs and manufactures electric vehicles and charging networks.
8+ YOE8+ years software engineering experience, BS/MS/PhD or equivalent, deep AI/ML expertise with language models, retrieval/embeddings, enterprise production deployments, security/compliance familiarity, strong communication.
2mo
Save
Mark Applied
Hide
Sr. AI Engineer, Special Programs
Palo Alto or Washington
HybridFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOE5+ years software engineering; strong Python; Bachelor's degree or equivalent; experience with AI/ML and large language models.
Python, APIs, AWS, GCP, Azure, Docker, Kubernetes
1mo
Save
Mark Applied
Hide
Lead AI Solutions Systems Engineer
Herndon or Sunnyvale or Grand Prairie or Atlanta or Mobile
HybridFull Time
Airbus
AirbusEuronext Paris: AIR: Designs and manufactures commercial and military aircraft and space systems.
10+ YOEMaster's in AI/CS or quantitative field, 10+ years in advanced analytics/AI with production-grade architecture and at least one generative/agentic AI system deployed; strong ML, data modeling, microservices, DevOps and cloud experience; authorized to work in the US.
SciKit, MLlib, PySpark, PyTorch, TensorFlow, Python, Langgraph, ADK (Google Enterprise Agent Platform (formerly Vertex AI)), Vertex AI, RAG, Streamlit, React, PostgreSQL, Firebase, BigQuery, GCP Cloud Run, Node, Typescript, Gemini CLI, Claude Code, Codex, Palantir Foundry, Skywise, Vertex AI Studio, Spacy, FlairNLP, Neo4j, Spanner, GKG, Github Copilot, Docker, Kubernetes, Gemini Enterprise, LLM Wiki
3w
Save
Mark Applied
Hide
Video AI Engineer
San Jose, California, United States
$138k-$275k/yr HybridFull Time
Zoom
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
PhD in EE/CS/Applied Math or Masters+3yrs or 7yrs equivalent; experience in image/video processing, neural rendering/generative/diffusion models, multithreaded programming; proficiency in C/C++, Objective-C, and Python.
C, C++, Objective-C, Python, BrightHire
1mo
Save
Mark Applied
Hide
Principal AI/ML Engineer
Sunnyvale or Austin or Detroit or Warren or Milford or Mountain View or United States
$296k-$424k/yr HybridFull Time
General Motors
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
MS or PhD in CS/Robotics/ML or related; experience leading technical teams delivering production ML systems; deep expertise in robotics planning/control, imitation or reinforcement learning, generative models, and large-scale ML; strong software engineering skills in Python and C++.
Python, C++
1mo
Save
Mark Applied
Hide
Senior Lead AI Engineer (GenAI Platform, Agentic Infrastructure)
New York or San Francisco or McLean or Cambridge or San Jose or Plano
$209k-$286k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's in CS/AI/EE/CE +6 years or Master's +4 years; 6+ years programming with Python/Go/Scala/Java; experience deploying scalable AI on cloud; LLM, inference, similarity search, VectorDBs, guardrails, model evaluation, and optimization experience; leadership and research literacy.
Python, Go, Scala, Java, C++, C#, Golang, AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, AWS, Google Cloud, Azure
1mo
Save
Mark Applied
Hide
Sr. Staff AI Engineer - On-Prem AI Infrastructure & Agentic Systems
San Jose, California, United States
$140k-$165k/yr OnsiteFull Time
SK hynix Memory Solutions America
SK hynix Memory Solutions AmericaKorea Exchange: 000660: Develops semiconductor controllers and firmware for enterprise data storage.
2+ YOE2+ years in AI/ML engineering with on-prem/private-cloud deployment, experience building agentic AI and RAG pipelines, model fine-tuning (LoRA/QLoRA), Python/Linux/Docker/Kubernetes proficiency, and familiarity with vector DBs and AI serving frameworks.
vLLM, TGI, Triton, Milvus, Qdrant, FAISS, Kubernetes, Helm, Docker, LangGraph, AutoGen, LoRA, QLoRA, Model Control Protocols (MCP), Python, Linux, Pinecone, Ollama, ONNX, TensorRT, GGUF, BabyAGI, LangSmith, Weights & Biases, Prometheus, Grafana, CI/CD