93 ai modeling engineer jobs at 36 companies in Newman, CA

1mo
Save
Mark Applied
Hide
Distinguished Engineer - AI
San Jose, California, United States
$266k-$396k/yr OnsiteFull Time
NetApp
NetAppNasdaq: NTAP: Provides intelligent data infrastructure for hybrid cloud environments.
15+ YOEExpert in AI inferencing and distributed systems at scale with 15+ years experience; hands-on with inference engines, model optimization, GPU/TPU orchestration, Kubernetes, RDMA/DPDK; strong architecture, communication, and mentorship skills.
TensorRT, vLLM, ONNX Runtime, Triton, RDMA, DPDK, Kubernetes, GPU, TPU
1w
Save
Mark Applied
Hide
Traffic Modeling System Engineer for Missile Tracking (Space Force)
Scottsdale or Chantilly or Huntsville or San Jose or Grand Forks
$150k-$166k/yr HybridFull Time
General Dynamics Mission Systems
General Dynamics Mission SystemsNYSE: GD: Engineers secure communication and mission-critical defense technology systems.
8+ YOEBachelor's in systems engineering or related field, 8+ years experience (or Master's+6), DoD Top Secret clearance with polygraph eligibility, U.S. citizenship, systems engineering and traffic modeling experience with satellite/ground networks, Agile preferred, cloud/AI/image processing knowledge preferred.
2w
Save
Mark Applied
Hide
Staff GenAI Engineer - Early PDN Modeling
San Jose, California, United States
$170k-$290k/yr OnsiteFull Time
Micron Technology
Micron TechnologyNASDAQ: MU: Manufacturer of semiconductor memory and data storage products.
5+ YOEBachelor's in EE/CE/CS or equivalent,5-7 years experience in power integrity/PDN/circuit or EDA workflows,production Python coding,experience applying ML/AI to engineering data.
Python, EDA
1w
Save
Mark Applied
Hide
Distinguished AI Engineer (Hybrid)
San Jose or Austin or Seattle or New York City
$343k-$433k/yr HybridFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
10+ YOEExtensive AI/ML leadership with deep learning, NLP, LLMs, large-scale model development and production deployment; strong communication and mentorship skills.
1mo
Save
Mark Applied
Hide
Distinguished Engineer - AI (San Jose, CA, US, 95128)
San Jose, California, United States
$266k-$396k/yr OnsiteFull Time
NetApp
NetAppNASDAQ: NTAP: Sells enterprise data storage and cloud management software.
15+ YOE15+ years building low-latency, fault-tolerant distributed systems and AI/ML inference platforms; expertise with inference engines, model optimization, storage for AI, RDMA/DPDK, and Kubernetes-based orchestration.
TensorRT, vLLM, ONNX Runtime, Triton, RDMA, DPDK, Kubernetes
1mo
Save
Mark Applied
Hide
Lead AI Engineer
McLean or San Jose or Richmond or New York City
$197k-$225k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's in CS/AI/EE/CE +4 yrs (or Master's +2 yrs); 4+ yrs programming with Python, Go, Scala, or Java; experience deploying scalable AI; familiarity with LLM inference, vector DBs, guardrails, and model optimization.
AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, Python, Go, Scala, Java, C++, C#, Golang, AWS, Google Cloud, Azure
1mo
Save
Mark Applied
Hide
Video AI Engineer
San Jose, California, United States
$138k-$275k/yr HybridFull Time
Zoom
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
PhD in EE/CS/Applied Math or Masters+3yrs or 7yrs equivalent; experience in image/video processing, neural rendering/generative/diffusion models, multithreaded programming; proficiency in C/C++, Objective-C, and Python.
C, C++, Objective-C, Python, BrightHire
1w
Save
Mark Applied
Hide
Fellow Software Engineer — AI Performance & Reliability
San Jose or Bellevue
$235k-$402k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
PhD or equivalent in AI/ML/CS, strong software engineering, experience profiling and optimizing ML models and AI workloads, proficiency in Python/C++, ML frameworks, customer-facing troubleshooting and performance analysis.
Python, C++, PyTorch, TensorFlow, JAX, ROCm, HIP, CUDA, Triton, XLA, MLIR, NCCL
21h
Save
Mark Applied
Hide
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization
San Jose or Mountain View
$170k-$351k/yr OnsiteFull Time
DiDi Autonomous Driving
DiDi Autonomous Driving: Develops Level 4 autonomous driving technology for shared mobility.
3+ YOEMaster's degree or higher in a technical field; 3+ years in HPC, AI infrastructure, model optimization, or embedded deployment; C++, Python, CUDA, OpenMP, inference engines, GPU architectures, and system profiling expertise.
C++, Python, CUDA, OpenMP, TensorRT, ONNX Runtime, vLLM, SGLang, TensorRT-LLM, NVIDIA Hopper, NVIDIA Thor, TGI, LightLLM, PyTorch, INT8, FP8, AWQ, LLaMA, Qwen, GPT
1mo
Save
Mark Applied
Hide
Sr. Staff AI Engineer - On-Prem AI Infrastructure & Agentic Systems
San Jose, California, United States
$140k-$165k/yr OnsiteFull Time
SK hynix Memory Solutions America
SK hynix Memory Solutions AmericaKorea Exchange: 000660: Develops semiconductor controllers and firmware for enterprise data storage.
2+ YOE2+ years in AI/ML engineering with on-prem/private-cloud deployment, experience building agentic AI and RAG pipelines, model fine-tuning (LoRA/QLoRA), Python/Linux/Docker/Kubernetes proficiency, and familiarity with vector DBs and AI serving frameworks.
vLLM, TGI, Triton, Milvus, Qdrant, FAISS, Kubernetes, Helm, Docker, LangGraph, AutoGen, LoRA, QLoRA, Model Control Protocols (MCP), Python, Linux, Pinecone, Ollama, ONNX, TensorRT, GGUF, BabyAGI, LangSmith, Weights & Biases, Prometheus, Grafana, CI/CD
2mo
Save
Mark Applied
Hide
Lead AI Engineer (Vision model customization, VLM)
New York City or McLean or San Jose or Cambridge
$197k-$246k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Provides credit card, banking, and auto loan services.
4+ YOEBachelor's degree plus 4 years or master's degree plus 2 years developing AI/ML algorithms or technologies, and 4 years programming with Python, Go, Scala, or Java.
AWS Ultraclusters, Hugging Face, VectorDBs, Nemo Guardrails, PyTorch, Python, Go, Scala, Java, AWS, Google Cloud, Azure, C++, C#, Golang
1mo
Save
Mark Applied
Hide
Senior AI / Machine Learning Engineer
San Francisco or New York City or San Jose or Seattle or Austin or Boston
$115k-$200k/yr HybridFull Time
Absentia Labs
Absentia Labs: AI-native toxicology platform accelerating drug safety and discovery.
5+ YOE5+ years ML industry experience, proven production-scale model training, expertise with LLMs/diffusion/GNNs, strong PyTorch skills, distributed training and data-pipeline experience, and solid software engineering practices.
PyTorch, GitHub
3mo
Save
Mark Applied
Hide
Staff AI Research Engineer
San Jose, California, United States
$190k-$240k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
MS or PhD in CS or Computer Engineering with strong emphasis on AI/ML; strong ML frameworks; Transformer and multimodal models expertise; production-ready research experience.
PyTorch, TensorFlow, Transformers, Multimodal models, Diffusion models
2w
Save
Mark Applied
Hide
Staff GenAI Engineer - Early PDN Modeling
San Jose, California, United States
$170k-$290k/yr OnsiteFull Time
Micron Technology
Micron TechnologyNASDAQ: MU: Designs and manufactures semiconductor memory and data storage solutions.
5+ YOEBachelor's in EE/CE/CS,5+ years in power integrity/PDN/circuit/physical/package/EDA workflows,production-quality Python coding,ML/AI applied to engineering workflows,understanding of PDN and IR drop analysis.
Python
2mo
Save
Mark Applied
Hide
Staff Software Engineer - AI
San Jose, California, United States
$184k-$230k/yr HybridFull Time
Cloudera
Cloudera: Provides hybrid data platforms for enterprise analytics and AI.
10+ YOE10+ years building scalable microservices; AI inference and Generative AI; foundation models; vector databases; Kubernetes; strong ownership and communication.
Go, Python, Node.js, Java, C#, Kubernetes, GRPC, SQL, Pinecone, Milvus, HuggingFace, Nvidia AI frameworks
1w
Save
Mark Applied
Hide
Principal AI/ML Research Engineer
United States or California or San Jose or Seattle or Portland or Boston or Chicago
$250k-$289k/yr RemoteFull Time
WEX
WEXNYSE: WEX: Provides global payment processing and business information management services.
12+ YOE12+ years in software or ML engineering, including 5+ years in applied AI/ML research. Requires advanced expertise in model architecture, algorithms, Python, deep learning frameworks, distributed AI, cloud, and ML platforms.
Python, PyTorch, TensorFlow, JAX, C++, Java, Ray, DeepSpeed, Megatron, Spark, AWS, Azure, SageMaker, MLflow, Databricks, LanceDB, Pinecone, Qdrant, Milvus, Transformer, Tabular Transformers, LLMs, RAG, LoRA, PEFT, Diffusion, Reinforcement Learning, RLHF, NeurIPS, ICML, KDD, ACL
2mo
Save
Mark Applied
Hide
AI/ML Engineer - Agentic
San Jose, California, United States
$137k-$277k/yr HybridFull Time
Hewlett Packard Enterprise
Hewlett Packard EnterpriseNYSE: HPE: Providing global edge-to-cloud infrastructure and IT solutions for businesses.
4+ YOESenior IC with 4-7 years experience; design, build, operate production-grade agentic platform; hybrid work model; Bachelor’s degree in CS/Engineering; Master’s preferred.
FastAPI, gRPC, Kafka, Redis, OpenSearch, Kubernetes, Docker, GitHub Actions, Jenkins, ArgoCD, Prometheus
1mo
Save
Mark Applied
Hide
Machine Learning Engineer - AI Compiler Optimization
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Proficient with AI compiler frameworks and GPU/NPU compilation optimization; experience with model import/conversion for PyTorch/TensorFlow and performance tuning for recommendation models.
Triton, MLIR, TVM, PyTorch, TensorFlow
1mo
Save
Mark Applied
Hide
Staff AI Inference and Acceleration Engineer
San Jose, California, United States
$180k-$275k/yr OnsiteFull Time
Figure
Figure: Develops autonomous humanoid robots for commercial and residential tasks.
8+ YOEMS/PhD or equivalent, 8+ years in hardware acceleration/ML systems, expertise in inference runtimes, quantization and pruning, profiling and benchmarking, model-to-hardware mapping, and strong C++/Python skills.
ONNX, TFLite, TVM, MLIR, TensorRT, Torch, SNPE/QNN, JAX, CUDA, ROCm, C++, Python
2mo
Save
Mark Applied
Hide
Advisory Enterprise Solutions Engineer, AI Orchestration
San Jose or Morrisville
$180k-$230k/yr OnsiteFull Time
Lenovo
LenovoHKSE: 992: Manufactures personal computers, mobile devices, and server infrastructure.
5+ YOEBachelor's in CS/AI/Business Analytics (or related), 5+ years in business liaison or external resource integration in AI/technology, deep understanding of enterprise AI (LLMs, multimodal models, agents), strong cross-cultural communication, and enterprise AI solution design experience.