80 optimization modeling jobs at 26 companies in Pacific Grove, CA

2w
Save
Mark Applied
Hide
Multimodal Model Training and Inference Optimization Engineer
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
M.S. or PhD in CS/EE/AI, experience optimizing model training and inference, proficiency in Python, C++, CUDA, PyTorch, Megatron, Deepspeed, distributed training, and knowledge of transformers/diffusion models.
Python, C++, CUDA, PyTorch, Megatron, Deepspeed
2w
Save
Mark Applied
Hide
Multimodal Model Training and Inference Optimization Engineer
San Jose, California, United States
$156k-$388k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
MS/PhD in CS/EE/AI or related, experience optimizing model training and inference, distributed training, Python/C++/CUDA, PyTorch/Megatron/Deepspeed, knowledge of transformers and diffusion models.
Python, C++, CUDA, PyTorch, Megatron, Deepspeed
3mo
Save
Mark Applied
Hide
SoC Power Analysis and Optimization Engineer
Beaverton or Cupertino or San Diego
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
3+ YOEBachelor's degree and 3+ years in SOC power analysis, optimization, and modeling; ML, Python, Verilog/SystemVerilog; ASIC/SOC design experience.
Python, Verilog, SystemVerilog, Machine Learning, Power modeling, Power analysis
1mo
Save
Mark Applied
Hide
Machine Learning Engineer 5 - Decisioning & Optimization
New York City or Seattle or Los Angeles or Los Gatos
$466k-$750k/yr OnsiteFull Time
Netflix
NetflixNASDAQ: NFLX: Provider of global streaming entertainment and video content.
7+ YOE7+ years software engineering experience with 3+ years on ML infrastructure or model serving; proficiency in Java, Python, or Scala; experience building high‑QPS, low‑latency model serving, feature serving, and model monitoring.
Java, Python, Scala, Chronon, Signal Service, JVM
1mo
Save
Mark Applied
Hide
Machine Learning Engineer 5 - Decisioning & Optimization
New York or Los Angeles or Los Gatos or Seattle
$466k-$750k/yr OnsiteFull Time
Netflix
NetflixNASDAQ: NFLX: Global video streaming and media production service.
7+ YOE7+ years software engineering; 3+ years ML infrastructure, model serving, or ML platform experience in ads/real-time decisioning; real-time model serving with sub-20ms latency; proficiency in Java, Python, or Scala; experience with ML serving frameworks and real-time feature pipelines; strong model monitoring and production readiness.
Java, Python, Scala, ML serving frameworks, feature stores, model registries
1mo
Save
Mark Applied
Hide
Sr. Staff Software Development Engineer - Collectives and Network optimization
San Jose, California, United States
$179k-$306k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Senior engineer with deep knowledge of network, NIC and GPU architecture, performance optimization and modeling, experience with AI frameworks (PyTorch, JAX, vLLM, SGLang) and ROCm; PhD or master's in CS/EE or related preferred; strong communication and leadership.
PyTorch, JAX, vLLM, SGLang, ROCm
1mo
Save
Mark Applied
Hide
Technical Lead, Front-End Power Optimization
San Jose, California, United States
$184k-$264k/yr OnsiteFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
8+ YOEBachelor’s or Master’s in Electrical/Computer Engineering with 6+ years (Master) or 8+ years (Bachelor); PhD with 3+ years; RTL design, power modeling; PrimePower RTL; scripting (Tcl/Python).
PrimePower RTL, RTL design tools, Tcl, Python
2w
Save
Mark Applied
Hide
Senior Principal Engineer (PE) / Subject Matter Expert (SME) – Quartus Timing Analysis & Optimization
San Jose, California, United States
$266k-$392k/yr OnsiteFull Time
Altera
Altera: Manufacturer of field-programmable gate arrays and programmable logic devices.
15+ YOEMS or PhD in CS/CE/EE,15+ years EDA/timing analysis experience,expertise in STA,timing closure/modeling,physical design optimization,and large-scale C++ development.
Quartus, Timing Analyzer, Fitter, Routing, C++
1mo
Save
Mark Applied
Hide
ML Engineer - Inference & Model Deployment
Cupertino, California, United States
$250k-$310k/yr OnsiteFull Time
Hiring.Cafe
Hiring.Cafe: An AI-powered job search engine and aggregator.
Experience deploying and optimizing deep learning models in production, multi-GPU inference, profiling/benchmarking model performance, inference optimization techniques, and cloud/distributed systems familiarity.
vLLM, TensorRT, SGLang, GPU
6d
Save
Mark Applied
Hide
Software Development Manager, LLM Inference Model Enablement, Neuron SDK
Cupertino, California, United States
$213k-$288k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
7+ YOE3+ MgmtManage engineering team to onboard and optimize LLMs for inference on Trainium; strong background in LLM architectures, model performance optimization, and inference techniques; experience with PyTorch and Neuron stack.
PyTorch, AWS Neuron, Neuron compiler, Neuron runtime
1mo
Save
Mark Applied
Hide
Model Distillation Engineer
San Jose, California, United States
$120k-$300k/yr OnsiteFull Time
Hark
Hark: A building multimodal AI models and next-generation hardware to create natural human-machine interfaces.
3+ YOE3+ years in model compression/distillation/quantization, strong fluency in PyTorch or TensorFlow, experience with PTQ/QAT and int8 conversion, hardware-aware optimization for constrained devices, and familiarity with audio/sequence model architectures.
PyTorch, TensorFlow, TFLite, ONNX Runtime, AIMET, Hexagon DSP, NPUs, Ambiq MCUs
3mo
Save
Mark Applied
Hide
3D Game Artist
Hong Kong or San Jose
HybridFull Time
Nex
Nex: Gaming system that turns body movement into interactive play.
3D modeling/rigging experience; Unity proficiency; optimize assets for performance; strong communication.
Unity, Shader Graph, 3D Modeling, Rigging, Particle Systems, In-Engine Implementation
1mo
Save
Mark Applied
Hide
Power Systems Research Scientist
Cupertino, California, United States
$175k-$235k/yr OnsiteFull Time
Gridmatic
Gridmatic: AI-powered platform for optimizing energy trading and battery storage.
Advanced degree in EE/power systems, strong power systems modeling and optimization background, experience with power flow, transmission/congestion analysis, large-scale optimization, and Python programming.
Python, PSS/E, PowerWorld, PSLF, PyTorch, JAX
3mo
Save
Mark Applied
Hide
Sr. Applied Scientist
San Jose or Seattle
$164k-$313k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
Master’s or Ph.D. in CS/ML; track record in mid-training of multimodal models; diffusion architectures; Vision-Language Models; scalable data pipelines; optimize inference.
Python, PyTorch, TensorFlow, Distributed Training, Vision-Language Models
3mo
Save
Mark Applied
Hide
Performance Engineer
San Jose, California, United States
$175k-$275k/yr OnsiteFull Time
Etched
Etched: Designs specialized AI chips optimized for transformer architectures.
Develop performance models for transformer architectures; profile DL workloads; optimize HW/SW co-design.
gem5, CUDA, GPUs, FPGA, CGRA
2w
Save
Mark Applied
Hide
Distinguished Applied Researcher
San Francisco or McLean or Cambridge or San Jose or New York
$306k-$381k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEPhD+4yrs or MS+6yrs in CS/EE/AI/math with deep learning and LLM research experience, track record of publications and large-model training, expertise in training optimization and model deployment.
PyTorch, AWS Ultraclusters, Huggingface, Lightning, VectorDBs, DeepSpeed, NeMo
3w
Save
Mark Applied
Hide
ML Researcher
San Jose, California, United States
$150k-$290k/yr OnsiteFull Time
Rivet Industries
Rivet Industries: Building integrated task systems for frontline industrial and defense operators.
5+ YOEBS + 5+ years (or MS + 2+ years) in ML research or applied ML engineering; proficiency in Python and C++; experience with PyTorch/TensorFlow, ML pipelines, model deployment, and model optimization for edge/embedded systems.
Python, C++, PyTorch, TensorFlow, TensorFlow Lite, CUDA, AWS, GCP, Azure, Docker, Kubernetes
2w
Save
Mark Applied
Hide
Super Sparks-校招-AI大模型应用研究员/工程师
Shanghai or San Jose
OnsiteFull Time
Nio
NioNYSE: NIO: Designs and manufactures smart premium electric vehicles.
Master's degree or above in AI/CS/automation/vehicle engineering/electronic information; deep learning and large-model experience; PyTorch/TensorFlow/ONNX proficiency; model optimization and embedded inference experience; embedded Linux/RTOS/QNX familiarity.
PyTorch, TensorFlow, ONNX, RT-2, OpenVLA, Groot, pi0
1mo
Save
Mark Applied
Hide
Machine Learning Engineer
San Jose or United States
$150k-$180k/yr RemoteFull Time
BetterHelp
BetterHelp: Provides online therapy and professional mental health counseling services.
3+ YOE3+ years building ML systems; strong Python; experience with NLP, LLMs, PyTorch/TensorFlow, SQL; model evaluation, fine-tuning, inference optimization, and production deployment experience.
Python, PyTorch, TensorFlow, SQL
3w
Save
Mark Applied
Hide
Aerodynamics Engineer
San Jose, California, United States
$160k-$220k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
10+ YOEBS/MS/PhD in aerospace or related field; 4–10+ years modeling full-vehicle rotorcraft/eVTOL aerodynamics depending on degree; expert fixed-wing and rotorcraft aerodynamics, optimization, surrogate modeling, Python, Git, experimental data processing, and strong technical communication.
Python, Git, CI/CD, MATLAB/Simulink, Fortran, C++, RCAS, CAMRAD2, Kriging, Neural Networks, High-Performance Computing (HPC)