80 optimization modeling jobs at 26 companies in Pacific Grove, CA
2w
Save
Mark Applied
Hide
2w
Multimodal Model Training and Inference Optimization Engineer
San Jose, California, United States
OnsiteFull Time
ByteDance: Developing AI-driven content platforms and mobile applications.
M.S. or PhD in CS/EE/AI, experience optimizing model training and inference, proficiency in Python, C++, CUDA, PyTorch, Megatron, Deepspeed, distributed training, and knowledge of transformers/diffusion models.
Multimodal Model Training and Inference Optimization Engineer
San Jose, California, United States
$156k-$388k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
MS/PhD in CS/EE/AI or related, experience optimizing model training and inference, distributed training, Python/C++/CUDA, PyTorch/Megatron/Deepspeed, knowledge of transformers and diffusion models.
New York City or Seattle or Los Angeles or Los Gatos
$466k-$750k/yrOnsiteFull Time
NetflixNASDAQ: NFLX: Provider of global streaming entertainment and video content.
7+ YOE7+ years software engineering experience with 3+ years on ML infrastructure or model serving; proficiency in Java, Python, or Scala; experience building high‑QPS, low‑latency model serving, feature serving, and model monitoring.
NetflixNASDAQ: NFLX: Global video streaming and media production service.
7+ YOE7+ years software engineering; 3+ years ML infrastructure, model serving, or ML platform experience in ads/real-time decisioning; real-time model serving with sub-20ms latency; proficiency in Java, Python, or Scala; experience with ML serving frameworks and real-time feature pipelines; strong model monitoring and production readiness.
Java, Python, Scala, ML serving frameworks, feature stores, model registries
Sr. Staff Software Development Engineer - Collectives and Network optimization
San Jose, California, United States
$179k-$306k/yrHybridFull Time
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Senior engineer with deep knowledge of network, NIC and GPU architecture, performance optimization and modeling, experience with AI frameworks (PyTorch, JAX, vLLM, SGLang) and ROCm; PhD or master's in CS/EE or related preferred; strong communication and leadership.
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
8+ YOEBachelor’s or Master’s in Electrical/Computer Engineering with 6+ years (Master) or 8+ years (Bachelor); PhD with 3+ years; RTL design, power modeling; PrimePower RTL; scripting (Tcl/Python).
Altera: Manufacturer of field-programmable gate arrays and programmable logic devices.
15+ YOEMS or PhD in CS/CE/EE,15+ years EDA/timing analysis experience,expertise in STA,timing closure/modeling,physical design optimization,and large-scale C++ development.
Hiring.Cafe: An AI-powered job search engine and aggregator.
Experience deploying and optimizing deep learning models in production, multi-GPU inference, profiling/benchmarking model performance, inference optimization techniques, and cloud/distributed systems familiarity.
Software Development Manager, LLM Inference Model Enablement, Neuron SDK
Cupertino, California, United States
$213k-$288k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
7+ YOE3+ MgmtManage engineering team to onboard and optimize LLMs for inference on Trainium; strong background in LLM architectures, model performance optimization, and inference techniques; experience with PyTorch and Neuron stack.
Hark: A building multimodal AI models and next-generation hardware to create natural human-machine interfaces.
3+ YOE3+ years in model compression/distillation/quantization, strong fluency in PyTorch or TensorFlow, experience with PTQ/QAT and int8 conversion, hardware-aware optimization for constrained devices, and familiarity with audio/sequence model architectures.
Gridmatic: AI-powered platform for optimizing energy trading and battery storage.
Advanced degree in EE/power systems, strong power systems modeling and optimization background, experience with power flow, transmission/congestion analysis, large-scale optimization, and Python programming.
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
Master’s or Ph.D. in CS/ML; track record in mid-training of multimodal models; diffusion architectures; Vision-Language Models; scalable data pipelines; optimize inference.
San Francisco or McLean or Cambridge or San Jose or New York
$306k-$381k/yrOnsiteFull Time
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEPhD+4yrs or MS+6yrs in CS/EE/AI/math with deep learning and LLM research experience, track record of publications and large-model training, expertise in training optimization and model deployment.
Rivet Industries: Building integrated task systems for frontline industrial and defense operators.
5+ YOEBS + 5+ years (or MS + 2+ years) in ML research or applied ML engineering; proficiency in Python and C++; experience with PyTorch/TensorFlow, ML pipelines, model deployment, and model optimization for edge/embedded systems.
NioNYSE: NIO: Designs and manufactures smart premium electric vehicles.
Master's degree or above in AI/CS/automation/vehicle engineering/electronic information; deep learning and large-model experience; PyTorch/TensorFlow/ONNX proficiency; model optimization and embedded inference experience; embedded Linux/RTOS/QNX familiarity.
BetterHelp: Provides online therapy and professional mental health counseling services.
3+ YOE3+ years building ML systems; strong Python; experience with NLP, LLMs, PyTorch/TensorFlow, SQL; model evaluation, fine-tuning, inference optimization, and production deployment experience.
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
10+ YOEBS/MS/PhD in aerospace or related field; 4–10+ years modeling full-vehicle rotorcraft/eVTOL aerodynamics depending on degree; expert fixed-wing and rotorcraft aerodynamics, optimization, surrogate modeling, Python, Git, experimental data processing, and strong technical communication.