79 performance modeling engineer jobs at 45 companies in San Rafael, CA

2w
Save
Mark Applied
Hide
AI Performance Modeling Engineer
Burlingame, California, United States
$180k-$225k/yr HybridFull Time
Quadric
Quadric: Designing licensable processor IP for on-device AI inference.
Strong Python and quantitative modeling skills; computer architecture knowledge; technical writing ability; AI inference or performance modeling expertise; BS, MS, PhD, or equivalent practical experience.
Python, C++, CUDA, Triton, gem5, Timeloop, MAESTRO, Accel-Sim
2w
Save
Mark Applied
Hide
SystemC Modeling Developer (Remote)
San Francisco, California, United States
RemoteFull Time
UST
UST: Global provider of digital transformation and IT services.
5+ YOEBachelor's or master's degree in a related field and 5+ years in SystemC/TLM, C/C++, computer architecture, pre-silicon simulation, virtual platforms, performance modeling, and debugging.
SystemC, TLM2.0, C, C++, ARM, BIOS, Synopsys Virtualizer, Cadence VSP, QEMU
3mo
Save
Mark Applied
Hide
Founding Engineer - ML Performance
San Francisco, California, United States
$250k-$395k/yr RemoteFull Time
uRun
uRun: Infrastructure cloud for interactive, stateful AI inference.
Hands-on CUDA, GPU optimization, and large-scale model inference experience; strong systems and performance engineering skills.
CUDA, GPU, NCCL, PyTorch, Triton, TensorRT, CUDA kernels
2mo
Save
Mark Applied
Hide
Performance Verification Engineer
Mountain View, California, United States
$200k-$350k/yr OnsiteFull Time
DensityAI
DensityAI: Designing custom AI hardware accelerators for large language models.
8+ YOEMaster's degree plus ~8 years experience in performance validation/verification of complex SoCs, performance modeling, debugging and correlation, tool-flow ownership, and collaboration with RTL designers and architects.
3w
Save
Mark Applied
Hide
PV Performance Engineer
San Francisco or New York City or Denver or Austin or Calgary or Toronto
$189k-$205k/yr RemoteFull Time
Intersect
IntersectNASDAQ: GOOGL: Develops-located data centers and renewable energy infrastructure.
4+ YOEBachelor's or Master's in engineering,4+ years PV/utility-scale solar experience,ASTM E2848 capacity testing expertise,advanced PV energy modeling,DC collection system knowledge,strong analysis and communication.
ASTM E2848
1w
Save
Mark Applied
Hide
Senior Staff Performance Codesign Engineer, TPU
Sunnyvale, California, United States
$240k-$333k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
12+ YOEBachelor's degree or equivalent practical experience, 12 years in computer or chip architecture or hardware-software co-design, and experience with performance modeling, simulation, or system analysis.
PyTorch, TensorFlow
4w
Save
Mark Applied
Hide
Principal Performance Architect
Mountain View or Austin or Raleigh or Hillsboro or Redmond
$143k-$275k/yr HybridFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
3+ YOEAdvanced degree in EE/CE/CS or equivalent experience, 3+ years technical engineering experience (depending on degree), experience with performance modeling and SoC architecture, strong Python/C/C++ skills, ability to pass security and export-control screening.
Python, C, C++
2mo
Save
Mark Applied
Hide
Performance & Capacity Engineering - Capacity Planning Optimization
Bellevue or Menlo Park or Boston or New York
$184k-$257k/yr OnsiteFull Time
Meta
MetaNASDAQ: META: Develops social networking platforms and virtual reality technologies.
8+ YOEBachelor's in CS/CE or equivalent; 8+ years experience in performance/software/optimization; expertise designing optimization models, LP solvers (Gurobi/Xpress), distributed systems, infrastructure operations, and coding (Python, R, Java, C/C++, PHP).
Python, R, Java, C, C++, PHP, Xpress, Gurobi
3d
Save
Mark Applied
Hide
Senior Engineer, ATS Performance Testing and Analysis
Danville or San Ramon
$122k-$194k/yr OnsiteFull Time
PG&E
PG&ENYSE: PCG: Provides natural gas and electric service in California.
5+ YOEBachelor's degree in engineering and 5 years of related experience required. Desired expertise includes thermal analysis, equipment testing, energy modeling, LabVIEW, Python, MATLAB, and a professional engineer license.
EnergyPro, LabVIEW, Python, MATLAB
2mo
Save
Mark Applied
Hide
Software Engineer, ML Performance Optimization
Foster City, California, United States
$192k-$257k/yr OnsiteFull Time
Zoox
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
4+ YOE4+ years total exp; 2+ years in large-scale model training or inference; PyTorch; GPU-accelerated inference; profiling tools; Python or C++.
PyTorch, TensorRT, NVIDIA Nsight, Python, C++
1mo
Save
Mark Applied
Hide
Design Verification Engineer -Performance Modelling
Mountain View, California, United States
$45k-$121k/yr OnsiteFull Time
Wipro
WiproNYSE: WIT: Global technology services and consulting for digital transformation.
5+ YOEBachelor's degree and 5+ years design verification experience with SystemVerilog/UVM, scripting (Python/Perl/Shell), simulator/debug tool experience, strong debugging and KPI-driven performance validation skills.
SystemVerilog, UVM, Python, Perl, Shell, VCS, Xcelium, Questa, Verdi, HVL
2mo
Save
Mark Applied
Hide
System Performance & AI Architect
Los Altos, California, United States
OnsiteFull Time
Majestic Labs
Majestic Labs: Developing memory-first AI server platforms for data centers.
10+ YOEPhD in EE/CS or related, 10+ years in performance modeling and system architecture for AI/hyperscale or semiconductor organizations; expertise in ISAs, cache coherence, memory systems, interconnects; proficiency in Modern C++ and Python; experience with NSight, ROCm, VTune.
Modern C++, Python, NSight, ROCm, VTune, HBM4, CXL, Network-on-Chip (NoC), GPUs, NPUs, SoCs, Instruction Set Architectures (ISA)
2mo
Save
Mark Applied
Hide
Distinguished Technologist - AI Model Performance Architect
Spring or Palo Alto
$190k-$274k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Manufacturer of personal computers, printers, and imaging devices.
12+ YOE12+ years experience in electrical/hardware design or related field; degree in electrical engineering or related (preferred); expertise in memory and AI model design, HW/SW co-design, workload optimization, FPGA, MATLAB, and systems design.
MATLAB, Oscilloscope, Field-Programmable Gate Array (FPGA), Schematic Capture
2mo
Save
Mark Applied
Hide
Distinguished Technologist - AI Model Performance Architect
Spring or Palo Alto or Houston
$190k-$274k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Produces personal computers, printers, and related digital imaging products.
12+ YOERecommended 4-year or graduate degree in electrical engineering or related field; typically 12+ years' experience in electrical/hardware design or related areas; expertise in memory and AI model design, HW/SW co-design, and systems optimization.
MATLAB, Oscilloscope, Field-Programmable Gate Array (FPGA), Schematic Capture
2mo
Save
Mark Applied
Hide
Distinguished Technologist - AI Model Performance Architect
Spring or Palo Alto or Houston
$190k-$274k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Manufactures personal computers, printers, and 3D printing hardware.
12+ YOE12+ years' experience in electrical/hardware design or related field; four-year or graduate degree in electrical engineering or related discipline preferred; expertise in memory and AI model design, hardware architecture, FPGA, MATLAB, oscilloscope, PCB and schematic capture.
MATLAB, Field-Programmable Gate Array (FPGA), Oscilloscope, Schematic Capture
3mo
Save
Mark Applied
Hide
AI Engineer, Model Quality and Performance
Sunnyvale, California, United States
OnsiteFull Time
Cerebras Systems
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
Experience building AI agents, strong math/statistics background, familiarity with Docker, Git, CI and automation stacks, plus tooling design and model evaluation/benchmarking experience.
Claude, Docker, Git, CI, EvalScope, lm-eval-harness
3mo
Save
Mark Applied
Hide
ML Inference Engineer
San Francisco, California, United States
OnsiteFull Time
Reactor
Reactor: Building infrastructure for real-time generative world models.
Strong expertise in ML engineering, PyTorch, CUDA, and high-performance inference; experience with diffusion models and low-latency systems.
PyTorch, TensorRT, TransformerEngine, Nsight, ONNX Runtime, CUDA
1mo
Save
Mark Applied
Hide
Staff ML Engineer, Generative Model Performance & Efficiency
Mountain View or New York City
$251k-$310k/yr OnsiteFull Time
Waymo
Waymo: Autonomous driving technology for ride-hailing and logistics.
5+ YOEMS/PhD in CS/ML/Robotics, 5+ years deep learning experience (Transformers, Diffusion, MoEs), proficiency with JAX/Flax and ML tooling, experience with model compression and profiling, strong Python and C++ skills.
JAX, Flax, TensorFlow, PyTorch, XLA, xprof, Perfetto, NVIDIA Nsight, Python, C++, Gemax, XManager, TPUs, GPUs
3w
Save
Mark Applied
Hide
CAE Engineer
Palo Alto, California, United States
OnsiteFull Time
Mind Robotics
Mind Robotics: A robotics developing physical AI systems for industrial deployment, focusing on real-world robot performance and scalability.
5+ YOE5+ years performing structural FEA on electromechanical systems; strong fundamentals in stress/strain, dynamics, and vibration; proficiency with FEA, MBD, and scripting; experience correlating models with test data.
Abaqus, Ansys, OptiStruct, HyperWorks, Nastran, LS-DYNA, Adams, Simpack, Python, MATLAB, FE-Safe, nCode DesignLife
1mo
Save
Mark Applied
Hide
Principal Machine Learning Engineer
Mountain View, California, United States
$278k-$417k/yr OnsiteFull Time
Unity
UnityNYSE: U: Provides software for creating real-time 3D interactive content.
8+ YOE4+ Mgmt8+ years software/ML engineering with 4+ years on-device/edge inference; production deployment of transformer/diffusion models; WebGPU/WGSL and GPU API performance tuning; proficiency with TypeScript/JavaScript and Python; leadership experience.
WebGPU, WebNN, WGSL, Metal, Vulkan, SPIR-V, D3D12, CUDA, Chrome, Dawn, PIX, Instruments, Snapdragon Profiler, Nsight, RenderDoc, ONNX Runtime Web, ONNX Runtime, Transformers.js, WebLLM, TensorFlow.js, CoreML, TFLite, ExecuTorch, TypeScript, JavaScript, Python, MLIR, TVM, IREE, XLA, wgpu