195 performance modeling engineer jobs at 78 companies in South San Francisco, CA

3w
Save
Mark Applied
Hide
AI Performance Modeling Engineer
Burlingame, California, United States
$180k-$225k/yr HybridFull Time
Quadric
Quadric: Private semiconductor IP licensor providing programmable AI processors for on-device inference to chip designers.
Strong Python and quantitative modeling skills; computer architecture knowledge; technical writing ability; AI inference or performance modeling expertise; BS, MS, PhD, or equivalent practical experience.
Python, C++, CUDA, Triton, gem5, Timeloop, MAESTRO, Accel-Sim
2mo
Save
Mark Applied
Hide
System Performance Modeling Engineer
Santa Clara, California, United States
$166k-$284k/yr HybridFull Time
Pensando Systems
Pensando SystemsNASDAQ: AMD: Pensando Systems was a private distributed-services platform serving cloud, enterprise, and edge infrastructure customers.
Develop high-performance ASIC/SoC models using C/C++, SystemC, and ARM Fast Models; strong debugging, scripting, and cross-functional collaboration; BSEE required, MSEE preferred.
C++, C, SystemC, ARM Fast Models, OMNeT, Boost, STL, Python, SystemVerilog
1mo
Save
Mark Applied
Hide
Graphics (GPU) Performance Modeling Engineer
Santa Clara, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designing and manufacturing consumer electronics, software, and digital services.
Perform advanced GPU performance modeling and exploration for next-generation GPU architectures; solve difficult technical challenges and collaborate with cross-disciplinary chip design teams.
3mo
Save
Mark Applied
Hide
CPU Performance Modeling Engineer (RISC-V)
Cupertino or Austin
$142k-$213k/yr OnsiteFull Time
Qualcomm Technologies, Inc.
Qualcomm Technologies, Inc.: Developing semiconductor, wireless, connectivity, automotive, AI, and computing technologies for device and enterprise customers.
0+ YOEBachelor's/MS/PhD in EE/CE/CS or related field with relevant hardware/software engineering experience; strong CPU microarchitecture knowledge; proficiency in C/C++ and scripting (Perl/Python); performance modeling experience.
C, C++, Perl, Python, RISC-V, RTL
1mo
Save
Mark Applied
Hide
Performance Modeling Architect – AI Systems
Durham or Santa Clara or Boston or Austin
$200k-$500k/yr OnsiteFull Time
Velaura AI
Velaura AI: Private AI compute infrastructure developing ultra-low-power silicon and software for data centers and Physical AI.
Experienced in computer/system architecture and performance modeling; building simulation or analytical models; strong programming skills (Python, C++); knowledge of CPUs/GPUs/accelerators; ability to analyze system bottlenecks.
Python, C++
2mo
Save
Mark Applied
Hide
Sr. Performance Modeling Architect
Santa Clara or Austin
$100k-$500k/yr HybridFull Time
Tenstorrent
Tenstorrent: Builds computers for artificial intelligence.
1+ YOEPhD preferred (MS considered) in a related field; 1+ years industry or research experience in CPU/core or microarchitecture; experience with Gem5, SST, SimpleScalar; programming in C++ and Python; strong processor subsystem knowledge.
Gem5, SST, SimpleScalar, C++, Python, RISC-V
3mo
Save
Mark Applied
Hide
Senior Performance Modeling Architect, CPU Fabric and LLC
Santa Clara, California, United States
$152k-$242k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOEDesign and maintain cycle-accurate performance models for CPU caches and interconnects; analyze bottlenecks; model coherency protocols; run benchmarks; collaborate with teams.
C++, SystemC, Python
1mo
Save
Mark Applied
Hide
Datacenter Compute SoC Performance Modeling Architect
San Jose or San Diego or Portland or Austin
$211k-$356k/yr OnsiteFull Time
MediaTek
MediaTekTaiwan Stock Exchange: 2454: Global fabless semiconductor providing system-on-chip solutions.
5+ YOE5+ years in CPU/SoC performance modeling, expert C++ and scripting, experience with datacenter workloads, strong computer-architecture knowledge and communication skills.
C++, SystemC, Python, Perl, gem5, Sniper, PCIe, CXL, DDR, LPDDR5, LPDDR6, HBM, AMBA CHI, AXI
1w
Save
Mark Applied
Hide
Inference Performance Engineer
San Francisco, California, United States
HybridFull Time
Adaption
Adaption: AI building adaptive intelligence that continually learns for industries, languages, and specialized workflows.
5+ YOE5+ years in ML systems, inference infrastructure, or performance engineering; model-serving expertise; Python and systems-language proficiency; and GPU performance experience with measurable cost or latency improvements.
vLLM, SGLang, TensorRT-LLM, Python, C++, Rust, CUDA, NCCL
2d
Save
Mark Applied
Hide
Architecture Modeling Engineer
Paris or Mountain View
OnsiteFull Time
Arago
Arago: Private photonic AI hardware building light-powered accelerators and software for energy-efficient AI workloads.
Requires strong mathematics, computer architecture, C/C++ and Python skills, performance-critical workload optimization, AI/ML workload knowledge, and proficient English.
C, C++, Python
3mo
Save
Mark Applied
Hide
Senior Performance Modeling Architect, CPU Fabric and LLC
Santa Clara, California, United States
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOEMaster’s or Ph.D. in computer, electrical, or computer science engineering, or equivalent experience, plus 5+ years in architecture. Requires C++/SystemC, Python, CPU microarchitecture, cache coherency, NoC, and memory systems expertise.
C++, SystemC, Python, SPEC, MLPerf, ISO 26262, CXL (Compute Express Link), HBM (High Bandwidth Memory)
3mo
Save
Mark Applied
Hide
Senior Performance Modeling Architect, CPU Fabric and LLC
Santa Clara, California, United States
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOEMaster's or PhD in CE/EE/CS (or equivalent) with 5+ years experience; deep knowledge of CPU microarchitecture, cache coherency, NoC topologies; experience with C++/SystemC and Python; benchmarking and performance analysis experience.
C++, SystemC, Python, SPEC, MLPerf
3mo
Save
Mark Applied
Hide
Founding Engineer - ML Performance
San Francisco, California, United States
$250k-$395k/yr RemoteFull Time
uRun
uRun: AI infrastructure helping model labs, builders, and research teams run real-time interactive video and stateful inference.
Hands-on CUDA, GPU optimization, and large-scale model inference experience; strong systems and performance engineering skills.
CUDA, GPU, NCCL, PyTorch, Triton, TensorRT, CUDA kernels
2mo
Save
Mark Applied
Hide
Performance Verification Engineer
Mountain View, California, United States
$200k-$350k/yr OnsiteFull Time
DensityAI
DensityAI: Semiconductor startup building full-stack AI accelerators for frontier-scale large language model inference.
8+ YOEMaster's degree plus ~8 years experience in performance validation/verification of complex SoCs, performance modeling, debugging and correlation, tool-flow ownership, and collaboration with RTL designers and architects.
3mo
Save
Mark Applied
Hide
Product Performance Engineer
San Jose, California, United States
$120k-$300k/yr OnsiteFull Time
Hark
Hark: Private AI building multimodal personal-intelligence systems and native hardware for consumers.
5+ YOE5+ years analyzing workload behavior and modeling system performance; BS in CS/EE; profiling/benchmarking/tracing tools; Python and systems language experience; translate analytics into architectural recommendations.
Python, C++, Rust, Profiling Tools, Benchmarking Tools, Tracing Tools
2w
Save
Mark Applied
Hide
Lead SoC Performance & Power Engineer
San Jose, California, United States
$204k-$250k/yr OnsiteFull Time
TYLsemi
TYLsemi: Built with chiplets, made for AI infrastructure
6+ YOEBS/MS in Electrical Engineering or related field and 6+ years in computer architecture, performance modeling, or microarchitecture configuration. Requires ESL tools, protocol, Python, and SystemC experience.
Platform Architect, Virtualizer, PCIe, UCIe, Ethernet, NoC, AMBA, AXI, CHI, Python, SystemC, PowerPoint
2w
Save
Mark Applied
Hide
Senior Staff Performance Codesign Engineer, TPU
Sunnyvale, California, United States
$240k-$333k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOG, GOOGL: Global technology specializing in internet-related services and products.
12+ YOEBachelor's degree or equivalent practical experience, 12 years in computer or chip architecture or hardware-software co-design, and experience with performance modeling, simulation, or system analysis.
PyTorch, TensorFlow
2mo
Save
Mark Applied
Hide
Staff Engineer, GPU Graphic Core Performance Verification
San Jose, California, United States
$168k-$252k/yr OnsiteFull Time
Samsung Electronics
Samsung ElectronicsKorea Exchange: 005930: Global leader in technology, semiconductors, and consumer electronics.
6+ YOE6+ years (BSc) or equivalent advanced degree experience in GPU performance verification, strong GPU architecture knowledge, C++/Python proficiency, experience with performance tests/profiling/automation, and ability to analyze/correlate performance across models, emulation, and silicon.
C++, Python, Linux, OpenGL, Vulkan, OpenCL, SystemC, Sparta, RTL, emulation
4d
Save
Mark Applied
Hide
Principal Engineer, Architecture & Performance Research Engineer for Data Center and Agentic AI CPU
San Jose, California, United States
$219k-$351k/yr OnsiteFull Time
Samsung Semiconductor
Samsung SemiconductorKorea Exchange (KRX): 005930: Global leader in semiconductor solutions including memory, system LSI, and foundry services.
8+ YOEMaster’s degree with 18+ years or PhD with 15+ years relevant experience; 8+ years in CPU architecture or performance engineering; expertise in CPU architectures, modeling, C/C++, Python, RTL, and silicon validation.
RISC-V, ARM, x86, gem5, C, C++, Python, RTL, RVV, SIMD
2mo
Save
Mark Applied
Hide
Performance & Capacity Engineering - Capacity Planning Optimization
Bellevue or Menlo Park or Boston or New York
$184k-$257k/yr OnsiteFull Time
Meta
MetaNASDAQ: META: Builds technologies that help people connect, find communities, and grow businesses.
8+ YOEBachelor's in CS/CE or equivalent; 8+ years experience in performance/software/optimization; expertise designing optimization models, LP solvers (Gurobi/Xpress), distributed systems, infrastructure operations, and coding (Python, R, Java, C/C++, PHP).
Python, R, Java, C, C++, PHP, Xpress, Gurobi