Pensando SystemsNASDAQ: AMD: Pensando Systems was a private distributed-services platform serving cloud, enterprise, and edge infrastructure customers.
Develop high-performance ASIC/SoC models using C/C++, SystemC, and ARM Fast Models; strong debugging, scripting, and cross-functional collaboration; BSEE required, MSEE preferred.
C++, C, SystemC, ARM Fast Models, OMNeT, Boost, STL, Python, SystemVerilog
Qualcomm Technologies, Inc.: Developing semiconductor, wireless, connectivity, automotive, AI, and computing technologies for device and enterprise customers.
0+ YOEBachelor's/MS/PhD in EE/CE/CS or related field with relevant hardware/software engineering experience; strong CPU microarchitecture knowledge; proficiency in C/C++ and scripting (Perl/Python); performance modeling experience.
Velaura AI: Private AI compute infrastructure developing ultra-low-power silicon and software for data centers and Physical AI.
Experienced in computer/system architecture and performance modeling; building simulation or analytical models; strong programming skills (Python, C++); knowledge of CPUs/GPUs/accelerators; ability to analyze system bottlenecks.
Tenstorrent: Builds computers for artificial intelligence.
1+ YOEPhD preferred (MS considered) in a related field; 1+ years industry or research experience in CPU/core or microarchitecture; experience with Gem5, SST, SimpleScalar; programming in C++ and Python; strong processor subsystem knowledge.
Senior Performance Modeling Architect, CPU Fabric and LLC
Santa Clara, California, United States
$152k-$242k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOEDesign and maintain cycle-accurate performance models for CPU caches and interconnects; analyze bottlenecks; model coherency protocols; run benchmarks; collaborate with teams.
MediaTekTaiwan Stock Exchange: 2454: Global fabless semiconductor providing system-on-chip solutions.
5+ YOE5+ years in CPU/SoC performance modeling, expert C++ and scripting, experience with datacenter workloads, strong computer-architecture knowledge and communication skills.
Adaption: AI building adaptive intelligence that continually learns for industries, languages, and specialized workflows.
5+ YOE5+ years in ML systems, inference infrastructure, or performance engineering; model-serving expertise; Python and systems-language proficiency; and GPU performance experience with measurable cost or latency improvements.
Senior Performance Modeling Architect, CPU Fabric and LLC
Santa Clara, California, United States
$152k-$288k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOEMaster’s or Ph.D. in computer, electrical, or computer science engineering, or equivalent experience, plus 5+ years in architecture. Requires C++/SystemC, Python, CPU microarchitecture, cache coherency, NoC, and memory systems expertise.
Senior Performance Modeling Architect, CPU Fabric and LLC
Santa Clara, California, United States
$152k-$288k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOEMaster's or PhD in CE/EE/CS (or equivalent) with 5+ years experience; deep knowledge of CPU microarchitecture, cache coherency, NoC topologies; experience with C++/SystemC and Python; benchmarking and performance analysis experience.
DensityAI: Semiconductor startup building full-stack AI accelerators for frontier-scale large language model inference.
8+ YOEMaster's degree plus ~8 years experience in performance validation/verification of complex SoCs, performance modeling, debugging and correlation, tool-flow ownership, and collaboration with RTL designers and architects.
Hark: Private AI building multimodal personal-intelligence systems and native hardware for consumers.
5+ YOE5+ years analyzing workload behavior and modeling system performance; BS in CS/EE; profiling/benchmarking/tracing tools; Python and systems language experience; translate analytics into architectural recommendations.
TYLsemi: Built with chiplets, made for AI infrastructure
6+ YOEBS/MS in Electrical Engineering or related field and 6+ years in computer architecture, performance modeling, or microarchitecture configuration. Requires ESL tools, protocol, Python, and SystemC experience.
GoogleNASDAQ: GOOG, GOOGL: Global technology specializing in internet-related services and products.
12+ YOEBachelor's degree or equivalent practical experience, 12 years in computer or chip architecture or hardware-software co-design, and experience with performance modeling, simulation, or system analysis.
Samsung ElectronicsKorea Exchange: 005930: Global leader in technology, semiconductors, and consumer electronics.
6+ YOE6+ years (BSc) or equivalent advanced degree experience in GPU performance verification, strong GPU architecture knowledge, C++/Python proficiency, experience with performance tests/profiling/automation, and ability to analyze/correlate performance across models, emulation, and silicon.
Principal Engineer, Architecture & Performance Research Engineer for Data Center and Agentic AI CPU
San Jose, California, United States
$219k-$351k/yrOnsiteFull Time
Samsung SemiconductorKorea Exchange (KRX): 005930: Global leader in semiconductor solutions including memory, system LSI, and foundry services.
8+ YOEMaster’s degree with 18+ years or PhD with 15+ years relevant experience; 8+ years in CPU architecture or performance engineering; expertise in CPU architectures, modeling, C/C++, Python, RTL, and silicon validation.