Quadric has created an innovative general purpose neural processing unit (GPNPU) architecture. Quadric's co-optimized software and hardware is targeted to run neural network (NN) inference workloads in a wide variety of edge and endpoint devices, ranging from battery operated smart-sensor systems to high-performance automotive or autonomous vehicle systems. Unlike other NPUs or neural network accelerators in the industry today that can only accelerate a portion of a machine learning graph, the Quadric GPNPU executes both NN graph code and conventional C++ DSP and control code.
Note: Our preference is for this internship to be based out of our Burlingame, California office. Candidates should be based in the Bay Area or able to relocate for the internship period and available to work on site.
Responsibilities:
Kernel Implementation and optimization: Implement onnx operator kernels that are not in SDK yet. Fully utilize Claude Code to optimize the performance.
This job has expired
This job posting is no longer active and is not accepting applications. Explore similar roles below!
Posted 2mo ago
AI Kernel Engineer Intern - Kernel Optimization
Quadric
Burlingame, California, United States
$45-$60/hrOnsiteTemporary, Internship
Responsibilities
- Kernel implementation
- Optimize performance
- Implement kernels
Requirements
- MS student in CS/CE or related fields
- Proficient in C/C++/Python
- Experience in kernel implementation and optimization
- Experience in performance profiling
Technical tools mentioned
C++PythonONNXProfiling Tools
Job description
About Quadric
Designing licensable processor IP for on-device AI inference.
Year founded
2016
Employees
75
Industries
Activities
Organization type
Private
Latest investment
Raised $30.00M Series C (2026) — led by ACCELERATE Fund
Subsidiaries
Headquarters
US