NVIDIA
Posted 2mo ago

Senior AI Software Engineer, Kernel Libraries

NVIDIA
Santa Clara or United States
$184k-$288k/yrRemoteFull Time
Responsibilities
  • developing kernels
  • designing abstractions
  • building runtimes
Requirements
  • Masters in CS/EE or equivalent experience
  • 6+ years in ML/DL systems
  • Strong Python and C/C++
  • Experience with deep learning frameworks
  • Inference engines
  • Runtimes, and GPU kernel development
Technical tools mentioned
PythonC/C++PyTorchJAXTensorFlowONNXvLLMSGLangMLCFlashInferFlash AttentionApache TVMMLIRCUDA C/C++cuTileTriton

Job description

We're looking for outstanding AI systems engineers to develop groundbreaking technologies in the inference systems software stack! We build innovative AI systems software to accelerate for AI inference. As a member of the team, you'll develop libraries, code generators, and GPU kernel technologies for NVIDIA's hardware architecture. This means designing and building things like new abstractions, efficient attention kernel implementations, new LLM inference runtimes components, and kernel code generators to accelerate large language models, agents, and other high-impact AI workloads.

What you'll be doing:

  • Innovating and developing new AI systems technologies for efficient inference

  • Designing, implementing, and optimizing kernels for high impact AI workloads

  • Designing and implementing extensible abstractions for LLM serving engines

  • Building efficient just-in-time domain specific compilers and runtimes

  • Collaborating closely with other engineers at NVIDIA across deep learning frameworks, libraries, kernels, and GPU arch teams

  • Contributing to open source communities like FlashInfer, vLLM, and SGLang

What we need to see:

  • Masters degree in Computer Science, Electrical Engineering, or related field (or equivalent experience); PhD are preferred

  • 6+ years (academic/ industry) experience with ML/DL systems development preferable

  • Strong experience in developing or using deep learning frameworks (e.g. PyTorch, JAX, TensorFlow, ONNX, etc) and ideally inference engines and runtimes such as vLLM, SGLang, and MLC.

  • Strong Python and C/C++ programming skills

Ways to stand out from the crowd:

  • Background in domain specific compiler and library solutions for LLM inference and training (e.g. FlashInfer, Flash Attention)

  • Expertise in inference engines like vLLM and SGLang

  • Expertise in machine learning compilers (e.g. Apache TVM, MLIR)

  • Strong experience in GPU kernel development and performance optimizations (especially using CUDA C/C++, cuTile, Triton, or similar)

  • Open source project ownership or contributions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until June 6, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

About NVIDIA

Designs graphics processing units and artificial intelligence hardware.

Similar jobs

AI Software Engineer roles near Santa Clara, California
1w
Save
Mark Applied
Hide
AI Software Engineer - HP IQ
San Francisco, California, United States
$127k-$175k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Produces personal computers, printers, and related digital imaging products.
1+ YOEMaster's degree in a relevant field, 1+ years of backend software development, proficiency in Python, Java, and C++, and experience with LLM integration, databases, and data streaming.
Python, Java, C++, SQL, NoSQL, PyTorch, TensorFlow Lite, ONNX Runtime, Core ML
1w
Save
Mark Applied
Hide
AI Software Engineer - HP IQ
San Francisco, California, United States
$127k-$175k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Manufactures personal computers, printers, and 3D printing hardware.
1+ YOEMaster's degree in a relevant field, 1+ year backend software development experience, Python, Java and C++ proficiency, LLM integration knowledge, and SQL/NoSQL database familiarity.
Python, Java, C++, SQL, NoSQL, PyTorch, TensorFlow Lite, ONNX Runtime, Core ML
2w
Save
Mark Applied
Hide
Senior AI Software Engineer
San Jose, California, United States
$256k-$299k/yr HybridFull Time
Zoom
ZoomNasdaq: ZM: Provides a cloud-based platform for video, voice, and collaboration.
8+ YOEBachelor's degree and 8 years experience in software engineering/AI including C++, Python, deep learning, ASR, TensorFlow/PyTorch, Git and CI/CD; strong algorithm and system design skills.
C++, Python, TensorFlow, PyTorch, Git, CI / CD, ASR, BrightHire
3w
Save
Mark Applied
Hide
Senior AI Software Engineer, AI Controls
Fremont or Salem or Pittsburgh
$187k-$292k/yr HybridFull Time
Agility Robotics
Agility Robotics: Develops bipedal humanoid robots for industrial warehouse automation.
4+ YOE4+ years developing and deploying RL policies for robotics; strong Python and PyTorch experience; experience with perception-in-loop control, sim-to-real, reward design, and deploying policies on real robots.
Python, PyTorch, Mujoco-Warp, Isaac
1mo
Save
Mark Applied
Hide
Senior Software Engineer AI
Pleasanton or New York City or Los Angeles
$145k-$182k/yr HybridFull Time
BlackLine
BlackLineNASDAQ: BL: Provides cloud-based financial close and accounting automation software.
3+ YOE3+ years programming experience (Python/Java/Scala), expertise with ML frameworks, production ML/LLM pipelines, cloud (GCP/AWS/Azure), CI/CD, observability, and governance for AI systems.
PySpark, FiveTran, Plaid, TensorFlow, PyTorch, scikit-learn, Airflow, Kubeflow, Vertex AI, MLflow, W&B, LangChain, LangGraph, ADK, Prometheus, Grafana, Newrelic, Docker, Kubernetes, Python, Java, Scala, Bash, GCP, AWS, Azure
1mo
Save
Mark Applied
Hide
AI software Engineer Project Intern (Transaction Platform) - 2026 Start (BS/MS)
San Jose, California, United States
OnsiteInternship
TikTok
TikTok: Global short-form video hosting and social media platform.
Currently pursuing BS/MS in Computer Science or related field; strong programming and software engineering fundamentals; hands-on experience with AI systems (IR, LLM apps, agents, RAG); familiarity with backend/frontend development; strong problem-solving and communication.
Java, React, GitLab, Meego, CI/CD
1mo
Save
Mark Applied
Hide
Staff AI Software Engineer, Siri Core Modeling
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Work on Siri core modeling to build a privacy-first, context-aware AI assistant across Apple platforms with strong user-centered design focus.
iOS, iPadOS, macOS, watchOS, visionOS
1mo
Save
Mark Applied
Hide
AI Software Engineer, Full-Stack (SF)
San Francisco, California, United States
$145k-$205k/yr HybridFull Time
Loop
Loop: AI-native data platform for supply chain and logistics.
2+ YOEWork in SF office 4+ days/week, 2 years software engineering experience, experience with TypeScript, GraphQL, Relay, React, Node.js, AWS; early-stage product experience and strong communication.
TypeScript, GraphQL, Relay, React, Node.js, AWS, CDK, Fargate, ECS, Prisma, NestJS, PostgreSQL, Kafka, Redis, Elasticsearch, Ant Design, Vite, Cursor, Codex, Claude, LLMs