NVIDIA
Posted 1w ago

Senior Infrastructure Software Engineer, TensorRT Edge-LLM

NVIDIA
Santa Clara or California
$184k-$288k/yrHybridFull Time
Responsibilities
  • building infrastructure
  • maintaining pipelines
  • monitoring systems
Requirements
  • Bachelor's degree or equivalent in computer science or computer engineering
  • 7+ years of experience
  • Python and C/C++ skills
  • CI systems
  • Cloud platforms, SCM, and build systems experience
Technical tools mentioned
PythonCC++CMakeGitLabGitHub ActionsKubernetesDockerJenkinsGitLab CIAWSGCPAzureGitPerforceMakeBazelTensorRTTensorRT-LLMvLLMSGLangUbuntuJetPackQNX

Job description

We are now seeking a Senior Infrastructure Software Engineer for NVIDIA TensorRT Edge-LLM!

NVIDIA's TensorRT Infrastructure group is seeking excellent software engineers to enable the next generation of edge AI. This is an outstanding chance to define the infrastructure/DevOps landscape for an emerging product. The mission is to develop scalable, modular infrastructure that streamlines development, builds, and tests across NVIDIA’s diverse set of platforms, from Drive AGX for autonomous vehicles to Jetson AGX for robotics and edge inference applications. You will work with autonomy to design and implement the best solutions and collaborate with external partners to achieve our goals. Join our technically diverse team of software engineers and infrastructure experts to design the systems that enable NVIDIA to stay ahead of the competition.

What you'll be doing:

  • Building and maintaining infrastructure from first principles needed to deliver TensorRT Edge-LLM

  • Maintaining CI/CD pipelines to automate the build, test, and deployment process and improve build and test bottlenecks

  • Configuring, maintaining, and building upon deployments of industry-standard tools (e.g. CMake, GitLab, GitHub Actions, Kubernetes, Docker, etc.)

  • Developing throughout the software stack, from the user experience and user interfaces down to the cluster layers

  • Monitoring and configuring embedded and desktop CPU and GPU systems to ensure high CI/CD reliability

  • Enable performing scans and handling of security CVEs for infrastructure components

What we need to see:

  • BS or equivalent experience or higher degree in Computer Science or Computer Engineering

  • 7+ years of proven experience

  • Strong programming skills in Python (or similar) and familiarity with modern C/C++ development

  • Experience setting up, maintaining, and automating continuous integration systems (e.g. Jenkins, GitHub Actions, GitLab CI)

  • Experience administering, monitoring, and deploying systems and services on GitHub and cloud platforms (e.g. AWS, GCP, Azure)

  • Fluency in SCM (e.g. Git, Perforce) and build systems (e.g. CMake, Make, Bazel)

Ways to stand out from the crowd:

  • Experience in defining and owning the DevOps strategy (design patterns, reliability and scaling) for a team or organization

  • Deep understanding of test automation infrastructure, framework, and test analysis

  • Familiarity with the development model for popular LLM frameworks and libraries such as TensorRT, TensorRT-LLM, vLLM, or SGLang

  • Experience with mobile/embedded/automotive platforms (e.g. Ubuntu, JetPack, QNX, or similar)

  • Track record of identifying useful new technologies and incorporating them into SW development flows

This is an opportunity to have a wide impact at NVIDIA by improving development velocity for our rapidly growing compute software projects. Are you creative, driven, and autonomous? Do you love a challenge? If so, we want to hear from you!


#LI-Hybrid

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 10, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

About NVIDIA

Designs graphics processing units and artificial intelligence hardware.

Similar jobs

Infrastructure Software Engineer roles near Santa Clara, California
1d
Save
Mark Applied
Hide
Infrastructure Software Engineer (Provisioning) - Apple Services Engineering
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Builds systems, infrastructure, and automation that provision hardware and support Apple services, including iTunes, iCloud, Siri, and Maps.
iTunes, iCloud, Siri, Maps
4d
Save
Mark Applied
Hide
Infrastructure Software Engineer, Fleet & Automation
Houston or New York City or San Francisco or Seattle
OnsiteFull Time
Nscale
Nscale: Vertically integrated AI infrastructure provider for high-performance computing.
5+ YOEBachelor's degree or equivalent experience, 5+ years building large-scale infrastructure applications, and expertise in Python, Linux, networking, distributed systems, and infrastructure tooling.
C, C++, Java, Python, Linux, TCP/IP, BGP, Ansible, Terraform, DCIMs, NetBox, OpenStack, MAAS, Ironic, IPMI, NVIDIA GPUs, InfiniBand, NCCL, SLURM, Prometheus, Grafana, OpenTelemetry, Kubernetes, Docker
1w
Save
Mark Applied
Hide
Member of Technical Staff - SWE Infrastructure
Palo Alto, California, United States
OnsiteFull Time
Ricursive Intelligence
Ricursive Intelligence: Automates semiconductor chip design using recursive artificial intelligence.
5+ YOERequires 5+ years of software engineering experience, with expertise in production systems, CI/CD, regression testing, cloud infrastructure, distributed systems, and automated deployment.
CI/CD
1w
Save
Mark Applied
Hide
Senior Infrastructure Software Engineer, TensorRT Edge-LLM
Santa Clara or United States
$184k-$288k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
7+ YOEBachelor's degree or equivalent, 7+ years of experience, Python and modern C/C++ skills, CI systems, cloud platforms, Git or Perforce, and build systems. Experience with infrastructure, DevOps, and automation preferred.
Python, C, C++, CMake, GitLab, GitHub Actions, Kubernetes, Docker, Jenkins, AWS, GCP, Azure, Git, Perforce, Make, Bazel, TensorRT, TensorRT-LLM, vLLM, SGLang, Ubuntu, JetPack, QNX
2w
Save
Mark Applied
Hide
Staff Infrastructure Software Engineer
Sunnyvale, California, United States
$210k-$314k/yr OnsiteFull Time
Carbon
Carbon: Manufacturer of industrial 3D printers and advanced polymer materials.
7+ YOE7+ years building and operating production cloud infrastructure with deep AWS or GCP expertise, Terraform, Kubernetes, Istio, CI/CD (Jenkins/GitHub Actions), Linux, and a scripting language (Python/Go/Bash).
AWS, GCP, Terraform, Kubernetes, Istio, Jenkins, GitHub Actions, Linux, Python, Go, Bash, Envoy, Bazel
3w
Save
Mark Applied
Hide
Infrastructure Software Engineer
Campbell, California, United States
$180k-$230k/yr RemoteFull Time
Camus Energy
Camus Energy: Software platform for managing renewable energy grid integration.
3+ YOE3+ years software engineering with infrastructure focus; Python 3, Kubernetes, GCP, CI/CD, observability (Prometheus,Grafana); familiarity with SQL; collaborative incident response and reliability practices.
Python 3, Kubernetes, GCP, CI/CD pipelines, Prometheus, Grafana, Bazel, Go, Node.js, SQL
2mo
Save
Mark Applied
Hide
Constellation Software Engineer, Infrastructure
Redwood Shores or Palo Alto
OnsiteFull Time
WindBorne Systems
WindBorne Systems: Operates smart weather balloons to provide global atmospheric data.
Experience building/shipping full-stack applications (Ruby on Rails, Postgres), operating 24/7 uptime systems with on-call responsibility, and designing low-latency data pipelines and robust infrastructure for flight operations and data processing.
Ruby on Rails, Postgres
2mo
Save
Mark Applied
Hide
Staff Software Engineer, Infrastructure
San Francisco, California, United States
$200k-$300k/yr HybridFull Time
F2
F2: AI-driven financial analysis platform for private market investors
7+ YOE7+ years building and scaling cloud infrastructure on AWS/GCP; expertise with containers, IaC (Terraform/Pulumi/CloudFormation), CI/CD, Postgres/Redis, message systems (Temporal/SQS/Kafka), SOC 2 and security practices.
AWS, GCP, ECS/Fargate, Python, Node, Temporal, Terraform, Pulumi, CloudFormation, GitHub Actions, Postgres, RDS, Supabase, Redis, SQS, Kafka, Kubernetes, GitOps, IaC, IAM, VPC, LLM