165 gpu jobs at 85 companies in Irvington, NJ

2w
Save
Mark Applied
Hide
GPU Performance Engineer
New York City, New York, United States
$165k-$300k/yr HybridFull Time
Two Sigma
Two Sigma: Systematic investment management and quantitative trading firm.
1+ YOEExpert CUDA and GPU architecture knowledge, C++ and Python skills, BS/MS in STEM, minimum 1 year relevant experience, GPU profiling and optimization experience.
CUDA, Nsight Systems, Nsight Compute, RAPIDS, CUTLASS, cuBLAS, TensorRT, NCCL, MPI, cuDF, C++, Python
1w
Save
Mark Applied
Hide
Senior Inference Engineer, GPU Kernel Optimization
Santa Clara or Austin or New York City or Seattle
$184k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOE6+ years industry experience; strong Python and C++; hands-on GPU profiling (CUPTI, NSYS, NCU); experience with LLM inference frameworks and GPU kernel optimization; advanced degree or equivalent experience.
Python, C++, CUPTI, NSYS, NCU, TRT-LLM, SGLang, vLLM, CUDA, CUTLASS, Triton, PTX, SASS, LLVM, MLIR, ptxas, FlashInfer
3mo
Save
Mark Applied
Hide
Infrastructure Engineer (GPU & Compute)
New York or San Francisco or Seattle
$180k-$200k/yr RemoteFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
NVIDIA DCGM, PXE, IPMI, Redfish, iDRAC, LiveCD, InfiniBand, NVLink, Linux, Python
2w
Save
Mark Applied
Hide
Engineering Manager, GPU Infrastructure
Toronto or San Francisco or New York City or London or Paris or Montreal
HybridFull Time
Cohere
Cohere: Provides enterprise-grade large language models and AI software platforms.
Experience managing engineering teams focused on GPU/ML infrastructure, Kubernetes, IaC, observability, and collaboration with AI researchers; strong communication and mentorship skills.
JAX, PyTorch, TensorFlow, Kubernetes, Prometheus, Grafana, Terraform, ArgoCD
1mo
Save
Mark Applied
Hide
Software Engineer, Compute (GPU)
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Kubernetes, Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, LLM APIs, Claude Code, Cursor
1w
Save
Mark Applied
Hide
Senior Inference Engineer, GPU Kernel Optimization
Santa Clara or Austin or New York City or Seattle
$184k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
6+ YOEMaster's/PhD or equivalent,6+ years industry experience,agentic AI systems experience,strong Python/C++,GPU profiling (CUPTI,NSYS,NCU),LLM inference frameworks,CUDA/CUTLASS/Triton and PTX/SASS familiarity.
Python, C++, CUPTI, NSYS, NCU, TRT-LLM, SGLang, vLLM, CUDA, CUTLASS, Triton, PTX, SASS, LLVM, MLIR, ptxas
17h
Save
Mark Applied
Hide
Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)
San Francisco or Seattle or New York City or United States
$250k-$485k/yr OnsiteFull Time
Perplexity
Perplexity: AI-powered search engine providing conversational answers with citations.
Deep Kubernetes and GPU cluster experience, multi-cloud orchestration, strong distributed systems fundamentals, systems-level coding in Go/Rust/C++, and experience with training and inference workloads.
Kubernetes, kubectl, NVIDIA, CUDA, InfiniBand, RoCE, CoreWeave, AWS, GCP, Go, Rust, C++, vLLM, SGLang, TensorRT-LLM, Slurm, Triton, RDMA, Prometheus, Grafana, Weights & Biases
1mo
Save
Mark Applied
Hide
Lead Software Developer – GPU-accelerated Free Energy Simulation and Machine Learning Methods
Piscataway, New Jersey, United States
$94k/yr HybridFull Time
Rutgers University
Rutgers University: Providing higher education degrees and conducting academic research.
PhD in biology or chemistry with postdoctoral experience using Amber; proficiency in Python, C++, CUDA/GPU programming, molecular simulation (alchemical free energy/quantum/ML), and high-performance computing; strong English communication.
Python, C++, CUDA, Amber, GPU
2mo
Save
Mark Applied
Hide
Failure Analysis Engineering Manager, GPU ASIC and PCBA Debug
Secaucus, New Jersey, United States
$129k-$221k/yr OnsiteFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
3+ MgmtProven people manager with 3+ years management experience; deep GPU ASIC and PCBA debug and failure analysis expertise; hands-on lab experience (oscilloscopes, logic analyzers); proficient in Python and shell; Windows/Linux experience; bachelor’s degree in EE/CE or related.
Python, Shell, Windows, Linux, Oscilloscope, Logic Analyzer
3mo
Save
Mark Applied
Hide
Head of Infrastructure, Stealth Edge AI Co
New York City, New York, United States
HybridFull Time
Montauk Capital
Montauk Capital: Investment firm building and funding climate technology companies.
Lead hardware and infrastructure buildout for edge AI compute; manage GPU infrastructure, data center deployments, and supply chain.
Linux, IPMI, Redfish, PXE, BMC, GPU, GPU scheduling, Automation, Monitoring, OOB management
2mo
Save
Mark Applied
Hide
Member of Research Staff, Reinforcement Learning, Voleon Securities
Berkeley or New York City
$250k-$275k/yr HybridFull Time
The Voleon Group
The Voleon Group: Quantitative investment management firm using machine learning strategies.
PhD-level coursework required; PhD preferred. Strong RL background, mathematical ability, Python production code, GPU computing, and interest in finance.
Python, GPU computing
1d
Save
Mark Applied
Hide
Machine Learning Performance Engineer
New York City, New York, United States
$300k/yr OnsiteFull Time
Jane Street
Jane Street: Proprietary quantitative trading and technology firm.
Experience in low-level systems programming and ML performance optimization, GPU/CUDA expertise, distributed training knowledge, and debugging/optimization tool experience.
CUDA, PTX, SASS, Tensor Cores, CUDA GDB, NSight Systems, NSight Compute, Triton, CUTLASS, CUB, Thrust, cuDNN, cuBLAS, NCCL, MPI, Infiniband, RoCE, GPUDirect, PXN, NVLink
2mo
Save
Mark Applied
Hide
Engineering Manager, Model Inference
San Francisco or New York or Pittsburgh
$220k-$270k/yr HybridFull Time
Abridge
Abridge: Automates medical documentation through AI-powered speech analysis
5+ YOE1+ Mgmt5+ years engineering with 1+ year in technical leadership; ML systems and inference experience; GPU, latency, throughput expertise; strong people leadership and collaboration.
PyTorch, TensorRT, vLLM, TensorFlow, GPU
3w
Save
Mark Applied
Hide
Campus AI Research Engineer (Full-Time)
Chicago or New York City
$250k-$300k/yr OnsiteFull Time
Jump Trading
Jump Trading: Global proprietary trading firm specializing in algorithmic and high-frequency strategies
Proficiency in Python/C++ and ML frameworks, expertise in GPU/accelerator programming, experience building large-scale AI/ML systems, strong communication and availability.
Python, C, C++, PyTorch, JAX, TensorFlow, CUDA, Triton, SYCL, ROCm
1mo
Save
Mark Applied
Hide
Member of Technical Staff, Research Engineer
New York City, New York, United States
$120k-$210k/yr OnsiteFull Time
Ataraxis AI
Ataraxis AI: Develops AI-native tools for cancer prognosis and treatment selection.
BS/MS/PhD in CS, ML, or statistics; strong ML and statistics foundation; proficiency in Python and PyTorch; GPU optimization and model deployment experience; strong research and communication skills.
Python, PyTorch, GPU
2w
Save
Mark Applied
Hide
Software Engineer - Control Systems
Princeton, New Jersey, United States
$150k-$235k/yr OnsiteFull Time
Logiqal
Logiqal: Developing modular quantum computing systems using neutral atom technology.
Hands-on instrumentation and software fundamentals; hardware-in-the-loop testing, data pipelines, distributed monitoring; experience integrating high-speed cameras, RF signal generators, GPUs, and Linux kernel customizations.
high-speed cameras, RF signal generators, GPUs, Linux kernel
1w
Save
Mark Applied
Hide
Senior Scientist, Computational Receptor Biology & Machine Learning - Long Island City, NY (82328)
Long Island City or San Diego
$132k-$158k/yr OnsiteFull Time
dsm-firmenich
dsm-firmenichEuronext Amsterdam: DSFIR: Produces ingredients for nutrition, health, and beauty products.
3+ YOEPhD or equivalent,3+ years academic or industry experience in ML/applied math/physics-inspired modeling,deep learning expertise,experience with PyTorch and/or JAX,and distributed GPU training for molecular modeling.
PyTorch, JAX, GPUs
2mo
Save
Mark Applied
Hide
Inference Performance Engineer
New York, New York, United States
HybridFull Time
Material
Material: Specialized inference cloud platform for high-performance AI workloads.
BS in CS/EE or related field; proficiency in Rust/Go/Python/C++; knowledge of concurrency, tail latency; experience with model serving; GPU/ASIC programming; low-precision inference; profiling and benchmarking.
Rust, Go, Python, C++, vLLM, TensorRT-LLM, llama.cpp, CUDA, ROCm, Triton, TGI, SGLang, Nsight, perf
2mo
Save
Mark Applied
Hide
Staff Software Engineer - AI Research Infrastructure
San Francisco or New York City
$199k-$270k/yr HybridFull Time
Databricks
Databricks: A unified platform for data analytics and artificial intelligence.
5+ YOEBS/MS or PhD in computer science; 5+ years of software engineering experience in distributed systems or infrastructure; strong experience with GPUs, clusters, and cloud platforms; proficient in systems languages and large-scale job orchestration.
C++, Rust, Go, Java, Scala, Kubernetes, Slurm, Ray, GPU, Cloud computing, Distributed systems
1mo
Save
Mark Applied
Hide
AMI Engineer (Platform and Infra) Singapore
Paris or Montreal or New York City or Singapore
OnsiteFull Time
Advanced Machine Intelligence
Advanced Machine Intelligence: Developing frontier world model-based artificial intelligence systems.
Bachelor's in Computer Science or equivalent; proficiency in Python; ability to design, run, and analyze experiments; knowledge of ML fundamentals and large-scale accelerator (GPU/TPU) training environments; deep learning framework experience preferred.
Python, PyTorch, JAX, GPU, TPU