155 gpu jobs at 81 companies in New York

1w
Save
Mark Applied
Hide
GPU Performance Engineer
New York City, New York, United States
$165k-$300k/yr HybridFull Time
Two Sigma
Two Sigma: Systematic investment management and quantitative trading firm.
1+ YOEExpert CUDA and GPU architecture knowledge, C++ and Python skills, BS/MS in STEM, minimum 1 year relevant experience, GPU profiling and optimization experience.
CUDA, Nsight Systems, Nsight Compute, RAPIDS, CUTLASS, cuBLAS, TensorRT, NCCL, MPI, cuDF, C++, Python
2mo
Save
Mark Applied
Hide
GPU Performance Engineer - Neural Reconstruction
United States or California or Texas or New York or Oregon or Missouri or Massachusetts
$224k-$431k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOEStrong background in GPU performance optimization, Python/C++, CUDA, PyTorch; experience profiling and benchmarking; ability to optimize end-to-end neural reconstruction workloads.
Python, C++, CUDA, PyTorch, Nsight Systems, Nsight Compute, NVTX, NumPy
3mo
Save
Mark Applied
Hide
Infrastructure Engineer (GPU & Compute)
New York or San Francisco or Seattle
$180k-$200k/yr RemoteFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
NVIDIA DCGM, PXE, IPMI, Redfish, iDRAC, LiveCD, InfiniBand, NVLink, Linux, Python
1w
Save
Mark Applied
Hide
Engineering Manager, GPU Infrastructure
Toronto or San Francisco or New York City or London or Paris or Montreal
HybridFull Time
Cohere
Cohere: Provides enterprise-grade large language models and AI software platforms.
Experience managing engineering teams focused on GPU/ML infrastructure, Kubernetes, IaC, observability, and collaboration with AI researchers; strong communication and mentorship skills.
JAX, PyTorch, TensorFlow, Kubernetes, Prometheus, Grafana, Terraform, ArgoCD
3w
Save
Mark Applied
Hide
Software Engineer, Compute (GPU)
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Kubernetes, Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, LLM APIs, Claude Code, Cursor
2mo
Save
Mark Applied
Hide
Head of Infrastructure, Stealth Edge AI Co
New York City, New York, United States
HybridFull Time
Montauk Capital
Montauk Capital: Investment firm building and funding climate technology companies.
Lead hardware and infrastructure buildout for edge AI compute; manage GPU infrastructure, data center deployments, and supply chain.
Linux, IPMI, Redfish, PXE, BMC, GPU, GPU scheduling, Automation, Monitoring, OOB management
2mo
Save
Mark Applied
Hide
Member of Research Staff, Reinforcement Learning, Voleon Securities
Berkeley or New York City
$250k-$275k/yr HybridFull Time
The Voleon Group
The Voleon Group: Quantitative investment management firm using machine learning strategies.
PhD-level coursework required; PhD preferred. Strong RL background, mathematical ability, Python production code, GPU computing, and interest in finance.
Python, GPU computing
2mo
Save
Mark Applied
Hide
Engineering Manager, Model Inference
San Francisco or New York or Pittsburgh
$220k-$270k/yr HybridFull Time
Abridge
Abridge: Automates medical documentation through AI-powered speech analysis
5+ YOE1+ Mgmt5+ years engineering with 1+ year in technical leadership; ML systems and inference experience; GPU, latency, throughput expertise; strong people leadership and collaboration.
PyTorch, TensorRT, vLLM, TensorFlow, GPU
3w
Save
Mark Applied
Hide
Campus AI Research Engineer (Full-Time)
Chicago or New York City
$250k-$300k/yr OnsiteFull Time
Jump Trading
Jump Trading: Global proprietary trading firm specializing in algorithmic and high-frequency strategies
Proficiency in Python/C++ and ML frameworks, expertise in GPU/accelerator programming, experience building large-scale AI/ML systems, strong communication and availability.
Python, C, C++, PyTorch, JAX, TensorFlow, CUDA, Triton, SYCL, ROCm
4w
Save
Mark Applied
Hide
Member of Technical Staff, Research Engineer
New York City, New York, United States
$120k-$210k/yr OnsiteFull Time
Ataraxis AI
Ataraxis AI: Develops AI-native tools for cancer prognosis and treatment selection.
BS/MS/PhD in CS, ML, or statistics; strong ML and statistics foundation; proficiency in Python and PyTorch; GPU optimization and model deployment experience; strong research and communication skills.
Python, PyTorch, GPU
5d
Save
Mark Applied
Hide
Senior Scientist, Computational Receptor Biology & Machine Learning - Long Island City, NY (82328)
Long Island City or San Diego
$132k-$158k/yr OnsiteFull Time
dsm-firmenich
dsm-firmenichEuronext Amsterdam: DSFIR: Produces ingredients for nutrition, health, and beauty products.
3+ YOEPhD or equivalent,3+ years academic or industry experience in ML/applied math/physics-inspired modeling,deep learning expertise,experience with PyTorch and/or JAX,and distributed GPU training for molecular modeling.
PyTorch, JAX, GPUs
2mo
Save
Mark Applied
Hide
Inference Performance Engineer
New York, New York, United States
HybridFull Time
Material
Material: Specialized inference cloud platform for high-performance AI workloads.
BS in CS/EE or related field; proficiency in Rust/Go/Python/C++; knowledge of concurrency, tail latency; experience with model serving; GPU/ASIC programming; low-precision inference; profiling and benchmarking.
Rust, Go, Python, C++, vLLM, TensorRT-LLM, llama.cpp, CUDA, ROCm, Triton, TGI, SGLang, Nsight, perf
1mo
Save
Mark Applied
Hide
Senior Solutions Engineer, AI Infrastructure
New York or United States or Israel
RemoteFull Time
VAST Data
VAST Data: Enterprise AI infrastructure and unified data platform.
8+ YOE8+ years hands-on infrastructure experience; strong Linux, distributed systems, storage, networking, Kubernetes, GPU/HPC knowledge; proven customer-facing solution design, PoC execution, debugging, and communication skills.
Linux, PAXOS, Raft, GPU, Kubernetes, Slurm, Ray, Spark, Lustre, Ceph, Weka, BeeGFS, GPFS, VAST, InfiniBand, RoCE, RDMA, CUDA, NCCL, DCGM, GPUDirect, NVIDIA, Mellanox
1mo
Save
Mark Applied
Hide
Software Developer
New York City, New York, United States
$250k-$600k/yr HybridFull Time
D. E. Shaw Research
D. E. Shaw Research: Developing custom supercomputers for biomolecular simulation and drug discovery.
Experience with Python, Linux/UNIX, GPU clusters, ML/AI frameworks (PyTorch, MLflow, vLLM, SGLang), and C++; managing scientific/chemical/ML data; interest in AI; open to all experience levels.
Python, PyTorch, MLflow, vLLM, SGLang, C++, GPU clusters, Linux/UNIX, LLMs
2mo
Save
Mark Applied
Hide
Staff Software Engineer - AI Research Infrastructure
San Francisco or New York City
$199k-$270k/yr HybridFull Time
Databricks
Databricks: A unified platform for data analytics and artificial intelligence.
5+ YOEBS/MS or PhD in computer science; 5+ years of software engineering experience in distributed systems or infrastructure; strong experience with GPUs, clusters, and cloud platforms; proficient in systems languages and large-scale job orchestration.
C++, Rust, Go, Java, Scala, Kubernetes, Slurm, Ray, GPU, Cloud computing, Distributed systems
1mo
Save
Mark Applied
Hide
AMI Engineer (Platform and Infra) Singapore
Paris or Montreal or New York City or Singapore
OnsiteFull Time
Advanced Machine Intelligence
Advanced Machine Intelligence: Developing frontier world model-based artificial intelligence systems.
Bachelor's in Computer Science or equivalent; proficiency in Python; ability to design, run, and analyze experiments; knowledge of ML fundamentals and large-scale accelerator (GPU/TPU) training environments; deep learning framework experience preferred.
Python, PyTorch, JAX, GPU, TPU
1mo
Save
Mark Applied
Hide
Machine Learning Engineer
Singapore or Shanghai or Hong Kong or New York City
HybridFull Time
Tower Research Capital
Tower Research Capital: Global quantitative trading firm developing automated algorithmic strategies.
2+ YOE2+ years building large-scale distributed systems; strong Python; experience with Linux HPC or cloud, GPU workloads, PyTorch/TensorFlow/JAX; distributed computing and data pipeline optimization; strong debugging and collaboration skills.
Python, Linux, HPC, GPU, PyTorch, TensorFlow, JAX, Docker, Kubernetes
1mo
Save
Mark Applied
Hide
Senior Applied Scientist, Amazon Brand Stores
Seattle or New York
$167k-$249k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
3+ YOEPhD or advanced degree in relevant field with applied ML experience; 3+ years building ML models for business, programming in Java/C++/Python, experience deploying LLMs on GPUs/Neuron/TPU, and advertising technology experience.
Java, C++, Python, GPU, Neuron, TPU, LLMs
1mo
Save
Mark Applied
Hide
AI Researcher, LLMs
London or New York City
$200k-$300k/yr OnsiteFull Time
Hudson River Trading
Hudson River Trading: A quantitative firm using technology to trade global financial markets.
2+ YOE2+ years training/post-training/evaluating LLMs at scale; experience with large-scale pretraining, post-training (SFT, RL), evaluation design, or distributed training/inference; strong research judgment and engineering skills (GPU/PyTorch).
PyTorch, GPU
1mo
Save
Mark Applied
Hide
Research Engineer, Code RL (Reinforcement Learning)
San Francisco or New York City
$500k-$850k/yr HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
Bachelor's degree or equivalent; strong software-engineering skills with deep Python expertise (async/concurrent); experience designing RL environments, building verifiers/eval harnesses, PyTorch and distributed training, CUDA/TPU/GPU performance optimization preferred.
Python, PyTorch, CUDA, TPU, GPU