50 software engineer gpu jobs at 33 companies in New York

1mo
Save
Mark Applied
Hide
Software Engineer, Compute (GPU)
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Kubernetes, Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, LLM APIs, Claude Code, Cursor
11h
Save
Mark Applied
Hide
Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)
San Francisco or Seattle or New York City or United States
$250k-$485k/yr OnsiteFull Time
Perplexity
Perplexity: AI-powered search engine providing conversational answers with citations.
Deep Kubernetes and GPU cluster experience, multi-cloud orchestration, strong distributed systems fundamentals, systems-level coding in Go/Rust/C++, and experience with training and inference workloads.
Kubernetes, kubectl, NVIDIA, CUDA, InfiniBand, RoCE, CoreWeave, AWS, GCP, Go, Rust, C++, vLLM, SGLang, TensorRT-LLM, Slurm, Triton, RDMA, Prometheus, Grafana, Weights & Biases
3mo
Save
Mark Applied
Hide
Infrastructure Engineer (GPU & Compute)
New York or San Francisco or Seattle
$180k-$200k/yr RemoteFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
NVIDIA DCGM, PXE, IPMI, Redfish, iDRAC, LiveCD, InfiniBand, NVLink, Linux, Python
1mo
Save
Mark Applied
Hide
Software Engineer, AI Labs
Edinburgh or New York City or Atlanta or San Francisco or Seattle
HybridFull Time
BlackRock
BlackRockNYSE: BLK: Provides investment management and financial technology services globally.
3+ YOE3+ years professional software engineering experience; strong Python and SQL skills; experience building, testing, deploying, and operating cloud-native applications, services, APIs, or data pipelines; strong engineering fundamentals and communication.
Python, SQL, Spark, Airflow, Dagster, Flyte, GPUs, TPUs, AWS Inferentia, CI/CD, AI coding assistants
3w
Save
Mark Applied
Hide
Senior Deep Learning Software Engineer, Inference
California or Texas or New York or Washington or Massachusetts
$152k-$288k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOEMaster's/PhD or equivalent experience, 5+ years software development, strong C/C++ skills, GPU/CUDA experience, DL inference optimization and production deployment experience.
CUTLASS, OAI TRITON, NCCL, CUDA, Python, C, C++, vLLM, SGLang, FlashInfer, PyTorch, NVSHMEM
2mo
Save
Mark Applied
Hide
Staff Software Engineer - AI Research Infrastructure
San Francisco or New York City
$199k-$270k/yr HybridFull Time
Databricks
Databricks: A unified platform for data analytics and artificial intelligence.
5+ YOEBS/MS or PhD in computer science; 5+ years of software engineering experience in distributed systems or infrastructure; strong experience with GPUs, clusters, and cloud platforms; proficient in systems languages and large-scale job orchestration.
C++, Rust, Go, Java, Scala, Kubernetes, Slurm, Ray, GPU, Cloud computing, Distributed systems
1mo
Save
Mark Applied
Hide
Staff+ Software Engineer, Inference Runtime
San Francisco or Seattle or New York City
$405k-$485k/yr HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
Senior IC with deep systems or ML infrastructure experience, hands-on performance profiling and optimization, accelerator ecosystem expertise (CUDA/TPU/Trainium), strong software engineering and cross-org alignment skills, and a relevant bachelor’s degree or equivalent.
Rust, Python, CUDA, XLA, Triton, NeuronX, AWS Neuron, Kubernetes, CI/CD
3w
Save
Mark Applied
Hide
Software Engineer - Systems
New York City or San Francisco or London
$200k-$275k/yr OnsiteFull Time
Spiral
Spiral: Building high-performance multimodal data infrastructure for AI systems.
5+ YOE5+ years systems programming experience, strong OS foundation, performance engineering, familiarity with Apache Arrow/Parquet/DataFusion/Clickhouse/DuckDB, PyTorch/CUDA understanding, Rust a bonus; able to work onsite in NYC, SF, or London.
Vortex, Apache Arrow, Parquet, DataFusion, Clickhouse, DuckDB, PyTorch, CUDA, Rust
3w
Save
Mark Applied
Hide
Senior Deep Learning Software Engineer, Inference
California or Texas or New York or Washington or Massachusetts
$152k-$288k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOEMasters/PhD or equivalent,5+ years software development,excellent C/C++ skills,CUDA and GPU programming experience preferred,experience optimizing/deploying DL inference,Python and performance profiling experience helpful.
CUTLASS, OAI Triton, NCCL, CUDA, vLLM, SGLang, FlashInfer, PyTorch, NVSHMEM, C/C++, Python
2w
Save
Mark Applied
Hide
Senior Software Engineer, Gemini Audio, DeepMind
Mountain View or New York
$174k-$253k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
5+ YOEBachelor's degree or equivalent experience; 5+ years software development with Python; 3+ years testing/maintaining/launching software; 1+ year software design; experience with compilers, codegen, runtimes, ML and performance optimization.
Python, GPU, TPU, OpenXLA, MLIR, vLLM, sglang, JAX, PyTorch, Accelerated Linear Algebra (XLA)
1mo
Save
Mark Applied
Hide
Software Engineer, Machine Learning Infrastructure - Gen AI
San Francisco or New York City or Illinois or Colorado or United States
$137k-$202k/yr OnsiteFull Time
DoorDash
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
4+ YOEB.S./M.S./PhD in CS or equivalent; 4+ years software engineering; strong backend fundamentals in Python and distributed systems; experience building production services, observability, and ML workflows.
Python, Kubernetes, AWS, GCP, GPUs, LLM
4w
Save
Mark Applied
Hide
Principal Software Engineer, Machine Learning Infrastructure
Palo Alto or Seattle or Los Angeles or New York or Bellevue
$235k-$414k/yr OnsiteFull Time
Snap
SnapNYSE: SNAP: Develops social media applications and augmented reality technology.
10+ YOE10+ years software development experience, technical leadership, distributed systems and ML inference platform expertise, strong software design and debugging skills, ability to operate highly-available systems at scale.
Tensorflow, PyTorch, Kubernetes, GPU, LLM inference, RPC
3w
Save
Mark Applied
Hide
Software Engineer Sr
Liverpool, New York, United States
$93k-$164k/yr OnsiteFull Time
Lockheed Martin
Lockheed MartinNYSE: LMT: Designs and manufactures global security and aerospace systems.
5+ YOEBachelor's in Computer Science, 5+ years C++ OOP experience, Linux, automated testing and static analysis, Git/JIRA/Confluence, ability to obtain SECRET clearance and US citizenship.
C++, Linux, Git, JIRA, Confluence, bash, Python, CUDA, UML
1mo
Save
Mark Applied
Hide
Senior Software Developer for Quantum-Centric Supercomputing
Yorktown Heights, New York, United States
$161k-$276k/yr OnsiteFull Time
IBM
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Master's degree required; strong proficiency in Python and a systems language (C, C++, or Rust); experience with modern software engineering practices, HPC/resource managers, EM/quantum ecosystems, and strong communication skills.
Python, Rust, C, OpenMP, MPI, GPU, Slurm, PBS, Grid Engine, LSF, Qiskit, OpenQASM, CUDA-Q, Linux, HPC, CI/CD
4w
Save
Mark Applied
Hide
Principal Software Engineer, Machine Learning Infrastructure
Palo Alto or Seattle or Los Angeles or New York City or Bellevue
$235k-$414k/yr OnsiteFull Time
Snap
SnapNYSE: SNAP: Provides visual messaging software and augmented reality wearable devices.
10+ YOE10+ years software development experience (or equivalent with advanced degree), technical leadership experience, expertise in large-scale distributed systems, RPC services, debugging, performance analysis, and mentorship.
Tensorflow, PyTorch, Kubernetes, GPU
1mo
Save
Mark Applied
Hide
Staff HPC Systems Software Engineer
New York City or United States
$225k-$275k/yr OnsiteFull Time
Nscale
Nscale: Vertically integrated AI infrastructure provider for high-performance computing.
Extensive experience designing and building Slurm-based HPC systems, strong software skills in Go or Python, deep knowledge of GPU infrastructure and HPC networking (InfiniBand, RoCE, RDMA), and experience integrating HPC with cloud-native platforms.
Slurm, Kubernetes, Go, Python, Kueue, InfiniBand, RoCE, RDMA
3w
Save
Mark Applied
Hide
Backend Software Engineer - Platforms
New York, New York, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
3+ YOEBS in CS or equivalent, experience with web apps, Unix/Linux, distributed systems, networking, and large systems; DICM/ITOM/ITSM experience required; GPU or big data (Hadoop/Kafka/Apache) experience valuable; Golang or Python development experience and production troubleshooting skills.
Golang, Python, Unix/Linux, Hadoop, Kafka, Apache, GPU, Discovery, Insights, and Configuration Management (DICM), IT Operation Management (ITOM), IT Service Management (ITSM)
2w
Save
Mark Applied
Hide
Senior AI Engineer
Chicago or Hong Kong or London or New York City or Singapore
$200k-$300k/yr OnsiteFull Time
DV Trading
DV Trading: Proprietary trading firm providing liquidity to global financial markets.
5+ YOE5+ years software engineering with strong Python; production fine-tuning/distillation of open-weight models; on-prem LLM serving, GPU infrastructure and Kubernetes experience; model evaluation and tooling.
Python, Llama, Qwen, Mistral, vLLM, TGI, Triton, Kubernetes, Hugging Face
2mo
Save
Mark Applied
Hide
Inference Performance Engineer
New York, New York, United States
HybridFull Time
Material
Material: Specialized inference cloud platform for high-performance AI workloads.
BS in CS/EE or related field; proficiency in Rust/Go/Python/C++; knowledge of concurrency, tail latency; experience with model serving; GPU/ASIC programming; low-precision inference; profiling and benchmarking.
Rust, Go, Python, C++, vLLM, TensorRT-LLM, llama.cpp, CUDA, ROCm, Triton, TGI, SGLang, Nsight, perf
1mo
Save
Mark Applied
Hide
Senior Machine Learning Platform Engineer
New York, New York, United States
$155k-$215k/yr HybridFull Time
Charlie Health
Charlie Health: Providing virtual intensive outpatient programs for mental health and recovery.
6+ YOE6+ years software engineering (2+ years ML infrastructure); strong Python; experience with cloud ML services (AWS SageMaker, GCP Vertex AI), Terraform/Pulumi, Kubernetes/ECS, GPU inference, observability, and CI/CD for ML.
Python, AWS SageMaker, GCP Vertex AI, Terraform, Pulumi, Kubernetes, ECS, Go, Rust, C++