50 software engineer gpu jobs at 33 companies in New York
1mo
Save
Mark Applied
Hide
1mo
Software Engineer, Compute (GPU)
San Francisco or New York or Austin or Seattle
$175k-$300k/yrOnsiteFull Time
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)
San Francisco or Seattle or New York City or United States
$250k-$485k/yrOnsiteFull Time
Perplexity: AI-powered search engine providing conversational answers with citations.
Deep Kubernetes and GPU cluster experience, multi-cloud orchestration, strong distributed systems fundamentals, systems-level coding in Go/Rust/C++, and experience with training and inference workloads.
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
Edinburgh or New York City or Atlanta or San Francisco or Seattle
HybridFull Time
BlackRockNYSE: BLK: Provides investment management and financial technology services globally.
3+ YOE3+ years professional software engineering experience; strong Python and SQL skills; experience building, testing, deploying, and operating cloud-native applications, services, APIs, or data pipelines; strong engineering fundamentals and communication.
Staff Software Engineer - AI Research Infrastructure
San Francisco or New York City
$199k-$270k/yrHybridFull Time
Databricks: A unified platform for data analytics and artificial intelligence.
5+ YOEBS/MS or PhD in computer science; 5+ years of software engineering experience in distributed systems or infrastructure; strong experience with GPUs, clusters, and cloud platforms; proficient in systems languages and large-scale job orchestration.
Anthropic: Developing safe and reliable artificial intelligence systems.
Senior IC with deep systems or ML infrastructure experience, hands-on performance profiling and optimization, accelerator ecosystem expertise (CUDA/TPU/Trainium), strong software engineering and cross-org alignment skills, and a relevant bachelor’s degree or equivalent.
Spiral: Building high-performance multimodal data infrastructure for AI systems.
5+ YOE5+ years systems programming experience, strong OS foundation, performance engineering, familiarity with Apache Arrow/Parquet/DataFusion/Clickhouse/DuckDB, PyTorch/CUDA understanding, Rust a bonus; able to work onsite in NYC, SF, or London.
5+ YOEBachelor's degree or equivalent experience; 5+ years software development with Python; 3+ years testing/maintaining/launching software; 1+ year software design; experience with compilers, codegen, runtimes, ML and performance optimization.
Software Engineer, Machine Learning Infrastructure - Gen AI
San Francisco or New York City or Illinois or Colorado or United States
$137k-$202k/yrOnsiteFull Time
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
4+ YOEB.S./M.S./PhD in CS or equivalent; 4+ years software engineering; strong backend fundamentals in Python and distributed systems; experience building production services, observability, and ML workflows.
Principal Software Engineer, Machine Learning Infrastructure
Palo Alto or Seattle or Los Angeles or New York or Bellevue
$235k-$414k/yrOnsiteFull Time
SnapNYSE: SNAP: Develops social media applications and augmented reality technology.
10+ YOE10+ years software development experience, technical leadership, distributed systems and ML inference platform expertise, strong software design and debugging skills, ability to operate highly-available systems at scale.
Lockheed MartinNYSE: LMT: Designs and manufactures global security and aerospace systems.
5+ YOEBachelor's in Computer Science, 5+ years C++ OOP experience, Linux, automated testing and static analysis, Git/JIRA/Confluence, ability to obtain SECRET clearance and US citizenship.
Senior Software Developer for Quantum-Centric Supercomputing
Yorktown Heights, New York, United States
$161k-$276k/yrOnsiteFull Time
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Master's degree required; strong proficiency in Python and a systems language (C, C++, or Rust); experience with modern software engineering practices, HPC/resource managers, EM/quantum ecosystems, and strong communication skills.
Nscale: Vertically integrated AI infrastructure provider for high-performance computing.
Extensive experience designing and building Slurm-based HPC systems, strong software skills in Go or Python, deep knowledge of GPU infrastructure and HPC networking (InfiniBand, RoCE, RDMA), and experience integrating HPC with cloud-native platforms.
ByteDance: Developing AI-driven content platforms and mobile applications.
3+ YOEBS in CS or equivalent, experience with web apps, Unix/Linux, distributed systems, networking, and large systems; DICM/ITOM/ITSM experience required; GPU or big data (Hadoop/Kafka/Apache) experience valuable; Golang or Python development experience and production troubleshooting skills.
Golang, Python, Unix/Linux, Hadoop, Kafka, Apache, GPU, Discovery, Insights, and Configuration Management (DICM), IT Operation Management (ITOM), IT Service Management (ITSM)
Chicago or Hong Kong or London or New York City or Singapore
$200k-$300k/yrOnsiteFull Time
DV Trading: Proprietary trading firm providing liquidity to global financial markets.
5+ YOE5+ years software engineering with strong Python; production fine-tuning/distillation of open-weight models; on-prem LLM serving, GPU infrastructure and Kubernetes experience; model evaluation and tooling.
Python, Llama, Qwen, Mistral, vLLM, TGI, Triton, Kubernetes, Hugging Face
Material: Specialized inference cloud platform for high-performance AI workloads.
BS in CS/EE or related field; proficiency in Rust/Go/Python/C++; knowledge of concurrency, tail latency; experience with model serving; GPU/ASIC programming; low-precision inference; profiling and benchmarking.
Charlie Health: Providing virtual intensive outpatient programs for mental health and recovery.
6+ YOE6+ years software engineering (2+ years ML infrastructure); strong Python; experience with cloud ML services (AWS SageMaker, GCP Vertex AI), Terraform/Pulumi, Kubernetes/ECS, GPU inference, observability, and CI/CD for ML.