60 hpc engineer jobs at 29 companies in Piscataway, NJ

1w
Save
Mark Applied
Hide
Cloud HPC Engineer
New York City or Jersey City
$125k-$140k/yr HybridFull Time
Capgemini
CapgeminiEuronext Paris: CAP: Global leader in consulting, digital transformation, and engineering services.
10+ YOERequires 10 years designing and operating massive-scale compute grids, AWS or GCP expertise, Docker and Kubernetes, C and Python, distributed systems, performance tuning, infrastructure as code, and a technical degree.
AWS, GCP, Docker, Kubernetes, C, Python
2mo
Save
Mark Applied
Hide
Staff HPC Systems Software Engineer
New York City or United States
$225k-$275k/yr OnsiteFull Time
Nscale
Nscale: Vertically integrated AI infrastructure and GPU cloud platform.
Extensive experience designing and building Slurm-based HPC systems, strong software skills in Go or Python, deep knowledge of GPU infrastructure and HPC networking (InfiniBand, RoCE, RDMA), and experience integrating HPC with cloud-native platforms.
Slurm, Kubernetes, Go, Python, Kueue, InfiniBand, RoCE, RDMA
2w
Save
Mark Applied
Hide
HPC Operations Engineer
New York City, New York, United States
$175k-$225k/yr HybridFull Time
Tower Research Capital
Tower Research Capital: Proprietary quantitative trading firm employing traders, engineers, researchers, and business-support staff to trade global financial markets.
2+ YOEBachelor's degree or equivalent practical experience, 2+ years supporting Linux production environments, Linux administration, technical troubleshooting, user support, and strong written communication.
Linux, RHEL, Ubuntu, Bash, Python, Slurm, HTCondor, LSF, NFS, LDAP, PXE, Ansible, Prometheus, Grafana
1mo
Save
Mark Applied
Hide
HPC Middleware Developer
Santa Clara or Illinois or Colorado or Boulder or California or Holmdel
$152k-$288k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOE5+ years programming in C/C++, 3+ years Linux experience, deep knowledge of InfiniBand and Ethernet, computer architecture, OS, and performance optimization; MSc or equivalent preferred.
C, C++, Linux, InfiniBand, Ethernet, MPI, RDMA
3w
Save
Mark Applied
Hide
HPC Middleware Developer
Santa Clara or Colorado or Illinois or Holmdel or California or Boulder
$152k-$288k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOERequires 5 years of C/C++ programming, 3 years in Linux environments and tools, networking protocol expertise, computer architecture and operating systems knowledge, performance optimization experience, and an MSc or equivalent experience.
C, C++, Linux, InfiniBand, Ethernet, MPI, RDMA
3w
Save
Mark Applied
Hide
Senior HPC Support Engineer, InfiniBand - NVLink
Westford or Durham or New York City or Santa Clara or Redmond
$108k-$207k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOE5+ years of customer support and debugging in large-scale networking or AI infrastructure; networking, Linux administration, containers, virtualization, cloud, InfiniBand, and GPU expertise required.
Linux, InfiniBand, NVLink, GPU, Claude, Codex, Cursor, IP, TCPDUMP, Wireshark, Open platforms, KVM, ESXi, AWS, OCI, RDMA/RoCEv2, NCCL, MPI, Slurm/SchedMD, Bash, Python, CCNP, CCIE, JNCIE-DC/ENT, RHCE, LFCS, NCP-AII/AIO/AIN
2w
Save
Mark Applied
Hide
AI & HPC Infrastructure Engineer
Albany or Arlington or Atlanta or Austin or Beaverton or Bentonville or Boston or Carmel or Charlotte or Chicago or Cincinnati or Cleveland or Columbus or Culver City or Denver or Des Moines or Detroit or Hartford or Houston or Irvine or Irving or Kirkland or Miami or Milwaukee or Minneapolis or Morristown or Mountain View or Nashville or New York City or Oklahoma City or Overland Park or Philadelphia or Pittsburgh or Raleigh or Redmond or Sacramento or San Diego or San Francisco or Scottsdale or Seattle or St. Louis or St. Petersburg or Walnut Creek or United States
$80k-$266k/yr HybridFull Time
Accenture
AccentureNYSE: ACN: Global professional services delivering 360° value.
5+ YOERequires 5+ years designing AI infrastructure, accelerated computing, clusters, orchestration, and automation; bachelor's degree or equivalent experience. Strong Kubernetes, Python, Terraform, and cloud expertise required.
Slurm, Run:ai, Kubernetes, NVIDIA Base Command Manager (BCM), NVIDIA NGC, NCCL, NVLink, CUDA, TensorRT-LLM, vLLM, SGLang, Triton Inference Server, NVIDIA Dynamo, llm-d, MLPerf, NCCL tests, fio, iperf, InfiniBand, Ethernet, SONiC, NVMe, NVMe-oF, VAST, Weka, DDN, AWS, Azure, GCP, VMware, Nutanix, Python, Terraform, Ansible, REST, OpenAPI, JSON, YAML, TensorFlow, PyTorch, JAX, Jupyter notebooks, Google Colab, MCP
2mo
Save
Mark Applied
Hide
Lead Systems Engineer (HPC)
Princeton, New Jersey, United States
$135k-$150k/yr OnsiteFull Time
Princeton University
Princeton University: Private nonprofit research university providing undergraduate and graduate education across humanities, sciences, engineering and public affairs.
10+ YOE10+ years managing research computing systems, strong Linux administration, scripting (bash/Python/Perl), SLURM experience, networking for HPC, and ability to lead technical projects.
Linux, bash, Python, Perl, SLURM, Globus
1mo
Save
Mark Applied
Hide
High Performance Computing (HPC) AI Engineer
New York, New York, United States
$101k-$140k/yr OnsiteFull Time
NYU Langone Health
NYU Langone Health: Academic medical center focused on patient care, research, and education.
5+ YOEBachelor's degree, 5+ years relevant experience; expertise in parallel storage, SMB/NFS, Linux/UNIX and bash, networking (Ethernet, Infiniband, fiber channel), scripting (Bash, Python, Lua), and HPC/AI optimization.
Linux/UNIX, Bash, Python, Lua, SMB, NFS, Infiniband, fiber channel, UltraViolet
1mo
Save
Mark Applied
Hide
Distinguished Engineer, Full Stack (Remote Eligible)
Cambridge or Richmond or New York City or McLean or United States
$245k-$335k/yr RemoteFull Time
Capital One
Capital OneNYSE: COF: A technology-driven bank providing diverse financial services.
7+ YOEBachelor's degree and 7+ years in software engineering, solution or enterprise architecture, distributed HPC/ML systems, cloud computing, and data engineering; advanced coding and AI-native engineering experience preferred.
Python, Java, Go, Scala, JavaScript, TypeScript, AWS, Microsoft Azure, Google Cloud
4w
Save
Mark Applied
Hide
Security engineer, detection and response
San Francisco or Seattle or New York City
$132k-$258k/yr HybridFull Time
Writer
Writer: Enterprise generative AI platform that helps businesses build and supervise secure AI agents.
3+ YOE3+ years in security operations, detection engineering, or incident response; experience securing AI/ML or distributed HPC infrastructure; strong programming and forensic skills; SIEM and detection experience.
Python, KQL, SPL, SIEM
2mo
Save
Mark Applied
Hide
Senior Specialist Field Engineer - Compute Infrastructure
Livingston or New York or Sunnyvale or San Francisco or Bellevue or Dallas
$188k-$275k/yr FieldFull Time
CoreWeave
CoreWeaveNasdaq: CRWV: Specialized cloud provider for large-scale AI and machine learning.
7+ YOEB.S. or equivalent experience, 7+ years in solutions architecture/field/infrastructure engineering or TAM for cloud/HPC; deep expertise with bare-metal GPU clusters, Linux, networking, InfiniBand/NVLink, PXE, and customer-facing technical leadership.
Kubernetes, Slurm, Python, Bash, Ansible, NCCL, ib_write_bw, InfiniBand, NVLink, NVIDIA HGX, GB200, Linux, PXE, BMC, BIOS, TCP/IP, Bare Metal as a Service (BMaaS)
1mo
Save
Mark Applied
Hide
AI Compute Engineer
Palo Alto or New York or San Francisco or Montreal
RemoteFull Time
Mistral AI
Mistral AI: Developer of open-weight and frontier AI models.
Strong Linux systems administration in large-scale/HPC or cloud; experience with job schedulers (Slurm), containers (Kubernetes), storage (Ceph, Lustre, NFS), automation (Ansible, Terraform), and scripting (Python, Bash).
Python, Bash, Ansible, Terraform, Slurm, Kubernetes, Ceph, Lustre, NFS, InfiniBand, Ethernet
2mo
Save
Mark Applied
Hide
Senior Network Engineer
New York City or San Francisco or Seattle or London or United States
$150k-$190k/yr RemoteFull Time
Lightning AI
Lightning AI: Private AI software and GPU-cloud helping developers and enterprises build, train, deploy, and run AI systems.
5+ YOE5+ years large-scale data center networking experience with Cumulus NOS, spine/leaf design, BGP/EVPN/VXLAN, HPC/GPU networking, automation (Python/Ansible/Terraform), and strong documentation skills.
Cumulus NOS, SONiC, Junos, Python, Ansible, Terraform, EVPN, VXLAN, BGP, RoCE, RDMA, InfiniBand, VPC, NFV, Direct Connect, Cloud Connect, NVIDIA Spectrum, NVIDIA Quantum, NVIDIA BlueField
2mo
Save
Mark Applied
Hide
Systems Engineer, Agents/Clusters/Supercomputers
New York, New York, United States
$250k-$800k/yr HybridFull Time
D. E. Shaw Research
D. E. Shaw Research: Private computational-biochemistry research developing supercomputers, software, and precisely targeted drugs for disease treatment.
Strong Linux fundamentals; experience with large-installation systems administration, HPC, Kubernetes, Python, distributed systems, and excellent communication skills. Candidates at all experience levels considered.
Kubernetes, Python, Linux, RDMA, GPU
2mo
Save
Mark Applied
Hide
Senior Cloud Support Engineer
New York or Denver
$145k-$175k/yr OnsiteFull Time
Crusoe
Crusoe: Vertically integrated AI infrastructure and energy.
5+ YOEBachelor's or 4+ years technical experience; 5+ years customer support experience; Linux CLI, Git, Kubernetes, Slurm, Terraform, Grafana; public cloud (AWS/Azure/GCP); HPC knowledge; on-call availability.
Zendesk, CLI, Linux, Git, Kubernetes, Slurm, Terraform, Grafana, AWS, Azure, GCP, Infiniband, RDMA, RoCE, SDN
1mo
Save
Mark Applied
Hide
Sr. Liquid Cooling Thermal Design & Validation Engineer
Secaucus, New Jersey, United States
$99k-$170k/yr OnsiteFull Time
AMD
AMDNASDAQ: AMD: Leader in high-performance computing, graphics, and visualization technologies.
Expertise in thermal design and validation for liquid-cooled data center or HPC platforms; CFD experience; mechanical CAD and lab testing skills; Python/MATLAB/LabVIEW for data analysis; bachelor’s or master’s in engineering preferred.
Icepak, FloTHERM, Ansys Fluent, Python, MATLAB, LabVIEW, Creo, SolidWorks, FEA, Open Compute Project (OCP)
2w
Save
Mark Applied
Hide
Senior Simulation Engineer - Electromagnetics Specialist
New York City, New York, United States
HybridFull Time
PhysicsX
PhysicsX: Private physics-AI software helping industrial engineering and manufacturing teams design and optimize hardware.
3+ YOERequires 3–5 years of commercial EM simulation experience after a master's or PhD, electromagnetic theory expertise, industry-standard EM simulation platform proficiency, and complex model development skills.
Ansys HFSS, Ansys Maxwell, CST Studio Suite, COMSOL Multiphysics, Altair FEKO, Altair Flux, Siemens Simcenter MAGNET, Python, MATLAB, Cadence Allegro, Altium Designer, Ansys SIwave, Machine Learning, Deep Learning, cloud platform, HPC, CAE
1mo
Save
Mark Applied
Hide
Distinguished Engineer, Full Stack (Remote Eligible)
Cambridge or Richmond or New York City or McLean or United States or New York
$245k-$335k/yr RemoteFull Time
Capital One
Capital OneNYSE: COF: A technology-driven bank providing diverse financial services.
7+ YOEBachelor's degree and 7+ years in software engineering, architecture, distributed HPC/ML systems, cloud computing, and data engineering; expertise with AI-native engineering and coding languages preferred.
AWS, Microsoft Azure, Google Cloud, Python, Java, Go, Scala, JavaScript, TypeScript, AI/ML frameworks, AIOps, CI/CD, SDLC, LLM
1mo
Save
Mark Applied
Hide
Campus AI Research Engineer – Deep Learning(Full-Time)
Chicago or New York City
$300k/yr OnsiteFull Time
Jump Trading
Jump Trading: Privately held proprietary trading firm using research and technology to trade global financial markets.
Advanced ML research experience with strong publication or open-source contributions; expertise in deep learning and language-modeling; strong software skills in Python/C++; experience with model training, HPC, and production integration.
C, C++, Python, CUDA, PyTorch, JAX, TensorFlow, HPC, ROCm