6,329 llm engineer jobs at 2,449 companies in United States

2mo
Save
Mark Applied
Hide
LLM Engineer (Visa Sponsorship)
San Francisco, California, United States
OnsiteFull Time
Spherecast
Spherecast: AI Supply Chain Manager for Omni-Channel CPG Brands
Hands-on experience with modern LLM APIs (Anthropic, OpenAI, DeepSeek, OpenRouter, Gemini, Moonshot), HuggingFace experience, production LLM agents, evaluation pipelines, prompting, and observability; bonus for self-hosted LLMs.
Anthropic, OpenAI, DeepSeek, OpenRouter, Gemini, Moonshot, HuggingFace, LLM APIs
2d
Save
Mark Applied
Hide
LLM Engineer
New York City, New York, United States
OnsiteFull Time
Clera
Clera: AI-powered talent agent matching candidates to startup roles.
3+ YOE3+ years shipping production LLM systems; experience with modern LLM APIs, agents, system design, evaluation pipelines, observability, model selection, and HuggingFace.
OpenAI, Anthropic, DeepSeek, Gemini, HuggingFace
4d
Save
Mark Applied
Hide
LLM Engineer
United States
$100k-$150k/yr RemoteFull Time
Bright Vision Technologies
Bright Vision Technologies: AI-powered enterprise automation and software development firm.
6+ YOEMaster’s or PhD in computer science, machine learning, or related field or equivalent experience; 6+ years of ML research and engineering experience; Python, PyTorch, LLM fine-tuning, distributed training, evaluation, and GPU cluster operations.
Python, PyTorch, FSDP, ZeRO, LoRA, QLoRA, RLHF, DPO
1w
Save
Mark Applied
Hide
LLM Engineer (Remote)
Louisville, Kentucky, United States
$81k-$142k/yr RemoteFull Time
Cognizant
CognizantNASDAQ: CTSH: Global professional services providing technology and consulting services.
7+ YOERequires 7+ years in software, AI, machine learning, or platform engineering; 2+ years deploying generative AI in production; Python, LLMs, RAG, MLOps, and AI/ML frameworks.
Python, LoRA, QLoRA, RAG, APIs
1w
Save
Mark Applied
Hide
Senior LLM Engineer
Austin or Fremont or Lisle
$195k-$255k/yr HybridFull Time
Molex
Molex: Electronic, electrical, and fiber-optic connectivity systems manufacturer serving automotive, industrial, medical, and data-center customers.
8+ YOERequires 8+ years of ML/AI engineering experience, hands-on LLM fine-tuning and deployment, strong Python and PyTorch, transformer expertise, RAG, embeddings, vector search, and agentic systems.
Azure OpenAI, Azure AI Foundry, LoRA, QLoRA, RLHF, DPO, PEFT, Python, PyTorch, Azure AI Search, vLLM, TensorRT-LLM, Azure Kubernetes Service, DeepSpeed, FSDP
2w
Save
Mark Applied
Hide
LLM Engineer (Remote)
Louisville or United States
$81k-$142k/yr RemoteFull Time
Cognizant
CognizantNASDAQ: CTSH: Global professional services providing technology and consulting services.
7+ YOERequires 7+ years in software, AI, machine learning, or platform engineering; 2+ years deploying generative AI in production; Python, AI/ML frameworks, LLMs, RAG, APIs, and MLOps experience.
Python, LoRA, QLoRA, Retrieval-Augmented Generation (RAG), APIs, MLOps, GPU, vector databases
1w
Save
Mark Applied
Hide
LLM Engineer (Remote)
Louisville, Kentucky, United States
$81k-$142k/yr RemoteFull Time
Cognizant
CognizantNASDAQ: CTSH: Global professional services providing technology and consulting services.
7+ YOERequires 7+ years in software, AI, machine learning, or platform engineering; 2+ years deploying generative AI in production; Python, LLMs, RAG, APIs, cloud/on-premises platforms, MLOps, and model optimization.
Python, LoRA, QLoRA, Retrieval-Augmented Generation (RAG), APIs
2mo
Save
Mark Applied
Hide
LLM Algorithmic Optimization Engineer
San Jose, California, United States
$143k-$186k/yr OnsiteFull Time
NIO
NIONYSE: NIO: Global smart electric vehicle and battery technology provider.
Master's or PhD in a relevant technical field; expertise in GPU/NPU optimization, LLM/VLM architectures, Python, C/C++, PyTorch, inference engines, ONNX, and distributed computing.
Large Language Models (LLMs), Open Neural Network Exchange (ONNX), GPU, NPU, Python, PyTorch, C, C++, Linux kernel, hypervisor
1mo
Save
Mark Applied
Hide
LLM Inference Engineer
San Francisco or United States
RemoteFull Time
NEAR AI
NEAR AI: Confidential AI infrastructure running verifiable workloads for enterprises, governments, and AI applications.
Expert in LLM inference and serving systems, optimizing throughput/latency/cost for open-source LLMs, deep GPU architecture knowledge, and experience with PyTorch, Triton, CUDA and inference engines like vLLM/SGLang/TensorRT.
SGLang, vLLM, TensorRT, PyTorch, Triton, CuTe, CUDA
2mo
Save
Mark Applied
Hide
Principal LLM Inference Engineer
Santa Clara, California, United States
$195k-$285k/yr HybridFull Time
d-Matrix
d-Matrix: Private AI infrastructure serving data centers with inference accelerators, networking, and software.
10+ YOEBachelor's in CS/EE (or equivalent) with 10+ years experience (Master/PhD with 6+ years preferred); strong Python and C/C++; experience optimizing LLM inference, quantization, batching, GPU kernel programming and contributor-level work on inference frameworks.
Python, C, C++, vLLM, SGLang, TensorRT-LLM, ONNX Runtime, CUDA, Triton, JAX
2w
Save
Mark Applied
Hide
LLM Application Engineer
United States
RemoteFull Time
Bjak
Bjak: Southeast Asia's leading online insurance platform and fintech.
Strong software engineering fundamentals, hands-on LLM or generative AI experience, production-quality coding, prompt and workflow design, evaluation experience, and problem-solving in ambiguous environments.
Python, OpenAI, PyTorch, JAX
2mo
Save
Mark Applied
Hide
AI LLM Engineer
Atlanta, Georgia, United States
$94k-$129k/yr HybridFull Time
Varian Medical Systems
Varian Medical Systems: American medical-device and software manufacturer serving hospitals and cancer clinics with radiotherapy and oncology-care solutions.
4+ YOEBachelor's degree in CS/Data Science or similar, 4+ years in AI/ML or data engineering, production experience with GenAI/LLM solutions, strong Python, Snowflake, SQL, and experience with enterprise analytics (Power BI/Fabric).
Python, Snowflake, SQL, LangChain, LangGraph, Semantic Kernel, Microsoft Copilot Studio, Azure AI Foundry, Azure Open AI, Databricks, Microsoft Power BI, Microsoft Fabric, Microsoft OneLake, Direct Lake, Microsoft Power Automate, Microsoft Power Apps, ServiceNow, Salesforce
2mo
Save
Mark Applied
Hide
AI LLM Engineer
Atlanta, Georgia, United States
$94k-$129k/yr HybridFull Time
Varian Medical Systems
Varian Medical Systems: American medical-device and software manufacturer serving hospitals and cancer clinics with radiotherapy and oncology-care solutions.
4+ YOE4+ years in AI/ML or data engineering; experience building GenAI/LLM solutions, RAG, embeddings, vector DBs; strong SQL, Python, cloud (Azure), and enterprise data platform experience (Snowflake, Power BI/Fabric).
Python, LangChain, LangGraph, Semantic Kernel, Microsoft Copilot Studio, Azure AI Foundry, Azure Open AI, Databricks, Snowflake, Microsoft Power BI, Microsoft Fabric, OneLake, Direct Lake, SQL, ServiceNow, Salesforce, Microsoft Power Automate, Microsoft Power Apps
1mo
Save
Mark Applied
Hide
Senior AI/LLM Engineer
United States or Canada
$148k-$201k/yr RemoteFull Time
Censys
Censys: Internet intelligence and cybersecurity software providing threat hunting and attack-surface management to governments and enterprises.
5+ YOE5+ years software engineering experience with 2+ years building AI-powered user-facing features; Python proficiency; experience with LLMs, RAG pipelines, prompt engineering, secure AI practices, CI/CD and test automation.
Python, RAGAS, LangSmith, retrieval-augmented generation (RAG), vector search, CI/CD
2mo
Save
Mark Applied
Hide
Principal Engineer - Context Engineering & LLM Optimization
Charlotte, North Carolina, United States
OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Global financial services and banking institution.
10+ YOE10+ years engineering experience, 5+ years designing enterprise systems, 2+ years with LLM/RAG/vector search; strong prompt, context, retrieval, and evaluation expertise; bachelor’s in CS/Engineering/Info Systems/Applied Math.
OpenAI, Azure OpenAI, Anthropic, Google Gemini, Meta Llama, Retrieval-Augmented Generation (RAG)
3w
Save
Mark Applied
Hide
Applied LLM Systems Engineer
Costa Mesa or Santa Ana
$112k-$149k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology developing AI-powered autonomous military systems.
5+ YOERequires 5+ years of software engineering, production AI delivery, Python proficiency, LLM optimization, evaluation and observability experience, and eligibility for Secret or higher US security clearance.
Python, Lattice OS, S1000D, DITA, MIL-STD-40051, CI/CD
2mo
Save
Mark Applied
Hide
AI/LLM Safety Engineer
Overland Park, Kansas, United States
RemoteFull Time
Propio Language Services
Propio Language Services: Private U.S. language-services provider delivering interpretation, translation, localization, and language technology to organizations worldwide.
4+ YOE4+ years building production software with ML/LLM security experience; production-grade Python coding, threat modeling, safety evaluation/red-teaming, and understanding of LLM risks (prompt injection, jailbreaks, data exfiltration).
Python, garak, PyRIT, promptfoo, Giskard, NeMo Guardrails, MITRE ATLAS, STRIDE, PASTA, OWASP LLM Top 10, NIST AI RMF, ISO/IEC 42001, EU AI Act, CI/CD
1mo
Save
Mark Applied
Hide
AI Security & LLM Engineer
Reston, Virginia, United States
OnsiteFull Time
AnaVation
AnaVation: Private federal IT contractor serving U.S. government agencies with intelligence, cloud, big-data, cybersecurity, and software-engineering services.
8+ YOEActive TS/SCI with CI polygraph, bachelor\u0002s and 8+ years relevant experience, IAT Level II and CSSP certifications required, proficiency in Rust or Python, experience with RAG pipelines, prompt injection defense, and token optimization.
Rust, Python, ChatGPT, Grok, Claude, Claude Code, Codex
3mo
Save
Mark Applied
Hide
Distributed LLM Inference Engineer
San Francisco or Palo Alto
$170k-$247k/yr HybridFull Time
Anyscale
Anyscale: AI infrastructure software helping developers and AI teams build, deploy, and scale machine-learning workloads with Ray.
Familiarity with running ML inference at large scale with high throughput and low latency; experience with PyTorch; solid understanding of distributed systems.
PyTorch, Ray, vLLM, TensorRT-LLM
2mo
Save
Mark Applied
Hide
LLM Researcher and Engineer
New York City, New York, United States
$450k-$1000k/yr HybridFull Time
D. E. Shaw Research
D. E. Shaw Research: Private computational-biochemistry research developing supercomputers, software, and precisely targeted drugs for disease treatment.
Expertise in large-scale ML systems, LLM architecture/training, multimodal learning, and strong Python programming; experience with distributed training, pre/post-training techniques, and model scaling.
Python