12 inference optimization engineer jobs at 3 companies in Virginia
2mo
Save
Mark Applied
Hide
2mo
Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform)
San Jose or San Francisco or New York City or Cambridge or McLean
$230k-$286k/yrOnsiteFull Time
Capital OneNYSE: COF: Provides credit card, banking, and auto loan services.
6+ YOEBachelor's plus 6 years or master's plus 4 years developing AI/ML technologies, and 6 years programming with Python, Go, Scala, or Java. Cloud AI deployment and team leadership are preferred.
Senior Lead AI Engineer (FM Hosting, LLM Inference)
New York or McLean or Cambridge or San Jose
$251k-$286k/yrOnsiteFull Time
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
6+ YOEBachelor's in CS/AI/EE/CE or related with 6+ years (or Master's with 4+ years); 6+ years programming with Python, Go, Scala, or Java; experience deploying scalable AI systems, LLM inference, similarity search, and optimization of training/inference.
Sr. Lead AI Engineer (Inference Optimization, FM hosting, AI Platform)
San Jose or San Francisco or New York City or Cambridge or McLean
$230k-$286k/yrOnsiteFull Time
Capital OneNYSE: COF: A diversified financial services providing banking and credit products.
6+ YOEBachelor's degree plus 6 years or master's degree plus 4 years developing AI/ML technologies; 6 years programming with Python, Go, Scala, or Java; cloud AI deployment experience preferred.
Innovative Defense Technologies: A defense technology that develops automated software and systems to rapidly deliver mission-critical capabilities to government customers.
5+ YOEDesign and deliver production-grade on‑prem AI/agent systems, integrate LLMs and tools, build RAG pipelines, optimize local inference, and obtain/maintain Secret clearance.
Senior Lead AI Engineer (GenAI Platform, Agentic Infrastructure)
New York or San Francisco or McLean or Cambridge or San Jose or Plano
$209k-$286k/yrOnsiteFull Time
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's in CS/AI/EE/CE +6 years or Master's +4 years; 6+ years programming with Python/Go/Scala/Java; experience deploying scalable AI on cloud; LLM, inference, similarity search, VectorDBs, guardrails, model evaluation, and optimization experience; leadership and research literacy.
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
6+ YOEDegree in CS/AI/EE or related plus 6+ years AI/ML engineering experience (or MS+4 yrs). Strong Python/Go/Scala/Java skills, cloud AI deployment experience, LLM/Inference and optimization expertise, leadership and communication skills.
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's in CS/AI/EE/CE plus 6+ years, or Master's plus 4+ years; 6+ years programming with Python, Go, Scala or Java; experience deploying scalable AI systems, LLM inference, vector databases, and optimization techniques; strong communication and leadership.
Senior Lead AI Engineer(MLX, Agentic AI, Gen AI platform Services)
San Jose or San Francisco or McLean or Cambridge or New York City
$230k-$286k/yrOnsiteFull Time
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
6+ YOE6+ years developing AI/ML systems (or 4+ with a master\u0002s), strong software and math foundations, experience with LLM inference, similarity search, and optimization, Python/Go/Scala/Java proficiency, cloud deployment experience.
Lead AI Engineer (AI Foundations, LLM Core and Agentic AI)
Cambridge or McLean or New York or San Jose or United States
$197k-$246k/yrOnsiteFull Time
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
2+ YOEBachelor's in CS/AI/EE/CE (4+ yrs) or Master's (2+ yrs). 4+ yrs programming in Python, Go, Scala, or Java. Experience with LLMs, model training/inference, vector search, cloud deployments, and AI system optimization.
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
7+ YOE7+ years designing and operating distributed applications; 5+ years customer-facing experience building production GenAI/ML systems; hands-on LLM, RAG, agentic workflows, prompt engineering, inference optimization; experience with AWS AI/ML services.