27 inference engineer jobs at 9 companies in Fredericksburg, VA
1d
Save
Mark Applied
Hide
1d
Machine Learning Performance Engineer - Offboard Training & Inference
Sunnyvale or Washington, D.C. or San Diego or Fort Walton Beach or Ann Arbor or London or Stuttgart or Munich or Stockholm or Bangalore or Seoul or Tokyo
$215k-$285k/yrOnsiteFull Time
Applied Intuition: Developing software and simulation infrastructure for autonomous vehicles.
ML performance engineering experience with distributed training, batch inference, GPU or accelerator optimization, Python, and C++ or another systems language; strong debugging and analytical skills required.
Senior Lead AI Engineer (FM Hosting, LLM Inference)
New York or McLean or Cambridge or San Jose
$251k-$286k/yrOnsiteFull Time
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
6+ YOEBachelor's in CS/AI/EE/CE or related with 6+ years (or Master's with 4+ years); 6+ years programming with Python, Go, Scala, or Java; experience deploying scalable AI systems, LLM inference, similarity search, and optimization of training/inference.
Senior Lead AI Engineer (FM Hosting, LLM Inference)
New York City or McLean or San Jose or Cambridge
$230k-$286k/yrOnsiteFull Time
Capital OneNYSE: COF: Provides credit card, banking, and auto loan services.
6+ YOEBachelor’s plus 6 years or master’s plus 4 years in AI/ML development; 6 years programming in Python, Go, Scala, or Java; cloud AI deployment and engineering leadership preferred.
Santa Clara or Washington or Texas or New York or Washington or Massachusetts
$184k-$357k/yrRemoteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
6+ YOE3+ MgmtRequires MS, PhD, or equivalent experience; 6+ years software development; 3+ years technical leadership or engineering management; C/C++, GPU programming, performance optimization, and production deep learning deployment.
Seekr: Transparent AI platform for enterprise and government decision-making.
8+ YOE8+ years building distributed systems and cloud-native AI infrastructure; strong Python and systems programming skills; Kubernetes, GPU inference, and platform engineering experience; leadership and architecture experience.
UnitedHealth GroupNYSE: UNH: Provides health insurance and technology-enabled health care services.
3+ YOEBachelor's in CS or quantitative field; 3+ years programming (Python, SQL, packaging), experience with observability, model registry/experiment tracking, responsible AI, GenAI/agents, serving/inference, and cloud platforms.
Bana Solutions: Provides software engineering, cybersecurity, and data solutions for government agencies.
6+ YOEOwnership of full production ML lifecycle, model serving for low-latency inference, monitoring/drift detection, automated retraining pipelines, Kubernetes/Docker deployment, strong Python and software engineering skills.
Booz Allen HamiltonNYSE: BAH: Provides technology and management consulting services to diverse organizations.
3+ YOE3+ years building production software in Python, integrating services/APIs and LLMs, implementing inference flows, writing integration tests, and knowledge of HTTP auth patterns and container/CI/CD tooling.
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
7+ YOE7+ years designing and operating distributed applications; 5+ years customer-facing experience building production GenAI/ML systems; hands-on LLM, RAG, agentic workflows, prompt engineering, inference optimization; experience with AWS AI/ML services.
W. R. BerkleyNYSE: WRB: Provides commercial property, casualty, and specialty insurance and reinsurance.
10+ YOE10+ years building and shipping production ML/AI systems; expert Python and ML engineering; experience with LLMs, MLOps, causal inference, and statistical methods; Bachelor's in quantitative field required; advanced degree preferred.