2,481 rag engineer jobs at 1,103 companies in United States
1mo
Save
Mark Applied
Hide
1mo
Senior RAG Engineer
New York or Norway or Sweden or Ireland or United States
RemoteFull Time
Newcode.ai: N AI legal-workflow platform serving law firms, Fortune 500 companies, and in-house legal teams.
5+ YOE5+ years building backend production systems with 2+ years shipping retrieval/RAG; strong Python and FastAPI; experience running vector indexes (Qdrant), embeddings, PostgreSQL/Redis, OCR-heavy ingestion and IR metrics.
InfosysNYSE: INFY: Global leader in next-generation digital services and consulting.
Experience with generative AI, LLMs and agentic frameworks; cloud AI/ML platforms (Azure,GCP,AWS); strong ML model development, data engineering, MLOps and analytics experience; bachelor’s degree or equivalent experience.
Principal Engineer - RAG Database & Embeddings Architect
Charlotte, North Carolina, United States
OnsiteFull Time
Bank of AmericaNYSE: BAC: Global financial services and banking institution.
10+ YOE10+ years software/data/platform engineering with 5+ years designing large-scale systems and 2+ years working with LLM/RAG/vector search. Experience with vector DBs, embeddings, retrieval optimization, and enterprise/regulatory environments.
Azure AI Search, Cosmos DB vector search, Pinecone, Weaviate, Milvus, OpenSearch, Elasticsearch, PostgreSQL/pgvector, OpenAI, Azure OpenAI, Cohere, Hugging Face
Principal Engineer - RAG Database & Embeddings Architect
Charlotte, North Carolina, United States
OnsiteFull Time
Bank of AmericaNYSE: BAC: Global financial services and banking institution.
10+ YOE10+ years engineering experience, 5+ years designing enterprise systems, 2+ years with LLM/RAG/vector search; strong vector DB, embeddings, retrieval, distributed systems, and API design skills; Bachelor’s in a technical field required.
Azure AI Search, Cosmos DB vector search, Pinecone, Weaviate, Milvus, OpenSearch, Elasticsearch, PostgreSQL/pgvector, OpenAI, Azure OpenAI, Cohere, Hugging Face
Lead Machine Learning Engineer - Agentic Models, LLM, RAG, GenAI
Santa Clara, California, United States
$193k-$258k/yrHybridFull Time
Eightfold AI: AI talent intelligence platform provider helping employers recruit, retain, develop, and deploy their workforce.
5+ YOESenior ML engineer with expertise in AI agents, LLMs, distributed systems; 5-7+ years of experience; strong Python and ML frameworks; AWS; Docker/Kubernetes; RAG/GenAI experience.
AI or ML Engineer - LLMs, RAG, LangChain, LangGraph
United States
OnsiteFull Time
UnitedHealth GroupNYSE: UNH: Diversified health care helping people live healthier lives.
5+ YOERequires 5+ years of software engineering experience, AI/ML or LLM application development, RAG, Agentic AI, Spring AI, LangChain, LangGraph, Java or Python, cloud, DevOps, and enterprise integration expertise.
AppleNASDAQ: AAPL: Designing and manufacturing consumer electronics, software, and digital services.
Software engineering role developing real-time search, information retrieval, query understanding, document ranking, machine learning systems, data generation, and evaluation.
Eightfold AI: AI talent intelligence platform provider helping employers recruit, retain, develop, and deploy their workforce.
6+ YOELead AI/ML engineer with 6+ years in ML, Gen AI, LLMs; strong Python, TensorFlow/PyTorch; AWS, Docker, Kubernetes; expert in agentic AI and distributed systems.
Variance: Private AI software building investigative agents for fraud, identity, risk, and compliance teams.
Experience shipping production systems, hands‑on LLM work (prompting, RAG, evals), strong software engineering, customer communication, and CS degree or equivalent.
Iris.ai: Norwegian AI platform helping regulated enterprises build, evaluate, and deploy reliable agentic RAG workflows.
5+ YOE5+ years in software development, 3+ years Python; backend web experience (Django/Flask), REST APIs, databases, cloud; ML systems interest; strong English; remote-friendly.
Tata Consultancy ServicesBSE: 532540: Global leader in IT services, consulting, and business solutions.
8+ YOE8+ years experience with Automation Anywhere, Python, LLMs/RAG, APIs, SQL Server, Power Platform; ability to design fault-tolerant production systems and work cross-functionally.
Automation Anywhere, Automation Anywhere A360, Python, LLM, RAG, Java, Jira, Microsoft Office, Microsoft Power Apps, Microsoft Power Automate, Power Platform, REST APIs, JSON, XML, Web Services, SQL Server, C#, .NET, JavaScript, VB.NET
Securian Financial: Mutual insurance holding offering life insurance, retirement plans, and investment products to North American customers.
Technical experience in AI, ML, cloud platform engineering, production MLOps, model hosting and monitoring; AWS preferred; knowledge of foundation models, vector databases, RAG, governance and regulated industries.
MaintainX: Modern maintenance and asset management software platform.
Applied GenAI experience with prompt engineering, structured outputs, RAG, evaluation datasets, production LLM features, multimodal inputs, and model quality, cost, and latency optimization.
NetSpeek: AI-powered SaaS platform helping enterprises and MSPs automate multi-vendor AV and unified communications environments.
5+ YOE5+ years ML/AI engineering experience; production LLM systems (RAG, agents), evaluation pipelines, vector DBs/embeddings, strong Python, and production reliability under latency/cost/observability constraints.
RAG, LLM, vector databases, Python, .NET, Cursor, Claude Code, GitHub Copilot, Lena
AdyenEuronext Amsterdam: ADYEN: Global financial technology platform for payments and financial services.
Product-minded architect with applied AI experience (LLMs, RAG), production-grade full-stack skills, security-by-design for financial data, strong Java and TypeScript/React/Vue expertise, and experience mentoring engineers.
Hippocratic AI: Healthcare technology developing non-diagnostic generative AI agents for health systems, payors, and pharmaceutical companies.
0+ YOEBachelor's in engineering/CS from a top university, 0–2+ years software engineering experience with production AI deployments, experience with LLMs/RAG, Python, cloud (AWS/GCP/Azure), EHR/FHIR/HL7 integrations, strong communication skills.
LLMs, RAG, Python, APIs, AWS, GCP, Azure, FHIR, HL7, EHR
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
6+ YOEBS/MS in a relevant engineering or computer science field, 6+ years building production software, Python and LLM application experience, RAG architectures, agentic workflows, and strong production engineering skills.