2,481 rag engineer jobs at 1,103 companies in United States

1mo
Save
Mark Applied
Hide
Senior RAG Engineer
New York or Norway or Sweden or Ireland or United States
RemoteFull Time
Newcode.ai
Newcode.ai: N AI legal-workflow platform serving law firms, Fortune 500 companies, and in-house legal teams.
5+ YOE5+ years building backend production systems with 2+ years shipping retrieval/RAG; strong Python and FastAPI; experience running vector indexes (Qdrant), embeddings, PostgreSQL/Redis, OCR-heavy ingestion and IR metrics.
Python, FastAPI, Qdrant, PostgreSQL, Redis
1mo
Save
Mark Applied
Hide
GEN AI, RAG Engineer
Alpharetta, Georgia, United States
OnsiteFull Time
Infosys
InfosysNYSE: INFY: Global leader in next-generation digital services and consulting.
Experience with generative AI, LLMs and agentic frameworks; cloud AI/ML platforms (Azure,GCP,AWS); strong ML model development, data engineering, MLOps and analytics experience; bachelor’s degree or equivalent experience.
SAS, R, Python, BigQuery, Hadoop, SageMaker, Snowflake, AWS, Azure, GCP, CI/CD, DevOps
2mo
Save
Mark Applied
Hide
Principal Engineer - RAG Database & Embeddings Architect
Charlotte, North Carolina, United States
OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Global financial services and banking institution.
10+ YOE10+ years software/data/platform engineering with 5+ years designing large-scale systems and 2+ years working with LLM/RAG/vector search. Experience with vector DBs, embeddings, retrieval optimization, and enterprise/regulatory environments.
Azure AI Search, Cosmos DB vector search, Pinecone, Weaviate, Milvus, OpenSearch, Elasticsearch, PostgreSQL/pgvector, OpenAI, Azure OpenAI, Cohere, Hugging Face
2mo
Save
Mark Applied
Hide
Principal Engineer - RAG Database & Embeddings Architect
Charlotte, North Carolina, United States
OnsiteFull Time
Bank of America
Bank of AmericaNYSE: BAC: Global financial services and banking institution.
10+ YOE10+ years engineering experience, 5+ years designing enterprise systems, 2+ years with LLM/RAG/vector search; strong vector DB, embeddings, retrieval, distributed systems, and API design skills; Bachelor’s in a technical field required.
Azure AI Search, Cosmos DB vector search, Pinecone, Weaviate, Milvus, OpenSearch, Elasticsearch, PostgreSQL/pgvector, OpenAI, Azure OpenAI, Cohere, Hugging Face
3mo
Save
Mark Applied
Hide
Lead Machine Learning Engineer - Agentic Models, LLM, RAG, GenAI
Santa Clara, California, United States
$193k-$258k/yr HybridFull Time
Eightfold AI
Eightfold AI: AI talent intelligence platform provider helping employers recruit, retain, develop, and deploy their workforce.
5+ YOESenior ML engineer with expertise in AI agents, LLMs, distributed systems; 5-7+ years of experience; strong Python and ML frameworks; AWS; Docker/Kubernetes; RAG/GenAI experience.
Python, TensorFlow, PyTorch, AWS, Docker, Kubernetes, Kafka, AWS SQS, LangGraph, CrewAI, AutoGen, Pinecone, pgvector, vLLM, TensorRT-LLM
2w
Save
Mark Applied
Hide
AI or ML Engineer - LLMs, RAG, LangChain, LangGraph
United States
OnsiteFull Time
UnitedHealth Group
UnitedHealth GroupNYSE: UNH: Diversified health care helping people live healthier lives.
5+ YOERequires 5+ years of software engineering experience, AI/ML or LLM application development, RAG, Agentic AI, Spring AI, LangChain, LangGraph, Java or Python, cloud, DevOps, and enterprise integration expertise.
LLMs, RAG, Spring AI, LangChain, LangGraph, AWS, Azure, Java, Spring Boot, Python, OpenAI, Azure OpenAI, Anthropic, Amazon Bedrock, Git, Docker, Kubernetes, SageMaker, Azure AI Foundry, Azure AI Search, Splunk, Dynatrace, Amazon CloudWatch, Azure Monitor, Azure Log Analytics, Pinecone, Weaviate, Milvus, Chroma, FAISS, OpenSearch, REST APIs, CI/CD
1w
Save
Mark Applied
Hide
Software Engineer, Search and RAG Applications
Seattle, Washington, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designing and manufacturing consumer electronics, software, and digital services.
Software engineering role developing real-time search, information retrieval, query understanding, document ranking, machine learning systems, data generation, and evaluation.
Siri, Spotlight, Safari, Lookup, Messages
3mo
Save
Mark Applied
Hide
Staff Machine Learning Engineer - Agentic Models, LLM, RAG, GenAI
Santa Clara, California, United States
$232k-$310k/yr HybridFull Time
Eightfold AI
Eightfold AI: AI talent intelligence platform provider helping employers recruit, retain, develop, and deploy their workforce.
6+ YOELead AI/ML engineer with 6+ years in ML, Gen AI, LLMs; strong Python, TensorFlow/PyTorch; AWS, Docker, Kubernetes; expert in agentic AI and distributed systems.
Python, TensorFlow, PyTorch, AWS, Docker, Kubernetes
3mo
Save
Mark Applied
Hide
Forward Deployed Engineer
Mountain View, California, United States
OnsiteFull Time
AI Fund
AI Fund: AI Fund is a privately held venture studio that-founds early-stage artificial-intelligence companies with entrepreneurs and partners.
Strong AI engineering with LLMs, agents, RAG; client-facing, collaborative; travel.
Claude Code, Codex, LLMs, AGENTS, RAG systems
1mo
Save
Mark Applied
Hide
Agent Engineer
San Francisco, California, United States
OnsiteFull Time
Variance
Variance: Private AI software building investigative agents for fraud, identity, risk, and compliance teams.
Experience shipping production systems, hands‑on LLM work (prompting, RAG, evals), strong software engineering, customer communication, and CS degree or equivalent.
LLMs, RAG
3mo
Save
Mark Applied
Hide
Senior AI Engineer
United States
$180k-$245k/yr RemoteFull Time
Cadence
Cadence: Private clinical AI and remote care delivering chronic disease monitoring and treatment support to older adults.
5+ YOE5+ years software engineering; 2+ years AI/ML in production; LLM APIs; RAG; end-to-end ownership.
LLM APIs, OpenAI, Anthropic, embedding models, vector stores, RAG, SFT, RLHF, LoRA
2mo
Save
Mark Applied
Hide
Senior Engineer, AI
New York City, New York, United States
$215k-$265k/yr OnsiteFull Time
CLARK
CLARK: Private European digital insurance broker helping consumers manage, compare, and improve coverage through technology and expert advice.
5+ YOE5+ years software engineering experience; production software systems; experience with LLMs, agents, RAG, workflow orchestration; strong backend/full-stack fundamentals and product sensibility.
LLMs, RAG, evals, agent frameworks
3mo
Save
Mark Applied
Hide
Senior AI Engineer - RAG & AI Agents
United States
RemoteFull Time
Iris.ai
Iris.ai: Norwegian AI platform helping regulated enterprises build, evaluate, and deploy reliable agentic RAG workflows.
5+ YOE5+ years in software development, 3+ years Python; backend web experience (Django/Flask), REST APIs, databases, cloud; ML systems interest; strong English; remote-friendly.
Python, Django, Flask, REST, PostgreSQL, Elasticsearch, OpenSearch, ChromaDB, Couchbase, Gensim, Torch, Transformers, spaCy, Autogen, AWS, Windsurf, Cursor, Amazon Q, Claude, Git, CI/CD
1mo
Save
Mark Applied
Hide
Engineer
Pittsburgh, Pennsylvania, United States
$100k-$105k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesBSE: 532540: Global leader in IT services, consulting, and business solutions.
8+ YOE8+ years experience with Automation Anywhere, Python, LLMs/RAG, APIs, SQL Server, Power Platform; ability to design fault-tolerant production systems and work cross-functionally.
Automation Anywhere, Automation Anywhere A360, Python, LLM, RAG, Java, Jira, Microsoft Office, Microsoft Power Apps, Microsoft Power Automate, Power Platform, REST APIs, JSON, XML, Web Services, SQL Server, C#, .NET, JavaScript, VB.NET
6d
Save
Mark Applied
Hide
AI Engineer (Hybrid)
Saint Paul, Minnesota, United States
$72k-$134k/yr HybridFull Time
Securian Financial
Securian Financial: Mutual insurance holding offering life insurance, retirement plans, and investment products to North American customers.
Technical experience in AI, ML, cloud platform engineering, production MLOps, model hosting and monitoring; AWS preferred; knowledge of foundation models, vector databases, RAG, governance and regulated industries.
AWS, MLOps, RAG
1w
Save
Mark Applied
Hide
Applied AI Engineer
Canada or San Francisco or Raleigh or Boston
HybridFull Time
MaintainX
MaintainX: Modern maintenance and asset management software platform.
Applied GenAI experience with prompt engineering, structured outputs, RAG, evaluation datasets, production LLM features, multimodal inputs, and model quality, cost, and latency optimization.
LLMX, Attachments, OCR, RAG
3mo
Save
Mark Applied
Hide
Senior AI Engineer
United States or Canada
RemoteFull Time
NetSpeek
NetSpeek: AI-powered SaaS platform helping enterprises and MSPs automate multi-vendor AV and unified communications environments.
5+ YOE5+ years ML/AI engineering experience; production LLM systems (RAG, agents), evaluation pipelines, vector DBs/embeddings, strong Python, and production reliability under latency/cost/observability constraints.
RAG, LLM, vector databases, Python, .NET, Cursor, Claude Code, GitHub Copilot, Lena
1mo
Save
Mark Applied
Hide
Staff Engineer - Intelligent Customer Experiences
San Francisco, California, United States
$258k-$367k/yr OnsiteFull Time
Adyen
AdyenEuronext Amsterdam: ADYEN: Global financial technology platform for payments and financial services.
Product-minded architect with applied AI experience (LLMs, RAG), production-grade full-stack skills, security-by-design for financial data, strong Java and TypeScript/React/Vue expertise, and experience mentoring engineers.
Java, TypeScript, React, Vue, LLMs, RAG
2mo
Save
Mark Applied
Hide
Forward Deployment Engineer
Menlo Park, California, United States
OnsiteFull Time
Hippocratic AI
Hippocratic AI: Healthcare technology developing non-diagnostic generative AI agents for health systems, payors, and pharmaceutical companies.
0+ YOEBachelor's in engineering/CS from a top university, 0–2+ years software engineering experience with production AI deployments, experience with LLMs/RAG, Python, cloud (AWS/GCP/Azure), EHR/FHIR/HL7 integrations, strong communication skills.
LLMs, RAG, Python, APIs, AWS, GCP, Azure, FHIR, HL7, EHR
1mo
Save
Mark Applied
Hide
Senior SOCD Applied AI Engineer
Santa Clara, California, United States
$168k-$311k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
6+ YOEBS/MS in a relevant engineering or computer science field, 6+ years building production software, Python and LLM application experience, RAG architectures, agentic workflows, and strong production engineering skills.
Python, Claude Code, OpenAI Codex, Cursor, LangChain, LlamaIndex, React, TypeScript, FastAPI, MCP (Model Context Protocol), CI/CD, RAG