NestifyTech Solutions
Posted 1y ago

LLM Engineer (Large Language Model Engineer)

NestifyTech Solutions
United States
₹2000k-₹40000k/yrRemoteFull Time
Responsibilities
  • Fine-tune models
  • Build pipelines
  • Design RAG pipelines
Requirements
  • Expert in NLP/ML with hands-on LLM experience
  • Strong Python and ML framework skills
Technical tools mentioned
Hugging FaceOpenAI APIsLangChainPyTorchTensorFlowPython

Job description

At NestifyTech Solutions, looking for a highly skilled LLM Engineer to develop, fine-tune, and integrate large language models (LLMs) into enterprise-grade applications. This role requires deep expertise in NLP, model architecture, prompt engineering, and deployment of LLMs at scale.

Responsibilities:

  • Fine-tune and optimize LLMs (e.g., GPT, LLaMA, Mistral) for specific business tasks.

  • Build pipelines for data preprocessing, training, evaluation, and inference.

  • Design retrieval-augmented generation (RAG) pipelines using vector databases.

  • Integrate LLMs with downstream systems using APIs or microservices.

  • Implement prompt engineering techniques for few-shot and zero-shot learning.

  • Benchmark model performance and apply techniques for cost optimization and latency reduction.

  • Ensure safety, fairness, and compliance in model outputs.

Required Qualifications:

  • Advanced degree (MSc/PhD preferred) in AI, Machine Learning, Computer Science, or related field.

  • 3+ years in ML/NLP roles; 1+ year hands-on with LLMs.

  • Strong experience with Hugging Face, OpenAI APIs, LangChain, or similar.

  • Deep understanding of Transformer architecture and attention mechanisms.

  • Familiarity with fine-tuning methods (LoRA, QLoRA, PEFT, RLHF).

  • Proficiency with Python, PyTorch or TensorFlow.

Preferred Qualifications:

  • Experience deploying models in production using cloud services (AWS/GCP/Azure).

  • Knowledge of vector stores (e.g., Pinecone, Weaviate, FAISS).

  • Contributions to open-source LLM frameworks or research publications.

About NestifyTech Solutions

Independent, enterprise-grade AI solutions provider.

Similar jobs

LLM Engineer roles
4mo
Save
Mark Applied
Hide
Principal LLM Engineer (Productionization)
United States
OnsiteFull Time
A5 Labs
A5 Labs: Develops AI and blockchain solutions for online gaming integrity.
Experience scaling LLM systems in production with strong infrastructure and deployment background.
6mo
Save
Mark Applied
Hide
Applied LLM Engineer
San Francisco, California, United States
$180k-$200k/yr HybridFull Time
Artos
Artos: AI-powered regulatory document automation platform for life sciences.
4+ YOEProduction-grade LLM system experience, strong backend Python (FastAPI/Django), AWS/cloud scaling, Terraform/Pulumi, multi-step LLM workflow and agent design; 4+ years experience preferred.
Python, FastAPI, Django, AWS, Terraform, Pulumi, React
6d
Save
Mark Applied
Hide
LLM Application Engineer
United States
RemoteFull Time
BJAK
BJAK: Online platform for insurance comparison and road tax renewal.
Strong software engineering fundamentals, hands-on LLM or generative AI experience, production-quality coding, prompt and workflow design, evaluation experience, and problem-solving in ambiguous environments.
Python, OpenAI, PyTorch, JAX
1w
Save
Mark Applied
Hide
LLM Backend Engineer Graduate (Applied Machine Learning) - 2027 Start
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Bachelor's or master's degree in computer science or related field; proficiency in Golang, Java, C++, or Python; software development experience; and knowledge of databases, networks, operating systems, and distributed systems.
Golang, Java, C++, Python, Kubernetes, Docker, Istio, Envoy, Service Mesh, Function Calling, MCP
2w
Save
Mark Applied
Hide
Senior AI/LLM Engineer
United States or Canada
$148k-$201k/yr RemoteFull Time
Censys
Censys: Maps the internet to discover and manage security risks.
5+ YOE5+ years software engineering experience with 2+ years building AI-powered user-facing features; Python proficiency; experience with LLMs, RAG pipelines, prompt engineering, secure AI practices, CI/CD and test automation.
Python, RAGAS, LangSmith, retrieval-augmented generation (RAG), vector search, CI/CD
1mo
Save
Mark Applied
Hide
AI/LLM Engineer
New York, New York, United States
$120k-$130k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
8+ YOEStrong Python engineering, experience with agent frameworks (LangChain/LangGraph), EMR/data platform integration (Snowflake, Databricks), 8+ years experience, bachelor in computer science.
Python, LangChain, LangGraph, Snowflake, Databricks
1mo
Save
Mark Applied
Hide
Senior/Staff LLM Application Engineer - Data Application
San Jose, California, United States
$213k-$450k/yr OnsiteFull Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Experience with data products and LLM application development, strong coding skills in Python, knowledge of prompt engineering, retrieval and benchmarking, and ability to analyze user feedback.
Python
1mo
Save
Mark Applied
Hide
Principal LLM Inference Engineer
Santa Clara, California, United States
$195k-$285k/yr HybridFull Time
d-Matrix: Develops high-performance semiconductor chips for generative AI inference.
10+ YOEBachelor's in CS/EE (or equivalent) with 10+ years experience (Master/PhD with 6+ years preferred); strong Python and C/C++; experience optimizing LLM inference, quantization, batching, GPU kernel programming and contributor-level work on inference frameworks.
Python, C, C++, vLLM, SGLang, TensorRT-LLM, ONNX Runtime, CUDA, Triton, JAX