
LLM Engineer (Large Language Model Engineer)
- Large Language Models
- Natural Language Processing
- GPT
- RAG
- Microservices
- AI
- Machine Learning
- Hugging Face
- OpenAI
- LangChain
- LoRA
- RLHF
- Python
- PyTorch
- TensorFlow
- AWS
- GCP
- Azure
- Pinecone
- Weaviate
- FAISS
At NestifyTech Solutions, looking for a highly skilledLLM Engineer to develop, fine-tune, and integrate large language models (LLMs) into enterprise-grade applications. This role requires deep expertise in NLP, model architecture, prompt engineering, and deployment of LLMs at scale.
Responsibilities:
Fine-tune and optimize LLMs (e.g., GPT, LLaMA, Mistral) for specific business tasks.
Build pipelines for data preprocessing, training, evaluation, and inference.
Design retrieval-augmented generation (RAG) pipelines using vector databases.
Integrate LLMs with downstream systems using APIs or microservices.
Implement prompt engineering techniques for few-shot and zero-shot learning.
Benchmark model performance and apply techniques for cost optimization and latency reduction.
Ensure safety, fairness, and compliance in model outputs.
Required Qualifications:
Advanced degree (MSc/PhD preferred) in AI, Machine Learning, Computer Science, or related field.
3+ years in ML/NLP roles; 1+ year hands-on with LLMs.
Strong experience with Hugging Face, OpenAI APIs, LangChain, or similar.
Deep understanding of Transformer architecture and attention mechanisms.
Familiarity with fine-tuning methods (LoRA, QLoRA, PEFT, RLHF).
Proficiency with Python, PyTorch or TensorFlow.
Preferred Qualifications:
Experience deploying models in production using cloud services (AWS/GCP/Azure).
Knowledge of vector stores (e.g., Pinecone, Weaviate, FAISS).
Contributions to open-source LLM frameworks or research publications.
LLM Engineer (Large Language Model Engineer) ยท Nestifytech Solutions Private Limited