Likeremote

Subscribe to the latest remote jobs:

  • Likeremote jobs on https://LinkedIn.com/
  • Likeremote jobs on https://telegram.org/
  • Likeremote jobs on Reddit.com
TI

Llama Developer (Generative AI / LLM Engineer)

Tranzeal Inc.
  • ๐Ÿ‡จ๐Ÿ‡ฆ Canada
  • Remote
  • 10 months ago
  • AI
  • RAG
  • Microservices
  • MLOps
  • Machine Learning
  • Python
  • FastAPI
  • LangChain
  • LlamaIndex
  • Hugging Face
  • FAISS
  • Pinecone
  • Weaviate
  • Chroma
  • LoRA
  • PyTorch
  • Docker
  • Kubernetes
  • CI/CD
  • AWS
  • GCP
  • Azure
  • vLLM
  • Ollama
  • Triton
Not scoredNo CV on file. Upload one and this job gets a score out of 100.Upload CV

Key Responsibilities
  • Build AI applications usingLlama (Llama 3 / Llama Stack / Llama API / local LLM inference).
  • Fine-tune and evaluate Llama models on proprietary and domain-specific datasets.
  • Implement Retrieval-Augmented Generation (RAG) pipelines using vector databases.
  • Develop conversational agents, copilots, or knowledge assistants for business workflows.
  • Optimize model performance via quantization, prompt engineering, and latency reduction.
  • Integrate LLM capabilities into back-end services, microservices, APIs, or cloud platforms.
  • Ensure compliance, safety, and responsible-AI standards for generated content.
  • Collaborate with Data Science, MLOps, and Product teams to deploy scalable AI products.

Required Skills & Experience
  • 3โ€“8+ years of software engineering or machine-learning experience.
  • Proven experience withLlama models (self-hosted or via Meta API).
  • Proficiency inPython (FastAPI, LangChain, LlamaIndex, Hugging Face ecosystem).
  • Experience withvector databases (FAISS, Pinecone, Weaviate, ChromaDB).
  • Strong understanding ofprompt engineering, model fine-tuning, LoRA / QLoRA.
  • Hands-on experience withGPU computing, PyTorch, Docker, Kubernetes.
  • Familiarity withMLOps practices: CI/CD for ML, model monitoring, logging.

Preferred / Nice to Have
  • Experience deploying models onAWS / GCP / Azure / on-prem GPU clusters.
  • Knowledge ofRAG architectures, knowledge graphs, and document parsing pipelines.
  • Understanding ofmodel safety, hallucination mitigation, red-team testing.
  • Experience withllama.cpp,vLLM,Ollama, orNVIDIA Triton.
  • Contributions to open-source LLMs or AI frameworks.

Llama Developer (Generative AI / LLM Engineer) ยท Tranzeal Inc.

Auto apply with Likeremote