
LLM Data Scientist/Algorithm Engineer (Fully Remote)
Binance
🇧🇩 Bangladesh | 🇨🇳 China | 🇮🇩 Indonesia | 🇮🇳 India | 🇮🇷 Iran | 🇯🇴 Jordan | 🇯🇵 Japan | 🇰🇷 South Korea | 🇰🇿 Kazakhstan | 🇱🇦 Laos | 🇱🇧 Lebanon | 🇱🇰 Sri Lanka | 🇲🇲 Myanmar | 🇲🇻 Maldives | 🇲🇾 Malaysia | 🇳🇵 Nepal | 🇴🇲 Oman | 🇵🇠Philippines | 🇵🇰 Pakistan | 🇸🇬 Singapore | 🇹🇠Thailand | 🇹🇷 Turkey | 🇹🇼 Taiwan | 🇻🇳 Vietnam | 🇾🇪 Yemen
Remote
48 months ago
- AI
- Large Language Models
- RAG
- vLLM
- LangGraph
- CrewAI
- AutoGen
- Natural Language Processing
48 months ago
We are seeking a highly skilled professional to join our team, focusing on advancing throughinnovative AI solutions.  The successful candidate willdevelop and refine Large Language Models (LLMs) to extract actionable insights, improve business decision-making, and optimize prompt design for more accurate outputs. Additionally, the role includes creating scalable and robust LLM/RAG frameworks tailored to customer service scheduling, fostering innovation and maintaining a competitive market edge. This role is 100% Remote, Work from Home based.
Responsibilities
- Own the full LLM pipeline from data preparation to production real case usage.
- Design, iterate and optimize prompts (zero-/few-shot, chain-of-thought, tool-calling, etc.) to maximize model utility and safety across products and languages.
- Build and maintain Retrieval-Augmented Generation (RAG) QA/search systems that connect to multi-source knowledge bases.
- Familiar with vLLM/SGLang inference architectures and have proven experience deploying and operating LLM services on multi‑GPU or cluster environments.
- Design, implement and operate multi‑agent LLM architectures (e.g. LangGraph, CrewAI, AutoGen) including task decomposition, agent orchestration, memory sharing and tool‑calling workflows.
- Develop evaluation pipelines (automatic metrics & human feedback) to measure prompt and model quality, bias, and hallucination rates.
- Collaborate with product and CS teams to integrate AI models into conversational Chatbot in different scenarios.
- Track cutting-edge research, author tech blogs, and keep improve current architecture.Â
Requirements
- Master’s Degree or higher in Computer Science, Data Science or related field..
- At least 2 years of deep-learning/NLP experience, including1+ year practical LLM work (SFT, DPO, RAG, quantization, inference optimization, etc.).
- Demonstratedprompt engineering & tuning expertise (few-shot design, structured prompting, prefix-/p-tuning, reward re-ranking, safety filtering).
- Practical experience building and deploying multi‑agent LLM workflows, with understanding of agent‑orchestrator patterns, shared memory, long‑horizon planning and guard‑rail design.
- Proficient in both English and Chinese communication for efficient cross team collaboration
Binance is committed to being an equal opportunity employer. We believe that having a diverse workforce is fundamental to our success.By submitting a job application, you confirm that you have read and agree to ourCandidate Privacy Notice.
LLM Data Scientist/Algorithm Engineer (Fully Remote) · Binance