Binance
LLM Applied Data Scientist (RAG/ NLP)
Remotefull timeunspecifiedfinance
Large Language Models (LLM) — Retrieval-Augmented Generation (RAG) — Natural Language Processing (NLP) — Reinforcement Learning — Python — Machine Learning — Reward Modeling — Multi-agent Systems
Description
About the Role We are seeking a highly skilled Research Scientist/Engineer to advance the reasoning and planning capabilities of large foundation models. In this role, you will enhance model performance across the entire development lifecycle—including data acquisition, supervised fine-tuning (SFT), reward modelling, and reinforcement learning—while driving innovations in reasoning and decision-making. You will synthesise large-scale, high-quality datasets through rewriting, augmentation, and generation techniques to strengthen foundation models during pretraining, SFT, and RL stages. A key part of the role involves solving complex tasks using System 2 thinking and applying advanced decoding strategies such as MCTS and A*. You will design and implement robust evaluation methodologies, teach models to interact with external tools, APIs, and code interpreters, and build agents and multi-agent systems capable of addressing sophisticated real-world problems.