Subject
2 entries
Instruction Following
Bookmarks
TART: Task-Aware Retrieval with Instructions
TART (Task-Aware Retrieval with Instructions) introduces BERRI, a dataset of ~40 retrieval tasks annotated with human-written task instructions, and trains a multi-task retrieval system that adapts its behavior based on explicit instructions. TART outperforms much larger models on BEIR by understanding the user's intent rather than just matching query-document similarity.
Training Language Models to Follow Instructions with Human Feedback (InstructGPT)
OpenAI's InstructGPT paper (2022) shows that a 1.3B model fine-tuned with RLHF on human preference data is preferred over raw GPT-3 at 175B — establishing that alignment via human feedback is more important than raw scale for following instructions. This is the foundational paper behind ChatGPT and the instruction-tuned era.
