🚀NEW LABGetting Started with Claude AgentsStart lab
Reinforcement Learning · Data

Following Length Constraints in Instructions

First page
Following Length Constraints in Instructions
Paper summary

presents an approach for how to deal with length bias and train instruction following language models that better follow length constraint instructions; fine-tunes a model using DPO with a length instruction augmented dataset and shows less length constraint violations and while keeping a high response quality.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack