Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling
Apple researchers have introduced the Length Value Model, a framework designed to enable fine-grained control over generation length in autoregressive models by integrating length modeling at the token level. This approach allows developers to explicitly manage output length during the training phase, which directly impacts inference costs and overall reasoning efficiency. By addressing the current lack of precise length control, this method provides a scalable way to optimize how models utilize computational resources during text generation.
Covered by 1 source
- AApple Machine Learning Blog↗2d ago