s/pmarcaLLM RESEARCH•May 1
7
votes
224
seen
LenVM 3B model hits 63% reasoning vs 6% at 200-token cap
Controlling exact output length just got treated like a first-class objective instead of a side constraint. LenVM models the remaining token budget as a per-token value function, which turns length into a dense training signal instead of something you hand-tune. In the arXiv paper, a 3B open model reports 63% reasoning accuracy at a 200-token cap versus 6% for a token-budget baseline, plus big gains on strict length-control benchmarks.
Timeline1
May 1
The LenVM paper was announced on arXiv, with claims of better precise length control and higher reasoning accuracy under fixed token budgets.
0 comments
May 1