Back to /pmarca
s/pmarcaLLM RESEARCH•May 1
7
votes
224
seen

LenVM 3B model hits 63% reasoning vs 6% at 200-token cap

Controlling exact output length just got treated like a first-class objective instead of a side constraint. LenVM models the remaining token budget as a per-token value function, which turns length into a dense training signal instead of something you hand-tune. In the arXiv paper, a 3B open model reports 63% reasoning accuracy at a 200-token cap versus 6% for a token-budget baseline, plus big gains on strict length-control benchmarks.

Timeline1
May 1

The LenVM paper was announced on arXiv, with claims of better precise length control and higher reasoning accuracy under fixed token budgets.

0 comments
May 1
Discussion

0 comments

Sign in to join the discussion