arXiv Machine Learning By Ankit Goyal, Jaideep Ray

Mind the Cap: Output-Budget Regimes Change the Measured Multilingual Reasoning Gap

Read the original on arXiv Machine Learning →

arXiv:2608. 04160v1 Announce Type: cross Abstract: Multilingual evaluations report accuracy at a single output-token cap, but languages need different numbers of tokens to express the same content, so the cap is a hidden experimental variable.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.