reasoning
Posts tagged “reasoning”.
-
AI Brief, 6 October 2026: two tokens that do what reinforcement learning does
A paper published on 5 October lifts Olmo-3-7B on MATH-500 from 42.2% to 77.9% by appending two tokens to the prompt, more than reinforcement learning gained on the same model, then proves the mechanism by editing the training corpus. Reflection AI announced a 501B open-weight model with no weights. And three independent Jev evaluations landed at once.