Test-time compute: paying for thinking instead of training
Compute spent thinking can substitute for compute spent training. That trade reorganised the field, and moved the bill onto you.
Read MoreIndependent AI news, with every claim linked to its primary source
Independent AI news, with every claim linked to its primary source
New models, benchmarks and research from the labs.
Compute spent thinking can substitute for compute spent training. That trade reorganised the field, and moved the bill onto you.
Read MoreAgents that work for five steps fall apart over fifty. The research says the problem is what they are carrying, not how clever they are.
Read MorePublishing weights used to mean you had lost the frontier. Now it means you want the distribution.
Read MoreDaybreak now runs Blue and Red tiers behind approval. The gate says more about the capability than the launch does.
Read MoreGoogle says its new entry-level model beat Anthropic and OpenAI across nine benchmarks. The cadence matters more than the scores.
Read MoreMuse Glimmer ships under Apache 2.0, sized for consumer hardware. Zuckerberg paired it with a 6,500-word case against closed AI.
Read MoreGrok 4.6 scored 61 on the AA Intelligence Index, level with GPT-5.6 Sol Max, while charging $6 per million output tokens against $30.
Read More