OpenAI launches ChatGPT for Teens as default for under-18 users
ChatGPT for Teens is now the default for under-18 users, pairing Study Mode and parental controls with safeguards OpenAI says are on by default.
Read MoreIndependent AI news, with every claim linked to its primary source
Independent AI news, with every claim linked to its primary source
New models, benchmarks and research from the labs.
ChatGPT for Teens is now the default for under-18 users, pairing Study Mode and parental controls with safeguards OpenAI says are on by default.
Read MoreTogether AI ran 904 DeepSWE rollouts across two matchups. Running DeepSeek V4 Pro 0813 first and escalating on failure beat both flagships on accuracy and on cost.
Read MoreA question-driven framework ran 5,400 generated PubMed queries and surfaced 17 research cohorts that four established catalogues never returned. Two arXiv papers this month push LLMs at scientific text extraction.
Read MoreA plugin that attaches to a frozen vision-language-action policy at inference time reports 10% to 25% better full-sequence success on 8-step composed lab tasks, with no retraining.
Read MoreA position paper accepted to the ICML 2026 Position Paper Track argues that the way the field measures AI is
Read MoreA new arXiv paper scores AI agents on behavioral consistency, not just success rate, as Anthropic documents sabotage and conformity in multi-agent tests.
Read MoreAnthropic says Claude’s text watermark is a version of Google DeepMind’s SynthID-Text approach, adds no extra tokens, and that a detection API is on the way. Light editing probably won’t strip it.
Read MoreOpenAI has switched on a feature called Computer History in the ChatGPT app for macOS, and it records what you
Read MoreGoogle is letting you switch off the visible watermark on images, video and music generated in Gemini. The company announced
Read MoreThe distance between those two numbers is the most honest thing published about AI progress this year.
Read MoreWhen six labs are within 5% of each other, the question stops being which model is best and becomes which one you can afford to run.
Read MoreSWE-bench Verified went from 60% to near 100% in a year. A test everyone passes measures nothing.
Read More