For all their remarkable fluency with language, code, and complex reasoning, the latest generation of AI systems — including GPT-5 — have stumbled on one of psychology's most enduring measures of basic cognition: the sustained attention test. Researchers found not a gradual decline but a near-total collapse in performance as cognitive demands increased, revealing that the ability to hold focus over time may rest on principles fundamentally different from those powering today's most capable models. The finding invites a quieter, more unsettling question beneath the headlines about AI progress:
Advanced AI Models Collapse on Classic Psychology Test, Raising Questions About Human-Level AI
Related Coverage
A woman was secretly filmed by someone wearing Meta's AI smart glasses in a viral prank video, raising concerns about we…
CBS News · Aug 21 Consumer groups urge FTC probe into AI firms' 'hoard-and-destroy' book practicesConsumer advocacy groups urge the FTC to investigate AI developers for allegedly buying, scanning, and destroying millio…
BBC News · Aug 21 Ofcom investigates Sky News over Farage family privacy claimsOfcom has launched an investigation into Sky News following harassment complaints by Reform UK leader Nigel Farage, who …
Pocket-lint · Aug 21 Amazon's Fire OS 16 Update Bypasses Fire Sticks EntirelyAmazon's new Fire OS 16 update will only launch on smart TVs, not Fire Sticks, as the company transitions all future sti…
Bias & Framing
No detailed analysis data available for this lens. Try re-running lenses from the admin panel.
Geopolitical Impact
AI capability gaps in sustained attention tasks have minimal geopolitical impact; primarily a technical limitation affecting AI development timelines rather than international power dynamics.
No direct shifts. Indirectly relevant: countries investing heavily in AI (US, China, EU) may experience delayed timelines for AGI-dependent strategic advantages, potentially extending current technological leadership windows.
Economic Lens
AI model limitations on attention tasks may delay human-level AI development, potentially affecting AI-dependent sectors and investment timelines in the technology industry.
Consumers may experience delayed deployment of fully autonomous AI systems in critical applications (healthcare, autonomous vehicles, financial services), potentially extending reliance on human oversight and hybrid AI-human solutions longer than anticipated.
Regulators may use these findings to justify more cautious AI deployment frameworks and extended testing requirements before approving AI systems for high-stakes applications. Could support arguments for stronger AI safety regulations and slower rollout timelines.