# Even GPT-5 Failed This Human Attention Test

> **Open Intelligence Dossier** · First detected: 2026-06-16 00:07 UTC · Category: Science

## Executive Summary
New scientific research reveals that even advanced models like GPT-5 struggle with the Stroop test as task length increases.

## Intelligence Brief
A recent scientific study has identified a significant limitation in the cognitive capabilities of advanced artificial intelligence. According to reports from SciTechDaily and Caliber.Az, these AI models, including the high-profile GPT-5, have failed a specific human attention test known as the Stroop test. The findings indicate that the performance of these models degrades specifically as the length of the task increases, suggesting a failure to maintain a consistent level of attention or processing accuracy over extended sequences of information. Coverage from Baku.ws emphasizes that scientists have now identified what they describe as the main weakness of artificial intelligence through this specific testing methodology.


SciTechDaily further highlights that the failure occurs even in the most advanced iterations of large language models, such as GPT-5, which were previously expected to handle such cognitive tasks with greater ease. Caliber.Az provides the specific detail that the failure is directly linked to the increasing length of the task, meaning the models struggle more as the sequence of the Stroop test grows longer. To understand why this matters, it is necessary to recognize that the Stroop test is a standard psychological measure used to evaluate human attention and processing speed. It typically requires a subject to ignore a distracting stimulus to focus on a specific task, such as naming the color of a word when the word itself spells a different color.


The fact that AI models fail this test as the task expands indicates a fundamental gap between synthetic processing and human-like attention mechanisms, which has broader implications for how AI handles complex, long-form cognitive instructions. Moving forward, observers will be looking for how researchers address this identified weakness in AI architecture. Based on the reports from Caliber.Az, Baku.ws, and SciTechDaily, the primary focus remains on the correlation between task length and model failure. Future developments will likely center on whether subsequent updates to GPT-5 or other models can overcome this specific attention barrier or if this limitation remains an inherent characteristic of current artificial intelligence design.

## Multi-Source Evidence Table
| Source Outlet | Headline | Verification URL |
|---|---|---|
| Caliber.Az | Study finds AI models fail Stroop test as task length increases | [Source Link](https://news.google.com/rss/articles/CBMilAFBVV95cUxOcGxxVy1KVWtBN0V3ejUzWGF4WlcwOURESGZ3V3F4Y2VvLW5vMlZweFZGV0NpanBnY0x1MmlJSDNLQkdlVXd4eElSX3lnUFBRb200MEVFVnp1N25Ebm9NZXlZMmh5Yi01WW9SVUdKZjJuU1NjTXRnMktjenJkVDBlTlRzMGxkT0VNM2hmRjBZMDRyd1BQ?oc=5) |
| Baku.ws | Scientists have identified the main weakness of artificial intelligence | [Source Link](https://news.google.com/rss/articles/CBMirwFBVV95cUxQVkFaZzdEbE1DYWhhT3paT2xQZFB6LXlhS2NLWFY1dnhnMzk2MUtfM3AzQ1NzdHlHLU1WM1ZlaHpyeGw4TDJMTDN1QW5mTXZBSVZsc1lQNHNnRG55QjlVdEwzNDJTUWhudmswQnhXQ2txUUVlNlZ6ai0xVEhONldNbzlQODdLRmpXOGh0aHVWSFJrSm85cUpvSXBwYnM3NjNPbkF4VktHUGE4SkE5dDZj?oc=5) |
| SciTechDaily | Even GPT-5 Failed This Human Attention Test | [Source Link](https://news.google.com/rss/articles/CBMieEFVX3lxTE53SHl3OUhxYmlkbXdqQnpkbnB5NEg5Yk0zcGx0eURKTVdfanp0cXo0QXRYSzIyTkVxejZDTk5XTGVxZzNsM01jWVJkOVVSRkM2b1lTbHR6SHNPZHpWaXhLMzRsejk3Um96SEQ5eGJBNjBXWmZtd0FlQw?oc=5) |

---
*Canonical Source: https://pulse.byoviral.com/trend/2026-06-16/even-gpt-5-failed-this-human-attention-test*
