25 models collectively lose to humans: AI still fails to understand many common life knowledge points
2026-10-08 19:33:43
According to CoinMeta and a report by OneMillion.AI, a new evaluation jointly conducted by research institutions scale, labs, and elorian under scale and ai shows that 25 multimodal models collectively lost to humans in the tests, being unable to understand many common aspects of life. The study covered 522 open-ended questions involving 288 images and 234 videos, with human participants achieving an accuracy rate of 93.1%. The top-ranked model gpt-6 astra only achieved an accuracy rate of 53.6%, while gpt-6.1 SOL and claude opus scored 46.6% and 44.6% respectively. The study found that 94% of the models' failures were related to missing key clues, misidentifying objects, or being unable to infer underlying relationships. In terms of social understanding, 21 models performed the worst, with video questions generally being more difficult than image questions. Despite increasing the amount of reasoning, the models still failed to match the performance of humans.
Source:Internet
This content is for market information only and does not constitute investment advice.
Follow HQYC official accounts to stay updated

Hot Articles
Refresh

Ethereum October 2026 Outlook: $5,000 or $2,500?
9h ago

Bitcoin October 2026 Outlook: Can the 19% Historical Gain Hold?
09-30 12:57

Dogecoin Price: Whales Buy $112M, Can DOGE Break $0.10?
09-29 12:48

What is Solidigm? Is Its $150B IPO Valuation a Bubble?
09-28 13:00

Is PAXG Stable? Is Gold-Backed Better Than Stablecoins?
09-24 18:04



