AINewsnow

Introducing Humanity’s Sixth Sense, a new benchmark testing intuitive visual reasoning from spatial and causal reasoning to social understanding. The gap is significant. Humans score 93.1%, while the strongest model, GPT-6-astra, reaches 53.6%. The median model scores just 30.9%.

Coverage of "Introducing Humanity’s Sixth Sense, a new benchmark testing intuitive visual reasoning from spatial and causal reasoning to social understanding. The gap is significant. Humans score 93.1%, while the strongest model, GPT-6-astra, reaches 53.6%. The median model scores just 30.9%." from 1 source, with a live timeline of who reported what and when.

Read the full story at r/singularity ↗

Timeline · 1 report

  1. 2026-10-08 15:27 · r/singularity
    Introducing Humanity’s Sixth Sense, a new benchmark testing intuitive visual reasoning from spatial and causal reasoning to social understanding. The gap is significant. Humans score 93.1%, while the strongest model, GPT-6-astra, reaches 53.6%. The median model scores just 30.9%.

More stories

  1. GPT-6 and Intelligent UI for everyone — OpenAI News
  2. Sharing AI progress in mathematics — OpenAI News
  3. Claude Pro vs ChatGPT Plus vs Copilot Premium: which one would you choose for this use case? — r/ChatGPTPro
  4. ChatGPT for Teens Is an ‘Unacceptable Risk,’ Watchdog Group Says — CNET AI
  5. OpenAI publishes 722 mathematical proofs & manuscripts — r/singularity
  6. How Oracle Uses ChatGPT Work to Transform Recruitment — OpenAI YouTube
  7. How Oracle turns days of work into minutes with ChatGPT and Codex — OpenAI News
  8. Pollo AI turns creative ideas into campaigns with OpenAI — OpenAI News

Get the daily brief of stories like this at 6:30 every morning →