AI models struggle with expert-level global history knowledge

Intro:

Researchers recently evaluated the ability of advanced artificial intelligence (AI) models to answer questions about global history using a benchmark derived from the Seshat Global History Databank. The study, presented at the Neural Information Processing Systems conference in Vancouver, revealed that the best-performing model, GPT-4 Turbo, achieved a score of 46% on a multiple-choice test, a marked improvement over random guessing but far from expert comprehension. The findings highlight significant limitations in current AI tools’ ability to process and understand historical knowledge, particularly outside well-documented regions like North America and Western Europe.

Eric W. Dolan PsyPost January 22, 2025 Link
  1. Home
  2. /
  3. Press
  4. /
  5. AI models struggle with...

© Peter Turchin 2023 All rights reserved

Privacy Policy