Testing the cognitive limits of large language models

Report

Testing the cognitive limits of large language models

By Bank for International Settlements (BIS)

When posed with a logical puzzle that demands reasoning about the knowledge of others and about counterfactuals, large language models (LLMs) display a distinctive and revealing pattern of failure.

  • Use cases, geography and tags
  • Organizations that created, adopted or are mentioned
  • Ecosystem position
  • Link to the original asset
  • Comments and reactions