
New Research Tests How Well AI Models Update Beliefs Over Time
Researchers created BayesBench to test how large language models update their beliefs with new evidence. The study reveals that AI models often struggle to adjust their reasoning as conversations progress.

