#research

Research

936 stories tagged Research · page 16 of 39

AI Training Flaw: How Data Sampling Can Ruin Models
research

AI Training Flaw: How Data Sampling Can Ruin Models

Researchers found that selecting training data can accidentally bias AI models, making them less accurate. This happens when the data used to verify the model is itself incomplete or skewed, leading to a breakdown in performance. This affects how AI systems learn and could impact everything from chatbots to medical diagnostics.

via ArXiv cs.AI#ai#research#data
AI Judges Flip Decisions 13.6% of the Time – Here's What That Means
research

AI Judges Flip Decisions 13.6% of the Time – Here's What That Means

Researchers found that AI judges used to rank other AI models often change their minds when given the same question repeatedly. This inconsistency could affect how we measure AI performance and trust public leaderboards. The study tested two OpenAI judge models across 29 tasks and found that pairwise preferences flipped an average of 13.6% of the time, with 28% of questions exceeding a 20% flip rate. The findings highlight the need for more reliable evaluation methods.

Theory of Mind Utility: A Formal Framework for AI to Infer Human Beliefs
research

Theory of Mind Utility: A Formal Framework for AI to Infer Human Beliefs

Researchers have introduced the Theory of Mind Utility (ToM-U), a formal mathematical framework that specifies how an AI system could infer others' beliefs by tracking who told them what, in what order, and how credible that information is. This is a theoretical model, not a built AI, and could guide future AI systems toward better understanding human social interactions.