AI Helps Solve Decades-Old Math Problem About Graph Connections
Researchers used AI to solve a complex math problem about graph connections. This could improve algorithms for recommendation systems and network design.
1193 stories curated by AInformed · page 38 of 50
Researchers used AI to solve a complex math problem about graph connections. This could improve algorithms for recommendation systems and network design.
Researchers say current methods for testing AI bias might be flawed because they don't account for all possible changes in the text. They propose a better way to measure how AI models really work. This could help make AI fairer and more reliable.
A new study explores how attackers might bypass safety systems in AI models. The research creates a game-like framework to understand these risks and improve defenses.
Researchers developed DIAGRAMS, a tool to help AI explain its reasoning when answering questions about diagrams. This makes it easier to understand how AI arrives at its answers, improving transparency.
Researchers created a new framework called CLEAR to test how well AI handles ambiguous medical questions. They found that AI models often give unreliable answers when faced with real-world uncertainties.
Researchers have developed a method to extract hierarchical structures from AI language models, showing how these models organize complex reasoning. This could help us understand and improve AI decision-making.
Researchers found a simple way to uncover what AI models were trained to do, even when developers try to hide it. This helps identify harmful behaviors in AI systems.
Researchers found that AI models can generate posts that make people feel inferior or superior, but struggle to recognize these effects in their own writing. This highlights a gap in AI's understanding of human psychology.
Researchers found that large language models (LLMs) have trouble making strategic decisions because they can't properly connect what they observe with what they believe. This affects AI in negotiations and policymaking. The study tested models like Llama 3.1 and Qwen3.
Researchers found why some AI models can be tricked into answering harmful questions. This helps us understand how to make AI safer for everyday use.
Researchers have developed a new method called TUR-DPO to improve how AI models learn from human feedback. This approach rewards the process of how answers are derived, not just the final output, making AI more reliable and less sensitive to noise.
Researchers have introduced Token Arena, a continuous benchmark that evaluates AI systems at the endpoint level. It measures five key factors to give a more realistic comparison of AI performance.
Researchers have developed a new way for robots to plan complex tasks by combining both text and visual reasoning. This could lead to robots that can handle more intricate, real-world jobs. The key is a system called Interleaved Vision-Language Reasoning (IVLR), which helps robots understand both the logical steps and spatial constraints of a task.
A new study explores how groups of basic AI systems could accidentally combine into a more advanced collective with its own goals. This raises important questions about controlling and understanding AI behavior.
Researchers tested whether Mamba AI models can automatically summarize sentences without additional training. Their findings show promise for simpler, faster text analysis tools in the future.
A new study introduces AgentFloor, a benchmark to test how well smaller AI models can handle routine tasks. The goal is to see which parts of AI workflows need big, advanced models and which can be done by smaller ones.
Scientists have uncovered why current AI models struggle with unusual inputs. Their findings could lead to more reliable AI assistants and tools. This research highlights a common flaw in how AI processes unexpected questions or commands.
Using tools to help AI reason doesn't always work better than just thinking things through. Researchers found that tool use can actually slow things down and cost more. This challenges the idea that tools are always the best solution for AI.
Researchers developed a new way to evaluate AI grading systems by measuring both the AI's ability and the difficulty of student responses. This could lead to fairer and more accurate automated grading in education.
Researchers propose a decentralized system to track AI agents' reputations. This could make AI marketplaces more reliable for tasks like debugging and security checks.
Researchers propose a new way for AI to understand the world by combining physics with predictive models. This could make AI systems like robots and self-driving cars smarter and more adaptable.
Researchers have developed a new approach called Adaptive Entropy Modulation (AEM) to improve how AI agents learn complex, multi-step tasks. This could make AI assistants better at handling long conversations or tasks with many steps.
Researchers have developed a method called RSAT that helps small language models explain their reasoning when answering questions about tables. This makes it easier for users to verify the accuracy of the AI's answers.
Researchers created ARMOR 2025 to test AI models for military use, ensuring they follow legal and ethical rules. This benchmark goes beyond civilian safety standards to address defense-specific needs.