
PEAR: New AI Debate System Reduces Bias and Improves Reliability
Researchers developed PEAR, a dynamic AI debate system that improves reliability by changing roles and reducing biases. This could make AI responses more trustworthy and consistent.
All 2,608 AI stories, newest first · page 33 of 109

Researchers developed PEAR, a dynamic AI debate system that improves reliability by changing roles and reducing biases. This could make AI responses more trustworthy and consistent.

Omio has integrated OpenAI's technology to create a conversational travel assistant. This innovation makes planning trips faster and more intuitive for users.

A new research paper introduces a reference architecture for 'agent skills' — reusable, externalized behavioral knowledge that LLM agents can discover, activate, and interpret at runtime. The framework formalizes how skills are bound to context and authority, interpreted by stochastic agents, and recorded as run evidence.

A new study examines how different AI reasoning strategies perform under varying conditions. The findings show that some methods hit performance limits even with more computing power. In plain English, this means AI problem-solving isn't as flexible as we thought.

A new paper finds that offline recommendations to mix AI models from different "families" for diversity may not hold up in real-time interactive settings—the very environments where multi-LLM systems are actually deployed. This discovery could change how multi-AI systems are designed.

Researchers propose a new language to clearly define how AI and humans should work together in software development. This could make teamwork between people and AI more predictable and reliable.

A new paper argues that the success of modern AI vindicates a modest form of associationism: the idea that learning is uniform, gradual, and driven by feedback. This could reshape how we think about both AI and human cognition.

Researchers developed LAGO, a system that helps AI robots plan tasks by understanding human language. This could make robots more useful in everyday settings. The framework predicts intermediate goals from language, reducing errors in planning.

Mistral's CEO suggests AI companies should pay a fee to use copyrighted material, sparking debate. This could impact how AI models are trained and priced in Europe.

Researchers have developed a novel AI system called MindAlign that can decode inner speech from fMRI brain signals, enabling open-ended text generation without task-specific fine-tuning. This breakthrough could aid communication for people who cannot speak, but it also raises significant privacy concerns.

Microsoft has integrated DALL-E 3 into Bing Image Creator, making it easier for users to generate high-quality images. This update enhances creativity and accessibility for everyday users.

IBM has launched CUGA, an open-source framework for creating AI agent applications. It includes two dozen working examples to help developers get started quickly.

Hugging Face demonstrated a proof-of-concept using local AI models—including Qwen2.5 and DeepSeek—to automatically triage pull requests on the OpenClaw repository, achieving over 80% accuracy in flagging valid contributions for review.

Jason Liu shares techniques for using AI to maintain context and manage complex projects over time. This approach can help both developers and non-technical users streamline their workflows.

GPT-5 Pro helped immunologist Derya Unutmaz solve a three-year-old puzzle about how T cells behave, specifically regarding an unusual subset of T cells. This breakthrough could accelerate research into cancer and autoimmune diseases.

Researchers created FirstPass, a dataset and AI model trained on real multi-round peer-review dialogues from Nature Communications. This could make AI evaluations of scientific papers more accurate by learning from actual editorial decision-making, not just stylistic mimicry.

Stockholm-based Fika Jobs has raised $4 million to build a video-first hiring platform that combines AI interview agents with short-form video profiles, creating a cross between LinkedIn and TikTok for job applications.

A new paper proposes a roadmap for AI that learns through interaction with complex environments, mimicking natural evolution by systematically removing human priors. The authors have released an open-source infrastructure called Darwin Mobile Agent to test this idea using mobile graphical user interfaces as a practical proxy for an open-ended world.

Researchers have developed AlphaMemo, an AI system that helps financial trading agents learn from their past decisions. This could make AI-driven trading more reliable and less prone to repeating errors.

Researchers created a new test showing AI vision systems struggle with languages written in multiple scripts. The study highlights how current AI models unfairly disadvantage billions of people who use different writing systems for the same language. However, the original source does not support testing this with tools like Google Lens or Microsoft Seeing AI, as those are not explicitly mentioned.

Researchers found that AI models can still make irrational decisions even when trained to align with human values. This 'rational value risk' means models may not always choose the best possible action, even if they understand what's valuable.

Some GitHub repositories have faced repeated takedown attempts but keep coming back due to public demand. These tools often fill critical gaps that users need. Here are 10 such repositories that have been repeatedly deleted and restored.

A developer created an AI voice assistant to help older adults play chess. It supports English, Spanish, and French, making the game more accessible.

A developer discovered a serious security flaw in his app months after launch. This highlights risks in the trend of quickly building apps with AI tools. Here's how to protect yourself.