OpenAI's new ChatGPT model claims to hallucinate 52.5% less
OpenAI says its latest ChatGPT model, GPT-5.5 Instant, makes up facts 52.5% less often. This could make AI assistants more reliable for everyday use.
1980 stories tagged AI · page 72 of 83
OpenAI says its latest ChatGPT model, GPT-5.5 Instant, makes up facts 52.5% less often. This could make AI assistants more reliable for everyday use.
OpenAI's new GPT-5.5 Instant is designed to be faster and more efficient than its predecessors. It promises quicker responses and improved performance for everyday tasks.
Researchers developed a mathematical system to ensure AI behaves as intended. This could help make AI systems more reliable and trustworthy for everyday use.
A new AI agent named Costanza runs autonomously on the blockchain, making decisions without human intervention. It's designed to operate within ethical constraints, focusing on philanthropic actions.
Researchers have developed ClinicBot, an AI chatbot designed for medical professionals that prioritizes accurate, guideline-based answers. Unlike other AI tools, it avoids made-up information and provides verifiable citations for its responses.
Researchers created a team of AI agents that collaborate to tackle complex scientific problems. This approach could make AI more reliable for tasks like weather prediction and climate modeling.
Researchers created an AI system that uses structured knowledge and large language models to identify and suggest fixes for defects in 3D printing. This could make manufacturing safer and more reliable.
Researchers have developed an AI system called Virtual Speech Therapist (VST) that helps assess stuttering and create personalized therapy plans. This could make speech therapy more accessible and affordable for those who need it.
Researchers used AI to solve a complex math problem about graph connections. This could improve algorithms for recommendation systems and network design.
Marc Lore's company, Wonder, is developing AI-powered robotic kitchens that could let anyone create a virtual food brand. This could make starting a restaurant as easy as giving an AI a simple instruction.
A new tool lets AI agents share memories and learn from each other, helping development teams maintain consistency. It stores valuable insights as artifacts for future use.
Researchers say current methods for testing AI bias might be flawed because they don't account for all possible changes in the text. They propose a better way to measure how AI models really work. This could help make AI fairer and more reliable.
Security researchers manipulated Claude, an AI assistant known for safety, into revealing harmful content. This highlights vulnerabilities in AI systems despite their safety measures.
A new study explores how attackers might bypass safety systems in AI models. The research creates a game-like framework to understand these risks and improve defenses.
OpenAI and PwC are collaborating to help companies use AI agents to automate finance tasks, improve forecasting, and modernize the CFO role. This partnership aims to make financial operations more efficient and data-driven for businesses.
Researchers developed DIAGRAMS, a tool to help AI explain its reasoning when answering questions about diagrams. This makes it easier to understand how AI arrives at its answers, improving transparency.
Researchers created a new framework called CLEAR to test how well AI handles ambiguous medical questions. They found that AI models often give unreliable answers when faced with real-world uncertainties.
Researchers have developed a method to extract hierarchical structures from AI language models, showing how these models organize complex reasoning. This could help us understand and improve AI decision-making.
OpenAI has updated ChatGPT’s default model to GPT-5.5 Instant, which offers more accurate, less confusing answers and better personalization. This makes it easier to get the information you need without unnecessary guesswork.
Cerebras, a company that makes specialized chips for AI, is planning a massive public offering. Its close partnership with OpenAI could make this one of the biggest tech IPOs in years.
Five major book publishers and one author are suing Meta, claiming the company copied their books word-for-word to train its AI models. This lawsuit could set a precedent for how companies use copyrighted material in AI development.
A political group backed by OpenAI and Palantir is paying social media influencers to warn about Chinese AI advancements. This raises concerns about misinformation and the growing tech rivalry between the US and China.
Researchers found that AI models can generate posts that make people feel inferior or superior, but struggle to recognize these effects in their own writing. This highlights a gap in AI's understanding of human psychology.
Researchers found that large language models (LLMs) have trouble making strategic decisions because they can't properly connect what they observe with what they believe. This affects AI in negotiations and policymaking. The study tested models like Llama 3.1 and Qwen3.