Archive

All 2,608 AI stories, newest first · page 2 of 109

AlphaAgent: Google DeepMind's Skill-Driven AI for Materials Science Literature Analysis
research

AlphaAgent: Google DeepMind's Skill-Driven AI for Materials Science Literature Analysis

Google DeepMind released AlphaAgent, a skill-driven AI framework that decouples retrieval-based question answering from paper-level report generation for materials science literature analysis. Unlike conventional RAG pipelines, AlphaAgent uses explicit skill contracts to handle composition, processing, characterization, and property relationships simultaneously.

AINTMA: Six AI Agents That Autonomously Manage Software Testing and Cloud Security
research

AINTMA: Six AI Agents That Autonomously Manage Software Testing and Cloud Security

Researchers from ArXiv cs.AI introduced AINTMA, a multi-agent AI architecture that uses six specialized agents to autonomously handle software test discovery, risk assessment, prioritization, execution, generative quality intelligence, and cloud security monitoring. The system aims to transform traditional test management into a self-improving quality intelligence ecosystem for distributed cloud environments.

via ArXiv cs.AI#ai#software#testing
AI Watermarks in Medical Texts: Study Finds Critical Failures That Risk Patient Safety
research

AI Watermarks in Medical Texts: Study Finds Critical Failures That Risk Patient Safety

A new study from ArXiv cs.AI reveals that AI watermarks, designed to track machine-generated text, often fail in medical contexts. Researchers tested five watermarking schemes across 11 large language models (LLMs) and 7 vision-language models (VLMs) on various medical tasks. The findings show that small token-level perturbations introduced by watermarks can cause significant semantic changes, potentially leading to misdiagnoses or other serious errors in clinical settings. The study underscores the urgent need for domain-specific watermarking methods tailored to high-stakes fields like healthcare.