
Researchers Define New Framework for AI's Emergent Strategic Risks
A new taxonomy identifies risks like deception and reward hacking in advanced AI systems. The framework aims to benchmark these behaviors as models grow more capable.
All 2,608 AI stories, newest first · page 86 of 109

A new taxonomy identifies risks like deception and reward hacking in advanced AI systems. The framework aims to benchmark these behaviors as models grow more capable.

OpenAI is reportedly working on a smartphone that replaces traditional apps with AI agents. The device could enter mass production by 2028, according to an analyst.

Hoop introduces an open-source control layer to safely manage AI interactions with production systems. It aims to bridge the gap between AI development and real-world deployment.

A new AI usage analytics tool acts as a proxy to enforce budget limits and redact personally identifiable information before requests reach the provider. It offers real-time cost tracking and immediate suspension when thresholds are breached.

A new paper highlights the risks of AI agents producing selectively chosen, publishable analyses that lack rigorous validation. The study calls for adversarial experiments to ensure scientific integrity.

A new paper on arXiv introduces a two-layer certification framework to evaluate AI-generated research. This system separates knowledge quality from human contribution, addressing gaps in current publication standards.

Researchers propose an artifact-based agent framework to improve adaptability and reproducibility in medical image processing workflows. This approach addresses critical needs for real-world clinical deployment.

Researchers introduce MolClaw, an autonomous AI agent that excels in drug molecule evaluation, screening, and optimization. The agent uses a three-tier hierarchical skill architecture to unify over 30 specialized tools, addressing key challenges in computational drug discovery.

Microsoft plans to invest $18 billion in Australia over the next decade to expand AI, cloud, and digital infrastructure. This move aims to strengthen the country's position in the global tech landscape and create thousands of jobs.

Microsoft and OpenAI have updated their partnership agreement, removing a key clause about artificial general intelligence (AGI). The revised deal maintains Microsoft as OpenAI's primary cloud partner but removes exclusivity and AGI-focused terms.

Microsoft and OpenAI have updated their partnership to simplify the agreement and provide long-term clarity. The move aims to support continued AI innovation at scale, with both companies reaffirming their commitment to advancing the field.

Researchers introduce Memanto, a new memory system for autonomous agents that uses typed semantic storage and information-theoretic retrieval. This approach aims to reduce computational overhead compared to traditional graph-based methods.

Researchers introduce a new benchmark, Math Takes Two, to evaluate whether language models truly understand math or just memorize patterns. The test focuses on emergent mathematical reasoning through communication, challenging models to construct abstract concepts from first principles.

Glama has open-sourced Lightport, an AI gateway that makes various LLM providers compatible with OpenAI's API. This move aims to support the MCP ecosystem and give back to the community.

OpenAI's GPT 5.5 system card details its capabilities, limitations, and ethical considerations. The document highlights improvements in reasoning and multimodality while acknowledging persistent challenges.

Mistral AI, a French AI startup, has achieved a $14 billion valuation by focusing on European markets and avoiding the US-centric approach of its competitors. This strategy highlights the growing importance of non-US AI players in the global market.

Elon Musk is suing OpenAI, accusing the company of abandoning its original mission to benefit humanity. The trial, starting April 27, could reshape the AI landscape and OpenAI's future.

China has blocked Meta's $2 billion acquisition of AI startup Manus, citing national security concerns. The move highlights growing tensions between Western tech giants and Chinese regulatory oversight.

Canva's AI tool Magic Layers has been automatically replacing 'Palestine' with 'Palestinian territories' in user designs. The company has apologized and is working to fix the issue.

Canonical, the developer of Ubuntu Linux, has outlined a year-long plan to integrate AI features into its popular Linux distribution. The initiative aims to enhance user experience and system capabilities with AI-driven tools.

Researchers developed an AI system that can reproduce social science findings using only paper descriptions and raw data. The method achieves deterministic, cell-level result comparisons without access to original code or outputs.

A new study reveals that leading AI models give culturally biased advice, aligning more with individualistic than collectivist values. The research highlights significant discrepancies between AI responses and real-world cultural norms.

The Hugging Face Blog highlights how open-source AI is transforming cybersecurity by fostering collaboration and innovation. Transparency in AI models is crucial for building trust and resilience against cyber threats.

A new study reveals previously unknown attack vectors in sandboxed AI agents, challenging assumptions about their security. The findings highlight the need for enhanced isolation techniques and continuous monitoring.