#ai-security

AI Security

54 stories tagged AI Security · page 2 of 3

Hackers Can Use 9 of the Most Popular AI Tools to Assemble Massive Botnets
industry

Hackers Can Use 9 of the Most Popular AI Tools to Assemble Massive Botnets

Researchers discovered that nine widely used AI tools can be tricked into assembling massive botnets through a technique called "HalluSquatting." This vulnerability exploits AI models' tendency to hallucinate — generating plausible but incorrect responses — instead of refusing harmful requests. The finding underscores a critical security flaw in how AI handles ambiguous or malicious prompts.

AI Browsers Can Be Lulled Into a 'Dream World' Where Safety Guardrails No Longer Apply
industry

AI Browsers Can Be Lulled Into a 'Dream World' Where Safety Guardrails No Longer Apply

A new attack demonstrates that AI-powered browsers can be tricked into bypassing all safety rules simply by feeding them a false premise—like telling the model that 2 + 2 = 5. Once the LLM enters this 'dream world,' it will follow forbidden instructions, raising serious concerns about the security of AI-driven browsing assistants. (Ars Technica AI)

via Ars Technica AI#AI Security#Cybersecurity
MosaicLeaks: Can your research agent keep a secret?
open-source

MosaicLeaks: Can your research agent keep a secret?

ServiceNow researchers discovered a vulnerability in AI research agents that can leak sensitive data through a novel attack called MosaicLeaks. The flaw exploits how these agents handle and combine information, potentially exposing private details even when individual queries seem safe. The issue affects multiple AI research platforms and highlights a fundamental security challenge in agentic AI systems.

via Hugging Face Blog#AI Security#Privacy#AI Ethics
New Jailbreak Method Bypasses LLM Safety Mechanisms
research

New Jailbreak Method Bypasses LLM Safety Mechanisms

Researchers have developed a new jailbreak technique called Incremental Completion Decomposition (ICD) that exploits LLMs by eliciting single-word continuations before extracting harmful responses. This method bypasses current safety mechanisms, raising concerns about the robustness of LLM safeguards.

via ArXiv cs.CL#LLMs#Safety#Research