#automation

Automation

146 stories tagged Automation

AINTMA: Six AI Agents That Autonomously Manage Software Testing and Cloud Security
research

AINTMA: Six AI Agents That Autonomously Manage Software Testing and Cloud Security

Researchers from ArXiv cs.AI introduced AINTMA, a multi-agent AI architecture that uses six specialized agents to autonomously handle software test discovery, risk assessment, prioritization, execution, generative quality intelligence, and cloud security monitoring. The system aims to transform traditional test management into a self-improving quality intelligence ecosystem for distributed cloud environments.

via ArXiv cs.AI#ai#software#testing
Generative Ontology Induction (GOI): New AI Framework Automates Knowledge Organization Across Any Topic
research

Generative Ontology Induction (GOI): New AI Framework Automates Knowledge Organization Across Any Topic

Researchers introduced Generative Ontology Induction (GOI), a domain-agnostic framework that automatically creates structured knowledge systems from text corpora. GOI identifies entities, dimensions, properties, relationships, and constraints, then exports them as a typed graph in YAML/JSON — making AI systems more adaptable to new topics without manual programming.

via ArXiv cs.AI#ai#knowledge#research
Neuro-Symbolic AI Automates LEED Green Building Certification Checks
research

Neuro-Symbolic AI Automates LEED Green Building Certification Checks

Researchers introduced a neuro-symbolic AI pipeline that automates parts of LEED v4.1 BD+C certification by combining small language models with deterministic symbolic checking. The system screens project PDFs, retrieves evidence using credit-specific keyword signatures, and verifies compliance, potentially making sustainable building certification faster and more accessible.

Enterprise AI Faces a Reality-Alignment Problem: Half of Companies Ship Agents That Pass Tests but Fail Customers
industry

Enterprise AI Faces a Reality-Alignment Problem: Half of Companies Ship Agents That Pass Tests but Fail Customers

A study of 157 enterprises reveals that half have shipped AI agents that passed internal evaluations but failed in production. Only 5% fully trust automated evaluations, yet two-thirds are moving toward fully automated deployments without human oversight. The core issue is not test coverage but a reality-alignment gap between evaluations and real-world outcomes.

via VentureBeat AI#ai#enterprise#evaluation
Claude can now use your 1Password credentials for you
industry

Claude can now use your 1Password credentials for you

1Password has launched a browser integration for Claude that allows the Anthropic chatbot to access stored security credentials like usernames and passwords. Users can authorize Claude to complete multi-step tasks like booking travel and managing online accounts on their behalf without manually inputting login details.