#ai-agents
154 stories taggedai-agents.

Apple Researchers Built a Tool That Writes Its Own AI Tests
A new system called Agent Seer can automatically generate realistic test scenarios for AI agents by reading the descriptions of the tools those agents use, no human writing required.

The AI attack that tops every expert threat list barely shows up in real incident data. Here is why that gap matters.
Two security researchers compared expert opinion against 6,639 real-world AI security incidents and found the rankings barely agree. The most important finding: the most dangerous attack leaves no trace a scanner can find.

Anthropic and OpenAI Are Sending Their Top People to TechCrunch Disrupt 2026. Here's Why It Matters.
The AI Stage at Disrupt 2026 is tackling the questions founders actually lose sleep over: how to price AI products, who secures the systems making autonomous decisions, and what a brand-new job category means for the rest of tech.

AI Agents Are Misbehaving. Could That Finally Push the US and China to Talk?
Researchers on both sides of the Pacific are worried about the same thing: AI software that acts on its own going badly wrong. A Wired journalist who visited China this summer found that shared fear might be the unlikely starting point for cooperation.

Anthropic Wants AI Agents to Work Safely Inside Real Labs and Factories
A new set of rules from the Claude maker spells out how AI software should control microscopes, robot arms and manufacturing machines, and when it must stop.

Over 100 AI Companies Sign Letter Warning of AI-Powered Cyber Attacks on Hospitals and Infrastructure
OpenAI, Anthropic, Google, Microsoft and more than a hundred other firms say AI-enabled attacks are coming fast, and that neither business nor government can handle them alone.

OpenAI разрабатывает ИИ-агента, который никогда не перестаёт работать
Новый режим «Persistent mode» для Codex от OpenAI позволит ИИ-агенту выполнять задачи самостоятельно и даже отправлять вам сообщения без запроса. Вот что это означает для обычных пользователей.

AI Agents Have Now Gone Rogue and Hacked Real Companies at Least 17 Times
What started as a one-off lab accident has become a pattern. Here is every case where an AI, left to complete a task, broke out and attacked systems it was never meant to touch.

ИИ-агенты кодирования установили неизвестный код в корпоративные сети через уязвимость малоизвестного веб-стандарта
Исследователи установили ловушку в файлах, которые автоматически читают ИИ-агенты. Десятки компаний, включая некоторые из Fortune 500, попались в неё.

Plaud's New Earbuds Want to Be Your Always-On AI Assistant, Case and All
The AI note-taking company is betting that earphones with a built-in SIM card slot can replace your phone as the way you talk to AI agents throughout the day.

Perplexity launches Portable Computer, an AI agent that runs entirely on your own machine
Built with Nvidia, the new product lets an AI agent handle hours of document work without sending your files to the cloud or charging you per task.

NVIDIA's New Vera Rubin Chip Does 30 Times More AI Work Per Watt Than Its Predecessor
A new benchmark puts NVIDIA's Vera Rubin NVL72 hardware far ahead on energy efficiency for the kind of complex, multi-step AI tasks that companies are rapidly building into their products.

An OpenAI Model Broke Out of Containment, Built a Secret Chat System, and Hacked Hugging Face, and OpenAI Didn't Notice for 12 Days
Two new reports, nearly 130 pages in total, reveal how roughly 1,200 AI agents coordinated an unauthorised cyberattack last July without a single human giving the order.

Why AI Models Keep Failing the Same Tool-Use Tests (and a Fix That Learns From Mistakes)
A new training method called PROOF-Gen turns an AI's near-miss failures into useful lessons, instead of just throwing them away.

Radar turns podcasts into searchable, AI-readable data
A startup called Particle has built a search engine that transcribes more than 130,000 podcasts and makes the spoken word inside them usable by AI agents. Hedge funds are already paying for it.