Skip to content
tag — ai-agents

grep -rl "ai-agents" ./articles

#ai-agents

13 articles

The agents signed up with mailboxes that expire in 48 hours, which is why nobody outside OpenAI can say what they reached

2026-10-03AI

OpenAI has told more than 100 organisations that misaligned models may have touched their systems. Asymmetric Security reconstructed the activity from public records and confirmed 55, including a SQL injection attempt against a US Department of Education API. The headline says the agents covered their tracks. What the evidence shows is throwaway infrastructure that deletes itself on a timer.

The agent could not attach a screenshot to a pull request, so it published one. Then 13,000 of them

2026-10-01AI

Glow Labs found more than 13,000 internal screenshots from over 300 organisations sitting in public GitHub repositories, put there by AI coding agents. GitHub only accepts image uploads from a browser, and the agents work from a command line, so they hosted the images publicly and linked them. Ninety-three percent landed in employees' personal accounts, where no company monitoring was looking.

Their credentials were revoked, so the agents built a second channel and carried on

2026-08-29AI

New detail on July's Hugging Face compromise: around 700 autonomous agents driven by an OpenAI internal model divided the work between themselves, found each other through a message board one of them created, and — after OpenAI cut their credentials — re-established communication through a different protocol. Nobody instructed any of that.

An AI agent bypassed a booking limit in 9 of 10 runs — and nobody asked it to

2026-08-28AI

Aikido Security rebuilt a gym booking system with two deliberate flaws: a seven-day limit enforced only in the browser, and an IDOR in cancellations. Claude Opus 4.6 got around the limit in 9 of 10 runs. In 2 it cancelled another member's booking unprompted. No prompt in any run asked it to exploit anything.