AHA red-team agent finds reusable flaws in Claude Code, Codex
New arXiv preprint introduces AHA, a system using one AI agent to auto-discover reusable vulnerabilities in production agents like Claude Code and Codex.
A newsletter breaking down AI research, technology, and Australian property in plain English
AI and technology, explained through what they actually mean. An AI-assisted newsroom under human editorial rules — every story cites its primary sources so you can check them yourself.
306 stories
New arXiv preprint introduces AHA, a system using one AI agent to auto-discover reusable vulnerabilities in production agents like Claude Code and Codex.
FactorDiff decomposes diffusion samples into pixel-level factors and routes each to the best expert, beating global weighting on ARC-AGI reasoning tasks.
AMT-X multi-turn red-teaming framework hit up to 100% attack success on six frontier AI models, exposing critical gaps in current LLM safety testing.
A field experiment across six companies found GPT-5's playful rewriting lifted emotional tone, with engagement gains flowing through an indirect path.
PromptGraph models LLM prompts as graphs to catch privacy leaks that existing sanitisers miss, affecting anyone sending sensitive data to cloud AI.
ARMOR-IMC, an arXiv preprint, hardens in-memory computing chips against manufacturing faults and power attacks without retraining the AI model.
EvoCUA-1.5 teaches AI agents to use computers through trial-and-error in sandbox environments, reaching 63.2% on OSWorld-Verified.
A new open benchmark with 500+ tools reveals even frontier AI models can't reliably read an image and act on it, with most failures traced to seeing, not
A new latency predictor knows when it's wrong, routing uncertain guesses to real hardware and slashing the cost of designing AI models for edge devices.
Vision-language models internally encode the right answer but misread it, a new arXiv study finds. A simple fix lifts counting accuracy 15.6 points.
Spotify and Queen Mary researchers found voice features like tone and pace correlate with audiobook engagement even after controlling for the title.
Quantum learning separation proved in July 2026: a quantum procedure predicts many-body dynamics no classical algorithm can match, affecting quantum ML.