Safety Sentry gives AI agents a third option: ask first
A new arXiv paper from ShanghaiTech replaces binary safe/unsafe guards with a three-way EXECUTE-ASK-REFUSE routing that cuts alert fatigue for anyone
A newsletter breaking down AI research, technology, and Australian property in plain English
A new arXiv paper from ShanghaiTech replaces binary safe/unsafe guards with a three-way EXECUTE-ASK-REFUSE routing that cuts alert fatigue for anyone
Harness Handbook from Tencent and four universities builds a behaviour-to-code map for agent harnesses that cuts planner tokens and improves edit
New arXiv paper finds automatically evolving agent scaffolding doesn't consistently outperform test-time scaling, with limited generalization to new tasks.
A new preprint tested 327 real-world agent skills and found vulnerabilities across the entire skill lifecycle, from admission to evolution.
An arXiv preprint introduces AutoSynthesis, a multi-agent system that turns a plain-English research question into a full meta-analysis report.
New STOCKTAKE benchmark: LLM agents detect up to 88% of hidden supply-chain failures but two of four models score below a blind baseline on action.
NVIDIA's Vera Rubin platform reframes AI economics around intelligence per dollar, making continuous post-training the central workload for agentic AI.
A study of 2,361 popular GitHub repos finds most generate just one or two AI-agent pull requests per quarter, with heavy use in small teams.
Oracle agent memory hits 93.8% on LongMemEval with 10.7x fewer tokens than flat-history baselines, per a new arXiv preprint on long-horizon AI agents.
A new arXiv survey formalises how AI agents update their own prompts, memory and tools with minimal human input, and what breaks when they do.
A Google Public Sector preprint proposes splitting geospatial AI between big-compute pretraining and expert fine-tuning, with LLMs as orchestrators.
New arXiv paper pairs world models with human preferences and justifications to train safe AI agents without risky trial-and-error deployment.