Agent-ready websites nearly double AI shopping agent success
A new arXiv framework lifts AI browser-agent task completion from 49% to 89% by restructuring pages for machine reading, hitting every e-commerce site
A newsletter breaking down AI research, technology, and Australian property in plain English
AI and technology, explained through what they actually mean. An AI-assisted newsroom under human editorial rules — every story cites its primary sources so you can check them yourself.
306 stories
A new arXiv framework lifts AI browser-agent task completion from 49% to 89% by restructuring pages for machine reading, hitting every e-commerce site
A new arXiv system called ob tracks author identity through AI data pipelines, letting trainers build precise forget sets when contributors demand removal
arXiv preprint proposes directly editing chain-of-thought steps in LLMs, claiming 25% better error correction and 40% token savings on STEM tasks.
An arXiv preprint details a five-stage AI upskilling framework. Three learners passed NVIDIA's Agentic AI exam, but evidence is thin and unpeer-reviewed.
NVIDIA's new on-device model lets Japanese robotics giants like FANUC and Yaskawa run vision and reasoning locally on edge hardware, adapting to new tasks
arXiv preprint reveals exact token cost of attributing AI text to a user, plus a window where generated text is provably machine-made but unattributable.
A new arXiv framework shows aggressive network path randomisation drops attack success to 4-20% and lifts throughput 30.9% for enterprise defenders.
A new arXiv preprint proposes a training-free detector that reads LLM hidden states to catch prompt injection attacks on purpose-specific agents, reporting
IBM Research found Claude Sonnet cost half as much as GPT-4.1 across 417 agent tasks, because caching matters more than sticker price for AI routing.
New E3 framework cuts AI agent cost 85% by estimating task difficulty before acting, matching 100% success on a 121-edit benchmark with code released.
A new arXiv preprint pairs Wiener-Hopf linear prediction with a lightweight CNN to detect synthetic speech and explain why it was flagged.
A new arXiv preprint found every agent tested invented non-existent skill names, creating a supply-chain attack path through open skill registries.