Adversarial prompting framework finds encoded attacks bypass AI safety
New arXiv preprint from Microsoft and Walmart researchers finds encoded prompts bypass AI safety filters most, raising the stakes for enterprise AI
A newsletter breaking down AI research, technology, and Australian property in plain English
Daily AI and technology news decoded in plain English — models, chips, agents and research, and what each development actually means for you and your business.
418 stories
New arXiv preprint from Microsoft and Walmart researchers finds encoded prompts bypass AI safety filters most, raising the stakes for enterprise AI
A new arXiv framework lifts AI browser-agent task completion from 49% to 89% by restructuring pages for machine reading, hitting every e-commerce site
A new arXiv system called ob tracks author identity through AI data pipelines, letting trainers build precise forget sets when contributors demand removal
arXiv preprint proposes directly editing chain-of-thought steps in LLMs, claiming 25% better error correction and 40% token savings on STEM tasks.
An arXiv preprint details a five-stage AI upskilling framework. Three learners passed NVIDIA's Agentic AI exam, but evidence is thin and unpeer-reviewed.
NVIDIA's new on-device model lets Japanese robotics giants like FANUC and Yaskawa run vision and reasoning locally on edge hardware, adapting to new tasks
arXiv preprint reveals exact token cost of attributing AI text to a user, plus a window where generated text is provably machine-made but unattributable.
A new arXiv framework shows aggressive network path randomisation drops attack success to 4-20% and lifts throughput 30.9% for enterprise defenders.
A new arXiv preprint proposes a training-free detector that reads LLM hidden states to catch prompt injection attacks on purpose-specific agents, reporting
IBM Research found Claude Sonnet cost half as much as GPT-4.1 across 417 agent tasks, because caching matters more than sticker price for AI routing.
New E3 framework cuts AI agent cost 85% by estimating task difficulty before acting, matching 100% success on a 121-edit benchmark with code released.
A new arXiv preprint pairs Wiener-Hopf linear prediction with a lightweight CNN to detect synthetic speech and explain why it was flagged.