DROPJ trains safe AI agents from human justifications
New arXiv paper pairs world models with human preferences and justifications to train safe AI agents without risky trial-and-error deployment.
The AI knowledge 99.9% of people never discover
Daily AI and technology coverage — models, chips, agents and research, and what each development actually means for you and your business.
552 stories
New arXiv paper pairs world models with human preferences and justifications to train safe AI agents without risky trial-and-error deployment.
Google Vids now lets subscribers generate and edit video from text prompts and create a talking avatar from a selfie, but access is limited to paid tiers
New arXiv preprint from Microsoft and Walmart researchers finds encoded prompts bypass AI safety filters most, raising the stakes for enterprise AI
A new arXiv framework lifts AI browser-agent task completion from 49% to 89% by restructuring pages for machine reading, hitting every e-commerce site
A new arXiv system called ob tracks author identity through AI data pipelines, letting trainers build precise forget sets when contributors demand removal
arXiv preprint proposes directly editing chain-of-thought steps in LLMs, claiming 25% better error correction and 40% token savings on STEM tasks.
An arXiv preprint details a five-stage AI upskilling framework. Three learners passed NVIDIA's Agentic AI exam, but evidence is thin and unpeer-reviewed.
NVIDIA's new on-device model lets Japanese robotics giants like FANUC and Yaskawa run vision and reasoning locally on edge hardware, adapting to new tasks
arXiv preprint reveals exact token cost of attributing AI text to a user, plus a window where generated text is provably machine-made but unattributable.
A new arXiv framework shows aggressive network path randomisation drops attack success to 4-20% and lifts throughput 30.9% for enterprise defenders.
A new arXiv preprint proposes a training-free detector that reads LLM hidden states to catch prompt injection attacks on purpose-specific agents, reporting
IBM Research found Claude Sonnet cost half as much as GPT-4.1 across 417 agent tasks, because caching matters more than sticker price for AI routing.