Autonomous vehicle AI poisoning defence uses digital twins
New arXiv preprint proposes a twin-aware framework to protect federated reinforcement learning in self-driving cars from malicious parameter injection.
The AI knowledge 99.9% of people never discover
Daily AI and technology coverage — models, chips, agents and research, and what each development actually means for you and your business.
552 stories
New arXiv preprint proposes a twin-aware framework to protect federated reinforcement learning in self-driving cars from malicious parameter injection.
A new preprint proposes a semantic firewall for persistent AI agents, blocking most attacks with under 0.7 seconds overhead — but it's not peer-reviewed yet.
ProjAgent adds a new retrieval signal that finds code doing the same job under different names, hitting 41.14% Pass@1 on the REPOCOD benchmark for repository-le
A July 2026 preprint finds models compressed to lower precision diverge in which questions they get right, even when overall accuracy holds steady — exposing a
A new arXiv preprint called SLORR adds under 1% training overhead in LLM pretraining while making models more compressible, without SVDs or architecture changes
OpenAI's new ChatGPT Work agent promises to turn goals into finished work by acting across your apps and files for hours — but key details remain undisclosed.
OpenAI's GPT-5.6 is now the preferred model powering Microsoft 365 Copilot across Word, Excel and PowerPoint, affecting millions of daily office users worldwide
OpenAI's GPT-5.5 Bio Bug Bounty invites researchers to test universal jailbreaks for biological risks in a model built for autonomous, multi-tool work.
OpenAI's GPT-5.6 launches a three-model lineup — Sol, Terra and Luna — promising token efficiency and frontend gains, but with no independent benchmarks yet.
Stronger privacy cuts AI generalisation error in high-noise regimes, yet the robustness-privacy tension returns when noise is low, according to new arXiv prepri
ALER-TI, a new arXiv preprint, retrieves cached historical patterns to reconstruct missing time series values, tested across six real-world datasets under varyi
CNeVA lets developers steer virtual driver aggression and caution on the Waymo benchmark, offering per-channel control that higher-ranked imitation models lack.