AI safety paper: the quiet failures nobody measures
AI safety discourse focuses on dramatic harms, but a new preprint argues the real danger is in quiet failures normalised by everyday workflows.
A newsletter breaking down AI research, technology, and Australian property in plain English
AI and technology, explained through what they actually mean. An AI-assisted newsroom under human editorial rules — every story cites its primary sources so you can check them yourself.
303 stories
AI safety discourse focuses on dramatic harms, but a new preprint argues the real danger is in quiet failures normalised by everyday workflows.
Australia's Wage Price Index rose 3.3% in the year to March 2026, but private sector pay slowed to 3.2% as new home loans fell 6.2%.
A July 2026 arXiv preprint introduces TPIPS, a text-prompted image similarity metric showing frontier vision-language models lag human perception.
New arXiv framework uses twelve reasoning angles to separate jokes from hate in memes, hitting 80.3% humor and 75.9% hate detection on two benchmarks.
Six LLMs tested across navigation, triage and finance tasks showed stable risk preferences, a hidden trait anyone deploying AI agents needs to understand.
BSB jailbreak tricks text-to-video models like Veo and Sora by hiding harmful intent between two safe frames, beating rivals by 18.6% in attack success.
OpenAI's new ChatGPT for Small Businesses program aims to help entrepreneurs build AI skills and automate daily work, but key details are missing.
GigaPath-Flash and GigaTIME-Flash shrink billion-parameter pathology models to run on 50x less compute while predicting tumor profiles from routine slides.
MOSAIC, posted on arXiv on July 21, lifts long-conversation accuracy 27 points and catches 66% of factual conflicts before they corrupt an agent's memory.
A neural network trained on 314,630 demonstrations controls 25 systems from underwater vehicles to chemical reactors, matching hand-tuned controllers in
NVIDIA's 102.4-terabit Spectrum-6 Ethernet switch is reaching gigascale AI factories first at CoreWeave, Microsoft and Nebius, doubling network capacity.
New arXiv preprint shows AI agents can be redirected by reordering true facts, with 83.3% success across GPT, Claude, Gemini, DeepSeek and Qwen.