Archive
Every article published on Not A Tech Guy, grouped by month: daily AI, technology and Australian property coverage.
573 stories
October 2026
- Self-driving AI cuts collisions by scoring every reachable path Technology & AI
- OpenAI trains AI agents on Ironclad contracts Technology & AI
- AI agent for chip verification hits 95% accuracy Technology & AI
- CourseChat AI tutor picks 8B model over bigger rivals Technology & AI
- MCP Python SDK trends on GitHub with 24,000 stars Technology & AI
- Multimodal AI models merge vision and text via two distinct pathways Technology & AI
- AI agent skill scanners evaded 97% by Huawei researchers Technology & AI
- Weight tying costs private LLM training 4.74 accuracy points Technology & AI
- OpenAI's 'eternal complement': AI's edge is execution Technology & AI
- OpenAI partners with America's SBDC for small business AI training Technology & AI
- CoreWeave ships NVIDIA Vera Rubin, Cognition reports 4.8x speedup Technology & AI
September 2026
- OpenAI's GPT-6.1 Sol targets Astra quality at one-fifth price Technology & AI
- LLM pipelines lose 40 accuracy points at interfaces Technology & AI
- Holo4 27B scores 61.7% on OSWorld 2.0, trails Opus 5.5 Technology & AI
- LLVM trends on GitHub as AI agents target compiler code Technology & AI
- AI security tool reverse-skill hits 38,000 GitHub stars Technology & AI
- GRPO step-level bias: GRAFT graph method claims gains on agent tasks Technology & AI
- ABS moves CPI release earlier, skips mid-2026 weight update Property & Economy
- Chinese AI text humanizer trends with 18,000 GitHub stars Technology & AI
- Rivet Actors hits GitHub trending with 20ms cold-start claim Technology & AI
- Pistis multimodal models blend distillation and RL training Technology & AI
- NobodyWho runs local LLMs across 6 app frameworks Technology & AI
- Lean Pool: AI agents formalize math, paper is one human page Technology & AI
- AI coding agent router cuts enterprise model spend 21% Technology & AI
- Hindsight agent memory repo gains 1,668 stars in a day Technology & AI
- Australia bans card surcharges from October 1 Technology & AI
- Adobe-Cornell: clean-data prediction improves diffusion generation Technology & AI
- AI agent rare-event estimation 800× more efficient Technology & AI
- Google DeepMind's Gemini 3.8 TTS clones a voice from 30 seconds Technology & AI
- GPT-6 Astra halves Parallel's research time and cost, OpenAI says Technology & AI
- Qualcomm says new chip runs 30B AI model on a phone Technology & AI
- New LLM backdoor hides attacks in logical reasoning Technology & AI
- Climate models get memory from simple equations, not neural nets Technology & AI
- RBA flags inflation risks may be materialising, reports say Property & Economy
- OpenAI Academy adds role-based learning paths for five audiences Technology & AI
- SimLife benchmark: AI agents watch 15 hours, can't learn your habits Technology & AI
- Needle: 8 MB AI model runs on phones and microcontrollers Technology & AI
- OpenAI launches Australian Youth Safety Blueprint for teens Technology & AI
- Claude Code TradingView bridge hits 6,528 GitHub stars Technology & AI
- Together AI score centering stabilizes RL for LLMs up to 30B Technology & AI
- OpenAI maps how workers use AI beyond job descriptions Technology & AI
- Claude Skills directory hits 75,000 GitHub stars Technology & AI
- NYU 7.4B model matches GPT-3 13B with 20× less compute Technology & AI
- OpenAI tells firms to measure AI ROI with its own tools Technology & AI
- OpenAI previews Sponsored Agents in ChatGPT ads push Technology & AI
- AI adoption model: faster progress can make full deployment worse Technology & AI
- MoneyPrinterTurbo hits 123,000 GitHub stars with one-click AI video Technology & AI
- OpenBMB VoxCPM2: 2B-parameter open-source TTS in 30 languages Technology & AI
- NVIDIA shifts AI factory metric to tokens per megawatt Technology & AI
- ABS GDP up 0.4% as construction costs and rents climb Property & Economy
- Australian home values fall as Sydney's top homes drop 10.7% Property & Economy
- Apple iOS 27 ships with overhauled Siri AI, English-only beta Technology & AI
- Karpathy-inspired CLAUDE.md file surges 333 stars in a day on GitHub Technology & AI
- KuaiRP researchers claim role-playing AI matches proprietary models at Technology & AI
- Solo developer's AI trading agent hits GitHub trending with 277 stars Technology & AI
- Few-shot learning approaches for NIDS in 21 reviewed studies Technology & AI
- RBA September 2026 decision: what the July data say Property & Economy
- OpenAI Habitat storage: 22M requests/sec for 1 billion users Technology & AI
- ChatGPT searches extinct genomes for new antibiotics Technology & AI
- OpenAI's Agents API puts Codex harness in the cloud Technology & AI
- NVIDIA says its robotaxi tech runs every major fleet Technology & AI
- Paul Christiano joins OpenAI board and safety committee Technology & AI
- AI agents can probe neural networks but misread their own data Technology & AI
- AI agents don't know when they fail: new method reads internal signals Technology & AI
- Show-Harness lets VLMs control robots with no extra training Technology & AI
- NVIDIA targets 2 GW of AI factory capacity in Australia by 2027 Technology & AI
- Rental kids move every 2.3 years as housing crisis bites Property & Economy
- SAREF ontology maps distributed AI across edge-fog-cloud continuum Technology & AI
- OpenAI claims AI solved Navier-Stokes problem Technology & AI
- OpenAI $5M grant funds research on AI and teens Technology & AI
- AI safety paper: refuse harmful prompts, keep benign ones Technology & AI
- Home Assistant trends on GitHub at 90,000 stars Technology & AI
- Banning personal AI at work barely cuts risk, study finds Technology & AI
- Hugging Face datasets trends on GitHub as AI data prep scales Technology & AI
- OpenWhispr: free voice dictation app hits GitHub trending Technology & AI
- Google Pics: AI image editing inside Docs and Slides Technology & AI
- Qwen, Mistral and Llama verify fake developer identity, study finds Technology & AI
- X For You feed algorithm open-sourced on GitHub Technology & AI
- OpenAI says coding agents reshape its own AI research Technology & AI
- OpenAI Codex CLI trends on GitHub with 121,803 stars Technology & AI
- ACToR retrieval lifts AI code generation 15% on CoderEval Technology & AI
- AI agents learn your standards, up to 20.9% better Technology & AI
- LoopX adds durable memory to Claude Code and Codex agents Technology & AI
- GPT-6 Astra: OpenAI's first model to hit Critical cyber level Technology & AI
- Google Fairwind Program: AI writes verified patches in minutes Technology & AI
- Environment evolution for terminal agents: 18-point gain Technology & AI
- Vercel Labs agent-browser hits 41,800 GitHub stars Technology & AI
- Voice agents drop 9.7% when instructions are merely implied Technology & AI
- When to trust LLM recommendations: four-tier framework Technology & AI
- New technique cuts AI agent wait time up to 45% Technology & AI
- GPT-6 Astra crosses OpenAI's Critical cybersecurity threshold Technology & AI
- Hermes Agent: 240k GitHub stars for self-improving AI Technology & AI
- Gilbert + Tobin deploys ChatGPT Enterprise firm-wide Technology & AI
- Prompt injection attack on AI agents jumps to 28% success Technology & AI
- AIMC dashboard flags recurring flaws in AI-generated science Technology & AI
- ToolSiphon drains 74% of data from LLM agent tools Technology & AI
- AI morbidity and mortality framework proposed for hospitals Technology & AI
- AWS Firecracker microVM tops 36,443 GitHub stars Technology & AI
- Synthetic data privacy is a claim, not a guarantee, researchers warn Technology & AI
- DiaSentinel AI agents screen diabetes risk on-premise Technology & AI
- OpenAI's Astra hits Critical cybersecurity threshold Technology & AI
- MMJailBench: prompt framing is top jailbreak risk across 16 AI models Technology & AI
- Open-source trust signals are breaking, study finds Technology & AI
- ComfyUI hits GitHub trending at 130,000 stars Technology & AI
- NVIDIA's $3.5B MediaTek deal extends AI from cloud to car Technology & AI
August 2026
- last30days skill hits 60,000 stars searching Reddit and X Technology & AI
- LLM security agents act but lack guardrails, review finds Technology & AI
- Osmantic ODS turns your PC into a private AI server Technology & AI
- GPT-4o leaks secrets 100% when prompt injection is reframed Technology & AI
- Addy Osmani's agent-skills hits 90,000 GitHub stars Technology & AI
- IBM Research open-sources AI policy schema for GenAI apps Technology & AI
- TradingAgents hits 100K stars with data-leak fix in v0.3.1 Technology & AI
- NVIDIA RTX Spark adds EA, Ubisoft games ahead of fall launch Technology & AI
- Claude Code swarm tool Ruflo hits 69,000 GitHub stars Technology & AI
- New system spawns Claude Code agents with pre-loaded memory Technology & AI
- wayfinder: map AI projects too big for one session Technology & AI
- New STAR metric catches AI translations that drop sentences Technology & AI
- Google Gemini Omni 1.1 Flash: 10x more scene context Technology & AI
- PyTorch hits GitHub trending at 102,613 stars Technology & AI
- DeepMind's double-blind AI test locks benchmarks in crypto box Technology & AI
- AI coding agent defense cuts malware severity 83% Technology & AI
- WebMCP-Phalanx blocks 80 of 80 prompt injection attacks in browser agents Technology & AI
- StepGuard blocks AI agent attacks 77% before they run Technology & AI
- LLM agents run controlled experiments on pharma simulations Technology & AI
- TradingAgents nears 100,000 GitHub stars with AI trading desk Technology & AI
- LanceDB: 11,000-star vector database trends on GitHub Technology & AI
- Vibe coding: security prompt halves AI app flaws Technology & AI
- Google Gemini CLI hits 106,000 stars with free 1,000-request daily tier Technology & AI
- LLMs hallucinate more under strict EU rules, study finds Technology & AI
- SRPO trains Qwen3-8B to fix its own errors, hits 73.3% on AIME'24 Technology & AI
- Distilled AI safety guard runs on CPU in 24ms, matches teacher Technology & AI
- MiroFish: 71,000-star AI prediction engine hits GitHub trending Technology & AI
- NVIDIA Groq 3 LPX hits full production at 3,400 tokens/second Technology & AI
- TokEval: tokenizer metrics predict AI model performance Technology & AI
- Vite still trending on GitHub with 82,454 stars Technology & AI
- Docling hits 65,000 stars, turns PDFs and video into AI-ready data Technology & AI
- Bioscience AI needs trust checks before lab action, preprint says Technology & AI
- BERT-LER: explainable AI reads 75 million health records Technology & AI
- LLM corrections usually die with each session, arXiv preprint says Technology & AI
- n8n hits 201k GitHub stars but isn't open source Technology & AI
- Nakama AI agent platform trends on GitHub at 259 stars Technology & AI
- Tencent AI-Infra-Guard: 5,200-star AI red team tool trends Technology & AI
- Neurosymbolic world model transfers tasks without retraining Technology & AI
- Claude Code v2.1.239 adds cost tracking, hits 142k stars Technology & AI
- FedLNS catches rogue clients in federated LLM training Technology & AI
- GxP-Agent hits 100% on clinical trial coding benchmark Technology & AI
- DeAR: AI agents reason peer-to-peer without a central boss Technology & AI
- Looped LLMs improve multi-step AI tool calling, study finds Technology & AI
- Agentic AI review on arXiv as OpenAI agent repo nears 29,000 stars Technology & AI
- LeakGauge detects AI context-leakage attacks, AUROC to 0.996 Technology & AI
- NVIDIA uses ChatGPT Work to scale internal expertise Technology & AI
- StagedWorkspace lifts AI agent office task scores by 34 points Technology & AI
- Alibaba's Wuying browser agent hits 65% on 38-step web tasks Technology & AI
- Agent memory boosts gpt-oss 16 points but does nothing for GLM-5 Technology & AI
- OpenAI adds safeguards to pace frontier AI model development Technology & AI
- Hugging Face adds multi-vector retrieval to sentence-transformers Technology & AI
- Euclid-Omni: AI for Olympiad geometry with far less compute Technology & AI
- EU, US and China AI rules diverge, new study warns Technology & AI
- AI lock-in is already happening, researchers warn Technology & AI
- OpenAI's Defender's Window: AI reshapes cyber defense Technology & AI
- AI generates synthetic health data across multiple tables Technology & AI
- Agentao open-source runtime governs AI agent tool use Technology & AI
- SearchAuditor fixes 32% of AI agent failures, benchmark shows Technology & AI
- Federated learning privacy error scaling cut from 4^b to 2^b Technology & AI
- Chiplet and AI chip-design security threats mapped in new preprint Technology & AI
- OlmoEarth Studio exports AI embeddings as GeoTIFFs Technology & AI
- VLA robots fail 100% under sticker attack, defense cuts to 26% Technology & AI
- Strands Robots: one agent records, trains, deploys robot skills Technology & AI
- Telegram Mini Apps: 59% contact hidden third parties Technology & AI
- AI agent attack inflates costs 92% without breaking tasks Technology & AI
- OpenAI names Dali Rajic Chief Revenue Officer Technology & AI
- AI agents self-evolve security defenses in new HARD framework Technology & AI
- ABS housing finance falls 5.4% as investors retreat Property & Economy
- ICML 2026: 23% of papers have a falsified or contested claim Technology & AI
- AI agent runs vertical farm, cuts grow cycle 35% Technology & AI
- OpenAI Ultrafast runs GPT-5.6 Sol at 14X speed on Cerebras Technology & AI
- GPT-5.6 builder guide pushes cheaper, faster AI agents Technology & AI
- Hardware monitor catches attacks software misses Technology & AI
- Reinforcement learning cuts AI training power violations 89% Technology & AI
- AI image models fingerprinted without watermarks Technology & AI
- AI agent rewrites its own code, hits 22% on DBpedia Technology & AI
- Liquid AI's 3B vision model jumps 54% on grounding Technology & AI
- Model ML runs finance work on GPT-5.6 Sol, outputs editable decks Technology & AI
- Self-evolving GUI agents improve click accuracy 7.4% after deployment Technology & AI
- AI models lose over 90% of safety signal in African languages Technology & AI
- Ephemeral coin tracing limits surveillance power in crypto Technology & AI
- AI interaction creates behavior no model shows alone Technology & AI
- Google AMIE medical AI matches doctors in video consults Technology & AI
- Smart meter cyberattack could cost $5,097 a day Technology & AI
- ABS building approvals rebound 7.2% but units still fall Property & Economy
- Google puts AI agent Ask Advisor inside Ads and Analytics Technology & AI
- NVIDIA Magpie TTS expands to 12 languages with open weights Technology & AI
- KnowPlan AI agents plan degrees with 99.5% certified accuracy Technology & AI
- Self-evolving AI agents stumble under real task streams Technology & AI
- Google Cloud scanner catches AI safety tampering in 10 of 14 models Technology & AI
- LLM agent: code-only verification flips goal abandonment 100% to 0% Technology & AI
- NVIDIA Cosmos 3 open model combines three physical AI skills Technology & AI
- Voice input degrades LLM agents more than typing, study finds Technology & AI
- LLM interpreter explains outputs with no extra API calls Technology & AI
- Agentic AI bottleneck is the CPU, not GPU, study finds Technology & AI
- NVIDIA joins NSF AI hubs to expand US university compute access Technology & AI
- XSec: self-explainable AI hits 97% accuracy in security Technology & AI
- DreamGuard stops risky AI agent actions in 25 ms Technology & AI
- SkillTrace audits LLM agent skill reuse at 0.938 AUROC Technology & AI
- MedUPS lifts medical AI next-step accuracy 11 points Technology & AI
- RAG study tests LLaMA, Mistral and Qwen to cut AI hallucinations Technology & AI
- Chained RLM architecture restarts LLM reasoning with fresh context Technology & AI
- GPT-5.6 Sol improved, free ChatGPT access expanded Technology & AI
- GPT-5.6 study: max reasoning effort, zero unauthorized tool calls Technology & AI
- Five cognitive gaps that break AI agents on long tasks Technology & AI
- Self-improving AI agents reward their own mistakes Technology & AI
- Video-DeepResearch 35B beats GPT-5 on video reasoning, 64% vs 52.5% Technology & AI
- XGBoost hits 98.62% malware detection accuracy in new preprint Technology & AI
- AI persona agents leak private traits, defenses fail Technology & AI
- TrainShield serves AI security lessons when phishing risk hits Technology & AI
- DeBERTa-Sentinel detects AI text at 98%, shows its reasoning Technology & AI
- Circles lifts telco ARPU 22% with OpenAI API and Codex Technology & AI
- AI agent evaluation ignores time: this preprint fixes it Technology & AI
- SkillBoost stops AI agents forgetting old skills Technology & AI
- EvoPINN: AI agent invents new neural network for physics Technology & AI
- TAPR auto-rewrites LLM prompts to lift benchmark accuracy Technology & AI
- AI science papers score 2.47 out of 5 in first AI peer-review test Technology & AI
- GLASS steers AI text style without retraining or retrieval Technology & AI
- Federated learning predicts machine failure without sharing data Technology & AI
- RBA survey: most Australians get interest rates backwards Property & Economy
- MalGuard maps code structure to catch malware hiding in bytes Technology & AI
- OpenAI model proves ten math results with Lean certificates Technology & AI
- SCAIR steers AI agents with enterprise schemas, no retraining Technology & AI
- SeT-Diff: first foundation model for supercomputer telemetry Technology & AI
- OpenAI pledges responsible AI in Europe as EU Act advances Technology & AI
- AI agents do 100 office tasks cheaper than humans, but not as well Technology & AI
- World model AI can be tricked into false safety Technology & AI
- OpenAI 'abundant intelligence' post lays out AI cost flywheel Technology & AI
- MTGuard paper targets malicious MCP tool use in AI agents Technology & AI
July 2026
- RSMeM: satellite AI agents learn from mistakes, 6% accuracy gain Technology & AI
- Change2Task turns pull requests into coding agent tasks at 79.6% Technology & AI
- AI agent trust model cuts telecom cascade from hours to real-time Technology & AI
- TraceCoder adds audit trail to AI-generated code Technology & AI
- AI search lifts e-commerce discovery to 80% at 30% cost Technology & AI
- 23 AI agents tested on breach response: zero passed Technology & AI
- Reinforced Dreamer fixes flaw in AI world model training Technology & AI
- Kernel Forge: AI agent writes CUDA kernels, 2.83x speedup Technology & AI
- Will the RBA Raise Rates in August 2026? Our Model Just Flipped to Hike Property & Economy
- LLMs flip answers 23% of the time when you rephrase the question Technology & AI
- ABS CPI release moves earlier from 2027, skips weight update Property & Economy
- OpenAI field report: AI agents accelerate genomics research Technology & AI
- Gemini API agents get execution hooks and 3.6 Flash default Technology & AI
- MIT preprint: fluid search cuts autonomous AI research cost Technology & AI
- Active Directory security thesis rethinks defence with decoys Technology & AI
- AI reasoning models fail from self-doubt, not capability gaps Technology & AI
- ISPCloak fools deepfake detectors by faking camera noise Technology & AI
- OpenAI research: ChatGPT users take on tasks across roles Technology & AI
- FineServe shows LLM serving traffic varies by model type Technology & AI
- OPTScientist: AI agents discover new transformer optimizer Technology & AI
- Paired sampling cuts variance in domain adaptation training Technology & AI
- Ex-Fuzzy library adds regression with 15 readable rules Technology & AI
- Adversarial Frontiers paper rethinks AI robustness testing Technology & AI
- Language models overgeneralise and break on nested rules, study finds Technology & AI
- HyGRL targets multi-entity RAG with hybrid graph reasoning Technology & AI
- CARGO beats trained LLM routers without training data Technology & AI
- LLM agents pass 100% JSON schema checks, still fail 1 in 5 orders Technology & AI
- AI agents turn robot videos into physics simulations Technology & AI
- JAXBench: curated TPU docs lift AI kernel correctness 5.8% to 37% Technology & AI
- LLM confidence scores fail basic coherence test Technology & AI
- Small AI models outperform LLM prompt guardrails, Fence says Technology & AI
- LLM steering vectors get probabilistic fix in new paper Technology & AI
- AI licence laundering: under 7% of restrictive terms survive Technology & AI
- NVIDIA GB300 chips built in Fort Worth, 500 jobs live Technology & AI
- PEARL 4B model beats DeepSeek 685B at optimization modeling Technology & AI
- DBMol designs drug molecules using AI structure prediction Technology & AI
- AI safety paper: the quiet failures nobody measures Technology & AI
- Australian wages growth steady at 3.3%, private sector cools Property & Economy
- Text-prompted visual similarity metric exposes AI blind spot Technology & AI
- MAR-12 AI detects harmful humor in memes at 80.3% accuracy Technology & AI
- LLMs hold stable risk attitudes across tasks, study finds Technology & AI
- Text-to-video jailbreak gains 18.6% over closest rival Technology & AI
- OpenAI launches ChatGPT small business program Technology & AI
- GigaPath-Flash cuts pathology AI compute 50x, keeps 97% accuracy Technology & AI
- New LLM agent memory system MOSAIC hits 89% accuracy Technology & AI
- One AI controller runs 25 different systems without retraining Technology & AI
- NVIDIA Spectrum-6 doubles AI factory network to 102.4 Tbps Technology & AI
- True-fact attack hijacks RAG agents 83% of the time Technology & AI
- 23 attack paths target AI agent memory, OS defenses miss some Technology & AI
- LaCache speeds diffusion LLMs 1.3X with training-free caching Technology & AI
- RLHF bias audit: rater mood may skew AI preference data Technology & AI
- AI models guess your real intent only 22-32% of the time Technology & AI
- Systemic AI risk paper reframes harms as externalities Technology & AI
- AlphaFold2 weights hide protein folding landscapes, study finds Technology & AI
- DSWorld cuts AI data science agent training 14x Technology & AI
- AI trust gap leaves firms unable to prove safety, arXiv paper says Technology & AI
- arXiv paper revives rule-based AI for transparency via optimization Technology & AI
- AI safety thresholds: preprint proposes common frontier standard Technology & AI
- Gasp paper proposes bridge-free cross-chain DEX rollup Technology & AI
- Prompt syntax shifts open LLM code security, preprint finds Technology & AI
- Byzantine fault tolerance breaks with AI agents, preprint warns Technology & AI
- CRAFT method finds why LLMs fail, then fixes them Technology & AI
- SciForge: open-source AI workbench for scientists Technology & AI
- 5G NR side-channel attack cuts video quality 50% Technology & AI
- ToolVerse trains AI agents across 4500 real-world tools Technology & AI
- OpenAI reports failures in long-running AI models Technology & AI
- NVIDIA Cosmos 3 Edge: 4B model runs robots at 15 Hz Technology & AI
- PagedWeight cuts MoE serving memory 72% with FP16 accuracy Technology & AI
- AI trustworthiness framework tracks model drift with auditable levels Technology & AI
- AI trust tools favour fairness over security, arXiv study finds Technology & AI
- NVIDIA connects AI agents to creative tools via MCP at SIGGRAPH Technology & AI
- Productivity grows, wages lag: Australia's housing squeeze Property & Economy
- LatentFlow conditions stochastic processes in seconds, no training Technology & AI
- New framework tests if frontier AI helps plan CBRN attacks Technology & AI
- Qubes OS: 80% of security advisories trace to upstream code Technology & AI
- Mycelium AI routes science context across human-agent teams Technology & AI
- Knowledgeless AI models cut hallucination, boost evidence use 25% Technology & AI
- OpenAI makes case for safe teen ChatGPT access Technology & AI
- Autoformalization paper calls for theory-level AI math Technology & AI
- Traffic forecasting: simple mixing matches attention at 0.14% gap Technology & AI
- MusicMark watermarks AI-generated music during creation, not after Technology & AI
- Smart greenhouse RL audit splits reward into climate signals Technology & AI
- GNSS spoofing detection framework hits 95% accuracy in 3GPP networks Technology & AI
- LoRA cascaded fusion preprint targets medical training AI Technology & AI
- LLM accuracy hides prediction flips from irrelevant context Technology & AI
- RISC-V post-quantum crypto extension hits 129x speedup Technology & AI
- TerraZero: zero-demo self-driving AI hits 1.3M steps/sec Technology & AI
- EVOQUANT lifts trading Sharpe from -0.30 to 0.54 Technology & AI
- LLM answers silently shaped by own values, study finds Technology & AI
- Quantum-safe anonymous certificates proposed ahead of NIST's 2030 ECC deadline Technology & AI
- WarpGuard: first joint CPU-GPU attestation, no new hardware Technology & AI
- NVIDIA NeMo Automodel trains Diffusers models with no code rewrites Technology & AI
- malcos auto-extracts CPU leakage contracts for x86 and ARM Technology & AI
- Federated learning protocol cuts communication 100x with one server Technology & AI
- MCP security scanner FlowGuard finds 523 flaws in 326 servers Technology & AI
- AI agent benchmark tests tool-switching when reliability shifts Technology & AI
- Wind and solar prediction method cuts compute cost 21% Technology & AI
- AI agent safety monitor cuts covert sabotage to zero Technology & AI
- Traffic center AI costs $34/month with mixed model portfolio Technology & AI
- Evo 2 probes detect AMR genes with 0.977 ROC-AUC Technology & AI
- DREA AI vulnerability detector cuts API cost 48x Technology & AI
- Medical AI scores hide flawed reasoning, preprint finds Technology & AI
- Safety Sentry gives AI agents a third option: ask first Technology & AI
- New PUF design cuts wearable crypto power 89% to 2.7 μW Technology & AI
- Graph tools lift small language model molecular prediction 74% Technology & AI
- AI uncertainty measures unified in new risk decomposition paper Technology & AI
- MOJO framework decodes brain signals with little labelled data Technology & AI
- Quantum ML preprint targets post-quantum cryptography resilience Technology & AI
- Harness Handbook maps AI agent behaviour to source code Technology & AI
- Google Search now connects your apps inside AI Mode Technology & AI
- GPT-4o audit finds 66% of solved problems have flawed reasoning Technology & AI
- AI agent harness evolution may not beat simple search Technology & AI
- Item response theory for AI benchmarks gives unreliable rankings Technology & AI
- AI agent skills carry security risks beyond prompt injection Technology & AI
- Vulnerability scanner divergence explained in new arXiv framework Technology & AI
- LLM framework automates adversary emulation at 84% success Technology & AI
- RL agent stabilises pendulum beyond Kapitza with Lyapunov reward Technology & AI
- LLM vulnerability detection hits 100% recall, then collapses on real code Technology & AI
- SWE-bench needs 90% of tasks for reliable agent benchmark results Technology & AI
- AI coding agents install malicious packages from README edits Technology & AI
- AutoSynthesis automates meta-analysis end-to-end with AI agents Technology & AI
- Retrain-free recommendation system serves new users in under 1ms Technology & AI
- OpenAI CFO introduces four-metric AI ROI scorecard Technology & AI
- LLM agent memory poisoning evades defenses, 1,227-case study finds Technology & AI
- ReBound caches past queries to cut differential privacy cost Technology & AI
- AI penetration testing needs behavioral rethink, preprint argues Technology & AI
- LLM agents spot 88% of supply-chain failures but can't act Technology & AI
- LLM security log analysis hit by 88% prompt injection attack rate Technology & AI
- AI security agent costs: offense scales, defense doesn't Technology & AI
- NVIDIA Vera Rubin targets intelligence per dollar for AI agents Technology & AI
- AI coding agents: median GitHub repo files just 1-2 PRs Technology & AI
- Google DeepMind's bioresilience plan spans 15 partnerships Technology & AI
- RAG translation system pushes LLMs past sentence-by-sentence limits Technology & AI
- Oracle agent memory cuts token use 10.7x Technology & AI
- Albanese fast-tracks data centres with grid power pledge Technology & AI
- Self-improving AI agents: survey maps how agents edit themselves Technology & AI
- PriEval-Protect scores hospital privacy risk with legal LLM Technology & AI
- Geospatial AI paper maps path from satellite data to agents Technology & AI
- DROPJ trains safe AI agents from human justifications Technology & AI
- Google Vids adds text-to-video editing and selfie avatars Technology & AI
- Adversarial prompting framework finds encoded attacks bypass AI safety Technology & AI
- Agent-ready websites nearly double AI shopping agent success Technology & AI
- OriginBlame cuts AI training over-deletion from 101x to 1.3x Technology & AI
- Editing AI reasoning chains cuts tokens 40%, correction 25% Technology & AI
- AI upskilling framework: 3 learners pass NVIDIA Agentic AI exam Technology & AI
- NVIDIA Cosmos 3 Edge puts 4B-parameter AI on Japanese robots Technology & AI
- AI watermarking paper maps exact token cost of tracing users Technology & AI
- Moving target defence drops attack success as low as 4% Technology & AI
- PVDetector catches prompt injection attacks with under 1% miss rate Technology & AI
- Claude Sonnet beats GPT-4.1 on cost despite higher token price Technology & AI
- AI agents burn 91% of tokens on simple edits, study finds Technology & AI
- Audio deepfake detector uses Wiener-Hopf math for explainable AI Technology & AI
- LLM agents hallucinate 36.9% of skill names, study finds Technology & AI
- Robot swarm AI: study exposes hidden geometry in collective behaviour Technology & AI
- AI agents may train your brain toward frustration, preprint says Technology & AI
- Spectrum can't predict when context helps time-series forecasting Technology & AI
- LLM rubrics for AI grading biased toward high scores Technology & AI
- EnCF data assimilation filter handles non-Gaussian observations Technology & AI
- OpenAI GPT-Red automates red teaming with self-play Technology & AI
- OpenAI proposes 'reverse federalism' for US AI safety rules Technology & AI
- LLM judges flip 85% of verdicts when given a reference answer Technology & AI
- DiffusionGemma transcribes speech in 8 parallel steps, 6.6% error rate Technology & AI
- Physics-based AI fall detection runs on 50,000 parameters Technology & AI
- Bulkhead uses LLMs to detect and patch container escape bugs Technology & AI
- MindReader: LLM tool makes replacement passwords harder to guess Technology & AI
- 30 prompts fine-tune AI to near-optimal energy storage control Technology & AI
- LLM judge bias found in hidden states, steerable in both directions Technology & AI
- MCP security scanners: fewer than half of alerts valid Technology & AI
- Prezta replaces security gateways with zero-knowledge proofs Technology & AI
- Quantum neural network backdoor attack adapts per input Technology & AI
- StoryTeller: training-free AI for film audio descriptions Technology & AI
- Transformer learning dynamics reduced to few coordinates Technology & AI
- LLM temperature setting steers ideological bias in RAG answers Technology & AI
- Brisbane property market slows as buyers take the wheel Property & Economy
- GPT-5.5 scores 66 on doctoral math proof benchmark Technology & AI
- AI reasoning scaffold helps one model, hurts another Technology & AI
- Federated learning car security built on flawed tests, 60-paper audit Technology & AI
- AHA red-team agent finds reusable flaws in Claude Code, Codex Technology & AI
- FactorDiff routes diffusion experts by pixel on ARC-AGI Technology & AI
- AMT-X multi-turn red teaming hits 100% attack success on frontier LLMs Technology & AI
- Playful AI email tone lifts positivity in 16,880-email field study Technology & AI
- PromptGraph turns LLM prompts into privacy graphs Technology & AI
- ARMOR-IMC preprint targets 50% accuracy loss in AI memory chips Technology & AI
- EvoCUA-1.5 hits 63.2% on OSWorld with online RL training Technology & AI
- Best AI agents fail 50% of visual tool tasks, Apple benchmark shows Technology & AI
- HiFi-LLP cuts neural architecture search time 8.6× Technology & AI
- AI vision models encode correct count but output wrong numbers Technology & AI
- Spotify study links narrator voice to audiobook appeal Technology & AI
- Quantum learning separation proved for many-body dynamics Technology & AI
- Solidity compiler: 25 hidden miscompilation bugs found Technology & AI
- AI music transcription scores 38% on new pop benchmark Technology & AI
- Dynamic Fréchet Regression predicts distributions, selects features Technology & AI
- SurvFM-RMST lets tabular AI models handle censored survival data Technology & AI
- Lean-QIT: machine-checked proofs for quantum information theory Technology & AI
- MedRealMM benchmark: AI matches doctors but fails on safety Technology & AI
- Knowledge graphs and XAI reshape urban mining audits Technology & AI
- Physics-constrained ML cuts combustion simulation cost 10x Technology & AI
- LLM agent coordination framework cuts communication 70x Technology & AI
- Malaika AI framework uses tri-grounded reasoning for malware analysis Technology & AI
- TrustX ARC rates AI agent risk across 12 dimensions Technology & AI
- LLM agent framework blocks hallucinated actions in industrial control Technology & AI
- KV-PRM cuts AI agent scoring cost 5,000x via cache reuse Technology & AI
- CogniConsole: LLM reliability tied to control, not capability Technology & AI
- ARCANA multi-agent framework targets ARC-AGI-2 reasoning Technology & AI
- LLM agents fail 70% of skill safety tests in SLBench study Technology & AI
- Agora preprint: auction-based routing for LLM agents Technology & AI
- Neural network backdoors evade detection even with full weight access Technology & AI
- New arXiv paper reframes AI output as representation, not fact Technology & AI
- Visual pretraining outperforms text-only for language AI Technology & AI
- New protocol forces AI scientist agents to show their work Technology & AI
- 4DR360 radar-camera AI gives self-driving 360° scene awareness Technology & AI
- LLM chip design survey maps path to autonomous EDA agents Technology & AI
- ReContext fixes LLM long-context blind spots without retraining Technology & AI
- LLM-as-a-Verifier hits 86.5% on Terminal-Bench V2 Technology & AI
- VAORA aligns vision-language model reasoning with physical actions Technology & AI
- NARAD voting protocol tallies 50,000 ballots in under one second Technology & AI
- Quantum stabilizer testing fails even with 99% memory Technology & AI
- Label-free AI sorts real cosmic signals from noise Technology & AI
- QCNN with path signature kernels targets time series Technology & AI
- Orcaella protocol: choose fast commit or 54% fault tolerance Technology & AI
- CamVLA: robots work after cameras move, no recalibration Technology & AI
- Agentic Data Environments: turning data into agent guardrails Technology & AI
- LLM personality traits mapped and controlled in weight space Technology & AI
- Encrypted AI inference: layer-parallel cuts bootstraps 2.65× Technology & AI
- xDECAF open-sources data flow security tool, 20+ models Technology & AI
- AI image editing loses spatial accuracy — here's why Technology & AI
- FootsiesGym: open-source fighting game AI benchmark Technology & AI
- Sampling lets AI models generalise to unseen sizes Technology & AI
- AI world model's 90% accuracy was a cheat — fix recovers 88% Technology & AI
- AI bionic limbs open new privacy attack surface Technology & AI
- XGBoost security classifiers' 0.98 robustness collapses to 0.36 Technology & AI
- Neurosymbolic AI paper fuses logic rules with energy models Technology & AI
- MPC custody paper makes post-quantum migration a key rotation Technology & AI
- GPT-5.5 leads EvoPolicyGym: top-two across all 16 environments Technology & AI
- DynaKRAG makes multi-hop RAG adaptive, hits 0.60 F1 on HotpotQA Technology & AI
- PPGNN brings personalized privacy to decentralized graph learning Technology & AI
- Hugging Face Hub kernels add code signing and trusted publishers Technology & AI
- SOAP beats Adam for faster molecular AI training Technology & AI
- English reasoning closes multilingual LLM uncertainty gap Technology & AI
- AI inference: resample-or-reroute policy beats 5 cost baselines Technology & AI
- PeTeR hardens probabilistic circuits without retraining Technology & AI
- TopoBrick forecasts building sensors without training data Technology & AI
- AI chain-of-thought monitoring backfires, raising harmful approvals 9.5% Technology & AI
- CodeTracer traces backdoored AI code completions to source Technology & AI
- Three hikes, then a pause: the public data lean HOLD for the RBA’s August decision Property & Economy
- LLM personas generate executable smart-home schedules Technology & AI
- i-EXAM maps network attack paths with AI explanations Technology & AI
- Signed QR codes could end quishing, preprint proposes Technology & AI
- Distributed AI training: MIM beats contrastive learning on non-IID data Technology & AI
- AI scientists score 27% on scientific lineage benchmark Technology & AI
- Chain-of-thought audit catches AI reasoning gaps Technology & AI
- CausalDS benchmark tests AI agents' causal reasoning Technology & AI
- New arXiv paper maps how AI-human trust gets exploited Technology & AI
- QANTIS quantum belief updates hold to 32 steps on IBM Heron Technology & AI
- Agentic AI outperforms single-LLM in insurance underwriting Technology & AI
- Medical AI survey tests 18 LLMs, finds general models beat specialists Technology & AI
- Infinity-Parser2 tops document parsing benchmarks at 87.6% Technology & AI
- AI agent harness blocks leaks, keeps full 120/120 utility Technology & AI
- EvoSOP lets AI agents write their own playbooks Technology & AI
- Scaling improves LLM social simulation — but not human biases Technology & AI
- LeRobot v0.6.0 ships world models with zero inference cost Technology & AI
- OptiAgent turns everyday language into solver-ready optimization code Technology & AI
- Indic AI paper proposes Culture Sensing to preserve linguistic heritage Technology & AI
- Crossroads unifies Bitcoin, Ethereum, Solana on one smart-contract layer Technology & AI
- Subspace LoRA defence cuts fine-tuning poisoning to 8% Technology & AI
- Printed text slows AI robots nearly 7x via overthinking attack Technology & AI
- New framework unifies privacy and adversarial defense in distributed AI Technology & AI
- Browser crypto wallets leak address links across 35M users Technology & AI
- Homomorphic encryption preprint optimises activation approximation Technology & AI
- arXiv preprint proposes LLM workflows as knowledge objects Technology & AI
- UniClawBench tests AI agents on 400 real-world tasks Technology & AI
- SolarChain-Eval finds AI energy agents cheat physics Technology & AI
- Autonomous vehicle AI poisoning defence uses digital twins Technology & AI
- TokenWall cuts AI agent attack rate to 12.5% Technology & AI
- ProjAgent hits 41% on REPOCOD with procedural code retrieval Technology & AI
- Quantized LLMs behave differently despite matching accuracy Technology & AI
- SLORR cuts LLM training overhead below 1% Technology & AI
- OpenAI's ChatGPT Work agent runs across your apps for hours Technology & AI
- GPT-5.6 becomes preferred model in Microsoft 365 Copilot Technology & AI
- OpenAI GPT-5.5 Bio Bug Bounty targets biological jailbreaks Technology & AI
- OpenAI GPT-5.6 arrives with Sol, Terra and Luna trio Technology & AI
- ABS building approvals fall 1.1% as units dive 10.4% Property & Economy
- Wages growth holds at 3.3% as new home loans fall 6.2% Property & Economy
- Strong privacy noise cuts AI generalisation error Technology & AI
- ALER-TI fills time series gaps using historical retrieval Technology & AI
- CNeVA traffic sim adds controllability top models lack Technology & AI
- ActionCache speeds robot AI 34× without retraining Technology & AI
- G-RRM neural solver hits 33× speedup on Sudoku — with a catch Technology & AI
- 6,588 GitHub repos tagged with NAICS industry codes Technology & AI
- LLM refusal neurons: causal audit finds no unique mechanism Technology & AI
- MedPMC: 11M clean medical images beat bigger AI datasets Technology & AI
- LLM hidden states predict answer confidence before generation Technology & AI
- MCP server security: taint flaws mitigated via tool descriptions Technology & AI
- ChatGPT clicks out in 5.2% of sessions, study finds Technology & AI
- SciReasoner AI model wins 67 of 86 science benchmarks Technology & AI
- Optimal control grows neural network layers where error peaks Technology & AI
- OpenAI finds flaws in SWE-bench Pro coding benchmark Technology & AI
- OpenAI AI Skills Jams bring hands-on AI to K-12 teachers Technology & AI
- Anthropic finds a 'global workspace' inside Claude Technology & AI
- Discrete diffusion models: new proof unifies four leading methods Technology & AI
- GaP multi-agent harness self-learns robot tasks across 8 benchmarks Technology & AI
- Language models don't just measure culture — they shape it Technology & AI
- Cyber incident response taxonomy maps 417 studies Technology & AI
- FDIFormer: Transformer AI matches expert grid attack detection Technology & AI
- HilEnT turns malware into images for AI detection Technology & AI
- Codex and Claude Code: one benchmark run isn't enough Technology & AI
- RBA warns supply shocks now strike more often Property & Economy
- Backdoor attack hides multiple triggers in speech AI models Technology & AI
- Bit2Watt: 1,000 GPUs can destabilise data centre power grids Technology & AI
- AI agent memory injection attack hits 87.5% success rate Technology & AI
- World Models: arXiv Paper Defines AI's Most Debated Concept Technology & AI
- Neural networks beat NTK by 100,000x on compositional tasks Technology & AI
- Cortex framework chains 32 robot skills for long-horizon tasks Technology & AI
- CompactionRL lifts AI coding agents 7 points on SWE-bench Technology & AI
- Sydney home prices drop $31,000 — low-deposit buyers at risk Property & Economy
- Sydney and Melbourne property prices falling — HSBC warns worse is coming Property & Economy
- Nvidia RTX Spark: 1-petaflop power for personal AI Technology & AI
- kNNGuard: 10x faster AI safety without retraining Technology & AI
- XGBoost Scored 0.98 Against Hackers. Its Explanations Still Collapsed. Technology & AI
- Masked image modelling beats contrastive learning in messy distributed AI training, Sydney-led team shows Technology & AI
- Training-free AI guardrail runs 10x faster using 50 prompts Technology & AI
- Mobile coverage map standard 2026: ACCC warns on misleading claims Technology & AI
- Australian house prices cool as new home loans fall 6.2% Property & Economy
- New EvoPolicyGym benchmark forces AI agents to improve policies under a tight feedback budget Technology & AI
- Trust boundary semantic gap: Why passing every security test is no longer enough Technology & AI
- VeriChat: Multi-agent AI claims 87.73% faithfulness in chip-security verification Technology & AI
- WorldSample: New AI Method Cuts Real-Robot Training by 59% With Synthetic Data Technology & AI
- ACCC welcomes mobile coverage map standard and warns telcos on misleading claims Technology & AI
- Why scattered data breaks AI agents — and why masked modelling beats contrastive learning Technology & AI
- First unified cyber incident response taxonomy maps 457 sources across 25 years Technology & AI
- Controllable AI drivers: how ‘soft gates’ let you steer virtual agents without breaking physics Technology & AI
- A 90.9% backdoor catch rate: why 'boring' constraints could fix AI coding agents Technology & AI
- This Two-Signal Audit Spots ‘Abliterated’ Open-Weight AI Checkpoints With 95% Accuracy Technology & AI
- Can We Map the Wiring Inside AI Agents? A New Preprint Says Yes Technology & AI
- HaloGuard 1.0 promises elite AI safety in a lightweight open-weights package Technology & AI
- Quantum memory constraints collapse the testing-learning gap Science
- LIB-TRAP: When Standard Cell Libraries Become the Attack Vector Technology & AI
- New Transformer architecture targets extreme events in time-series forecasting Technology & AI
- EvoVuln and the $50 smart contract audit: too good to be unverified? Technology & AI