Tencent AI-Infra-Guard: 5,200-star AI red team tool trends
Tencent Zhuque Lab's open-source tool scans AI agents, MCP servers and LLM jailbreaks. It has 5,200+ stars and a README that explicitly asks for them.
A newsletter breaking down AI research, technology, and Australian property in plain English
Tencent Zhuque Lab's open-source tool scans AI agents, MCP servers and LLM jailbreaks. It has 5,200+ stars and a README that explicitly asks for them.
An August 2026 arXiv preprint splits RL world models so reward prediction uses only symbolic state, letting agents switch tasks without further training.
A new arXiv preprint proposes screening federated LLM updates by tracking normalisation-layer changes, beating six baselines under 40% attack.
GxP-Agent encodes regulatory steps as a graph, turning 0% failure into 100% structural match on a new FDA-pilot clinical trial benchmark.
Cambridge researchers find looped language models improve at multi-step API tool calling, with adaptive inference offering the best compute-performance
A new arXiv preprint introduces LeakGauge, a detector that flags when LLMs leak confidential context, tested across 11 models with AUROC to 0.996.
Euclid-Omni couples LLMs and vision models with a formal geometry solver to match state-of-the-art on Olympiad proofs using less compute.
New arXiv preprint introduces Agentao, a local-first runtime that separates what LLM agents propose from what they can execute, with open-source code on
New benchmark of 1,243 failed agent runs shows even GPT-5.5 auditors fix only 26.6%, with SearchAuditor at 32.3%, a warning for agent teams.
An arXiv preprint maps security threats across chiplet systems and LLM-driven chip design tools, directly affecting fabless semiconductor teams adopting
OpenAI has appointed Dali Rajic as Chief Revenue Officer to lead its global revenue organization and help businesses turn AI into measurable value.
A new preprint proposes HARD, a framework where LLM agents automatically build and improve their own runtime defenses from observed failures.