TraceCoder adds audit trail to AI-generated code
A new arXiv preprint tracks every repair attempt in LLM code generation, giving developers a replayable audit trail for AI-written code.
A newsletter breaking down AI research, technology, and Australian property in plain English
A new arXiv preprint tracks every repair attempt in LLM code generation, giving developers a replayable audit trail for AI-written code.
A new open-source tool from University of Michigan researchers lets any PyTorch model get automatic CUDA kernel speedups without manual GPU programming.
Headline inflation fell to 3.8% on 29 July, yet services and non-tradables inflation both rose and the trimmed mean did not move. Our ten-driver framework crossed its hike boundary by 0.03 — with the four strongest objections to that reading set out in full.
LLM answers change on over 23% of questions when wording shifts, and standard accuracy scores hide the reliability gap from anyone deploying AI.
The June 2026 quarter CPI landed on 29 July with two structural changes: earlier monthly releases from February 2027 and no mid-year weight update
OpenAI's July 28 field report says AI coding agents speed up genomics discovery, but offers no benchmarks or named scientists.
Gemini API managed agents now run 3.6 Flash by default and let developers intercept tool calls with custom hooks, changing how agentic workflows run on
An MIT-linked arXiv preprint argues autonomous research systems are graded on final results but ignore how much compute they burn getting there, and
An arXiv thesis from Adelaide proposes decoy placement and admin feedback loops for AD security, but admits the underlying maths is intractable.