GigaPath-Flash cuts pathology AI compute 50x, keeps 97% accuracy
GigaPath-Flash and GigaTIME-Flash shrink billion-parameter pathology models to run on 50x less compute while predicting tumor profiles from routine slides.
A newsletter breaking down AI research, technology, and Australian property in plain English
GigaPath-Flash and GigaTIME-Flash shrink billion-parameter pathology models to run on 50x less compute while predicting tumor profiles from routine slides.
NVIDIA's 102.4-terabit Spectrum-6 Ethernet switch is reaching gigascale AI factories first at CoreWeave, Microsoft and Nebius, doubling network capacity.
NVIDIA's new 4-billion-parameter open model generates 32 robot actions per inference on Jetson Thor, bringing real-time reasoning to edge devices for the
A new preprint claims PagedWeight dynamically quantizes MoE model weights at runtime, saving 72% GPU memory and lifting throughput 1.94× for AI inference.
A new simulator from Applied Intuition trains driving policies with zero human demonstrations, topping the InterPlan benchmark on a single GPU.
Researchers from NVIDIA and Oxford propose WarpGuard, first to jointly verify CPU and GPU code at runtime, targeting safety-critical embedded systems.
NVIDIA and Hugging Face shipped an open-source tool that fine-tunes video and image diffusion models across GPU clusters with no checkpoint conversion
A new arXiv framework cuts transportation management AI costs 97% by mixing open-source and closed APIs, with a greedy heuristic that picks the cheapest
NVIDIA's new on-device model lets Japanese robotics giants like FANUC and Yaskawa run vision and reasoning locally on edge hardware, adapting to new tasks
Bulkhead, a new arXiv preprint, uses multi-agent LLMs to automatically find and fix path traversal vulnerabilities in containers running AI workloads.