OpenAI research: ChatGPT users take on tasks across roles
OpenAI's latest research says ChatGPT users are taking on tasks across roles, reshaping job boundaries. Here's what that means for workers.
The AI knowledge 99.9% of people never discover
Daily AI and technology coverage — models, chips, agents and research, and what each development actually means for you and your business.
551 stories
OpenAI's latest research says ChatGPT users are taking on tasks across roles, reshaping job boundaries. Here's what that means for workers.
FineServe, a new dataset from a commercial LLM marketplace, reveals serving traffic varies fundamentally by model architecture and task type.
OPTScientist uses four AI agents to discover optimizer algorithms for transformer pretraining, including a new reduced-state matrix optimizer called RS-MR.
An arXiv preprint proposes pairing data into quadruplets to cut gradient variance in unsupervised domain adaptation, improving target accuracy on three
New preprint extends Ex-Fuzzy with interpretable regression, hitting R² of 0.86 across 10 benchmarks using 10 to 15 human-readable rules.
Adversarial Frontiers, a new arXiv preprint, argues single-budget AI robustness rankings are unstable and proposes a frontier-based evaluation framework.
Language models build shared rules instead of storing individual facts, KAIST AI finds, and the overgeneralisation affects teams fine-tuning on nested
HyGRL, a new arXiv paper from Beijing Institute of Technology, blends text with knowledge graphs to answer multi-entity questions that defeat standard RAG.
A new arXiv preprint from Purdue and Princeton shows LLM query routing can work without training data, using model self-agreement as a signal.
New OrderBench benchmark runs 2,400 calls across four open models and finds schema validity alone doesn't guarantee semantic correctness.
Agentic Real2Sim uses vision-language agents to convert real robot recordings into simulatable twins, aiming to cut the labor cost of training robots in
JAXBench, the first TPU benchmark for AI kernel generation, shows curated docs lift correctness from 5.8% to 37.3% and reach 1.6x speedup over XLA.