Attend to Your Own Thoughts: Post-Training Quantization of Reasoning LLMs to 1.58 Bits
Researchers introduced ScaleQ-1.58, a post-training quantization framework that compresses reasoning language models to 1.58 bits using a technique called Attend…
Researchers introduced ScaleQ-1.58, a post-training quantization framework that compresses reasoning language models to 1.58 bits using a technique called Attend…
Microsoft’s Agent Framework has reached general availability for its runtime components. The framework provides a supported harness that handles function…
Supabase released Evals, an Apache-2.0 licensed benchmark framework for evaluating AI coding agents on authentic engineering tasks. The system executes…
OpenAI describes GPT-Live, its third-generation voice system, which removes the traditional turn-detector component and instead streams audio continuously through a…
Hugging Face contributor sergiopaniego published a guide describing a framework for training coding agents using TRL’s AsyncGRPO algorithm together with…
Simon Willison released version 0.32 of his LLM command-line tool and Python library on August 4, 2026, adding support for…
Researchers introduce RestoreKV, a technique that addresses the accuracy loss large language models suffer under aggressive key-value cache eviction by…
Researchers Şuayp Talha Kocabay, Talha Rüzgar Akkuş, and Kamer Ali Yüksel of aiXplain propose ARCHead, a method for compressing the…
A team including Sitong Gong, Caixin Kang, and Yifei Huang presents GROVE, a training-free framework that organizes streaming video into…
Researchers propose TurnSight, a framework for multi-turn tool-integrated reasoning that addresses the credit-assignment problem by deriving training supervision from an…
Researchers including Wanshun Su, Yang Shi, and Feihu Liu, working across Northwestern Polytechnical University, Peking University, Alibaba Group, and Tsinghua…
A community project trained NanoColibri-Instruct, a 2.7-billion-parameter mixture-of-experts model, for a total cost of roughly $180-260 using a “relay pretraining”…