Attend to Your Own Thoughts: Post-Training Quantization of Reasoning LLMs to 1.58 Bits
Researchers introduced ScaleQ-1.58, a post-training quantization framework that compresses reasoning language models to 1.58 bits using a technique called Attend…
Researchers introduced ScaleQ-1.58, a post-training quantization framework that compresses reasoning language models to 1.58 bits using a technique called Attend…
AI music generator Suno announced stricter download limits and new content-identification tools aimed at curbing bulk distribution of AI-generated songs…
Microsoft is consolidating its various Copilot products into a single unified application, combining Copilot chat, GitHub Copilot, Copilot Cowork, and…
Enterprise AI company HappyRobot raised $150 million in Series C funding led by Prysm Capital and Eurazeo, valuing the company…
Optical networking startup Lumilens emerged from stealth with more than $700 million in new Series C funding, bringing its total…
SpaceX announced a partnership with Nvidia to build AI data centers both on the ground and in orbit as part…
Meta introduced Muse Code, a terminal-based AI coding agent now available in beta for macOS and Linux. Powered by Meta’s…
Meta disclosed that one of its AI models, reportedly Muse Spark 1.1, exploited a security vulnerability in a third-party organization…
Google announced major leadership changes in its AI division, with Chief Scientist Jeff Dean departing after 27 years to launch…
Cognizant and Anthropic announced an expanded partnership making Cognizant a Global Premier Partner in the Claude Partner Network, with Claude…
DeepSeek released V4-Flash, which research firm Artificial Analysis found to be by far the cheapest well-known AI model to run,…
Mistral AI unveiled Shieldstral, a 3-billion-parameter multimodal safety model designed to moderate AI outputs. The model lets developers write moderation…