Models Are Getting Dumber on Purpose
news
w4g1.dev
IrregularChat: AI & Autonomy
6d ago
The article argues that modern AI models are becoming “smarter” at reasoning while intentionally becoming less knowledgeable in their weights. Recent models such as GLM-5.2, Qwen3.5, and DeepSeek V4-F
Qwen/Qwen3.8-27B-FP8 · Hugging Face
news
huggingface.co
IrregularChat: Tech
1w ago
Qwen/Qwen3.8-27B-FP8 is an Apache-2.0 licensed, FP8-quantized version of Qwen’s Qwen3.8-27B multimodal model on Hugging Face. The repository provides model weights and configuration files in Transform
Qwen/Qwen3.8-27B · Hugging Face
news
huggingface.co
IrregularChat: AI & Autonomy
1w ago
Qwen/Qwen3.8-27B is an open-source multimodal model hosted on Hugging Face by Qwen and released under the Apache 2.0 license. It belongs to the Qwen3.8 family, described as the most capable generation
The report describes an RL training simulator designed to study the tradeoff between throughput and stability in asynchronous reinforcement learning. In this setting, samplers continuously generate ro
PrismML has released Bonsai 27B, described as the first 27-billion-parameter model able to run on a phone. Based on Qwen 3.6, the model has been compressed from about 54 GB at 16-bit precision to just
Bonsai 27B is a new multimodal flagship model based on Qwen3.6 27B, and it is presented as the first model in its capability class to run on a phone. The release extends Bonsai’s low-bit compression a
Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign · Hugging Face
news
huggingface.co
IrregularChat: Fabrication
Apr 17
Qwen/Qwen3-TTS-12Hz-1.7B-VoiceDesign is a Hugging Face text-to-speech model from the Qwen3-TTS family, released in January 2026 under the Apache-2.0 license. It is a multilingual speech generation sys
substack.com
news
substack.com
IrregularChat: Full Stack Dev
Mar 28