Back to Entity Graph

🏢 SGLang Organization

7 articles First seen: May 3, 2026 Last seen: 1w ago
Activity Timeline (90 days)
Co-occurring Entities
Articles (7)
Qwen/Qwen3.8-27B-FP8 · Hugging Face
Qwen/Qwen3.8-27B-FP8 is an Apache-2.0 licensed, FP8-quantized version of Qwen’s Qwen3.8-27B multimodal model on Hugging Face. The repository provides model weights and configuration files in Transform
Qwen/Qwen3.8-27B · Hugging Face
Qwen/Qwen3.8-27B is an open-source multimodal model hosted on Hugging Face by Qwen and released under the Apache 2.0 license. It belongs to the Qwen3.8 family, described as the most capable generation
Smaller, faster, safer: running Kimi and GLM at scale
Cloudflare’s Workers AI serves large open models such as Moonshot’s Kimi K-series and Z.ai’s GLM on GPUs close to users, but these models are difficult to host efficiently because of memory limits. To
moonshotai/Kimi-K3 · Hugging Face
Kimi K3 is Moonshot AI’s flagship open-weight multimodal model, hosted on Hugging Face as `moonshotai/Kimi-K3`. It is described as the company’s most capable model and the world’s first open 3T-class
The 4-bitter Lesson: Balancing Stability and Performance in NVFP4 RL
The report describes an RL training simulator designed to study the tradeoff between throughput and stability in asynchronous reinforcement learning. In this setting, samplers continuously generate ro
thinkingmachines/Inkling · Hugging Face
Inkling is a general-purpose multimodal model released by Thinking Machines Lab with open weights under the Apache 2.0 license. It is designed to handle text, image, and audio inputs and generate text
Keep the Tokens Flowing: Lessons from 16 Open-Source RL Libraries
The article examines why synchronous reinforcement learning (RL) training is inefficient at scale and how the open-source ecosystem has responded. In modern post-training, especially with long reasoni
🏠Portal 📰Links Q&A 📅Events 💼Jobs