QUASAR: Lowering the Loss Floor of Quantization-Aware Training with Loss-Aware Reconstruction Paper • 2608.13966 • Published 22 days ago • 3
Qwen3.8 Collection Qwen3.8 Unsloth quants including Qwen3.8-27B! Run and train Qwen3.8 with the Unsloth Desktop app. • 9 items • Updated 7 days ago • 70
The FID Lottery: Quantifying Hidden Randomness in Generative-Model Evaluation Paper • 2606.20536 • Published Jun 18 • 13
view article Article Introducing North Mini Code: Cohere’s First Model For Developers CohereLabs • Jun 9 • 85
view article Article Fine-tune FLUX.2 [klein] with a LoRA under 60 minutes black-forest-labs • Jun 4 • 27
Qwen3.5 Collection Qwen3.5 is Qwen's new model family including Qwen3.5 Small: 0.8B, 2B, 4B, 9B and Qwen3.5 Medium: 35B-A3B, 27B, 122B-A10B and 397B-A17B. • 25 items • Updated 10 days ago • 167
view article Article GGML and llama.cpp join HF to ensure the long-term progress of Local AI +4 ggerganov, ngxson, allozaur, lysandre, victor, julien-c • Feb 20 • 510
SLA2: Sparse-Linear Attention with Learnable Routing and QAT Paper • 2602.12675 • Published Feb 13 • 59
Hibiki-Zero Collection Streaming speech translation without the need for word-level alignments • 4 items • Updated Jul 15 • 4