DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF Image-Text-to-Text • 27B • Updated 3 days ago • 518k • 432
QUASAR: Lowering the Loss Floor of Quantization-Aware Training with Loss-Aware Reconstruction Paper • 2608.13966 • Published 27 days ago • 4
GLM-5.3-Flash · Alis MLX Collection Rewritten 2026-08-30: router/mHC/KDA aux kept unquantized. 4bit=QUASAR-init (KL -8.5% vs RTN). ~29-33 tok/s on M3 Ultra. Receipts in each repo. • 4 items • Updated 11 days ago
avlp12/Kimi-K2.7-Code-Alis-MLX-Dynamic-3.6bpw-VLM Image-Text-to-Text • 1T • Updated 13 days ago • 814 • 2