DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF Image-Text-to-Text • 27B • Updated 13 days ago • 2.04M • 1.56k
QUASAR: Lowering the Loss Floor of Quantization-Aware Training with Loss-Aware Reconstruction Paper • 2608.13966 • Published Aug 14 • 6
GLM-5.3-Flash · Alis MLX Collection Rewritten 2026-08-30: router/mHC/KDA aux kept unquantized. 4bit=QUASAR-init (KL -8.5% vs RTN). ~29-33 tok/s on M3 Ultra. Receipts in each repo. • 4 items • Updated Aug 30