Unsloth 发布 GLM-5.2 的 GGUF 量化版本,支持在 llama.cpp 和 Unsloth Studio 本地运行,2-bit 量化后模型从 1.51TB 压缩至 238GB,保留约 82% 准确率 · AIWatch