Qwen3.6 35B A3B 无审查异端版原生MTP完整保留发布 KLD 0.0015, 10/100拒绝率 完整19个MTP保留 支持Safetensors、GGUF、NVFP4、NVFP4 GGUF和GPTQ-Int4格式
摘要
社区发布的Qwen3.6 35B A3B无审查变体版本,完整保留19个MTP张量,支持多种格式包括Safetensors、GGUF、NVFP4和GPTQ-Int4。
llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved: [https://huggingface.co/llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved](https://huggingface.co/llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved) llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-GGUF: [https://huggingface.co/llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-GGUF](https://huggingface.co/llmfan46/Qwen3.6-27B-uncensored-heretic-v2-Native-MTP-Preserved-GGUF) llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-NVFP4-Experts-Only-GGUF: [https://huggingface.co/llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-NVFP4-Experts-Only-GGUF](https://huggingface.co/llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-NVFP4-Experts-Only-GGUF) llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-NVFP4-Experts-Only: [https://huggingface.co/llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-NVFP4-Experts-Only](https://huggingface.co/llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-NVFP4-Experts-Only) llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-GPTQ-Int4: [https://huggingface.co/llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-GPTQ-Int4](https://huggingface.co/llmfan46/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-GPTQ-Int4) 应大家要求发布了,所有版本均已确认完整保留MTP张量数量。附带基准测试。所有模型可在此处查看:[HuggingFace-LLMFan46](https://huggingface.co/llmfan46/models) \*所有版本均已验证完整保留MTP张量。在Safetensors格式中,Qwen3.6-35B-A3B的MTP张量显示为19个条目,因为\`gate\_up\_proj\`存储为融合张量。在GGUF格式中,该融合张量拆分为独立的gate/up专家张量,因此相同的MTP组件显示为20个条目。数量因格式而异,但MTP张量均已完整保留。
相似文章
DavidAU/Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF
这是一个社区微调并合并的Qwen 3.5 9B版本,声称在关键基准测试中超越更大模型,支持多token预测,以GGUF格式提供。
Qwen 3.6 35B GGUF:跨GPU和CPU的NTP vs MTP量化结果
ByteShape发布了Qwen 3.6 35B GGUF的NTP和MTP变体量化,并在多个GPU和CPU上进行了详细基准测试,发现更大的量化模型通常优于较小的模型,MTP以内存为代价提供了GPU速度提升。
Qwen3.6-27B Uncensored Aggressive 发布,附带 K_P 量化!
社区释出去除安全拒答的 Qwen3.6-27B,并以专为 llama.cpp 与 LM Studio 优化的 K_P GGUF 量化格式打包。
@support_huihui: 新的MTP-GGUF:huihui-ai/Huihui-Qwen3.6-27B-abliterated-MTP-GGUF 这是Qwen/Qwen3.6-27B的无审查版本,通过abliteration创建...
huihui-ai在Hugging Face上发布了Qwen3.6-27B的无审查GGUF版本,通过abliteration创建。
Qwen3.8-27B Uncensored Aggressive 已发布,支持 K_P quants 和 HauhauCS FastMTP (TG 提升高达 3.02 倍)!
Qwen3.8-27B Uncensored Aggressive AI 模型现已发布,采用 K_P 量化和 HauhauCS FastMTP 技术,带来显著的性能提升,无内容过滤限制并支持多模态功能。