@YRSM_Simon:120 t/s!干得好,@UnslothAI

X AI KOLs Following 工具

摘要

Unsloth AI 宣布推出 DSpark,使 DeepSeek-V4-Flash GGUF 模型在本地运行速度提升约 1.4–2 倍,达到 120 tokens/s,且准确率无变化。

120 t/s!干得好,@UnslothAI
查看原文
查看缓存全文

缓存时间: 2026/08/07 04:50

120 t/s!干得漂亮,@UnslothAI

Unsloth AI (@UnslothAI): DeepSeek-V4-Flash 现在可以通过 DSpark 在本地以 2 倍速度运行!⚡️

DSpark 能让 V4-Flash-0731 的 GGUF 文件生成速度提升约 1.4–2 倍,且精度不变。

DeepSeek-V4-Flash-0731 可以达到 120 tokens/s。

GGUFs: https://t.co/ac6uOI8mZA 指南:

相似文章

@YRSM_Simon: 疯狂

X AI KOLs Timeline

DeepSeek-V4-Flash-DSpark 在 4 块 RTX PRO 6000 GPU 上实现了 328 tok/s 的单次推理和 1.7k tok/s 的批量吞吐量。