llama-cpp-comparison

标签

Cards List
#llama-cpp-comparison

Ninfer 和 RTX 5090 配合 3.8 27B 模型让我喜极而泣,效果太棒了。

Reddit r/LocalLLaMA · 2天前

一位用户报告称,使用 Ninfer 工具在 NVIDIA RTX 5090 GPU 上运行 Qwen 3.8B 模型时,获得了高令牌吞吐量,显著优于 llama.cpp。

0 人收藏 0 人点赞
← 返回首页

提交意见反馈