transformers-update

Tag

Cards List
#transformers-update

GGUFs in transformers natively!

Reddit r/LocalLLaMA · 7h ago

Hugging Face announces native support for GGUF files in the transformers library, allowing easier use of quantized models with PyTorch tooling and performance comparable to llama.cpp.

0 favorites 0 likes
#transformers-update

Transformers now runs llama.cpp quants

Hugging Face Blog · yesterday Cached

Hugging Face's transformers library now supports GGUF models from llama.cpp, enabling efficient local inference on consumer hardware through familiar APIs.

0 favorites 0 likes
← Back to home

Submit Feedback