llama-cpp-integration

Tag

Cards List
#llama-cpp-integration

Transformers now runs llama.cpp quants

Hugging Face Blog · yesterday Cached

Hugging Face's transformers library now supports GGUF models from llama.cpp, enabling efficient local inference on consumer hardware through familiar APIs.

0 favorites 0 likes
← Back to home

Submit Feedback