Run GGUF Models Directly in Transformers with llama.cpp Support
Hugging Face adds direct execution support for llama.cpp GGUF models in transformers, starting with local Qwen3.5 inf ...
News on open-weight models you can run locally
Hugging Face adds direct execution support for llama.cpp GGUF models in transformers, starting with local Qwen3.5 inf ...