New Models

Prism ML Releases Ternary-Bonsai-2-27B-gguf-dev Build

Prism ML has released a dev/test GGUF build for Ternary-Bonsai-2-27B, based on Qwen3.8-27B, for llama.cpp Q2_0 packin ...

New Models

Intern-S2-397B GGUF Quantized Models Released by bartowski

Discover bartowski/Intern-S2-397B-GGUF, a multimodal foundation model quantized for local inference using llama.cpp.

New Models

Orion-26B-A4B-v1.1 GGUF Quants Released by bartowski

Explore bartowski’s GGUF quantizations for Orion-26B-A4B-v1.1, featuring multimodal support, performance tables ...

New Models

Signal-3.8-27B-GGUF: Faster and More Token-Efficient

Discover Signal-3.8-27B-GGUF, a minimally invasive fine-tune of Qwen3.8-27B offering lower latency, reduced token usa ...

New Models

bartowski/nex-agi_Nex-N2.5-mini-GGUF: Specs and Hardware

Explore the GGUF quantization of nex-agi’s multimodal agent model Nex-N2.5-mini by bartowski, including hardwar ...