Orion-26B-A4B-v1.1 GGUF Quants Released by bartowski
Explore bartowski’s GGUF quantizations for Orion-26B-A4B-v1.1, featuring multimodal support, performance tables ...
Tencent Releases Simple Attention Sparsification for Qwen3
Tencent has released Simple Attention Sparsification (SAS) checkpoints for Qwen3 models, enabling efficient context p ...
Comfy-Org Releases YuE2 for ComfyUI: Open Music Generation
Comfy-Org has released Comfy-Org/YuE2, packaging the open-weight music generation model YuE2-3B for local execution v ...
DeepSeek-V4.1-Flash Uncensored FP8 Released
dealignai has released dealignai/DeepSeek-V4.1-Flash-UNCENSORED-FP8 on Hugging Face, an abliterated version of the 55 ...
Signal-3.8-27B-GGUF: Faster and More Token-Efficient
Discover Signal-3.8-27B-GGUF, a minimally invasive fine-tune of Qwen3.8-27B offering lower latency, reduced token usa ...
bartowski/nex-agi_Nex-N2.5-mini-GGUF: Specs and Hardware
Explore the GGUF quantization of nex-agi’s multimodal agent model Nex-N2.5-mini by bartowski, including hardwar ...
Together AI Expands Fine-Tuning Service with New Features
Together AI expands its fine-tuning service with new models, live metrics tracking, finer controls, and price reducti ...
Pantheon-Reasoning-26B-A4B-1.1-V2 GGUF Quantizations
Download bartowski’s GGUF quantizations for Gryphe Pantheon-Reasoning-26B, a Gemma 4 roleplay and reasoning mod ...
Edge0-35B-A3B-Preview: Sparse MoE for Phone-Class Memory
Edge0 released Edge0-35B-A3B-preview, a 35B sparse MoE model running on phone-class memory using streaming inference. ...
Lightricks Releases Multiple LoRA and IC-LoRA Adapters for LTX-2.5
Lightricks has released multiple LoRA and IC-LoRA adapters for Lightricks/LTX-2.5, enabling character consistency, ci ...