New Models

Tencent Releases Simple Attention Sparsification for Qwen3

Tencent has released Simple Attention Sparsification (SAS) checkpoints for Qwen3 models, enabling efficient context p ...

New Models

DeepSeek-V4.1-Flash Uncensored FP8 Released

dealignai has released dealignai/DeepSeek-V4.1-Flash-UNCENSORED-FP8 on Hugging Face, an abliterated version of the 55 ...

New Models

Signal-3.8-27B-GGUF: Faster and More Token-Efficient

Discover Signal-3.8-27B-GGUF, a minimally invasive fine-tune of Qwen3.8-27B offering lower latency, reduced token usa ...

New Models

bartowski/nex-agi_Nex-N2.5-mini-GGUF: Specs and Hardware

Explore the GGUF quantization of nex-agi’s multimodal agent model Nex-N2.5-mini by bartowski, including hardwar ...

New Models

Pantheon-Reasoning-26B-A4B-1.1-V2 GGUF Quantizations

Download bartowski’s GGUF quantizations for Gryphe Pantheon-Reasoning-26B, a Gemma 4 roleplay and reasoning mod ...

New Models

Edge0-35B-A3B-Preview: Sparse MoE for Phone-Class Memory

Edge0 released Edge0-35B-A3B-preview, a 35B sparse MoE model running on phone-class memory using streaming inference. ...

New Models

DeepSeek-V4.1-Flash Released: 484.6B MoE Model on HF

deepseek-ai has released DeepSeek-V4.1-Flash on Hugging Face, a multimodal MoE model supporting up to 1M tokens. Lear ...

New Models

Nex-AGI Announces Agent Model Nex-N2.5-Pro, Weights Coming Soon

Nex-AGI announces the Nex-N2.5 agent model family. Weights for mini, Pro, and Max are available on Hugging Face, with ...

New Models

Nex-AGI Releases Open-Weight Model Nex-N2.5-mini for Long Tasks

Discover the specs, performance, and hardware requirements for Nex-N2.5-mini, an open-weight multimodal model by Nex- ...

New Models

OpenBMB Releases On-Device 2B Model MiniCPM5-2B-GGUF

Discover OpenBMB’s MiniCPM5-2B-GGUF, a 2B-class open-weight model for on-device use. Learn specs, benchmarks, h ...