New Models

Signal-3.8-27B-GGUF: Faster and More Token-Efficient

Discover Signal-3.8-27B-GGUF, a minimally invasive fine-tune of Qwen3.8-27B offering lower latency, reduced token usa ...

New Models

bartowski/nex-agi_Nex-N2.5-mini-GGUF: Specs and Hardware

Explore the GGUF quantization of nex-agi’s multimodal agent model Nex-N2.5-mini by bartowski, including hardwar ...

Technical Reports

Together AI Expands Fine-Tuning Service with New Features

Together AI expands its fine-tuning service with new models, live metrics tracking, finer controls, and price reducti ...

New Models

Pantheon-Reasoning-26B-A4B-1.1-V2 GGUF Quantizations

Download bartowski’s GGUF quantizations for Gryphe Pantheon-Reasoning-26B, a Gemma 4 roleplay and reasoning mod ...

New Models

Edge0-35B-A3B-Preview: Sparse MoE for Phone-Class Memory

Edge0 released Edge0-35B-A3B-preview, a 35B sparse MoE model running on phone-class memory using streaming inference. ...

Image, Video and Audio

Lightricks Releases Multiple LoRA and IC-LoRA Adapters for LTX-2.5

Lightricks has released multiple LoRA and IC-LoRA adapters for Lightricks/LTX-2.5, enabling character consistency, ci ...

Technical Reports

NVIDIA Announces BioNeMo Inference Runtime for Boltz-2

NVIDIA released NVIDIA BioNeMo Inference Runtime (BioIR) to accelerate biomolecular structure prediction inference on ...

Technical Reports

NVIDIA NIM Optimization Boosts Nemotron 3 Ultra Throughput

NVIDIA announces that NIM 2.0.12 optimization for Nemotron 3 Ultra achieves up to 2.5x higher system throughput on 4x ...

Image, Video and Audio

YuE2-3B Music Generation Model Released for Lyrics & Styles

m-a-p releases YuE2-3B, an open music generation model using AR-NAR Mixture-of-Transformers to generate 48kHz stereo ...

New Models

DeepSeek-V4.1-Flash Released: 484.6B MoE Model on HF

deepseek-ai has released DeepSeek-V4.1-Flash on Hugging Face, a multimodal MoE model supporting up to 1M tokens. Lear ...