NVIDIA Announces BioNeMo Inference Runtime for Biomolecules
NVIDIA released NVIDIA BioNeMo Inference Runtime (BioIR) to accelerate biomolecular structure prediction inference on ...
NVIDIA Nemotron 3 Ultra NIM Achieves 2.5x Throughput on 4xB200
NVIDIA announces that NIM 2.0.12 optimization for Nemotron 3 Ultra achieves up to 2.5x higher system throughput on 4x ...
YuE2-3B Music Generation Model Released for Lyrics & Styles
m-a-p releases YuE2-3B, an open music generation model using AR-NAR Mixture-of-Transformers to generate 48kHz stereo ...
DeepSeek-V4.1-Flash Released: 552B MoE Multimodal Model
deepseek-ai has released DeepSeek-V4.1-Flash on Hugging Face, a multimodal MoE model supporting up to 1M tokens. Lear ...
NVIDIA Dynamo Details EPD Disaggregation for Multimodal
Learn how NVIDIA Dynamo uses Encode-Prefill-Decode (EPD) disaggregation to accelerate multimodal model serving, reduc ...
Goodfire Traces Olmo Behavior with Ai2 Post-Training Stack
Goodfire uses Ai2’s open post-training stack to trace and predict unwanted behaviors in Olmo models, showcasing ...
Nex-AGI Releases Agent Model Nex-N2.5-Pro
Nex-AGI announces the Nex-N2.5 agent model family. Weights for mini, Pro, and Max are available on Hugging Face, with ...
Nex-AGI Releases Open-Weight Long-Task Model Nex-N2.5-mini
Discover the specs, performance, and hardware requirements for Nex-N2.5-mini, an open-weight multimodal model by Nex- ...
Mistral AI Raises €3B in Series D to Accelerate Open-Weight AI
Mistral AI secures €3 billion in Series D funding to expand frontier research and continue open-weight model developm ...
OpenBMB Releases MiniCPM5-2B-GGUF for On-Device AI
Discover OpenBMB’s MiniCPM5-2B-GGUF, a 2B-class open-weight model for on-device use. Learn specs, benchmarks, h ...