Technical Reports

TensorRT Edge-LLM Accelerates MLPerf Agentic Benchmark

NVIDIA tests TensorRT Edge-LLM with Qwen3.6-27B on Jetson AGX Thor, achieving a 6.4x speedup over llama.cpp in MLPerf ...

Companies and Funding

The 2026 AI Inference Hardware Revolution and Local LLM Impact

An overview of reports on the 2026 AI inference hardware shift, new memory architectures, chip combinations, and pote ...