TensorRT-LLM
Also known as: TensorRT
0stories this week
3last 30 days
3all time
Timeline
- Oct 1, 2026 · Tutorial / explainer · 1 sourceBuild Local AI Apps with C++ and NVIDIA TensorRT RTX SamplesDo Inference Now...Adding AI models to local applications requires a portable model format, a reliable runtime, and acceleration that works across target systems.
- Sep 29, 2026 · Open-source release · 1 sourceAI Native by Design: Lessons Learned from Building NVIDIA TensorRT Model ConnectParallel work, model-family isolation, reversible changes, and GPU-backed validation shaped an open source project designed around coding agents NVIDIA TensorRT...Parallel work, model-family isolation, reversible changes, and GPU-backed validation shaped an open source project designed around coding agents NVIDIA TensorRT Model Connect is an open source collection of AI model reference implementations in C++, built on top of NVIDIA TensorRT.
- Sep 29, 2026 · Open-source release · 1 sourceNVIDIA/TensorRT-LLM v1.3.0rc29Expose Nemotron-H vision-language LoRA configuration for supported inference paths #19151