Tin mới
Generate text with guided decoding
Control generated text using logits processor
Generate text with multiple LoRA adapters
Runtime Configuration Examples
Run LLM-API with pytorch backend on Slurm
Run trtllm-bench with pytorch backend on Slurm
Run trtllm-serve with pytorch backend on Slurm
Curl Chat Client For Multimodal
OpenAI Chat Client for Multimodal
Openai Completion Client For Lora
OpenAI Completion Client with JSON Schema
Deployment Guide for Nemotron v3 (Ultra & Super) on TensorRT LLM - Blackwell & Hopper Hardware
Deployment Guide for DeepSeek R1 on TensorRT LLM - Blackwell & Hopper Hardware
Deployment Guide for Llama3.3 70B on TensorRT LLM - Blackwell & Hopper Hardware
Deployment Guide for Llama4 Scout 17B on TensorRT LLM - Blackwell & Hopper Hardware
Deployment Guide for GPT-OSS on TensorRT-LLM - Blackwell Hardware
Deployment Guide for Qwen3 on TensorRT LLM - Blackwell & Hopper Hardware
Deployment Guide for Qwen3.8 MoE and Qwen3.5 MoE on TensorRT LLM - Blackwell Hardware
Deployment Guide for Kimi K2 Thinking on TensorRT LLM - Blackwell
Deployment Guide for Kimi K3 on TensorRT LLM - Blackwell
Deployment Guide for GLM-5 on TensorRT LLM - Blackwell Hardware
Deployment Guide for MiniMax-M3 on TensorRT LLM
CPU Affinity configuration in TensorRT LLM