~ / skills / software-engineering / tensorrt-llm

tensorrt-llm

Optimizes LLM inference with NVIDIA TensorRT for maximum throughput and lowest latency. Use for production deployment on NVIDIA GPUs (A100/H100), when you need 10-100x faster…

Industry
software-engineering
License
Unverified
Source repo
NousResearch/hermes-agent · ★ 214,858
Source file
optional-skills/mlops/tensorrt-llm/SKILL.md
View full SKILL.md on GitHub →

不会安装?看中文图文教程 →