tensorrt-llm
Optimizes LLM inference with NVIDIA TensorRT for maximum throughput and lowest latency. Use for production deployment on NVIDIA GPUs (A100/H100), when you need 10-100x faster…
- Industry
- software-engineering
- License
- Unverified
- Source repo
- NousResearch/hermes-agent · ★ 214,858
- Source file
- optional-skills/mlops/tensorrt-llm/SKILL.md
不会安装?看中文图文教程 →