TensorRT-LLM
NVIDIA's toolkit for high-performance large language model inference.
The Linkredibles take
Why it's worth knowing.
Optimize and deploy LLM inference workloads on NVIDIA GPUs.
Topics
#llm#inference#nvidia#gpu#cuda#performance
Open source repository
Explore the source behind TensorRT-LLM.
Read the source, explore the project, and see how it is being developed on GitHub.
Keep exploring
You might also like.
ai→
MiroFish
A Simple and Universal Swarm Intelligence Engine, Predicting Anything. 简洁通用的群体智能引擎,预测万物
★ 74,570Python
ai agent-skills→
Addy Osmani Agent Skills
Production-grade engineering skills for AI coding agents covering planning, implementation, testing, review, simplification, and shipping.
★ 98,956JavaScript
ai→
ag2
AG2 (formerly AutoGen): The Open-Source AgentOS.Join us at: https://discord.gg/sNGSwQME3x
★ 4,955Python
