tensorrt-llm
Optimize LLM inference on NVIDIA GPUs using TensorRT-LLM with FP8/INT4 quantization.
npx skills add https://github.com/AlexKoncept/omnia-hub --skill tensorrt-llm-alexkoncept
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: tensorrt-llm Source: https://github.com/AlexKoncept/omnia-hub/tree/main/HERMES/optional-skills/mlops/tensorrt-llm Command: npx skills add https://github.com/AlexKoncept/omnia-hub --skill tensorrt-llm-alexkoncept