speculative-decoding
Accelerate LLM inference using draft models, Medusa heads, and lookahead decoding.
npx skills add https://github.com/Orchestra-Research/AI-research-SKILLs --skill speculative-decoding-orchestra-research
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: speculative-decoding Source: https://github.com/Orchestra-Research/AI-research-SKILLs/tree/main/19-emerging-techniques/speculative-decoding Command: npx skills add https://github.com/Orchestra-Research/AI-research-SKILLs --skill speculative-decoding-orchestra-research