jetson-speculative-decoding

Add EAGLE-3 or draft-model speculative decoding to a Jetson vLLM server when TPOT is the bottleneck.

Install

openclaw skills install @nvidia/jetson-speculative-decoding