LINKED LIST
[txt mode]
▸
Fast Inference from Transformers via ...
home
explore
|
log in
Fast Inference from Transformers via Speculative Decoding
arxiv.org · 2026-09-25 · 0 upvotes
log in
to save, upvote or flag this.
─── In 0 lists ─────────────────────────────────────────
(not in any lists yet)
related read on:
*
Accelerating Gemma 4: faster inference with multi-token prediction drafters
(from the discussion)