LINKED LIST [txt mode] ▸ Lossless LLM compression for efficien...
home explore | log in

Lossless LLM compression for efficient GPU inference via dynamic-length float

arxiv.org · first added by @hn_wayback · 2026-10-06 · 1 upvotes

log in to save, upvote or flag this.


─── In 0 lists ─────────────────────────────────────────

(not in any lists yet)


─── Discussions ────────────────────────────────────────

* Lossless LLM compression for efficient GPU inference via dynamic-length float
411 pts · 117 comments · node