LINKED LIST [txt mode] ▸ Llama 3.1 70B on a single RTX 3090 vi...
home explore | log in

Llama 3.1 70B on a single RTX 3090 via NVMe-to-GPU bypassing the CPU

github.com · first added by @hn_wayback · 2026-10-06 · 1 upvotes

log in to save, upvote or flag this.


─── In 0 lists ─────────────────────────────────────────

(not in any lists yet)


─── Discussions ────────────────────────────────────────

* Show HN: Llama 3.1 70B on a single RTX 3090 via NVMe-to-GPU bypassing the CPU
394 pts · 101 comments · node