← The log

lyogavin/airllm

★ 35,284 stars

Jupyter NotebookLLMs

Covered in week 2026-W40 · the 2026-10 report

AirLLM 70B inference with single 4GB GPU

02 Oct 2026 · 0:25
AirLLM Runs a 70B Model on a 4GB GPU. No Quantization.
Watch on YouTube →

View on GitHub