← The log

MakazhanAlpamys/Soup

★ 6,100 stars

★ 4,845 at coverage → ★ 6,100 (+1,255)

PythonLLMsCLI & terminalMachine learning

Soup turns the pain of LLM fine-tuning into a simple workflow. One config, one command, done.

View on GitHub Get the badge

07 Sep 2026 · 0:24
Soup: The Free Alternative to OpenAI Fine-Tuning
Watch on YouTube →

Transcript

Fine-tuning an LLM usually means SSH and config hell. Soup streams the frozen 8B base one decoder layer at a time onto a 4 GB card. One YAML config and soup train run QLoRA with auto quantization — 119.6 tokens per second, 3.32 GB peak. Layer streaming is still opt-in beta, but fine-tuning now happens on your own GPU, no cloud.