layer-streaming

Tag

Cards List
#layer-streaming

Show HN: Fine-tune an 8B model on a 4 GB laptop GPU

Hacker News Top ↗ · 2026-08-04 Cached

Soup is an open-source CLI that simplifies LLM fine-tuning and post-training with a single command, enabling QLoRA-based training on consumer GPUs with as little as 4 GB VRAM via layer streaming. The latest version adds preference losses like DPO, ORPO, SimPO, and KTO without doubling memory requirements.

0 favorites 0 likes
← Back to home

Submit Feedback