Soup

Soup

Fine-tune LLMs from one YAML

Description

Fine-tuning an LLM means SSH into servers, endless config and out-of-memory errors. Soup turns fine-tuning and post-training into one YAML and one command, with layer streaming that trains an 8B model on a 4GB laptop GPU.

It supports SFT, DPO, LoRA and QLoRA, exports GGUF for Ollama, and has a web UI.

Features



One config:YAML for everything.

Low VRAM:8B on 4GB via layer streaming.

Methods:SFT, DPO, LoRA and QLoRA.

GGUF export:Use it in Ollama.