From GitHub
- GitHub stars
- 23
- Last code update
- 2026-08-04
Minimal LLM post-training experiments supporting SFT, DPO, and GRPO on an 8GB GPU.
Why it made the shelf
This is an open-source developer tool and library for running LLM post-training on consumer hardware.
On 2 boards
Official signals
Sourced facts from the places where this tool ships.
From GitHub
Behind the tool
People, backing, growth, and ownership. Every claim links to its source.
Source review current
Public trail
1 reference from 1 publisher.