Hackernews
new
show
ask
jobs
Show HN: Minimal LLM Post-Training Experiments on an 8GB GPU (SFT, DPO, GRPO)
14 points
posted 8 hours ago
by popopanda
(github.com)
No comments yet