News

Show HN: Minimal LLM Post-Training Experiments on an 8GB GPU (SFT, DPO, GRPO)

1 min read
Source: Hacker News
HN score 6,comments 0