Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

    Minimal LLM Post-Training Experiments on an 8GB GPU (SFT, DPO, GRPO) · Birbla