Minimal LLM Post-Training Experiments on an 8GB GPU (SFT, DPO, GRPO)github.com21 pointsby popopanda0 commentsSharePost on XLinkedInCopy post