Show HN: Minimal LLM Post-Training Experiments on an 8GB GPU (SFT, DPO, GRPO)

  • Posted 7 hours ago by popopanda
  • 12 points
https://github.com/pochenai/nano-llm-posttraining

1 comments

    Loading..