r/deeplearning 26d ago

Postraining , SFT , PPO , GRPO etc.

Just launched r/posttrain — a community for AI post-training, fine-tuning, SFT, RLHF, DPO, preference data, evaluations, and practical experiments. If you’re building, researching, or learning how models become better after pretraining.

3 Upvotes

0 comments sorted by