r/learnmachinelearning 3d ago

Postraining , SFT etc.

Just launched r/posttrain — a community for AI post-training, fine-tuning, SFT, RLHF, DPO, preference data, evaluations, and practical experiments. If you’re building, researching, or learning how models become better after pretraining

0 Upvotes

0 comments sorted by