Post by Kevin Weil 🇺🇸 on X
Kevin Weil 🇺🇸@kevinweil
XDay 2 of the 12 days of OpenAI! 🎁
Today something for developers: we're introducing Reinforcement Fine Tuning, a new model customization technique for o1/o1-mini. RFT produces expert models in specific domains—and they're 🤯 good, with as few as a couple dozen examples.
218 likes14 repliesPosted Dec 6, 2024