Post by Andrew Curran on X
Andrew Curran@AndrewCurran_
XAmazing numbers. Phi-3 is topping GPT-3.5 on MMLU at 14B. Trained on 3.3 trillion tokens. They say in the paper 'The innovation lies entirely in our dataset for training - composed of heavily filtered web data and synthetic data.'

177 likes10 repliesPosted Apr 23, 2024