Shannon Sands@max_paperclipsfr though, how undertrained ARE models? How much improvement is still left on the table just from scaling data and quality? Do we eventually see Qwen 4-27b at Fable level? It's honestly kinda crazy. We're so far from Chinchilla now it's madOpens with a question
Shannon Sands@max_paperclips"Hans, I've finished training an LLM, but it stops in the 1930s! This will teach us great things, I wonder how it's beliefs will develop, how well it predicts the future?" "That's great Fritz, but did you remember the RLHF?" "Oh no"Opens with a number