instruction tuning
4 statements across 2 episodes · 3 bullish · 0 bearish · 2 people on the record · first statement Jan 11, 2024 by Nathan Lambert · across every show →
Everything said about instruction tuning, oldest first
Jan 11, 2024 positive
Lambert: Instruction tuning is more important than RLHF for most practitioners
“I think for most people, instruction tuning is probably still more important in their day-to-day life. I think instruction tuning works very well. You can write samples by hand that make sense. You can get the model to learn from them. You could do this with v…”
Jan 11, 2024 neutral
Lambert: The vast majority of instruction tuning data remains simple Q&A
“There's much more, like there's surely kind of more tricky things that people do, but I still think the vast majority of it is question and answer. It's like, please explain this topic to me, generate this thing for me. That hasn't changed that much this year.…”
Jan 11, 2024 positive
Lambert: Scaling from 7B to 70B parameters fixes nuance and repetition
“I think the things that people see now is like the small models don't really handle nuance as well, and they could be more repetitive if, even if they have really good instruction tuning, but if you take that kind of seven to seventy billion parameter jump, li…”
Jun 11, 2024 bullish