Yi Tay: UL2 is a 20B encoder-decoder model using T5 pre-training data
Yi Tay · The 10,000x Yolo Researcher Metagame — with Yi Tay of Reka · Jul 5, 2024 · at 14:54
Yi Tay, co-founder of Reka and former Google Brain researcher, explains the technical architecture and release background of Google's UL2 language model.
“So I think UL-II is an encoder, decoder, 20 B model. I think when we got it approved, it was, like kind of, you know, it was released as, like, kind of, like, the big brother of T-Five, you know, kind of like, okay, we updated T-Five with, like, a new objective, and train this new model into DBM, we want to, and it uses the same, like, pre-training data set and everything, right”
quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →