GPT-NeoX, every mention
3 scenes · ← back to GPT-NeoX
tap a year for its mentions
every year anyone Eugene Cheah 6
Verbatim, from the transcripts: the passages where GPT-NeoX comes up
How Zyphra went all-in on AMD + Why Devs feel faster with AI but are slower — with Quentin Anthony
- ▶ 29:54 unnamed speaker Lead development on a model training framework called GPT NeoX, which is like a model, a Megatron deep speed style framework for pre-training models on HPC systems.
RWKV: Reinventing RNNs for the Transformer Era
- ▶ 32:11 Eugene Cheah And this part has already been benchmarked against GPT-NeoX in the paper, 3 times in the scene
- ▶ 46:33 Eugene Cheah Uh, I think one thing significant for GPT Neo X was that, uh, it was one of the major models that had everything fully documented and they like, why they make this change in the architecture and so on and so forth. 3 times in the scene