why aren't all 8 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Assertion Not checkable as stated
Tan: GPT-4 wiped out YC startups' RL fine-tuning lead over GPT-3.5
“You know, we've also had YC companies where, ah, they had something that beat OpenAI, ah, you know, GPT-III.V and they were doing fine-tuning with RL, but then, ah, yeah, GPT-IV and then, ah, GPT-IV came out and, ah, you know, basically blew their fine-tuning …”
Insight
Tan: Any GPT-4 prompt workflow can be duplicated with custom fine-tuning
“Anything you do with those prompts, you can get your own model to do with a little bit more training.”
Opinion
Heller: GPT technology is underhyped because GPT-4 operates at postgraduate level
“So this is gonna sound probably insane to most people who hear it, but I think the GPT technology is under hyped because what we see in GPT-IV specifically as a model, and I'm assuming the same will be true as new models come out in the state of the art advanc…”
Opinion
Blomfield: GPT-4 Underperformed on Coding by Misimplementing and Asking Too Many Questions
“I tried GPT-IV just a couple of days ago, and honestly, I wasn't yet as impressed. It just came back with me with too many questions and actually got the implementation wrong too many times.”
Assertion Not checkable as stated
Friedman: Diode Failed on GPT-4 but Worked on o1 With Same Prompts
“The other thing I think is interesting about this example is like during the batch before a one came out, diode had tried to do this with GPT four. Oh, and it just flat out didn't work. And then they basically tried the same thing, the same prompts, but fed it…”
Assertion Supported
Tan: OpenAI enabled internal model distillation as a developer lock-in strategy
“OpenAI itself has now enabled distillation internal to its own API. So you can use O-one, you can use even GPT-IV or IV-O to distill it down into a much cheaper model that's internal to them, like GPT-IV, IV-O-Mini. And that's sort of their, you know, lock-in …”
Disclosure
Heller: Casetext received early access to GPT-4 in summer 2022
“Because we were so focused on large language models and were researching deeply in this space, we got really early access to GPT-IV. Like summer, 20, 22.”
Assertion Supported
Heller: Casetext received GPT-4 access six months before release
“When we got the chance to work at GPT-IV, maybe about six or so months before it was publicly released,”