The Ledger

Every statement that passed quotation and attribution checks. Mix any filter with any other: certainty 1/5, debate potential 5/5, or both at once.

clear all ✕

why aren't all 13 resolved? a statement only gets an assessment when the public record can support or contradict it. opinions and what-ifs never can, and 0 checkable ones are still open, waiting for their date. predictions held up or didn't; assertions are supported or contradicted. on every card: ▮▮▮▮▮ certainty · ▮▮▮▮▮ debate potential. speakers are clickable

Prediction Not checkable as stated
Royzen: Fine-tuned open source models will beat proprietary in 2024
“So I think that even if a delta exists, in twenty-twenty-four, the delta between proprietary and open source won't be large enough that a startup like us, with a lot of data that we've collected, can take the data that we have, fine-tune an open source model, …”
Michael Royzen Nov 3, 2023 ▶ 38:25 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Not checkable as stated
Royzen: GPT-4 was trained on HumanEval, proving data contamination
“GPT-IV itself has been trained on human eval, and we know this because GPT-IV is able to predict the exact doc string in many of the problems. I've seen it predict, like, the specific example values in the doc string, which is extremely improbable for it to ju…”
Michael Royzen Nov 3, 2023 ▶ 41:31 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Supported
Royzen: Phind built the first internet-scale LLM RAG search in 2022
“And to the best of my knowledge, I think that's the first example that I'm aware of a LLM search engine model that's effectively connected to, like, a large enough index that I would consider, like, an internet scale. So, so I think we were the first to releas…”
Michael Royzen Nov 3, 2023 ▶ 16:07 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Prediction Not checkable as stated
Phind CEO: Future Programming Will Just Be Problem Solving, Delegating Implementation to AI
“In the future, you know, in the future, I think programming is just going to be really just the problem solving. Like you come up with an idea, you come up with like the basic design for the algorithm in your head, and you just tell the AI, hey, just like, jus…”
Michael Royzen Nov 3, 2023 ▶ 25:41 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Not checkable as stated
Royzen: Users switch to Phind when ChatGPT-4 fails on code
“What really shocks us is that a lot of the people who do that they're coming from ChatGPT. So they tried it in ChatGPT with ChatGPT-IV. It didn't work. Maybe it required like some multi-step reasoning. Maybe it required to like, Some internet context or someth…”
Michael Royzen Nov 3, 2023 ▶ 27:07 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Prediction Not checkable as stated
Royzen: The leap from GPT-4 to GPT-5 will be smaller
“I think that GPT-IV, my hypothesis is that the jump from four to 4.5, or four to five, will be smaller than the jump from Three to four.”
Michael Royzen Nov 3, 2023 ▶ 37:38 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Not checkable as stated
Royzen: Training on code unlocked general spatial and temporal reasoning
“We've seen emerging capabilities in the find model, whereby training it on high quality code, it can actually, like, reason better. It went from not being able to solve like, World problems where like riddles where like with like temporal and like low, like pl…”
Michael Royzen Nov 3, 2023 ▶ 43:51 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Partly supported
Royzen: Phind Leads BigCode Leaderboard by 10 Points in Multi-Language Code
“All of our models are at the top of the big code leaderboard by far. It's not close, particularly in languages other than Python. We have a 10 point gap between us and the next best model on Java, JavaScript, I think C-sharp multilingual.”
Michael Royzen Nov 3, 2023 ▶ 41:03 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Partly supported
Royzen: BigScience's T-Zero predated InstructGPT in large-scale instruction tuning
“I think T-Zero is the first model that did large-scale instruction tuning from diverse data sources in the fall of twenty-twenty-one. This is before InstructGPT. This is before Flan T-Five, which came out in twenty-twenty-two. This is, I think, the very, very …”
Michael Royzen Nov 3, 2023 ▶ 14:42 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Not checkable as stated
Royzen: Paul Graham personally chose the company name 'Phind'
“Paul Graham actually picked it for us.”
Michael Royzen Nov 3, 2023 ▶ 30:01 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Supported
Royzen: INT8 quantization offers storage optimization without guaranteed inference speedups
“But with int eight, there's not necessarily a Speed increase. It's just the storage optimization.”
Michael Royzen Nov 3, 2023 ▶ 1:05:46 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Supported
Royzen: Quantized LLMs currently underperform unquantized baselines in quality
“So we have these great quantization libraries that, you know, for the most part are able to get the size down with not that much quality loss, but there is some, like the quantized models currently are actually worse than the non-quantized ones.”
Michael Royzen Nov 3, 2023 ▶ 1:05:53 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Not publicly verifiable
Royzen: NVIDIA Built Custom FasterTransformer Feature for Phind
“They actually implemented a custom feature for us in Faster Transformer which is one of their libraries... They implemented streaming generation for T-Five-based models, which we were running at the time up until we switched to GPT in In February, March of thi…”
Michael Royzen Nov 3, 2023 ▶ 1:04:04 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.