Michael Royzen

Co-Founder & CEO, Phind · 1 appearance on the record.

computed by AI from the episodes · how this works → · full disclaimer →

Michael Royzen is the co-founder and CEO of Phind. He builds AI search tools to help developers solve technical questions and implement code.

21statements → 13claims → 5claims resolved → 60%fully supported → 3.76/5average certainty → 2.48/5average debate potential → ≈4.0/5argument clarity, estimated →

3 supported 2 partly supported 0 contradicted 1 not yet assessed 7 not checkable as stated how the 13 claims stand · each chip opens the sources

3 predictions · 10 assertions · 3 opinions · 3 insights · 2 disclosures · every statement was checked. The predictions and assertions are the 13 claims: statements the public record can support or contradict. 5 are resolved, 1 is not yet assessed, and 7 name no date, number or outcome precise enough to check. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.

The record, in short

What the tape says about how Michael argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

Their most notable supported claim

Assertion Supported
Royzen: Phind built the first internet-scale LLM RAG search in 2022
“And to the best of my knowledge, I think that's the first example that I'm aware of a LLM search engine model that's effectively connected to, like, a large enough index that I would consider, like, an internet scale. So, so I think we were the first to releas…”
Michael Royzen Nov 3, 2023 ▶ 16:07 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind

Expressed certainty vs assessment result

none yet certainty 1
none yet certainty 2
75% certainty 3
100% certainty 4
50% certainty 5

weighted support: a fully supported claim counts one, a partly supported claim counts half. Each filled bar is clickable and opens exactly those claims; "none yet" means nothing said at that certainty level has resolved yet

How they sound: not measured why? →

We measure speaking style by listening to the audio itself, and a fair number needs at least 2,000 words from one person on tape we have measured. There is too little of Michael Royzen on measured tape to publish a rate. This says nothing about how they speak.

Everything Michael Royzen said on Latent Space that made the record, most notable first. Filter by type, assessment or year in the ledger →

Opinion
Royzen: AI dev tools do not need to own the IDE
“Somewhere where I disagree with him is that you need to own the IDE. I think like he made kind of some good points about, you know, not having platform risk in the long term, but some of the, you know, features that were mentioned, like suggesting diffs, for e…”
Michael Royzen Nov 3, 2023 ▶ 24:29 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Prediction Not checkable as stated
Royzen: Fine-tuned open source models will beat proprietary in 2024
“So I think that even if a delta exists, in twenty-twenty-four, the delta between proprietary and open source won't be large enough that a startup like us, with a lot of data that we've collected, can take the data that we have, fine-tune an open source model, …”
Michael Royzen Nov 3, 2023 ▶ 38:25 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Not checkable as stated
Royzen: GPT-4 was trained on HumanEval, proving data contamination
“GPT-IV itself has been trained on human eval, and we know this because GPT-IV is able to predict the exact doc string in many of the problems. I've seen it predict, like, the specific example values in the doc string, which is extremely improbable for it to ju…”
Michael Royzen Nov 3, 2023 ▶ 41:31 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Supported
Royzen: Phind built the first internet-scale LLM RAG search in 2022
“And to the best of my knowledge, I think that's the first example that I'm aware of a LLM search engine model that's effectively connected to, like, a large enough index that I would consider, like, an internet scale. So, so I think we were the first to releas…”
Michael Royzen Nov 3, 2023 ▶ 16:07 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Prediction Not checkable as stated
Phind CEO: Future Programming Will Just Be Problem Solving, Delegating Implementation to AI
“In the future, you know, in the future, I think programming is just going to be really just the problem solving. Like you come up with an idea, you come up with like the basic design for the algorithm in your head, and you just tell the AI, hey, just like, jus…”
Michael Royzen Nov 3, 2023 ▶ 25:41 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Not checkable as stated
Royzen: Users switch to Phind when ChatGPT-4 fails on code
“What really shocks us is that a lot of the people who do that they're coming from ChatGPT. So they tried it in ChatGPT with ChatGPT-IV. It didn't work. Maybe it required like some multi-step reasoning. Maybe it required to like, Some internet context or someth…”
Michael Royzen Nov 3, 2023 ▶ 27:07 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Prediction Not checkable as stated
Royzen: The leap from GPT-4 to GPT-5 will be smaller
“I think that GPT-IV, my hypothesis is that the jump from four to 4.5, or four to five, will be smaller than the jump from Three to four.”
Michael Royzen Nov 3, 2023 ▶ 37:38 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Insight
Royzen: Large context windows outperform RAG chunking for code
“Like, I think it's generally been shown that if you have the space to just put The raw files inside of a big context window. That is still better than chunking and retrieval. It just is.”
Michael Royzen Nov 3, 2023 ▶ 39:24 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Not checkable as stated
Royzen: Training on code unlocked general spatial and temporal reasoning
“We've seen emerging capabilities in the find model, whereby training it on high quality code, it can actually, like, reason better. It went from not being able to solve like, World problems where like riddles where like with like temporal and like low, like pl…”
Michael Royzen Nov 3, 2023 ▶ 43:51 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Opinion
Royzen: Smart Lens worked better than Google Lens v1 at launch
“It worked so well that it actually worked better than Google Lens which released its V-One around the same time.”
Michael Royzen Nov 3, 2023 ▶ 3:01 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Opinion
Royzen: AI code startups are ultimately competing with ChatGPT
“Really who everyone's competing with is ChatGPT which only has, like, that one web interface, and, like, ChatGPT is really the bar.”
Michael Royzen Nov 3, 2023 ▶ 23:08 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Partly supported
Royzen: Phind Leads BigCode Leaderboard by 10 Points in Multi-Language Code
“All of our models are at the top of the big code leaderboard by far. It's not close, particularly in languages other than Python. We have a 10 point gap between us and the next best model on Java, JavaScript, I think C-sharp multilingual.”
Michael Royzen Nov 3, 2023 ▶ 41:03 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Insight
Royzen: Users prefer text-to-image generation over image captioning
“I was also looking into image captioning, where like you give a model an image, and then it tells you what's in the image, but it turns out that what people want is the exact opposite. People want to give a description of an image, and then have the AI generat…”
Michael Royzen Nov 3, 2023 ▶ 4:20 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Partly supported
Royzen: BigScience's T-Zero predated InstructGPT in large-scale instruction tuning
“I think T-Zero is the first model that did large-scale instruction tuning from diverse data sources in the fall of twenty-twenty-one. This is before InstructGPT. This is before Flan T-Five, which came out in twenty-twenty-two. This is, I think, the very, very …”
Michael Royzen Nov 3, 2023 ▶ 14:42 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Not checkable as stated
Royzen: Paul Graham personally chose the company name 'Phind'
“Paul Graham actually picked it for us.”
Michael Royzen Nov 3, 2023 ▶ 30:01 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Insight
Royzen: NVIDIA Remains Cloud-Agnostic Because It Wins Regardless
“At NVIDIA, They know that they're going to win regardless. So they don't care where you get the GPUs from. They're like, they're truly neutral, unlike various sales reps that you might encounter at various like clouds and, you know, hardware companies, et cete…”
Michael Royzen Nov 3, 2023 ▶ 1:02:39 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Supported
Royzen: INT8 quantization offers storage optimization without guaranteed inference speedups
“But with int eight, there's not necessarily a Speed increase. It's just the storage optimization.”
Michael Royzen Nov 3, 2023 ▶ 1:05:46 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Supported
Royzen: Quantized LLMs currently underperform unquantized baselines in quality
“So we have these great quantization libraries that, you know, for the most part are able to get the size down with not that much quality loss, but there is some, like the quantized models currently are actually worse than the non-quantized ones.”
Michael Royzen Nov 3, 2023 ▶ 1:05:53 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Disclosure
Phind Plans Native Code Interpreter Features to Recursively Iterate on Code
“And Replit is great, and people use that feature. But yeah, I think there's more we can do in terms of, like, having something a bit closer to code interpreter where it's able to run the code and then, like, recursively iterate on it.”
Michael Royzen Nov 3, 2023 ▶ 49:32 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Assertion Not publicly verifiable
Royzen: NVIDIA Built Custom FasterTransformer Feature for Phind
“They actually implemented a custom feature for us in Faster Transformer which is one of their libraries... They implemented streaming generation for T-Five-based models, which we were running at the time up until we switched to GPT in In February, March of thi…”
Michael Royzen Nov 3, 2023 ▶ 1:04:04 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind
Disclosure
Royzen: Phind focuses on developer reasoning engines over general search
“So as, I think there's always an opportunity for us to become more general if we wanted. But We've been along this path of, like, what is the best, most advanced reasoning engine that's connected to your code base, that's connected to the internet, that we can…”
Michael Royzen Nov 3, 2023 ▶ 20:37 Beating GPT-4 with Open Source Models - with Michael Royzen of Phind

Appearances (1)

EpisodeDateSpeaking time
Beating GPT-4 with Open Source Models - with Michael Royzen of Phind Nov 3, 2023 1h 2m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.