Jungwon Byun, co-founder of Elicit, explains how engineer Charlie rapidly implemented Constitutional AI to fine-tune open-source models for faithful paper summarization.
“At the start of twenty-twenty-three, Anthropik kind of launched their constitutional AI paper and within a few days, I think four days, he had basically implemented that in production, and then we had it in-app, like, a week or so after that, and he has since kind of contributed to major improvements, like cutting, cutting costs down to, like, a 10th of what they were”
quote is from the automated transcript, cleaned for reading:
filler sounds and stutters are removed, nothing is rephrased. names can be misheard
(the analysis reads context, assessments check outside sources). how →
More from Jungwon Byun
Insight
Byun: Supervising step-by-step AI reasoning makes models far easier to evaluate
“The importance of supervising the process of AI systems, not just the outcomes. And so a big part of how, then, like, how Elicit is built is, We're very intentional about not just throwing a ton of data into a model and training it and then saying, cool, here'…”
Jungwon ByunApr 11, 2024▶ 8:12Supervise the Process of AI Research — with Jungwon Byun and Andreas Stuhlmüller of Elicit
Opinion
Byun: GPT-3 was a qualitative shift, while GPT-4 was an extension
“I think GPT-III was a big change because it kind of said, oh, now is the time to build to you that we can use AI to build these tools. And then GPT-IV was maybe a little bit more of an extension of GPT-III. It felt less like a level, GPT-III over GPT-II was li…”
Jungwon ByunApr 11, 2024▶ 26:47Supervise the Process of AI Research — with Jungwon Byun and Andreas Stuhlmüller of Elicit
Insight
Byun: Foundational models will not commoditize Elicit due to deep workflow specialization
“I think about this a lot in the context of moats. People are like, oh, what's your moat? What happens if GPT-V comes out? It's like, if GPT-V comes out, there's still like all of this other space that we can go into. And so I think being really obsessed with t…”
Jungwon ByunApr 11, 2024▶ 28:57Supervise the Process of AI Research — with Jungwon Byun and Andreas Stuhlmüller of Elicit
AssertionNot checkable as stated
Byun: LLM self-reported uncertainty is reasonably well-calibrated in production
“We found it to be pretty calibrated. There varies on the model.”
Jungwon ByunApr 11, 2024▶ 44:29Supervise the Process of AI Research — with Jungwon Byun and Andreas Stuhlmüller of Elicit
Insight
Byun: Highly structured, reproducible research workflows are uniquely amenable to automation
“Because it's so structured and designed to be reproducible, it's really amenable to automation. So that's kind of the one, the workflow that we want to automate first.”
Jungwon ByunApr 11, 2024▶ 18:00Supervise the Process of AI Research — with Jungwon Byun and Andreas Stuhlmüller of Elicit
Insight
Byun: Research, not probability modeling, is the bottleneck in forecasting
“The thing that's blocking people from making interesting predictions about important events in the world is less kind of on the probabilistic side and much more on the research side.”
Jungwon ByunApr 11, 2024▶ 20:46Supervise the Process of AI Research — with Jungwon Byun and Andreas Stuhlmüller of Elicit
Made with StarZero
Turn any episode into a week of clips.
This entire site, over 200 episodes transcribed, diarized, checked and made playable,
runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the
moments worth sharing, cuts them, captions them, and reframes them for every feed.
We use essential cookies to make the site work. With your permission we
also use analytics cookies (Google Analytics and Mixpanel) to understand
usage and improve StarZero. See our Cookie Policy.