Prediction certainty 3/5 debate potential 3/5

Alberti: AI App Companies Must Encode Product Needs Into RL Feedback

Silas Alberti · DeepWiki: The GitHub Encyclopedia · May 21, 2025 · at 31:16

Silas Alberti of Cognition discusses the emerging platform shift where foundation model providers offer reinforcement learning fine-tuning tools to enterprise developers.

0:00 / 0:23exact quote · 23.1s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“I think that's almost how I view like the future of application layer companies. Cause I mean, yeah, you see like the different, the labs are also now creating these like RL platforms and you can soon like customize models with RL on your personal, like on your company's use cases. So a lot of what like application layer companies have to figure out is like how to encode their product and their like customer needs into RL grading and RL feedback.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Silas Alberti

Insight
Alberti: Small numbers of manually curated evals beat large eval sets
“I think usually like small numbers of very high quality evals are the way to go and like we curate them manually and make sure they're really good.”
Silas Alberti May 21, 2025 ▶ 17:39 DeepWiki: The GitHub Encyclopedia
Insight
Alberti: Wiki pages are a better abstraction than pure RAG for codebase search
“If you just do like pure, like rack on like such a big base code files, it'll just be like pretty bad at a certain point, you know, on a single code base. Sure. I can see it, but Tens of thousands of code bases. It's tougher. But I think actually the wiki page…”
Silas Alberti May 21, 2025 ▶ 26:29 DeepWiki: The GitHub Encyclopedia
Insight
Alberti: Multi-Turn RL Enables Aggressive Code Optimization Over Single-Turn Models
“Basically the single-turn model that was just trained on, like, getting the best result after one turn. It would basically be a little bit, like, too careful, because it couldn't risk writing, like, non-compiling code, whereas, like, the multi-turn model would…”
Silas Alberti May 21, 2025 ▶ 30:19 DeepWiki: The GitHub Encyclopedia
Prediction Not checkable as stated
Alberti: DeepWiki compute spend may break $1M soon
“So we think we should be now at like the high, like tens of thousands of repos. So I guess you can like do the rough map. Like we might be approaching quite significant numbers of compute spend here. I mean maybe we'll like break the million soon.”
Silas Alberti May 21, 2025 ▶ 3:46 DeepWiki: The GitHub Encyclopedia
Insight
Alberti: Folder structure graphs are poor representations of codebase architecture
“There are some, the simplest one is a folder structure graph, but unfortunately, it's, like, actually a pretty bad one because people sometimes just, you know, like, put all the components in this folder and all the servers in this folder.”
Silas Alberti May 21, 2025 ▶ 8:28 DeepWiki: The GitHub Encyclopedia
Disclosure
Alberti: Cognition built its own in-house vector database
“This culture of like building everything in-house. It's also kind of funny, like we, there's like a rack component to this. And I've been actually, I personally advocate, I can just like use a vector database and make life easy. But then there's a really stron…”
Silas Alberti May 21, 2025 ▶ 10:46 DeepWiki: The GitHub Encyclopedia
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.