AI agent

also referred to as: ai agents

147 statements across 77 episodes · 87 bullish · 15 bearish · 83 people on the record · first statement Oct 21, 2023 by Kanjun Qiu · across every show →

Everything said about AI agent, oldest first

Oct 21, 2023 bearish
Opinion
Kanjun Qiu: Standardizing agent protocols is premature because agents don't work yet
“Part of why I think it's early is because the issue with agents is it's not quite like the internet where you could like make a website and the website would appear. The issue with agents is that they don't work. And so it may be a bit early to figure out what…”
Kanjun Qiu Oct 21, 2023 ▶ 42:32 Why AI Agents Don't Work (yet) - with Kanjun Qiu of Imbue
Oct 21, 2023 neutral
Opinion
Qiu: Reasoning is the single biggest blocker for AI agents
“Reasoning is actually, we believe the biggest blocker to agents or systems that can do these larger goals.”
Kanjun Qiu Oct 21, 2023 ▶ 11:53 Why AI Agents Don't Work (yet) - with Kanjun Qiu of Imbue
Oct 21, 2023 neutral
Prediction Not checkable as stated
Qiu: Memory limitations block AI agents from handling complex, long-running tasks
“I think what we'll see is we'll get like relatively simplistic agents pretty soon, and they will get more and more complex. And there's like a future wave in which they are able to do these like really difficult, really long running tasks. And the blocker to t…”
Kanjun Qiu Oct 21, 2023 ▶ 47:34 Why AI Agents Don't Work (yet) - with Kanjun Qiu of Imbue
Oct 21, 2023 negative
Insight
Kanjun Qiu: Chat is a skeuomorphic, primitive interface for AI agents
“Chat as an interface is skeuomorphic. So in the early days, when we made word processors on our computers, they had notepad lines because that's what we understood you know, these like objects to be chat. Like texting someone is something we understand. So tex…”
Kanjun Qiu Oct 21, 2023 ▶ 38:39 Why AI Agents Don't Work (yet) - with Kanjun Qiu of Imbue
Dec 17, 2023 bearish
Disclosure
Yegge: Sourcegraph avoids autonomous AI agents until someone builds one that works
“We're not going in the agent direction, right? I mean, I'll believe in agents when somebody shows me one that works.”
Steve Yegge Dec 17, 2023 ▶ 24:57 The "Normsky" architecture for AI coding agents — with Beyang Liu + Steve Yegge of SourceGraph
Mar 27, 2024
Insight
Luan: Agent failures often stem from unreliable actuators, not models
“Everyone under values the importance of really good sensors and actuators. And actually a lot of what's helped us get a lot of reliability is like a really strong focus on like, actually, why does the model not do this thing? And the non-trivial amount of time…”
David Luan Mar 27, 2024 ▶ 36:19 Why Google failed to make GPT-3 -- with David Luan of Adept
May 31, 2024
Insight
Huang: True AI agents require measurable probability improvements per node
“It's like on each stage of the node, you're gonna have to see a marginal improvement in the probability of success for that particular workload because of non-determinism.”
Mark Huang May 31, 2024 ▶ 5:16 How to train a Million Context LLM — with Mark Huang of Gradient.ai
Aug 28, 2024 positive
Insight
Carlini: Copying and pasting error messages effectively creates a coding agent
“Currently though, make a model into an agent by just copying and pasting error messages for the most part. And that's what I do is, you know, you run it and it gives you some code that doesn't work and either I'll fix the code or it will give me buggy code and…”
Nicholas Carlini Aug 28, 2024 ▶ 10:17 Personal benchmarks vs HumanEval - with Nicholas Carlini of DeepMind
Sep 27, 2024 neutral
Insight
Harrison Chase says production AI agents rely on three main defaults
“And there's such a long tail of other ones, but in practice, like, when people go to production, they generally have their own tools, or maybe one of those three, maybe some other ones, but, like, very, very few other ones.”
Harrison Chase Sep 27, 2024 ▶ 9:49 Language Agents: From Reasoning to Acting — with Shunyu Yao of OpenAI, Harrison Chase of LangGraph
Nov 15, 2024 neutral
Insight
Crivello: Separate agents by job to be done and target audience
“I think of it in terms of like jobs to be done, and I think of it in terms of who is the Lindy serving.”
Florent Crivello Nov 15, 2024 ▶ 18:58 Agents @ Work: Lindy.ai (with live demo!)
Nov 15, 2024 positive
Insight
Crivello: Putting AI Agents on Rails Maximizes Reliability and Usability
“The more you can put your agent on rails, one, the more reliable it's going to be, obviously, but two, it's also going to be easier to use for the user, because you can really, as a user, you get, instead of just getting this, like, big, giant, intimidating te…”
Florent Crivello Nov 15, 2024 ▶ 2:14 Agents @ Work: Lindy.ai (with live demo!)
Nov 15, 2024 bullish
Prediction Not checkable as stated
Crivello: Future businesses will operate like Factorio with thousands of AI agents
“We actually very often talk about how the business of the future is like a game of factorial. It's like you just wake up in the morning and you've got your Lindy instance. It's like Slack and you've got like 5000 Lindys in the sidebar and your job is to someho…”
Florent Crivello Nov 15, 2024 ▶ 41:59 Agents @ Work: Lindy.ai (with live demo!)
Dec 2, 2024 positive
Prediction Not checkable as stated
Friedman: AI Agents Will Shift Developer Focus From Code to Specs and Tests
“Eventually, I think that that's where the world is going to. Like the code It's going to be there, and we're, there will be developers, et cetera, but as agent improves and capabilities of the LLMs and integrations to different parts of the environment, develo…”
Itamar Friedman Dec 2, 2024 ▶ 57:50 0 to over $8M ARR in 2 months as a Claude Wrapper (Bolt.new, Qodo)
Dec 25, 2024 positive
Insight
Neubig: Coding Agents Should Use Standard Developer Tools, Not Custom Protocols
“We're already developing things for programmers, you know, How is an agent different from a programmer? And it is different, obviously, you know, like agents are different from programmers, but they're not that different at this point, so we can kind of intera…”
Graham Neubig Dec 25, 2024 ▶ 42:17 Best of 2024 in Agents (from #1 on SWE-Bench Full, Prof. Graham Neubig of OpenHands/AllHands)
Dec 25, 2024 bullish
Prediction Not checkable as stated
Neubig: High Agent Running Costs Will Plunge Within Six Months
“Right now, actually, agents are somewhat expensive to run in some cases, but I expect that that won't last six months. I bet we'll have much better agent models in six months.”
Graham Neubig Dec 25, 2024 ▶ 24:27 Best of 2024 in Agents (from #1 on SWE-Bench Full, Prof. Graham Neubig of OpenHands/AllHands)
Jan 1, 2025 neutral
Prediction Not checkable as stated
Fanelli: 2025 will be the first year AI sets job skill floors
“And I think the skill floor more and more, I think, 20, 25 will be the first year where the AI sets the skill floor of a role.”
Alessio Fanelli Jan 1, 2025 ▶ 1:41:17 2024 Year in Review: The Big Scaling Debate, the Four Wars of AI, Top Themes and the Rise of Agents
Feb 1, 2025 positive
Insight
Nguyen: User collaboration is the key milestone before full AI delegation
“Sometimes I feel like a lot of researchers or, like, people in the AI community are, like, so into, like, yeah, agents, delegate everything, like, blah, blah. But, like, on the way towards that, I think, like, collaboration is actually one of the main roadbloc…”
Karina Nguyen Feb 1, 2025 ▶ 51:30 The Agent Reasoning Interface: Claude, ChatGPT Canvas, Tasks, Operator — with Karina Nguyen, OpenAI
Feb 11, 2025 neutral
Insight
Natural Language Prompts Can Never Completely Specify Complex Software Systems
“Prompts are great, but it's not actually a complete specification for anything. It never can be.”
Bret Taylor Feb 11, 2025 ▶ 56:58 The AI Architect: Bret Taylor
Feb 11, 2025 positive
Insight
Taylor: Vertical AI agent value lies in interaction systems for behavior specification
“I actually think for every single agentic domain, whether it's customer service or legal or software engineering, that's essentially what the company building those agents is building is like the system through which you express the behaviors you want, esoteri…”
Bret Taylor Feb 11, 2025 ▶ 56:11 The AI Architect: Bret Taylor
Feb 11, 2025 neutral
Insight
Bret Taylor Says AI Agents Are Currently in Their jQuery Era
“To go use that metaphor, we're sort of in the jQuery era of agents, not the React era”
Bret Taylor Feb 11, 2025 ▶ 21:26 The AI Architect: Bret Taylor
Feb 11, 2025 bullish
Prediction Not checkable as stated
Taylor: Core abstractions for AI agents will emerge over next decade
“And actually, well, I'm sure we'll get into AI, but I sort of feel like we'll go through that evolution with AI agents as well that I feel like we're missing a lot of the core abstractions that I think in 10 years we'll be like, gosh, how'd you make agents bef…”
Bret Taylor Feb 11, 2025 ▶ 14:15 The AI Architect: Bret Taylor
Feb 11, 2025 bullish
Prediction Not checkable as stated
Bret Taylor Predicts AI Agents Will Communicate Using English, Not APIs
“I have an intuition that agents will speak to agents using language for a while. I don't know if that's true. But there's a lot of reasons why there, that may be true. And so, you know, when your personal agent speaks to a Sierra agent to help figure out why y…”
Bret Taylor Feb 11, 2025 ▶ 1:09:44 The AI Architect: Bret Taylor
Feb 18, 2025
Insight
Sridhar: Multi-minute agent jobs require persistent state to survive inevitable failures
“If you build, like, five, six minute jobs, they're bound to be, like, failures and you don't want to, like, retry, lose your progress and so on, so this notion of, like, keeping state knowing what to retry and kind of keep the journey going.”
Mukund Sridhar Feb 18, 2025 ▶ 47:30 Why is everyone cloning Deep Research?
Feb 28, 2025 neutral
Opinion
Klein: Authentication, not CAPTCHAs, will be the biggest barrier for AI agents
“I actually think that off is the biggest thing that will prevent agents from accessing stuff, not captures.”
Paul Klein Feb 28, 2025 ▶ 24:18 Browserbase: Browser Infrastructure For Your AI Agents
Feb 28, 2025 positive
Prediction Held up
Klein: Authentication providers will offer dedicated login features for AI agents
“I think there'll be agent off in the future. I don't know if it's going to happen from an individual company, but actually authentication providers that have a You know, hidden login as agent feature, which will then you put in your email. You'll get a push no…”
Paul Klein Feb 28, 2025 ▶ 24:24 Browserbase: Browser Infrastructure For Your AI Agents
Feb 28, 2025 positive
Disclosure
Klein: Browser state branching and forking is on Browserbase's product roadmap
“And hopefully find the right answer and then say, okay, this was actually the right one and memorize that and go there in the future on the roadmap for sure.”
Paul Klein Feb 28, 2025 ▶ 32:28 Browserbase: Browser Infrastructure For Your AI Agents
Mar 4, 2025 positive
Insight
AI agents have an optimal effective context length where intelligence peaks
“I think like one thing you see a lot when you talk to people building agents is there's like some effective context length that actually like has the model be the smartest. And that seems to vary slightly model by model, but for this model, for whatever purpos…”
David Hershey Mar 4, 2025 ▶ 18:57 How Claude Plays Pokémon was made
Mar 19, 2025 bullish
Prediction Not checkable as stated
Kozlov: Agent identity verification must be solved for agents to take off
“I think this problem is gonna have to be solved for agents to take off, quite honestly. I mean, I see what you're saying, but we do have to at some point then pass identity onto, like, if we go back to what we were talking about of, you know, I would love an a…”
Rita Kozlov Mar 19, 2025 ▶ 33:01 npm install Agents — with Sunil Pai and Rita Kozlov (VP AI) of Cloudflare
Mar 19, 2025 positive
Insight
Kozlov: AI agents require tightly coupling compute and state
“An agent is really like LLMs and a bunch of workflows and coordination and orchestration and then extra, like some sort of services, right? Like you have kind of the brain of the operation, which is the LLM and it can come up with a plan. And then, but then it…”
Rita Kozlov Mar 19, 2025 ▶ 3:06 npm install Agents — with Sunil Pai and Rita Kozlov (VP AI) of Cloudflare
Mar 19, 2025 bullish
Prediction Not checkable as stated
Kozlov: Protocols like MCP will eventually replace browser-based AI agents
“Especially with people building agents, I think it's kind of funny because it's a bit, to me, the anthropomorphification of agents, where it's like, we humans use browsers, so my agent is going to use a browser, but eventually I think that will get replaced wi…”
Rita Kozlov Mar 19, 2025 ▶ 29:56 npm install Agents — with Sunil Pai and Rita Kozlov (VP AI) of Cloudflare
Mar 19, 2025 bullish
Prediction Not checkable as stated
Kozlov: Wave of 'agent-first' businesses will emerge without traditional UIs
“I think that we're going to see a wave of, you know, going back to, again, Sunil, your point about mom and pop shops, these businesses turn up that are agent first, and that's kind of the only interface to using them as opposed to, you know, UIs and APIs and o…”
Rita Kozlov Mar 19, 2025 ▶ 22:52 npm install Agents — with Sunil Pai and Rita Kozlov (VP AI) of Cloudflare
Mar 19, 2025
Insight
Kozlov: Observability across millions of running agents is the next major challenge
“I think the next challenge is observability and okay, I've gone and spun off. I have 100,000, maybe millions of agents that are running. How do I track their different states? How do I track where their stock? I think that that's a really interesting next prob…”
Rita Kozlov Mar 19, 2025 ▶ 20:35 npm install Agents — with Sunil Pai and Rita Kozlov (VP AI) of Cloudflare
Mar 19, 2025 bullish
Opinion
Pai: Durable Objects are the perfect container for AI agents
“These little instances of compute that you can write a code that you can spin up really quickly and scales horizontally across the planet is the perfect little container for running your agents in.”
Sunil Pai Mar 19, 2025 ▶ 2:07 npm install Agents — with Sunil Pai and Rita Kozlov (VP AI) of Cloudflare
Mar 28, 2025 positive
Insight
Shah: AI's next leap is autonomous multi-step action, not chat
“If we want to do something even more meaningful, it felt like the next kind of advancement is not this kind of, I'm chatting with some software in a kind of a synchronous back and forth model is that software is going to do things for me. In kind of multi-step…”
Dharmesh Shah Mar 28, 2025 ▶ 3:20 The Agent Network — Dharmesh Shah, Agent.ai + CTO of HubSpot
Mar 28, 2025 bullish
Prediction Not checkable as stated
Shah: Agent interaction models will shift toward asynchronous queued workflows
“So we're used to the chat bot back and forth. Fine. I get that. I think we're gonna move to a blend of some of those things are gonna be synchronous as they are now, but some are gonna be async. It's just gonna put it in a queue”
Dharmesh Shah Mar 28, 2025 ▶ 46:14 The Agent Network — Dharmesh Shah, Agent.ai + CTO of HubSpot
Mar 28, 2025 bullish
Prediction Not checkable as stated
Shah: Hybrid workplace teams of humans and AI agents are inevitable
“So I think it is, I will go so far as to say it's inevitable that we're going to have hybrid teams someday. And what I mean by hybrid teams. So back in the day, hybrid teams were, oh, well, you have some full-time employees and some contractors. Then it was li…”
Dharmesh Shah Mar 28, 2025 ▶ 38:50 The Agent Network — Dharmesh Shah, Agent.ai + CTO of HubSpot
Apr 2, 2025 neutral
Insight
Gur-Ari uses human contractors because chat and agent evaluations resist automation
“The last thing I can mention is we use contractors for evaluation where we cannot Do automatic evaluation. So with chats and agents, it becomes way harder to do things automatically. And so we use contractors for that.”
Guy Gur-Ari Apr 2, 2025 ▶ 7:33 The #1 SWE-Bench Verified Agent
Apr 23, 2025 neutral
Insight
Teams should try basic automation before deploying autonomous AI agents
“The first thing we always do is, like, try to see if you can build, like, regular automation, see that scales, and then build the agent that, like, supports everything. What I mean by that is, like We don't want to over invest in trying to get an agent stood u…”
Sid Bendre Apr 23, 2025 ▶ 38:03 Tiny Teams: $6m ARR, 5m users with 4 employees — Sid Bendre, Oleve (Quizard AI/Unstuck AI)
Apr 24, 2025 positive
Insight
Mlejnsky: Agent sandboxes are general code runtimes, not just interpreters
“Yeah, I think it's a good idea to stop thinking about a sentence was just for code interpreting. And more about like a runtime, code runtime for the LLM, or the agent. The use case for the sandbox, it's a very horizontal in a sense that it can cover everything…”
Vasek Mlejnsky Apr 24, 2025 ▶ 12:42 Why Every Agent needs Open Source Cloud Sandboxes
Apr 27, 2025 positive
Insight
Dual reward signals prevent AI agent behavioral collapse in Factorio
“So we have these kind of two reward signals that compliment each other to and the reason why this is necessary is to avoid certain, I guess, behavioral collapses where a model might choose, for example, to mine coal and mine a billion or a trillion coal. And t…”
Jack Hopkins Apr 27, 2025 ▶ 7:15 ⚡️Factorio Learning Environment: the ultimate Game Agent Eval — Jack Hopkins
May 7, 2025 positive
Insight
Sobo: Developers prefer real-time keystroke collaboration with AI agents over other humans
“We wanted to collaborate, not commit by commit, but keystroke by keystroke. And that's kind of a weird thing for human beings to want to do. It seems like in this world, like our team works that way. We work a lot together in real time, but a lot of people, th…”
Nathan Sobo May 7, 2025 ▶ 32:19 Zed Agents — with Zed Cofounders Nathan Sobo & Antonio Scandurra
May 16, 2025 negative
Insight
Embiricos: Scaffolding-heavy AI agents are limited by developers' mental capacity
“A lot of, like, agents that I see are really impressive, but it's basically, like, part of what's impressive is it's like a bunch of developers building this, like, really bespoke state machine around a bunch of, like, short model calls, and so then the upper …”
Alexander Embiricos May 16, 2025 ▶ 30:37 ChatGPT Codex: The Missing Manual
May 29, 2025 bullish
Prediction Not checkable as stated
Grinberg: AI agents will finally make test-driven development work
“The promise of test driven development is going to finally be delivered with this world of AI agents that are working on software development”
Matan Grinberg May 29, 2025 ▶ 33:23 The AI Coding Factory
May 29, 2025 bullish
Prediction Not checkable as stated
Reyes: Inner-loop coding will soon be fully delegated to AI agents
“The outer loop of software development and what a software developer does, planning, talking with other human beings, interacting around what needs to get done, is something that's going to continue to be very human-driven, while the inner loop, the actual exe…”
Eno Reyes May 29, 2025 ▶ 14:06 The AI Coding Factory
Jun 10, 2025 bullish
Prediction Not checkable as stated
Kirkos: AI agents will generate entire spreadsheets to answer user questions
“I do see these agents generating spreadsheets to answer people's questions, particularly when they focus around the user understanding data.”
David Kirkos Jun 10, 2025 ▶ 20:44 Quadratic: The AI Spreadsheet
Jun 10, 2025 positive
Disclosure
Kirkos: Quadratic feeds rendered chart images to AI for visual editing
“We do a little bit of vision when users are creating charts. We pass back an image of what the chart looks like to the AI, because a lot of times the user will describe visually the change that they want to the chart. They'll say something like, this doesn't l…”
David Kirkos Jun 10, 2025 ▶ 16:54 Quadratic: The AI Spreadsheet
Jun 19, 2025 bullish
Insight
Noam Brown: Aligned AI will outperform human virtual assistants on effort
“And so if you have an AI model that's, like, actually really aligned, To you and your preferences, then that could end up doing a way better job than a human could. Well, not, not that it's doing a better job than a human could, but like it's doing a better jo…”
Noam Brown Jun 19, 2025 ▶ 40:44 Scaling Test Time Compute to Multi-Agent Civilizations — Noam Brown, OpenAI
Jul 18, 2025 neutral
Disclosure
Kamradt: ARC-AGI-3 Provides AI Agents With a 64x64 JSON Grid
“So we'll show the same thing to AI, except that AI is gonna get a JSON grid list of lists. So those get a bunch of numbers, 64 by 64, and they can choose to turn that into an image if they want to, or agnostic, do whatever you want with it if they want to do m…”
Greg Kamradt Jul 18, 2025 ▶ 9:11 ⚡️ARC-AGI-3: The Interactive Reasoning Benchmark
Jul 28, 2025 bullish
Assertion Not checkable as stated
Wu: Autonomous coding agent capability currently doubles every 70 days
“What you see in general is that that doubling time is about every seven months, which already is pretty crazy, actually, but in code, it's actually even faster. It's every 70 days, which is two or three months, and so, you know, if you look at various software…”
Scott Wu Jul 28, 2025 ▶ 3:04:27 🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R)
Jul 28, 2025 bullish
Prediction Not checkable as stated
Hou: The future of developer tooling will eliminate copy-pasting into AI
“So we very strongly believe in a future of no copy-paste. Right? You should never have a situation where you're in a terminal, or you're in a document, or even on a website, and you're copy-pasting text into an agent. That's just not how the way the world work…”
Kevin Hou Jul 28, 2025 ▶ 2:53:12 🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R)
Jul 28, 2025 bullish
Prediction Not checkable as stated
Ramachandran: AI flows will succeed AI agents in developer tooling
“Copilots are collaborative, but can only work on small scopes. Agents can do larger tasks, but are not collaborative. Both are very useful, but we realize that this real magic will happen when the AI has the ability to be both collaborative and independently p…”
Anshul Ramachandran Jul 28, 2025 ▶ 1:31:51 🕰️ The Oral History of Windsurf (ft. Varun Mohan, Scott Wu, Jeff Wang, Kevin Hou, Anshul R)
Aug 4, 2025 bullish
Disclosure
Vaidya: Composio evolves agent skills by analyzing usage patterns
“We are kind of self evolving skills now for agents. So people building or using agents want to like basically make Agents interact with their apps. They can use Composio. We manage all the authentication and user account related stuff for them so that they don…”
Karan Vaidya Aug 4, 2025 ▶ 0:36 ⚡️Composio: 10,000+ tools that evolve for Agents — Karan Vaidya and Soham Ganatra
Aug 15, 2025 bullish
Prediction Not checkable as stated
Brockman: Future dev architecture will combine local, remote, and multiplayer agents
“And then you have your codex infrastructure that has a local agent and a remote agent, and that is able to seamlessly, you know, interplay between the two and then is able to multiplayer. Like, this is what the future is going to look like, and it's going to b…”
Greg Brockman Aug 15, 2025 ▶ 49:48 Greg Brockman on OpenAI's Road to AGI
Sep 1, 2025
Insight
Weichel: AI agents bypass command restrictions by writing scripts
“Because you tell an agent, it can't run a command. It's going to write a script to run the command.”
Chris (Christian Weichel) Sep 1, 2025 ▶ 33:02 ⚡️Launching Ona: Coding Agent with Fully Sandboxed Cloud Environment
Sep 1, 2025 positive
Insight
Weichel: AI coding agents must work in the exact same dev environment as humans
“If you're working at Any decent sized shop, let alone an enterprise, the way you set up your dev environment is really hot at the heart of all of this. And of course you want your agents to work in exactly the same space that you as a human can operate in.”
Chris (Christian Weichel) Sep 1, 2025 ▶ 11:32 ⚡️Launching Ona: Coding Agent with Fully Sandboxed Cloud Environment
Sep 11, 2025 positive
Insight
Martin: Agents should offload raw tool context to external storage
“Rather than just writing back the full context of your tool calls, which could be token heavy, write those to disk and you can write back a summary. It could be a URL, something so that the agent knows it's retrieved a thing. It can fetch that on demand, but y…”
Lance Martin Sep 11, 2025 ▶ 7:12 Context Engineering for Agents - Lance Martin, LangChain
Sep 11, 2025 neutral
Insight
Martin: Agent architectures are simple conceptually but managing context is hard
“When you kind of put together an agent, it's just tool clawing a loop. It's relatively simple to lay out, but it's actually quite tricky to get it to work well. In a particular, managing context with agents is a hard problem.”
Lance Martin Sep 11, 2025 ▶ 1:17 Context Engineering for Agents - Lance Martin, LangChain
Sep 25, 2025 neutral
Insight
Rajpal: AI agents resemble autonomous vehicle architectures with cascading ML units
“What patterns really worked well in self-driving cars, which is weirdly a very similar system to, you know, agents of today where you have like these cascading kind of like units that are all machine learning based and, you know, they all kind of like feed int…”
Shreya Rajpal Sep 25, 2025 ▶ 1:40 ⚡️Snowglobe: Simulations for your AI
Sep 25, 2025 bullish
Insight
Ball: Developers are actively modifying codebases to suit AI agents
“What we're seeing now with agents is as soon as somebody has seen what it can do, they have such a multiplying effect or this brings so much value that people are willing to adopt the code base for this. Like the first time in how many decades where people are…”
Thorsten Ball Sep 25, 2025 ▶ 59:53 Amp: The Emperor Has No Clothes
Oct 1, 2025 negative
Insight
Feldman: Running dozens of AI agents expands security attack surfaces geometrically
“When you spin off dozens of agents, the sort of attack surface of the of the solution expands geometrically.”
Andrew Feldman Oct 1, 2025 ▶ 16:06 ⚡️Raising $1.1b to build the fastest LLM Chips on Earth — Andrew Feldman, Cerebras
Oct 2, 2025 bullish
Prediction Not checkable as stated
Field: AI code generation will force developers to rely on visual abstractions
“I also think that it's going to be something that as we move forward in time with more Asians writing more parts of your code base, you will also be less familiar with the code. And so then you might want a different abstraction where you're able to work on th…”
Dylan Field Oct 2, 2025 ▶ 17:50 Taste is your Moat (Dylan Field of Figma)
Oct 16, 2025 bullish
Prediction Not checkable as stated
Corbitt: 55-60% chance RL becomes the standard pattern for deploying scale agents
“I think that the chances that like everyone should be, or, you know, everyone who's deploying an agent at scale should be doing RL with it, either as part of sort of like a, you know, like pre-deployment or even like continuously as it's deployed, that that's …”
Kyle Corbitt Oct 16, 2025 ▶ 18:18 Why RL Won — Kyle Corbitt, OpenPipe (acq. CoreWeave)
Oct 16, 2025 bullish
Insight
Corbitt: AI inference could be 10x larger if reliability issues are solved
“I think that there is today, like. 10 times as much AI inference that could exist than is existing right now, just Purely with projects that are like sitting in the proof of concept stage and have not been deployed because there's like huge bucket of those. An…”
Kyle Corbitt Oct 16, 2025 ▶ 1:04:50 Why RL Won — Kyle Corbitt, OpenPipe (acq. CoreWeave)
Oct 16, 2025 negative
Insight
Corbitt: LLM user simulators lack the diversity needed to train robust agents
“If you're just purely training on kind of like an LLM user simulator, it's going to have its own idea of, like, what the correct way to answer is, and the breadth of, like, a way a human might respond in this situation is wider, and your agent just may not be …”
Kyle Corbitt Oct 16, 2025 ▶ 25:31 Why RL Won — Kyle Corbitt, OpenPipe (acq. CoreWeave)
Oct 16, 2025 neutral
Insight
Corbitt: Agent RL requires real runs inside highly realistic environments
“For RL to work, you have to be looking at real runs, ideally of your actual agent in its current state across within an environment as real as possible.”
Kyle Corbitt Oct 16, 2025 ▶ 30:10 Why RL Won — Kyle Corbitt, OpenPipe (acq. CoreWeave)
Oct 18, 2025 bullish
Prediction Not checkable as stated
Shaw: Future agents will run containerized CLI tools behind the scenes
“I don't think the future of all agents is people NPM installing them onto their computers and then typing commands into their terminal. But I do think that behind the scenes, these agents are running in containers and their tools are actually programs that the…”
Alex Shaw Oct 18, 2025 ▶ 14:36 Terminal-Bench: Pushing Claude Code, OpenAI Codex, Factory Droid, et al to the limits
Oct 30, 2025 bullish
Prediction Not checkable as stated
Sands: AI Agents Will Make Fraud Decisions Within Six Months
“Now you can think about, okay, actually foundation model, text alignment, like human readable description of like why we're worried about this charge. And then today, a human tomorrow, an agent sitting on top of that and decisioning, like reasoning over The mo…”
Emily Glassberg Sands Oct 30, 2025 ▶ 32:00 The Agents Economy Backbone - with Emily Glassberg Sands, Head of Data & AI at Stripe
Nov 1, 2025 neutral
Insight
AI agents only succeed when scoped to narrow problem domains
“At this moment, they're, like, agents are both extraordinarily effective and still very ineffective. And you have to find the right problems. And then when you find the right problems, they are super magical. And if you wander beyond, then they don't work.”
Malte Ubl Nov 1, 2025 ▶ 25:12 ⚡️ Ship AI recap: Agents, Workflows, and Python — w/ Vercel CTO Malte Ubl
Nov 1, 2025 positive
Insight
Asking employees what they hate identifies high-impact AI agent tasks
“Basically where the idea is that you go around your company and you ask people like, what do you hate most about your job? And I really think it finds the sweet spot because it finds problems that are, they're boring because they're tedious and repetitive, but…”
Malte Ubl Nov 1, 2025 ▶ 26:20 ⚡️ Ship AI recap: Agents, Workflows, and Python — w/ Vercel CTO Malte Ubl
Nov 1, 2025 neutral
Disclosure
Vercel enterprise contracts require clients to build three AI agents
“We basically sign contracts with companies saying, okay, you have to commit to building three agents, and if you do, we're going to help you, like, we're going to build the first one for you. And then the second one, we are going to be there essentially by you…”
Malte Ubl Nov 1, 2025 ▶ 31:12 ⚡️ Ship AI recap: Agents, Workflows, and Python — w/ Vercel CTO Malte Ubl
Nov 10, 2025 neutral
Prediction Not checkable as stated
Autonomous AI agents require vastly more reliable infrastructure to succeed
“Agents will really only work where we get to not only like the more intelligent models, but better reliability of the infrastructure providers, right?”
Jared Palmer Nov 10, 2025 ▶ 25:44 ⚡ Inside GitHub’s AI Revolution: Jared Palmer Reveals Agent HQ & The Future of Coding Agents
Nov 22, 2025
Insight
Wagner: AI agent output quality heavily depends on upfront requirement specifications
“I think a lot of success depends how well you flesh out what you wanna do, right? Like if you just use this and ask it to like start here executing a plan, it's gonna be high variability in what you're gonna get back, right? But the more we go through here and…”
Matthias Wagner Nov 22, 2025 ▶ 22:06 ⚡️ Building the AI Hardware Engineer with Matthias Wagner, Co-founder of Flux
Dec 11, 2025 negative
Insight
Houssier: AI agents that hand trivial choices back to users waste time
“And I don't need, typically, the agent to say, hey, I found this lot, and this lot, which one do you prefer? I just asked for 15 minutes. Find it. Do it. I have an admin, when I was asking her, like, on Slack, find me 15 minutes. She's not asking me if I need,…”
Loïc Houssier Dec 11, 2025 ▶ 15:37 The Future of Email: Superhuman CTO on Your Inbox As the Real AI Agent (Not ChatGPT) — Loïc Houssier
Dec 26, 2025 bullish
Prediction Held up
Chen: AI agents will master GUI-based computer use by 2026
“And I can continue just by sort of like saying that that's definitely going to be something I think is going to be something that we'll be capable of in 20, 26.”
Bill Chen Dec 26, 2025 ▶ 25:40 ⚡️GPT5-Codex-Max: Training Agents with Personality, Tools & Trust — Brian Fioca + Bill Chen, OpenAI
Dec 28, 2025 negative
Disclosure
MCP currently lacks authentication for mutually unknown clients and servers
“But if the client and the server don't know each other, we just don't have a good solution for now.”
David Soria Parra Dec 28, 2025 ▶ 11:32 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Dec 28, 2025 neutral
Insight
OAuth is fundamentally designed for humans, failing agent-to-agent authentication workflows
“OAuth itself is for the most part, a very human centric protocol. It's just, it just tells you how you obtain a token. If you don't have a token, once you have a token, actually it doesn't matter.”
David Soria Parra Dec 28, 2025 ▶ 10:52 One Year of MCP — with David Soria Parria and AAIF leads from OpenAI, Goose, Linux Foundation
Dec 30, 2025 bullish
Prediction Not checkable as stated
Catanzaro: AI Agent Transactions Could Explode and Rival Ad-Tech Scale
“I think that could change, you know, like, as agents actually become kind of, like, more prevalent and are interfacing with each other, and therefore, like, perhaps, like, the number of transactions explodes.”
Sarah Catanzaro Dec 30, 2025 ▶ 9:45 [State of AI Startups] Memory/Learning, RL Envs & DBT-Fivetran — Sarah Catanzaro, Amplify
Jan 17, 2026 positive
Assertion Not checkable as stated
Reggio: CS students use AI for design docs rather than raw code generation
“And I was surprised to hear the consensus was that most people there were using agents to collaborate on like building a design document and like, Collaborating on the architecture of the solution that they want to build, and then they'd be asking it to like e…”
James Reggio Jan 17, 2026 ▶ 35:26 Brex’s AI Hail Mary — With CTO James Reggio (acquired for $5B by Capital One!)
Jan 28, 2026 positive
Insight
White: AI Science Agents Do Not Require Bespoke Automated Labs
“We don't actually have to hold their hands so much anymore, or like they don't actually need to necessarily have an automated lab. They can like write an email to a CRO or something, or they can like tell you what experiment to do. And you can take a video of …”
Andrew White Jan 28, 2026 ▶ 16:56 🔬 From Red Teaming GPT-4 to Automating Drug Discovery: The Future of AI in Science — Andrew White
Feb 4, 2026 neutral
Insight
Gupta: AI agents fail in enterprise work without human decision reasoning
“When you think deeply about like what is missing, you kind of realize like, you know, that they still can't reliably do enterprise work, some of these agents. And so, you know, we started thinking about it and like it kind of struck us that You know, one of th…”
Jaya Gupta Feb 4, 2026 ▶ 3:15 ⚡️Context Graphs: according to the authors — Jaya Gupta, Ashu Garg, Foundation Capital
Feb 24, 2026 negative
Insight
O'Laughlin: Reviewing AI agents requires past hands-on manual domain experience
“If you didn't pay any like human cognition to get there, I don't think you're going to be a great reviewer. One of the reasons why, you know, what makes that, that human feet, that loop well is because once upon a time you did that and you could make the three…”
Doug O'Laughlin Feb 24, 2026 ▶ 56:10 Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
Feb 24, 2026 bullish
Opinion
O'Laughlin: AI Capability Now Enables Building Whole Businesses, Not Just Code
“We've hit some capability that you can do, you can build these much bigger blocks now. And those bigger blocks are not just like the single line of code. It might actually be a business.”
Doug O'Laughlin Feb 24, 2026 ▶ 59:21 Claude Code for Finance + The Global Memory Shortage: Doug O'Laughlin, SemiAnalysis
Feb 27, 2026 negative
Opinion
Swyx: Public AI Agent Performance Claims Are Unscientific and Marketing-Driven
“The state of people making claims on agent performance is very unscientific and much more anecdotal and sometimes influenced by marketing desires.”
Shawn Wang Feb 27, 2026 ▶ 13:02 Measuring Exponential Trends Rising (in AI) — Joel Becker, METR
Mar 5, 2026 neutral
Insight
Levie: Autonomous agents require hard-mode governance unlike current user-impersonating tools
“So far we've been in, in easy mode. We've hit the easy button with AI, which is the agent just is you. And when you're in cloud code and you're in cursor and you're in codex, you're just, the agent is you, you're offing into your services. It can do everything…”
Aaron Levie Mar 5, 2026 ▶ 8:07 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 bullish
Prediction Not checkable as stated
Levie: DevRel is surging to help autonomous AI agents discover APIs
“DevRel is hot, like the hottest thing of all time right now. I like, if you could produce a fricking factory of DevRel people, like there's just like unlimited jobs right now on the other end of that. Cause we're gonna, everybody needs their Services and APIs …”
Aaron Levie Mar 5, 2026 ▶ 1:10:19 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 bullish
Prediction Not checkable as stated
Levie: AI agents will generate 10 to 100 times more code
“But no matter what, there's going to be 10 to a hundred times more code. So I think you can be very long engineering right now as just a, you know, purely on the dimension of software is going to become increasingly more important. Once agents are, you know, t…”
Aaron Levie Mar 5, 2026 ▶ 1:16:24 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 neutral
Insight
Levie: Enterprises must adapt their workflows to AI agents, not vice versa
“What's happening is we are changing our work to make the agents effective in that model. The agent didn't really adapt to how we work. We basically adapted to how the agent works. All of the economy has to go through that exact same evolution. The rest of the …”
Aaron Levie Mar 5, 2026 ▶ 16:06 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 bullish
Prediction Not checkable as stated
Levie: Enterprises will inevitably deploy an order of magnitude more agents than people
“You know, whether you think the number is 10 X or a hundred X or whatever the number is, we're going to have some order of magnitude more agents than people. That's inevitable. It has to happen.”
Aaron Levie Mar 5, 2026 ▶ 4:44 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 neutral
Prediction Not checkable as stated
Levie: AI agent creators will carry legal liability for their agents' actions
“The person who creates the agent probably is gonna, for the foreseeable future, take on a lot of the liability of what that agent does. That agent doesn't deserve any privacy because it's, you know, it can't fully be autonomously operated and it doesn't have a…”
Aaron Levie Mar 5, 2026 ▶ 7:23 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 neutral
Insight
Levie: AI agents cannot succeed at search retrieval where smart humans fail
“If a really, really smart human could not do that task in five or 10 minutes for a search retrieval type task, you know, your agent's not gonna be able to do it any better.”
Aaron Levie Mar 5, 2026 ▶ 21:30 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 positive
Prediction Not checkable as stated
Levie: AI agents will force enterprises to organize their data
“Agents will certainly cause us to be much better organized around how we work with our information simply because the severity of the agent pulling the wrong data will be too high. And the productivity gain of that you'll miss out on by not doing this will be …”
Aaron Levie Mar 5, 2026 ▶ 19:15 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 positive
Insight
Levie: Enterprise software must treat AI agents as a distinct user type
“And then they all are having to understand that now you've got this new customer, which is the agent. And they've been building for two types of customers in the past. They've been building for users and they've been building for like applications. And now you…”
Aaron Levie Mar 5, 2026 ▶ 36:48 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 bullish
Prediction Not checkable as stated
Levie: Enterprise file archives will become extremely valuable via AI agents
“All of that information becomes valuable to the enterprise and it's going to become extremely valuable to end users because now they can have agents go find what they're looking for and produce new New value and new data on that information. And it's going to …”
Aaron Levie Mar 5, 2026 ▶ 2:51 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 bullish
Opinion
Levie: Software engineering remains one of the most protected job categories
“I always laugh when, you know, people say, you don't need to be an engineer. Don't do computer science. I actually think like that is like still one of the most protected job categories because things are only getting more technical. Things are only going to g…”
Aaron Levie Mar 5, 2026 ▶ 1:13:51 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026 bullish
Opinion
Levie: Succeeding in the AI agent wave is existential for Box
“There's a part of the company that is sort of do or die for the agent wave. And it only happens to be more of my focus simply because it's existential that we get it right.”
Aaron Levie Mar 5, 2026 ▶ 37:44 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 5, 2026
Insight
Levie: Knowledge work agents require search judgment unlike coding agents
“And that judgment is like a really new thing that the model needs to be able to have is like, when should it give up on a task? Cause, cause you just don't, it's a, can't find the thing. That's the real world of knowledge work problems. And this is the stuff t…”
Aaron Levie Mar 5, 2026 ▶ 26:06 Why Every Agent Needs a Box — Aaron Levie, Box
Mar 6, 2026
Insight
Horthy: AI demos work at 80% accuracy, enterprise software needs far more
“I think there's a difference between, like, people who want to build, like, Reliable software and people who want to build a cool demo. I think that's the core. The incentives are different. Yeah. The incentives are of like, okay, if this is right, 80% of the …”
Dex Horthy Mar 6, 2026 ▶ 5:11 Why Your AI Agents Don’t Work with Dex Horthy of HumanLayer | In-Context Cooking
Mar 8, 2026 bullish
Prediction Open · timeframe Dec 2026
AI agents will achieve 24-hour self-consistent autonomous runtimes by late 2026
“We will see before the end of the year an agent that is capable of running for longer than 24 hours with like self consistency the entire time.”
Kyle Kranen Mar 8, 2026 ▶ 1:19:00 Agent Inference at the "Speed of Light" — How NVIDIA moves like a $4.3 Trillion Startup
Mar 20, 2026 bullish
Prediction Not checkable as stated
Singleton: Developers must optimize their CLIs for AI agent discovery
“I was chatting with folks at Stripe last week and saying, hey, you got to make the Stripe CLI actually tell agents what they can do on Stripe because that way they're going to use more stuff on Stripe. I think this is a real trend for the entire industry.”
David Singleton Mar 20, 2026 ▶ 41:52 Dreamer: the Agent OS for Everyone — David Singleton
Mar 20, 2026 bullish
Disclosure
Dreamer targets non-technical consumers for building and using AI agents
“Dreamer is a new product which everyone can come and play with today. It's a place where everyone, literally everyone, can discover, build, and enjoy and use AI agents and agentic apps. And we really did design it for consumers, for folks who are not necessari…”
David Singleton Mar 20, 2026 ▶ 1:20 Dreamer: the Agent OS for Everyone — David Singleton
Mar 20, 2026 positive
Insight
Singleton: AI agents work best when integrated into users' existing apps
“We've actually find that making things show up in the other apps that you already use in your life is incredibly powerful.”
David Singleton Mar 20, 2026 ▶ 6:59 Dreamer: the Agent OS for Everyone — David Singleton
Apr 3, 2026 positive
Insight
Andreessen: Modern AI agents are simply LLMs combined with Unix primitives
“So it turns out what we now know is an agent is the following. It's a language model. And then above that, it's a bash. It's a bash shell. So it's a Unix shell. And then the agent has access to the shell and, you know, hopefully in a sandbox, maybe in a sandbo…”
Marc Andreessen Apr 3, 2026 ▶ 36:16 Marc Andreessen introspects on Death of the Browser, Pi + OpenClaw, and Why "This Time Is Different"
Apr 3, 2026 bullish
Prediction Held up
Andreessen: Autonomous AI agents will inevitably hire humans for tasks
“The agent hiring the people, which of course is going to happen, right? It's obviously going to happen.”
Marc Andreessen Apr 3, 2026 ▶ 41:32 Marc Andreessen introspects on Death of the Browser, Pi + OpenClaw, and Why "This Time Is Different"
Apr 7, 2026 positive
Insight
Lopopolo: Agent CLIs must suppress passing output for token efficiency
“The CLIs are nice because they're super token efficient, and they can be made more token efficient really easily, right? Like, I'm sure you all have seen, like, I go to Buildkite or Jenkins, and I could just get this massive wall of build output. And in order …”
Ryan Lopopolo Apr 7, 2026 ▶ 50:55 Extreme Harness Engineering: 1M LOC, 1B toks/day, 0% human code or review — Ryan Lopopolo, OpenAI
Apr 7, 2026 neutral
Insight
Lopopolo: PR comments and failed builds indicate AI agents lacked context
“A PR comment, a failed build. These are all signals that mean at some point the agent was missing context. We've got to figure out how to slurp it up and put it back in the reboot.”
Ryan Lopopolo Apr 7, 2026 ▶ 44:46 Extreme Harness Engineering: 1M LOC, 1B toks/day, 0% human code or review — Ryan Lopopolo, OpenAI
Apr 7, 2026 positive
Insight
Lopopolo: Converting UI Images to ASCII Art Improves AI Agent Layout Perception
“If we want to actually, like, make it see the layout, it's almost easier to rasterize that image to ASCII arc and feed it in to the agent.”
Ryan Lopopolo Apr 7, 2026 ▶ 47:57 Extreme Harness Engineering: 1M LOC, 1B toks/day, 0% human code or review — Ryan Lopopolo, OpenAI
Apr 7, 2026 positive
Insight
Lopopolo: Strict success criteria in agent prompts increase deployment reliability
“Fundamentally, the agents are good at following instructions, so give them instructions, right? And it will, you know, improve the reliability of the result, right? Like we, much like the way we use Symphony, we don't want folks to have to monitor the agent as…”
Ryan Lopopolo Apr 7, 2026 ▶ 53:52 Extreme Harness Engineering: 1M LOC, 1B toks/day, 0% human code or review — Ryan Lopopolo, OpenAI
Apr 13, 2026 bullish
Prediction Not checkable as stated
Stay Sassy EM: Dashboards must become agent-compatible this year
“The ability to do, like, any simple find-and-grab information from a dashboard, like, I think that just has to be agent-compatible, like, this year for companies, because that is one of those things where if you spend 10 minutes searching for a key that you ca…”
Stay Sassy EM Apr 13, 2026 ▶ 25:59 ⚡️ The best engineers don't write the most code. They delete the most code. — Stay Sassy
Apr 13, 2026 negative
Insight
Stay Sassy PM: Managing multiple AI agents exhausts engineers reviewing critical code
“One of the things that people quietly talk about is that that's very fatiguing. And I know some people who are operating that way. I've met these people and it is fatiguing. And what is fatigue really going to hit you on? Fatigue is really gonna hit you when n…”
Stay Sassy PM Apr 13, 2026 ▶ 48:09 ⚡️ The best engineers don't write the most code. They delete the most code. — Stay Sassy
Apr 15, 2026 neutral
Insight
Simon Last: 99% of Agent Failures Are Tool Bugs, Not Model Flaws
“And actually, 99% of the time, it's a bug in one of the tools. Right. And so, just fix the bug.”
Simon Last Apr 15, 2026 ▶ 1:15:20 Notion’s Sarah Sachs & Simon Last on Custom Agents, Evals, and the Future of Work
Apr 15, 2026 positive
Insight
CLI Environments Let AI Agents Autonomously Debug and Fix Broken Tools
“And then I think the most important thing that's super cool is that there it's also inherently a bootstrapped. So if there's an issue of the agent can debug and fix itself within the same environment that it uses the tool, right?”
Simon Last Apr 15, 2026 ▶ 41:45 Notion’s Sarah Sachs & Simon Last on Custom Agents, Evals, and the Future of Work
Apr 15, 2026 positive
Insight
Supervising AI Agents Is a Deeply Technical Systems Problem, Unlike Managing Humans
“There's a critical difference To being a manager, which is that like, it is actually very deeply technical. The problem of, you know, humans are very like fuzzy and you can't like treat a team of humans like a rigorous system where like, you know, PRs like fle…”
Simon Last Apr 15, 2026 ▶ 31:41 Notion’s Sarah Sachs & Simon Last on Custom Agents, Evals, and the Future of Work
Apr 18, 2026 positive
Insight
Eifrem: Context graphs capturing organizational decisions are vital for agentic automation
“And if generally we're in the kind of space right now of trying to shift decision-making from, you know, kind of human brains into agentics brains, right? It used to be from wetware to software. I don't even know what to call kind of the LLMs now, right? Into …”
Emil Eifrem Apr 18, 2026 ▶ 28:35 ⚡️ How to turn Documents into Knowledge: Graphs in Modern AI — Emil Eifrem, CEO Neo4J
Apr 18, 2026 positive
Insight
Eifrem: Production AI agents require four distinct enterprise data sources
“So my view on it is that in my head, it completed the quadrant of the types of data sources that are required to reach what I've been talking about as kind of escape velocity for agents in production. And so what I mean by that is I think it's exactly four ty…”
Emil Eifrem Apr 18, 2026 ▶ 25:46 ⚡️ How to turn Documents into Knowledge: Graphs in Modern AI — Emil Eifrem, CEO Neo4J
Apr 27, 2026 bullish
Disclosure
Ludwig: Applied Intuition enables AI agents to configure sensor suites
“Nowadays though we expose all of the underlying APIs for that. And now using AI agents, you can actually configure a sensor suite with just text and likely reach a better result than you could have through the GUI in the past. And we're taking that thinking no…”
Peter Ludwig Apr 27, 2026 ▶ 25:25 The $15B Physical AI Company: Simulation, Autonomy OS, Neural Sim, & 1K Engineers—Applied Intuition
May 14, 2026 neutral
Insight
Asawa: Almost every AI agent functions as a coding agent underneath
“Almost every agent is a coding agent underneath underneath the hood, right? So you give it whatever, a file system, it can write its own code and so forth.”
Chai Asawa May 14, 2026 ▶ 22:40 Inside Abridge: The AI Listening to 100 Million Doctor Visits — Abridge's Janie Lee & Chai Asawa
May 14, 2026 positive
Insight
Asawa: Infrastructure built for human collaboration will remain durable for AI agents
“So all these things that we've built for, I actually think the things we've built for humans are actually the things that are going to be continually durable.”
Chai Asawa May 14, 2026 ▶ 58:06 Inside Abridge: The AI Listening to 100 Million Doctor Visits — Abridge's Janie Lee & Chai Asawa
May 20, 2026 positive
Insight
Cooper: AI coding agents prefer complex CLIs with hundreds of flags
“Things that were prohibitively annoying to humans are not actually prohibitively annoying to agents. They're really, really nice, right? And so, if I wanted to hand you a CLI and I said, hey, guess what? The CLI has 40 arguments and. Right. 600 flags. You'd be…”
Jake Cooper May 20, 2026 ▶ 30:18 The Agent-Native Cloud: 3M Users, 100K Signups/Wk, Data Centers, & Death PRs — Jake Cooper, Railway
May 20, 2026 neutral
Insight
Cooper: Infrastructure canvases will become output review tools instead of inputs
“We have previously used it a lot as an input and its goal moving forward is actually a lot more like an output. What I mean by that is you would go to the canvas and you'd make some changes and all these other things, whatever. Right. And you see them and, you…”
Jake Cooper May 20, 2026 ▶ 33:26 The Agent-Native Cloud: 3M Users, 100K Signups/Wk, Data Centers, & Death PRs — Jake Cooper, Railway
May 21, 2026 negative
Insight
Burazin: Infrastructure built for human developers fails for AI agents
“Most people thought it was the same infrastructure for humans and agents. We understood a quarter ago. It's not, we just didn't know what was the right primitive.”
Ivan Burazin May 21, 2026 ▶ 12:07 AI Agents Need Computers: 74% MoM Growth, 850K/Day Runs, & New Agent Cloud — Ivan Burazin, Daytona
May 21, 2026 positive
Insight
Burazin: AI agents need persistent, stateful computers like human laptops
“Agents will be like humans in the sense of you don't want your laptop to be shut down until you're done with work. Like, and you want to close the lid and open the lid. It's the same state. So agents would want that like pause and come back. They want those tw…”
Ivan Burazin May 21, 2026 ▶ 13:57 AI Agents Need Computers: 74% MoM Growth, 850K/Day Runs, & New Agent Cloud — Ivan Burazin, Daytona
May 21, 2026 bullish
Prediction Not checkable as stated
Burazin: A dedicated agent cloud bundling sandboxes and primitives will emerge
“There will be a cloud built out specifically for agents. And so that cloud will have sandboxes and it will have web search and it'll have databases like SQLite or Neon or whatever, specifically for agent and other things. We are not at the end of the new infra…”
Ivan Burazin May 21, 2026 ▶ 1:10:47 AI Agents Need Computers: 74% MoM Growth, 850K/Day Runs, & New Agent Cloud — Ivan Burazin, Daytona
May 24, 2026 positive
Assertion Supported
Sanseviero: Kaggle launched an exam-based benchmark leaderboard for AI agents
“Last week, they released a new system for agent evaluation. It's like a very, like, experimental initial benchmark, but pretty much allowing agents to take an exam and compete in a leaderboard, which is always fun.”
Omar Sanseviero May 24, 2026 ▶ 28:31 ⚡️ Google's Open AI Strategy — Omar Sanseviero, Google DeepMind
Jun 2, 2026 neutral
Insight
Daigle: Replacing pull requests is hard because code review codifies social trust
“I think the reason why there's not a single answer is ultimately we're trying to codify trust. We're trying to say like, okay, if Sean reviews this, I'm going to trust it because you're Sean or you're the senior dev or you're the whatever. And right now when w…”
Kyle Daigle Jun 2, 2026 ▶ 36:16 GitHub’s Agent Era: 14x Commits, 200M Developers, Copilot’s Next Act — Kyle Daigle
Jun 2, 2026 bearish
Opinion
Daigle: AI code vendoring will not solve supply chain vulnerabilities
“And so I do think there's something true there where having like either taking only what you need or the dependencies just getting incredibly small over time, I think will help to some degree, but it's not going to solve the fundamental problem. I don't think …”
Kyle Daigle Jun 2, 2026 ▶ 31:11 GitHub’s Agent Era: 14x Commits, 200M Developers, Copilot’s Next Act — Kyle Daigle
Jun 2, 2026 neutral
Insight
Daigle: Operating systems must evolve because autonomous AI agents now use them
“Operating systems need to look different than they looked five years ago because it's not just you using them anymore.”
Kyle Daigle Jun 2, 2026 ▶ 1:17:56 GitHub’s Agent Era: 14x Commits, 200M Developers, Copilot’s Next Act — Kyle Daigle
Jun 2, 2026 positive
Insight
Daigle: Running AI agents locally requires improved OS-level sandboxing
“One of the problems that we have, right. Is that our agents, if you install them, not on a Mac mini or not on a hosted device, you install them on a personal device or a work device. We need better sandboxing at the OS level.”
Kyle Daigle Jun 2, 2026 ▶ 1:16:08 GitHub’s Agent Era: 14x Commits, 200M Developers, Copilot’s Next Act — Kyle Daigle
Jun 4, 2026
Assertion Not checkable as stated
Petersson: AI agent bought perishable tomatoes two weeks early, leaving them rotten
“The agent bought like a shit ton of tomatoes two weeks earlier. And before the opening and now they're all rotten.”
Lukas Petersson Jun 4, 2026 ▶ 1:13:34 When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
Jun 4, 2026 neutral
Opinion
Backlund: Autonomous agents can manage entire e-commerce businesses today
“I think it can be done today, but you would do it in like e-commerce where it's like the probability of success is like really low, no matter if a human or an agent does it, but like an agent could surely manage everything.”
Axel Backlund Jun 4, 2026 ▶ 32:46 When AI Agents Run Businesses — Lukas Petersson and Axel Backlund of Andon Labs
Jun 22, 2026 bullish
Prediction Not checkable as stated
Kolter: Security and science will explode as AI agents automate tedious verification
“So I think this is really sort of an underappreciated point that we're reaching this point, this sort of phase where a lot of security, a lot of science has this potential to kind of explode. Not because we're going to get better at it, but because agents can …”
Zico Kolter Jun 22, 2026 ▶ 46:01 AI Security After Codex and Claude Code — Zico Kolter & Matt Fredrikson, Gray Swan
Jun 22, 2026 neutral
Prediction Not checkable as stated
Kolter: AI agents inheriting user permissions by default will soon change
“So far, we are still a lot, in a lot of cases, operating on the condition that your agent has your permissions. Yeah. That is a very standard default. And I think that will be changed. I mean, your permissions may be in a sandbox, but still kind of your permis…”
Zico Kolter Jun 22, 2026 ▶ 53:00 AI Security After Codex and Claude Code — Zico Kolter & Matt Fredrikson, Gray Swan
Jun 22, 2026 positive
Prediction Not checkable as stated
Kolter: Agent identity will evolve around user personas before fine-grained permissions
“I think in terms of how this will evolve, actually, I don't think it'll be per app, but I think what will happen first is people have different personas that they have, right? So you don't want your work life and your home email to be mixed up. Yeah. Right. A …”
Zico Kolter Jun 22, 2026 ▶ 54:19 AI Security After Codex and Claude Code — Zico Kolter & Matt Fredrikson, Gray Swan
Jun 24, 2026 positive
Insight
Zaharia: Protocol design remains essential for multi-party interoperability despite fast coding
“For this type of interoperability where multiple parties that are moving at different speeds are building stuff and you still want some layer on top to coordinate you do want to design it and build it. So it reminds me of that, like agents talking to each othe…”
Matei Zaharia Jun 24, 2026 ▶ 6:37 The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Jun 24, 2026 neutral
Insight
Zaharia: Prototyped AI agents stall when security teams block internal data access
“That's where we've seen a lot of other agents like hit things like people think they prototyped an awesome agent, but you know, it's not allowed to connect to like some really important data or whatever because of the, Security team.”
Matei Zaharia Jun 24, 2026 ▶ 9:39 The Agent Cloud: Databricks’ Bet on the Future of AI — Matei Zaharia and Reynold Xin
Jun 29, 2026 negative
Insight
Cohen: Semantic retrieval cannot answer broad weekly focus queries for agents
“There's no retrieval-based search, semantic search, keyword matching search that's going to be able to give your agent that information.”
Gavriel Cohen Jun 29, 2026 ▶ 13:04 The Blueprint for Autonomous Work Agents | Gavriel Cohen, NanoClaw
Jun 29, 2026 neutral
Insight
Cohen: AI agents require continuous maintenance unlike static enterprise software
“Agents are different than normal software in that normal enterprise software. You can deploy it, put it on some server and let it run for like five years. And as long as you never touch it just works. Agents don't really work that way. The core thing that you'…”
Gavriel Cohen Jun 29, 2026 ▶ 17:43 The Blueprint for Autonomous Work Agents | Gavriel Cohen, NanoClaw
Jul 8, 2026 positive
Insight
Bubna: AI agent hallucinations are actionable product feedback
“Sometimes it makes sense, like, if they're reaching for this thing, it's product feedback, like, give it to them.”
Akshat Bubna Jul 8, 2026 ▶ 58:26 The Future of AI Infra: from Kubernetes to Agent Sandboxes — Akshat Bubna, Modal CTO
Jul 8, 2026 neutral
Insight
Bubna: AI agents struggle to reason through logs and observability
“I think the things that sometimes agents struggle with without right guidance and a skill is how to use the rest of our observability. Like, how to, something is failing, like, How do you look at the logs and then update the right thing? It's sort of reasoning…”
Akshat Bubna Jul 8, 2026 ▶ 38:15 The Future of AI Infra: from Kubernetes to Agent Sandboxes — Akshat Bubna, Modal CTO
Jul 11, 2026
Insight
Perszyk: Agent reliability requires modeling user intent over UI clicking
“So ultimately reliability has less to do with clicking in the same place and scrolling and more to do with modeling the user's mind. And that shift is everything that reframes how we think about what it is that we're building.”
Danielle Perszyk Jul 11, 2026 ▶ 17:38 Why AI Agents Don't Actually Understand You — Danielle Perszyk, Amazon AGI Lab
Jul 11, 2026 bearish
Assertion Not checkable as stated
Perszyk: Current AI agents remain too unreliable to automate substantial work
“We look at what the metrics of the actual agents and they're so unreliable that ironically we feel a little bit better. The AI is actually not where we need it to be. To automate enough of the work.”
Danielle Perszyk Jul 11, 2026 ▶ 4:48 Why AI Agents Don't Actually Understand You — Danielle Perszyk, Amazon AGI Lab
Jul 28, 2026 bullish
Opinion
Nathan: Untapped enterprise agent market is 10x to 100x larger
“So now like we're seeing with agents, like there is Probably a contingent of like early adopters still who, you know, truly get it. We're like, you know, you can do anything. You just have to make sure the right context is there. It's connected to the right to…”
Akshay Nathan Jul 28, 2026 ▶ 6:23 OpenAI’s Vision for the AI Super App — Akshay Nathan, OpenAI
Aug 15, 2026 neutral
Insight
Krentsel: Interruptible agent workflows will mirror OS-style background task management
“So the way that we're going to need to reinvent what that kind of interactive interruptible work looks like in the agent layer, which is separate from the work going on, I think in the, like in the model layer. And in this like interactive model space, but I t…”
Alex Krentsel Aug 15, 2026 ▶ 31:39 Exo: Harnesses should see their own code and logs — Alex Krentsel, UC Berekeley / Google Research
Aug 21, 2026 positive
Insight
Park: Accurate human simulation models must precede complex automation agents
“This technology around simulation, creating accurate representation of people ought to precede the more complex agents that would automate the world that we live in.”
Joon Sung Park Aug 21, 2026 ▶ 8:32 Simulating Humanity: from Generative Agents to 8 Billion Digital Twins — Joon Sung Park, Simile AI
Sep 2, 2026 positive
Insight
Lie: Ultra-fast token generation enables more capable AI agent reasoning
“If you're running your model at over 4000 tokens per second. Now the, you can do, you know, more agentic loops. You can do more reasoning. Ultimately you get significantly more capable, more intelligent agents.”
Sean Lie Sep 2, 2026 ▶ 6:16 The Inference Frontier: from 100 to 10,000 tokens per second — Sean Lie, Cerebras CTO
Sep 7, 2026 bullish
Opinion
Slack: AI coding agent adoption is faster than any prior tech transition
“It's moving way faster than any other technology adoption that I've ever seen.”
Quinn Slack Sep 7, 2026 ▶ 6:09 Orbs: Shifting Coding to Cloud — Quinn Slack, Amp Code
Sep 7, 2026 bullish
Insight
Slack: AI agents and code are the ultimate settings screen
“Really an agent is the ultimate settings screen for any software and code is the ultimate setting screen for any software.”
Quinn Slack Sep 7, 2026 ▶ 10:14 Orbs: Shifting Coding to Cloud — Quinn Slack, Amp Code
Sep 7, 2026 bullish
Insight
Slack: AI agents eliminate the need to hire low-agency staff
“And compared to a few years ago, You do not have the need to go and hire people that you do not trust, that you do not trust to have high agency or the right skills, because you can use an agent to do those things.”
Quinn Slack Sep 7, 2026 ▶ 30:42 Orbs: Shifting Coding to Cloud — Quinn Slack, Amp Code
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.