Augment Code

company on 3 shows · 11 statements across 1 episodes · said 15 times in 4 episodes since 2025

Latent Space 11 20VC 3 Big Technology 1

Mentions by year, every show

tap a year for its mentions
008215320252026episodesmentions
02320252026episodes it came up in
002.51.55320252026episodesmentions per episode

Latent Space 1120VC 3Big Technology 1

every mention on every show, scene by scene, with the transcript →

11 statements about Augment Code, every show

LATENT SPACE Assertion Supported
Gur-Ari: Augment Code Achieved #1 on SWE-Bench Verified
“We just made number one on Sweepbench. So for us Sweepbench has been a useful tool for exploring how can we get the most out of agents. And so we were able to get the best result on Sweepbench verified right now.”
Guy Gur-Ari Apr 2, 2025 ▶ 1:24 The #1 SWE-Bench Verified Agent
LATENT SPACE Disclosure
Augment's #1 SWE-Bench agent uses off-the-shelf models, unlike their custom product
“The generation models for Sweebench, it's all off the shelf models. For the product, it includes our own custom trained models that help the model with, that help the agent with code-based understanding.”
Guy Gur-Ari Apr 2, 2025 ▶ 1:42 The #1 SWE-Bench Verified Agent
LATENT SPACE Assertion Not checkable as stated
Sequential Thinking MCP outperformed Claude 3.7 native reasoning mode in evaluations
“We tried reasoning mode as well with the new three seven. And we didn't see that much of a bump in performance. We don't know if this is something that's code specific or not. I don't have an insight. We tried both and yeah, sequential thinking worked better.”
Guy Gur-Ari Apr 2, 2025 ▶ 3:57 The #1 SWE-Bench Verified Agent
Engineers should begin AI feature evaluation with 10-sample interactive notebooks
“So I would say the process that I like to follow is in the beginning when developing a feature, come up with a curated set of samples. Could be as small as 10, 10 samples that you run through, and then yeah, it's all notebooks basically. You run through the sa…”
Guy Gur-Ari Apr 2, 2025 ▶ 6:10 The #1 SWE-Bench Verified Agent
Gur-Ari uses human contractors because chat and agent evaluations resist automation
“The last thing I can mention is we use contractors for evaluation where we cannot Do automatic evaluation. So with chats and agents, it becomes way harder to do things automatically. And so we use contractors for that.”
Guy Gur-Ari Apr 2, 2025 ▶ 7:33 The #1 SWE-Bench Verified Agent
LATENT SPACE Disclosure
Augment Code to ship multi-minute codebase orientation agent for thorough analysis
“We also have a feature that I think has not shipped yet, but we are planning to ship, which is a more thorough orientation. So this is something that will run for several minutes probably and try to do a pretty thorough job of trying to figure out You know, wh…”
Guy Gur-Ari Apr 2, 2025 ▶ 11:38 The #1 SWE-Bench Verified Agent
LATENT SPACE Prediction Not checkable as stated
Running multiple parallel agents will unlock most value for software developers
“Being able to run multiple agents and not just one is going to be the way to unlock Honestly, most of the value out of these agents, and that's what we're working toward.”
Guy Gur-Ari Apr 2, 2025 ▶ 13:56 The #1 SWE-Bench Verified Agent
LATENT SPACE Assertion Supported
Augment Code runs inside the Cursor editor and already has active users
“Yes, you can use augment inside of cursor. We actually have quite a few users doing that.”
Guy Gur-Ari Apr 2, 2025 ▶ 15:26 The #1 SWE-Bench Verified Agent
LATENT SPACE Assertion Supported
Augment Code supports MCP alongside native GitHub, Linear, and Notion integrations
“Within augment, we have we have both MCP support for complete extensibility of tools, but we, we've also built in a few first party integrations and serve them as tools to the agent. So we have GitHub linear notion that, that we use internally, and then we hav…”
Guy Gur-Ari Apr 2, 2025 ▶ 18:17 The #1 SWE-Bench Verified Agent
LATENT SPACE Prediction Not checkable as stated
Developers will eventually spend 80% of their time controlling agents outside IDEs
“I think in the future, at some point, my guess is that the IDE is going to become less of the focal point and more like an app that you can launch when you need to dig in deeper, but you spend most of your time away from it. So maybe 80% of your time is in a w…”
Guy Gur-Ari Apr 2, 2025 ▶ 24:18 The #1 SWE-Bench Verified Agent
LATENT SPACE Assertion Supported
Augment Code open-sourced the SWE-bench implementation that reached number one
“We actually open sourced our implementation of Sweebench. So if you're curious how we got to number one, you'll be able to go see all the details of how we did it.”
Guy Gur-Ari Apr 2, 2025 ▶ 31:13 The #1 SWE-Bench Verified Agent

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.