The Exchanges

Every argument clarity score on this site is built from rows on this page. Each question and answer was assessed with names hidden, the host's own answers included, on four things from 1 to 5: directness (does it answer the question asked), coherence (do the ideas follow), precision (concrete details and clear references), compression (says a lot per word). The weighted mix (30/30/25/15) is the exchange score. A person's published score averages their exchange scores on raw tape only, at least 8 of them, shrunk toward the cohort mean. Full method →

Itamar Friedman no published score: only 2 usable exchanges on raw tape, and a fair score needs 8+ record → ← everyone

Every exchange below was scored with names hidden, four dimensions each from 1 to 5. An exchange's score is 0.30·directness + 0.30·coherence + 0.25·precision + 0.15·compression. The published score averages the raw tape exchange scores and shrinks small samples toward the cohort mean, so five great answers can't beat twenty good ones. Produced feed rows count only toward coarse estimates, never toward a full score.

clear all ✕
2exchanges match
2on raw tape
0redirected or not addressed
Answered raw tape D 5 · C 4 · P 4 · Cm 4 4.30

Q see something similar with Factory AI also doing like droids. They all have special purpose doing things, but people don't really want general purpose agents, right? The last time you were here, we talked about AutoGBT, the biggest thing of This year, not really relevant anymore. And I think it's mostly just because, like, when you give me a general purpose agent, I don't know what to do with it.

A Yeah, I totally agree with that. We're seeing it for a while, and I think it will stay like that despite the computer use, et cetera, that supposedly can just replace us, and it could, you can just, like, prompt it to be, hey, now be a QA, you know, or be a QA person or, or developer. I still think that there's a few reasons why you see, like, a dedicated agent. Again, I'm a bit more focused, uh, like, my head is more on, you know, Complex software for big teams and enterprise, etc. And even, you know, think about permissions, and what are the data sources, and just the same way you manage permissions for users. Like, developers, you probably want to have dedicated guardrails and dedicated approvals for agents. I intentionally, like, touched a point that not many people think about. And of course, then, what you can think of, like, maybe there's different tools, tool use, etc. But just the first point by itself is a good reason why you want to have different agents.

AI assessment note: “Yeah, I totally agree with that. We're seeing it for a while”

Partly raw tape D 3 · C 2 · P 2 · Cm 2 2.30

Q of an impact does that have on your performance? Like, you know, is most of the work you're doing actually figuring out environment and like the libraries that, because I'm sure they're using Outdated version of languages. They're using outdated libraries. They're using forks that have not been on the public internet before. How much of the work that you're doing is like there versus like at the LLM level?

A One of the reasons I, I was asking about, uh, uh, you know, what are the steps to, to break things down? Because it really matters, uh, like what's a tech stack, uh, how complicated the software is. It's hard to figure it out when you're dealing with, uh, Real world, uh, any environment of enterprise as a city when I'm, like, uh, while maybe sometimes, like, uh, I think you do enable, like, and, and Bolt, like, to install stuff, but it's quite of, like, a controlled environment, and, uh, that's, that's a good thing to do, because then you narrow it down, and it's easier to, to make things, uh, work. So definitely there are two dimensions, uh, I think, actually, spaces. One is, is the fact just, like, installing our software. Without yet, like, uh, doing anything. Making it work, just installing it, because we work with Enterprise and Fortune, et cetera, many of them want on-prem solution, so.

AI assessment note: “there are two dimensions, uh, I think, actually, spaces. One is, is the fact just, like, installing”

page 1
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.