The Exchanges

Every argument clarity score on this site is built from rows on this page. Each question and answer was assessed with names hidden, the host's own answers included, on four things from 1 to 5: directness (does it answer the question asked), coherence (do the ideas follow), precision (concrete details and clear references), compression (says a lot per word). The weighted mix (30/30/25/15) is the exchange score. A person's published score averages their exchange scores on raw tape only, at least 8 of them, shrunk toward the cohort mean. Full method →

Morgan Linton no published score: only 1 usable exchange on raw tape, and a fair score needs 8+ record → ← everyone

Every exchange below was scored with names hidden, four dimensions each from 1 to 5. An exchange's score is 0.30·directness + 0.30·coherence + 0.25·precision + 0.15·compression. The published score averages the raw tape exchange scores and shrinks small samples toward the cohort mean, so five great answers can't beat twenty good ones. Produced feed rows count only toward coarse estimates, never toward a full score.

clear all ✕
1exchanges match
1on raw tape
0redirected or not addressed
Answered raw tape D 5 · C 5 · P 4 · Cm 4 4.60

Q Totally. Yeah. And I think like, I guess one question I, I have is like, is one Is one model better for being more of a beginner, non-technical live coder, or, you know, doesn't really matter?

A Yeah, it's a good question. I mean, I think the fair answer would be probably Codex, because Codex, ah, edged out Opus four, six a little bit on some of those coding benchmarks, and is, is kind of known for writing better production code. Probably codex in that way. Um, at the same time, one of the downsides, and like I said, this is, I could only do this in a totally balanced way because they're so different. Um, you know, at the same time for a vibe coder, knowing when to interject and stop codex and say, oh, wait, you're doing this this way. Uh, can you instead look at doing it this way? They're probably not going to know how to do that, right? And so that's where maybe Opus four, six is better where you could say, okay, spin up four or five agents and let them work with each other. Right?

AI assessment note: “I think the fair answer would be probably Codex, because Codex”

page 1
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.