The Exchanges

Every argument clarity score on this site is built from rows on this page. Each question and answer was assessed with names hidden, the host's own answers included, on four things from 1 to 5: directness (does it answer the question asked), coherence (do the ideas follow), precision (concrete details and clear references), compression (says a lot per word). The weighted mix (30/30/25/15) is the exchange score. A person's published score averages their exchange scores on raw tape only, at least 8 of them, shrunk toward the cohort mean. Full method →

João Moura no published score: only 1 usable exchange on raw tape, and a fair score needs 8+ record → ← everyone

Every exchange below was scored with names hidden, four dimensions each from 1 to 5. An exchange's score is 0.30·directness + 0.30·coherence + 0.25·precision + 0.15·compression. The published score averages the raw tape exchange scores and shrinks small samples toward the cohort mean, so five great answers can't beat twenty good ones. Produced feed rows count only toward coarse estimates, never toward a full score.

clear all ✕
1exchanges match
1on raw tape
0redirected or not addressed
Answered raw tape D 4 · C 4 · P 4 · Cm 3 3.85

Q As you're, uh, as you're putting that in, I have a question around one of the things that I noticed you change was adding, I think it was like a senior writing correspondent or whatever. You added the word senior. How, how important is, like, does that actually affect the outputs of your crew if you, if you write senior?

A To some extent, yes. I mean, not necessarily senior alone, right? But because these agents, they're kind of like impersonating roles. You do get different behaviors depending on, you do get different behaviors depending on how you give them like a different role. So you can actually replicate this on ChatGPT. Like if you ask ChatGPT to like give you an assessment on a stock, it's going to give you kind of like a report. But if you say, uh, behave as an FCC, kind of like a proven kind of like a stock analyst, it's going to give you a completely different answer, even though all the remaining of the The prompt is still the same. So, uh, it's one way that you can steer the model to behave in a certain way if you really want to go there. So, uh, here you can see, like, uh, I just run it just for one execution, just one new one model, but you can see what would be your task scores, your crew overall score, your execution time. So it can actually compare many models in this, and you can run this like many different times and see what happens. And if you want to take this up a notch on the enterprise side of things that again, there is a free tier, so you can try it out. There is, let me see one of the agents that I did run. Tests for maybe the PR review. There is a way where you can compare. Yeah, there you go. You can compare many models. So you can see here, GPT, Faro Mini, and the…

AI assessment note: “To some extent, yes. I mean, not necessarily senior alone, right?”

page 1
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.