The Exchanges

Every argument clarity score on this site is built from rows on this page. Each question and answer was assessed with names hidden, the host's own answers included, on four things from 1 to 5: directness (does it answer the question asked), coherence (do the ideas follow), precision (concrete details and clear references), compression (says a lot per word). The weighted mix (30/30/25/15) is the exchange score. A person's published score averages their exchange scores on raw tape only, at least 8 of them, shrunk toward the cohort mean. Full method →

Satya Nadella no published score: only 1 usable exchange on raw tape, and a fair score needs 8+ record → ← everyone

Every exchange below was scored with names hidden, four dimensions each from 1 to 5. An exchange's score is 0.30·directness + 0.30·coherence + 0.25·precision + 0.15·compression. The published score averages the raw tape exchange scores and shrinks small samples toward the cohort mean, so five great answers can't beat twenty good ones. Produced feed rows count only toward coarse estimates, never toward a full score.

clear all ✕
1exchanges match
1on raw tape
0redirected or not addressed
Answered raw tape D 5 · C 4 · P 4 · Cm 4 4.30

Q think, uh, I, first of all, I have to congratulate you on basically building a Frontier Neo Lab inside of Microsoft in two years. Um, I'm wondering, you know, you have all this AI strategy that you're rolling out. What do you know now that you wish you would tell yourself two years ago, two or three years ago, three years for the Jensen partnership, two years for, uh, MAI?

A Yeah, I mean, I think the, the thing that I reflect quite a bit, right, which is sort of obviously I got into all this when I got excited by the, The scaling laws paper and, you know, when, you know, even the OpenAI partnership came about when those folks said, hey, we're going to really throw a lot of computer transformers, uh, and they've helped, right? The thing that I always look back and say, wow, these things, um, do have capability that they're climbing up. I mean, this, you know, this crude way of saying it is intelligence is log of compute kind of works. Now, what I think we underestimated perhaps is the real world complexity of deploying these so that they actually deliver the value in the real world, right? So the outcomes as measured by any benchmark is interesting, important, but the true eval is when people out there are able to do unique things that they only can value. And it's very measurable, right? That I wish we had sort of even like had more in our consciousness, right? Which is as an industry, because right now I think When people say, wow, I don't want a token max. It's an artifact of us not having taught ourselves as an industry that we are using tokens to create value every step of the way. So I think that's kind of what I wish we had gotten there, but I'm glad we are here.

AI assessment note: “what I think we underestimated perhaps is the real world complexity of deploying these”

page 1
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.