Raza Habib

Member of Technical Staff, Anthropic · 1 appearance on the record.

computed by AI from the episodes · how this works → · full disclaimer →

founderexecutiveengineerscientist@RazRazcle ↗LinkedIn ↗humanloop.com ↗

Raza Habib co-founded and served as CEO of Humanloop, a Y Combinator-backed LLM evaluation and observability platform, before joining Anthropic. He holds a PhD in Machine Learning from University College London and previously worked at Monolith AI and Google.

5statements → 1claims → 1claims resolved → 3.4/5average certainty → 2.6/5average debate potential → ≈4.5/5argument clarity, estimated →

1 supported 0 partly supported 0 contradicted how the 1 claim stands · each chip opens the sources

1 assertion · 2 opinions · 1 insight · 1 what if · every statement was checked. The predictions and assertion are the 1 claim: statements the public record can support or contradict. 1 is resolved. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.

The record, in short

What the tape says about how Raza argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

Their most notable supported claim

Assertion Supported
LangChain's LangSmith is fundamentally focused on monitoring chains and agents
“Their LangSmith product probably has a little bit of overlap with us, but it is not, you know, it has some overlap that is fundamentally, I think focused more around monitoring chains and agents.”
Raza Habib Nov 22, 2023 ▶ 22:37 How to Ship Reliable GenAI Apps: Humanloop CEO on LLM Observability, RAG & Rapid Iteration

How they sound: speaking style how? →

272 words/min while actually speaking · 1.7 um and uh per 1k words

No argument clarity score for Raza Habib: no usable question→answer exchanges on raw tape (a fair score needs 8+). We do not score a sample that small. Roundtable and news formats yield far fewer direct exchanges than interviews. Their coarse estimate from 7 exchanges on the produced feed is ≈4.5/5, shown at half point precision because the sample is small and produced tape scores higher.

Measured by listening to 5,916 words across 1 episode, but every recording we have of Raza Habib is the aired feed, and an editor cleaned that audio before release. Some of the hesitation was cut before we ever heard it, so read these as floors: the true rates are at least this high. These are measurements of speaking style, not scores. How it's measured →

Everything Raza Habib said on the MAD Podcast that made the record, most notable first. Filter by type, assessment or year in the ledger →

What-if
Restricting GPT-2 access would have cost years of AI progress
“It would have been really sad for the world, I think, if at the point of GPT-II, we had decided, hey, you know what? This is too dangerous. No one can have access. Cause we would have missed out on all of the past two or three years of incredible progress.”
Raza Habib Nov 22, 2023 ▶ 30:18 How to Ship Reliable GenAI Apps: Humanloop CEO on LLM Observability, RAG & Rapid Iteration
Opinion
Habib advocates regulating AI end-use cases rather than banning base models
“And so I'm generally in favor of finding ways to regulate end use cases and make the malicious use of the models be regulated and then punish that very strongly and make people responsible for the end outcomes that I am for banning the underlying technology.”
Raza Habib Nov 22, 2023 ▶ 30:39 How to Ship Reliable GenAI Apps: Humanloop CEO on LLM Observability, RAG & Rapid Iteration
Opinion
AI models will cross a capabilities threshold requiring restricted access
“Like there is a capabilities threshold above which it is correct that we would want to restrict access.”
Raza Habib Nov 22, 2023 ▶ 31:52 How to Ship Reliable GenAI Apps: Humanloop CEO on LLM Observability, RAG & Rapid Iteration
Insight
LLM observability must be tightly connected to prompt engineering tools
“Because prompt engineering allows you to intervene really quickly, I think it's very important to have evaluation and observability very closely connected to the prompt engineering tools.”
Raza Habib Nov 22, 2023 ▶ 12:08 How to Ship Reliable GenAI Apps: Humanloop CEO on LLM Observability, RAG & Rapid Iteration
Assertion Supported
LangChain's LangSmith is fundamentally focused on monitoring chains and agents
“Their LangSmith product probably has a little bit of overlap with us, but it is not, you know, it has some overlap that is fundamentally, I think focused more around monitoring chains and agents.”
Raza Habib Nov 22, 2023 ▶ 22:37 How to Ship Reliable GenAI Apps: Humanloop CEO on LLM Observability, RAG & Rapid Iteration

Appearances (1)

EpisodeDateSpeaking time
How to Ship Reliable GenAI Apps: Humanloop CEO on LLM Observability, RAG & Rapid Iteration Nov 22, 2023 25m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.