AI Security Researcher

2 appearances on the record.

computed by AI from the episodes · how this works → · full disclaimer →

founderengineerscientist

6statements → 5claims → 2claims resolved → 3.5/5average certainty → 3/5average debate potential →

2 supported 0 partly supported 0 contradicted 1 not yet assessed 2 not checkable as stated how the 5 claims stand · each chip opens the sources

5 assertions · 1 insight · every statement was checked. The predictions and assertions are the 5 claims: statements the public record can support or contradict. 2 are resolved, 1 is not yet assessed, and 2 name no date, number or outcome precise enough to check. Everything else (opinions, insights, what ifs, disclosures) can never be settled by the record, so it carries no assessment.

The record, in short

What the tape says about how AI argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

Their most notable supported claim

Assertion Supported
Webster: DeepSeek returns CCP propaganda or refusals on Tiananmen Square
“If you ask it about Tiananmen Square or whatever, it will either give you a refusal or it will give you like this long diatribe of The CCP party line, like, you know, nothing happened. We believe in harmony in China and blah, blah, blah.”
AI Security Researcher Feb 28, 2025 ▶ 3:37 How to use DeepSeek safely

Expressed certainty vs assessment result

none yet certainty 1
none yet certainty 2
100% certainty 3
100% certainty 4
none yet certainty 5

weighted support: a fully supported claim counts one, a partly supported claim counts half. Each filled bar is clickable and opens exactly those claims; "none yet" means nothing said at that certainty level has resolved yet

How they sound: speaking style how? →

285 words/min while actually speaking · 25.7 um and uh per 1k words

No argument clarity score for AI Security Researcher: no usable question→answer exchanges on raw tape (a fair score needs 8+). We do not score a sample that small. Roundtable and news formats yield far fewer direct exchanges than interviews.

Measured by listening to the audio itself: 5,369 words across 2 episodes of raw-level tape, transcribed verbatim with every um and uh kept, each one attributed only where the alignment onto our timed stream is unambiguous. These are measurements of speaking style. We do not rank them: across this corpus, fluency and argument quality are nearly uncorrelated (ρ≈0.2), and smooth talking does not signal clear thinking. How it's measured →

Everything AI Security Researcher said on the a16z Podcast that made the record, most notable first. Filter by type, assessment or year in the ledger →

Insight
Ayrey: Filtering API keys from training risks degrading AI data science skills
“So if we do our reinforcement learning and we skew it towards code snippets that's generating that don't have API keys, inadvertently, we may be training this thing to behave less like a data scientist. And then we lose the entire discipline of data science in…”
AI Security Researcher Feb 28, 2025 ▶ 9:48 Avoiding vulnerabilities in AI code
Assertion Supported
Webster: DeepSeek returns CCP propaganda or refusals on Tiananmen Square
“If you ask it about Tiananmen Square or whatever, it will either give you a refusal or it will give you like this long diatribe of The CCP party line, like, you know, nothing happened. We believe in harmony in China and blah, blah, blah.”
AI Security Researcher Feb 28, 2025 ▶ 3:37 How to use DeepSeek safely
Assertion Not checkable as stated
Webster: DeepSeek performs 20% worse than GPT on jailbreak benchmarks
“On our benchmarks, it performs about 20% worse.”
AI Security Researcher Feb 28, 2025 ▶ 4:43 How to use DeepSeek safely
Assertion Not checkable as stated
Webster: DeepSeek's safety is on par with early GPT-3.5 from 2023
“Qualitatively, what we see is performance on par with GPT 3.5, which is to say, you know, in 2023, when open AI launched GPT, there were there were a bunch of like zero day, really simple jailbreaks and deep seek is essentially susceptible to all of those.”
AI Security Researcher Feb 28, 2025 ▶ 4:56 How to use DeepSeek safely
Assertion Open · timeframe Feb 2026
Ayrey: Data scientists leak API keys more frequently than SREs
“Data scientists leak out API keys and passwords more often than site reliability engineers.”
AI Security Researcher Feb 28, 2025 ▶ 9:18 Avoiding vulnerabilities in AI code
Assertion Supported
Ian Webster: DeepSeek uses a separate system for political censorship
“For Deep Seek specifically, there was the part that limited speech about politically sensitive topics in China. So this is stuff like, you know, Taiwan or Tiananmen Square, that kind of thing. And it's pretty clear that that was Basically a separate system fro…”
AI Security Researcher Feb 28, 2025 ▶ 3:02 How to use DeepSeek safely

Appearances (2)

EpisodeDateSpeaking time
Avoiding vulnerabilities in AI code Feb 28, 2025 14m
How to use DeepSeek safely Feb 28, 2025 10m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 1,000 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.