The Ledger

Every statement that passed quotation and attribution checks. Mix any filter with any other: certainty 1/5, debate potential 5/5, or both at once.

clear all ✕

why aren't all 8 resolved? a statement only gets an assessment when the public record can support or contradict it. opinions and what-ifs never can, and 0 checkable ones are still open, waiting for their date. predictions held up or didn't; assertions are supported or contradicted. on every card: ▮▮▮▮▮ certainty · ▮▮▮▮▮ debate potential. speakers are clickable

Assertion Supported
Attackers can poison the LAION dataset simply by purchasing expired domains
“Here's this new dataset. It is being distributed in such a way that anyone in the world can buy domains that let you then inject arbitrary images in the dataset.”
Nicholas Carlini Aug 28, 2024 ▶ 55:14 Personal benchmarks vs HumanEval - with Nicholas Carlini of DeepMind
Assertion Supported
Carlini extracted production models from Google and OpenAI with legal permission
“We ran the attack that let us, yeah, stole several of OpenAI's models. With their permission... We notified everyone who was vulnerable to this attack. Some Google models were vulnerable. Some open AM models were vulnerable. There were one or two other people …”
Nicholas Carlini Aug 28, 2024 ▶ 57:21 Personal benchmarks vs HumanEval - with Nicholas Carlini of DeepMind
Assertion Supported
Prompting ChatGPT to repeat a word indefinitely leaks verbatim training data
“One of my co-authors, Milad was working on some other random experiments, and he figured out that if you prompt ChatGPT to repeat a word forever, then it will repeat the word many, many, many times in a row, and then like explode and like just start doing rand…”
Nicholas Carlini Aug 28, 2024 ▶ 1:02:16 Personal benchmarks vs HumanEval - with Nicholas Carlini of DeepMind
Assertion Supported
Carlini: ChatGPT emitted verbatim 50+ word sequences from internet training data
“And what I can say is that the output of the model was a verbatim, at least 50 word in a row match. To some other document that appeared on the internet previously.”
Nicholas Carlini Aug 28, 2024 ▶ 1:03:03 Personal benchmarks vs HumanEval - with Nicholas Carlini of DeepMind
Assertion Supported
Carlini published a paper proving C's printf function is Turing-complete
“A while ago as part of a research paper, I was able to show that in C, if you call into print def. It's Turing complete, like printf, you know, like which, like, you know, you can print numbers or whatever, right?”
Nicholas Carlini Aug 28, 2024 ▶ 3:28 Personal benchmarks vs HumanEval - with Nicholas Carlini of DeepMind
Assertion Supported
Stanford researchers detect benchmark contamination by testing evaluation question ordering
“There's a paper by Tatsu at a Stanford. Where they check if the order that the specific questions happen to be in matters. And if the answer is yes, then you probably trained on it because the order of the questions is arbitrary and shouldn't matter.”
Nicholas Carlini Aug 28, 2024 ▶ 51:47 Personal benchmarks vs HumanEval - with Nicholas Carlini of DeepMind
Assertion Supported
Carlini: Every image in LAION-400M is pulled from live domains
“Every image gets pulled from a live domain.”
Nicholas Carlini Aug 28, 2024 ▶ 52:38 Personal benchmarks vs HumanEval - with Nicholas Carlini of DeepMind
Assertion Supported
Carlini built a gate-level CPU emulation for the IOCCC
“I have A very fun gate level emulation of an old CPU that runs, like, fully precisely, and it's a fun kind of thing.”
Nicholas Carlini Aug 28, 2024 ▶ 37:12 Personal benchmarks vs HumanEval - with Nicholas Carlini of DeepMind
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.