personally identifiable information

also referred to as: pii

3 statements across 2 episodes · 3 bullish · 0 bearish · 1 people on the record · first statement Dec 31, 2025 by Mark Bissell · across every show →

Everything said about personally identifiable information, oldest first

Dec 31, 2025 positive
Assertion Supported
Bissell: Rakuten uses interpretability in production to scrub PII from customer chats
“From Goodfire's perspective, you know, we, so like one of our partners, Rakuten, is deploying an interpretability based tool in production with one of their language agents. This is a really cool use case where if you, What they needed to do was take chats bet…”
Mark Bissell Dec 31, 2025 ▶ 8:59 [State of MechInterp] SAEs in Production, Circuit Tracing, AI4Science, "Pragmatic" Interp — Goodfire
Dec 31, 2025 bullish
Insight
Bissell: Probing internal model features matches LLM-as-a-judge quality at 500x lower cost
“If you ask that model, try to like use it as an LLM as a judge, it's not very good. But if you probe its mind and you sort of detect when the features related to personally identifiable information are firing, that gets you the highest recall of anything. It's…”
Mark Bissell Dec 31, 2025 ▶ 9:53 [State of MechInterp] SAEs in Production, Circuit Tracing, AI4Science, "Pragmatic" Interp — Goodfire
Feb 5, 2026 positive
Assertion Not checkable as stated
Bissell: SAE-Based Approach Proved Most Generalizable for PII Detection
“Although in the PII instance, I think we're into SAE, an SAE based approach actually did prove to be the most generalizable.”
Mark Bissell Feb 5, 2026 ▶ 18:25 Goodfire AI’s Bet: Interpretability as the Next Frontier of Model Design — Myra Deng & Mark Bissell
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.