Promptfoo

product on 3 shows · 5 statements across 2 episodes · said 30 times in 3 episodes since 2024

Latent Space 28 the Y Combinator Startup Podcast 2 the a16z Podcast

Mentions by year, every show

tap a year for its mentions
0015130220242025episodesmentions
01220242025episodes it came up in
007.5115220242025episodesmentions per episode

Latent Space 28the Y Combinator Startup Podcast 2

2025 29 mentions in 2 episodes 15 per episode
2024 1 mention in 1 episode

every mention on every show, scene by scene, with the transcript →

5 statements about Promptfoo, every show

Webster: AI evaluation tools are table-stakes commodities facing a feature-parity bloodbath
“I think evals are our table stakes. I think that they're a commodity and everyone should be doing them. And yes, there are companies that are doing great in the eval space, but To me, it just seemed like a bloodbath, you know, like we would just be, had a grea…”
Ian Webster Oct 24, 2025 ▶ 6:13 Breaking AI to Fix It: Ian Webster's Journey from Discord's Clyde to Promptfoo's $18M Series A
LATENT SPACE Prediction Not checkable as stated
Webster: Meaningful AI red teaming will require internal tracing and observability
“I think especially where, where things are headed, like with more complex rags and agents and so forth, you're going to have to have some type of observability or like internal tracing in order to have, to do meaningful automated red teaming.”
Ian Webster Oct 24, 2025 ▶ 14:21 Breaking AI to Fix It: Ian Webster's Journey from Discord's Clyde to Promptfoo's $18M Series A
Webster: AI guardrails are a commodity and easy to build
“Building a business on guardrails really scares me because there are so many incumbents that can come in and eat your lunch. And it's like, It's not actually that hard to build a guardrail. I don't know if I'm gonna make people angry by saying that maybe some …”
Ian Webster Oct 24, 2025 ▶ 32:20 Breaking AI to Fix It: Ian Webster's Journey from Discord's Clyde to Promptfoo's $18M Series A
a16z Assertion Not checkable as stated
Webster: DeepSeek performs 20% worse than GPT on jailbreak benchmarks
“On our benchmarks, it performs about 20% worse.”
AI Security Researcher Feb 28, 2025 ▶ 4:43 How to use DeepSeek safely
a16z Assertion Supported
Webster: DeepSeek hard-censors 85% of politically sensitive topics in Promptfoo benchmark
“We did a benchmark on Chinese politically sensitive topics that, that found that about 85% of those topics in our test set were hard censored.”
Ian Webster Feb 28, 2025 ▶ 7:47 How to use DeepSeek safely

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.