Assertion Supported
Rajpal: Anthropic Claude models had regressions from serving architecture changes
“Anthropix kind of cloud models kind of had a regression, right? Because they changed to a new serving architecture.”
Insight
Rajpal: Foundation models rarely generate toxic outputs without explicit jailbreaks
“Most of the stuff that the frameworks will recommend is actually stuff that the model providers are already working on. So toxicity, unless you're doing, unless somebody is very explicitly trying to jailbreak what you've built, you know, you won't run into the…”
Insight
Rajpal: AI simulations should prioritize product KPIs over generic safety metrics
“So I would actually say that like a lot of the things to simulate are more aligned with like product KPIs or product metrics that actually make Whatever AI system you're building very sticky, rather than, you know, focusing more on, like, traditional safety se…”
Insight
Rajpal: Fine-tuning open-source models on synthetic data closes proprietary capability gaps
“Not out of the box, but with a lot of that fine tuning and that the training, et cetera, you are able to kind of close the gap and even have better performance on metrics.”
Insight
Rajpal: AI agents resemble autonomous vehicle architectures with cascading ML units
“What patterns really worked well in self-driving cars, which is weirdly a very similar system to, you know, agents of today where you have like these cascading kind of like units that are all machine learning based and, you know, they all kind of like feed int…”
Insight
Rajpal: Simulations reveal which failure modes actually require runtime guardrails
“And then, you know, in simulation, figure out, you know, what is actually robust, what isn't, and then the stuff that isn't robust is the stuff that you need guardrails for.”
Insight
Rajpal: Product managers already act as AI persona engineers
“Interestingly, there are already persona engineers, and we call them like product managers, basically, you know. So your product managers are already thinking about, okay, I've built this, you know, model or this chatbot or this agent. Who are the personas? Wh…”
Opinion
Rajpal: Pure ChatGPT test conversations lack diversity and realism
“Compared to, let's say, you were asking, like, ChatGPT to generate, you know, these, like, conversations for you. They all kind of have that ChatGPT vibe, and this ends up looking, you know, very diverse and very grounded in, like, your use case and your data.”
Opinion
Rajpal: Historically conservative US banks are becoming very AI-forward
“Even, especially in the US, right, like, a lot of banks that you would think would be, like, historically maybe more conservative, like, maybe not the earliest technology adopters, like, they are very, like, AI forward and tech forward.”
Assertion Not checkable as stated
Rajpal: Most AI audio applications are a 'voice sandwich' around text
“I think like most audio applications today are like a text sandwich, or sorry, a voice sandwich with like kind of text in the middle.”
Assertion Not checkable as stated
Rajpal: General-Purpose AI Simulators Now Replace Manually Crafted Simulation Systems
“And with Snowglobe, a big kind of idea is that for the first time in history, we can actually have, you know, a general purpose simulation system, right? Like simulation systems have existed, but, and, you know, we saw them like extensively in self-driving and…”
Insight
Rajpal: Running massive AI simulations yields diminishing returns
“So I think I think there's definitely I guess a diminishing kind of like returns. You know phenomena with, like, running simulations that are absolutely massive.”
Assertion Supported
Rajpal: Waymo had 20 million real-world miles versus 20 billion in simulation
“Like Waymo had twenty million miles in the real World driving, but twenty billion miles in simulation.”
Assertion Not checkable as stated
Rajpal: Simulation testing revealed an early partner's real failure was over-refusal
“Organizations that we're, we were, we had as our design partner, we you know, they were like, oh, we're very worried about toxicity, and we want toxicity guardrails, and we did all of this testing for them in production, and toxicity was actually not a real co…”
Disclosure
Rajpal: Reusable persona libraries are Snowglobe's top requested feature
“Today, all personas are net new, but this is our number one requested feature, which is I want to be able to, you know, like maybe this, maybe some product leader already has a set of like personas that they want to test again. So I want to bring those, be abl…”