Assertion Supported AI assessment confidence: 95% certainty 4/5 debate potential 1/5

Bhatawdekar: Construction companies use AI to create RFP proposals from drawings

Ameya Bhatawdekar · The “Messy State” of AI & How to Fix It | Ameya, Braintrust CTO · May 15, 2026 · at 6:24

Ameya Bhatawdekar, Field CTO at Braintrust, discusses real-world enterprise AI agent use cases.

0:00 / 0:17exact quote · 17.6s
▶ Watch the full episode on YouTube → 720p mp4 · rendered on demand · StarZero watermark
“I've seen systems where construction companies are able to now put together effective proposals using complex engineering drawings, architectural plans, specifications to submit proposals for new RFPs.”

quote is from the automated transcript, cleaned for reading: filler sounds and stutters are removed, nothing is rephrased. names can be misheard (the analysis reads context, assessments check outside sources). how →

More from Ameya Bhatawdekar

Insight
Bhatawdekar: Gen AI Systems Require Observability Feedback Loops for Evals
“So when you're building Gen AI systems, you really want that feedback loop of observability that helps you build better evals, that helps you ship better AI.”
Ameya Bhatawdekar May 15, 2026 ▶ 4:49 The “Messy State” of AI & How to Fix It | Ameya, Braintrust CTO
Insight
Bhatawdekar: AI systems require full reasoning traces to evaluate response quality
“These AI systems need to log the entire trace of how the AI reasoned on the initial input. What were the tool calls it made? How did it interact with the LLMs? How did it sort of ultimately generate the response? And did that response actually meet the user in…”
Ameya Bhatawdekar May 15, 2026 ▶ 10:51 The “Messy State” of AI & How to Fix It | Ameya, Braintrust CTO
Insight
Bhatawdekar: Span-level scorers pinpoint errors in AI agent execution
“You can define those as deterministic functions, you know, implemented in code, or you can use LLM as judges, but then you can evaluate like, how did each span perform? And that can give you a fairly good way to zero in on problematic areas of your agents.”
Ameya Bhatawdekar May 15, 2026 ▶ 22:24 The “Messy State” of AI & How to Fix It | Ameya, Braintrust CTO
Insight
Bhatawdekar: Rigorous evals are existential for AI apps built with 'vibe coding'
“When you're building these intelligent agentic applications using Vibe Coding evals almost become existential. You know, that's the only way you have a high degree of confidence that what you've built is going to work well.”
Ameya Bhatawdekar May 15, 2026 ▶ 24:53 The “Messy State” of AI & How to Fix It | Ameya, Braintrust CTO
Insight
Ameya Bhatawdekar: Enterprise AI quality comes from surrounding engineering, not just models
“These intelligent systems, these AI systems are not just a model, right? There's a lot of layering that happens on top of these models. These systems have to really deliver specific capabilities or specific experiences that help people do certain specific task…”
Ameya Bhatawdekar May 15, 2026 ▶ 28:02 The “Messy State” of AI & How to Fix It | Ameya, Braintrust CTO
Insight
Bhatawdekar: AI models natively handle orchestration, replacing complex external engineering frameworks
“People built these very fancy frameworks and systems that were fairly complex and complicated to improve the orchestration capability of the model. And, you know, there were some very impressive engineering feats that happened as a result of that. But now the …”
Ameya Bhatawdekar May 15, 2026 ▶ 44:58 The “Messy State” of AI & How to Fix It | Ameya, Braintrust CTO
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 300 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.