Tool Calling

topic on 5 shows · 10 statements across 9 episodes

Latent Space No Priors the MAD Podcast Big Technology 20VC

10 statements about Tool Calling, every show

20VC Insight
Weitzman: Sierra's core value is tool calling, unlike ElevenLabs
“If you use a tool like Sierra, the wedge right now is voice, but the important part is tool calling. 11 Labs lets you do some tool calling, but that's not the bread and butter.”
Cliff Weitzman Sep 5, 2026 ▶ 54:40 How to Build Your Own Data Center & Why Every Startup Should Do It
MAD Prediction Not checkable as stated
Wolf: Monitoring Tool Calls Will Soon Be Insufficient for AI Safety
“As we deploy, how we use this modeling, very complex, long-term, like parallel setup, I think it's going to be harder to just say, I can look at the tools and I know if it's doing something great or not.”
Thomas Wolf Aug 6, 2026 ▶ 27:55 “OpenAI’s Model Hacked Us” - Hugging Face’s Thomas Wolf
NO PRIORS Prediction Not checkable as stated
Herzig: Most enterprise agents will rely on tool calling, not GUI
“I still believe for the most part it will, The majority will live with tool calling, right? And agents running in the background and so on, right? Because you also don't, you know, maybe want to have the browser open all the time. Okay, we can do this with hea…”
Philipp Herzig Apr 23, 2026 ▶ 21:27 SAP: Bringing the ‘Operating System’ of a Company into the AI Era with CTO Philipp Herzig
MAD Insight
Tool calling reduces LLM hallucinations by outsourcing memory retrieval tasks
“And that is very, very powerful because I think this is one of the Ways you can mitigate not totally mitigate, but let's say reduce hallucinations because then the LM suddenly doesn't have to remember everything anymore.”
Sebastian Raschka Jan 29, 2026 ▶ 48:09 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
MAD Assertion Supported
GPT OSS benchmarks demonstrate 1.2x capability jump when tool calling is enabled
“And also you can actually go to the GPT OSS release block, and they did have benchmarks to show how the performance on the benchmarks is with the same model with tool called enabled and disabled. And you can actually see there is, I mean, it's not like two tim…”
Sebastian Raschka Jan 29, 2026 ▶ 49:36 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
LATENT SPACE Assertion Not checkable as stated
Wagner: Reliable LLM tool calling in late 2023 unlocked functional agents
“And I would say like that tool calling, I think really only became reliable about a year ago, like last October, November, when the models that shipped were suddenly really good at calling tools. And I think that then changed everything, right? And you see tha…”
Matthias Wagner Nov 22, 2025 ▶ 5:24 ⚡️ Building the AI Hardware Engineer with Matthias Wagner, Co-founder of Flux
Frontier LLMs remain unreliable at realistic multi-turn tool calling
“Our last leaderboard is saying that models are really great at tool calling. So it's like safe, but they actually not, right? They're making mistakes and this is going to recalibrate the expectation of the users that Be careful because they're still not perfec…”
Pratik Bhavsar Jul 14, 2025 ▶ 33:22 ⚡️Ranking Agentic LLMs — Pratik Bhavsar, Galileo
BIG TECHNOLOGY Prediction Not checkable as stated
Roy: Tool Calling Will Be the Next Model Improvement Battleground
“Tool calling is going to become, like, one of the big next battlegrounds in terms of model improvement”
Ranjan Roy Jul 7, 2025 ▶ 41:47 $100 Million AI Engineers, Vending Machine Claude, Legend Of Soham
BIG TECHNOLOGY Prediction Not checkable as stated
Roy: Tool calling will be the next decisive battleground in AI
“The next great battle in AI is tool calling. That's where we're going to see the maximum amount, like actually bringing these models and agent and agentic AI. That's all that matters is the ability for an action to understand its concept context and then take …”
Ranjan Roy Jun 14, 2025 ▶ 24:14 Sam Altman’s Gentle Singularity, Zuck’s AI Power Play, Burning Of The Waymos
LATENT SPACE Assertion Supported
Harrison Chase says OpenAI recommends adding a thought field to tool schemas
“I think open AI even recommended, like when you're doing tool calling, it's sometimes helpful to put like a thought field in the tool along with all the actual acquired arguments and then have that one first. So it fills out that first and then, and that's, th…”
Harrison Chase Sep 27, 2024 ▶ 12:10 Language Agents: From Reasoning to Acting — with Shunyu Yao of OpenAI, Harrison Chase of LangGraph

← every entity, every show

Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.