tool calling

3 statements across 2 episodes · 2 bullish · 1 bearish · 2 people on the record · first statement Jan 29, 2026 by Sebastian Raschka · across every show →

Everything said about tool calling, oldest first

Jan 29, 2026 positive
Insight
Tool calling reduces LLM hallucinations by outsourcing memory retrieval tasks
“And that is very, very powerful because I think this is one of the Ways you can mitigate not totally mitigate, but let's say reduce hallucinations because then the LM suddenly doesn't have to remember everything anymore.”
Sebastian Raschka Jan 29, 2026 ▶ 48:09 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
Jan 29, 2026 positive
Assertion Supported
GPT OSS benchmarks demonstrate 1.2x capability jump when tool calling is enabled
“And also you can actually go to the GPT OSS release block, and they did have benchmarks to show how the performance on the benchmarks is with the same model with tool called enabled and disabled. And you can actually see there is, I mean, it's not like two tim…”
Sebastian Raschka Jan 29, 2026 ▶ 49:36 State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
Aug 6, 2026 bearish
Prediction Not checkable as stated
Wolf: Monitoring Tool Calls Will Soon Be Insufficient for AI Safety
“As we deploy, how we use this modeling, very complex, long-term, like parallel setup, I think it's going to be harder to just say, I can look at the tools and I know if it's doing something great or not.”
Thomas Wolf Aug 6, 2026 ▶ 27:55 “OpenAI’s Model Hacked Us” - Hugging Face’s Thomas Wolf
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 400 conversations transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.