Ben Hillock

1 appearance on the record.

computed by AI from the episodes · how this works → · full disclaimer →

Ben discussed his experiences using the O1 artificial intelligence model on the podcast. He explored its capabilities, prompt engineering techniques, and practical challenges in AI coding applications.

8statements → 0claims → 0claims resolved → 3.38/5average certainty → 2.38/5average debate potential → ≈4.0/5argument clarity, estimated →

2 opinions · 6 insights · every statement was checked. none of them is a claim the record can settle: opinions, insights and disclosures never carry an assessment.

The record, in short

What the tape says about how Ben argues and how the claims held up. Everything they said, and everything said about them, is in the tabs below.

How they sound: not measured why? →

We measure speaking style by listening to the audio itself, and a fair number needs at least 2,000 words from one person on tape we have measured. There is too little of Ben Hillock on measured tape to publish a rate. This says nothing about how they speak.

Everything Ben Hillock said on Latent Space that made the record, most notable first. Filter by type, assessment or year in the ledger →

Insight
Hylak: Do not micromanage OpenAI o1's step-by-step reasoning when prompting
“What I found that worked really well was not telling the model how to think about it. If that makes any sense, like almost more like you're actually interfacing with a person, which is like scary in its own right. But no, if you're working with a like a cowork…”
Ben Hillock Jan 17, 2025 ▶ 8:33 OpenAI o1 isn’t a chat model (and that’s the point)
Insight
Hylak: Users willing to wait five minutes for AI will wait an hour
“I think that it's like the number of tasks that you're willing to wait, you know, like 3:05 minutes for it. It's like probably pretty similar to the number of tasks you're willing to wait like an hour for, which is interesting.”
Ben Hillock Jan 17, 2025 ▶ 16:40 OpenAI o1 isn’t a chat model (and that’s the point)
Opinion
Hylak: Claude.com yields better coding results than Cursor, which loops frequently
“I actually generally get better results out of Claude. It's in, you know, just Claude.com versus Cursor a lot of times which is interesting, like, I love being able to apply code, et cetera, But as far as just like, a lot of times I find it getting stuck in so…”
Ben Hillock Jan 17, 2025 ▶ 26:06 OpenAI o1 isn’t a chat model (and that’s the point)
Insight
Hylak: OpenAI o1 struggles to match personal tone and writing styles
“I think that I've had a very hard time getting it to actually write stuff. I know that I've heard of people using it for writing where it's like processing diffs, more like providing critiques or feedback, but at least for myself, I haven't found a good way to…”
Ben Hillock Jan 17, 2025 ▶ 10:24 OpenAI o1 isn’t a chat model (and that’s the point)
Insight
Hylak: Hidden reasoning tokens create an information asymmetry between OpenAI and developers
“I think that what makes a one even trickier than other models is that there is actually an asymmetric miss to how well open AI understands the model and how well we, for example, the fact that like reasoning tokens are hidden, right? So there's all this stuff …”
Ben Hillock Jan 17, 2025 ▶ 12:27 OpenAI o1 isn’t a chat model (and that’s the point)
Opinion
Hylak: OpenAI o1 is OpenAI's most capable yet hardest model to use
“We're finding that like, oh, one is the most capable model. I think that opening eye has made. And it's also, I think the hardest to use.”
Ben Hillock Jan 17, 2025 ▶ 15:00 OpenAI o1 isn’t a chat model (and that’s the point)
Insight
Hylak: AI development is returning to multi-step chaining of specialized models
“It feels like we're, like, going back Into a time where having these separate steps with almost like varying levels of intelligence becomes increasingly important, like this idea of chaining.”
Ben Hillock Jan 17, 2025 ▶ 20:32 OpenAI o1 isn’t a chat model (and that’s the point)
Insight
Hylak: Non-engineers form fixed negative mental models after one failed AI test
“Most people try something once and if it doesn't work, they like They have this like very fixed mental model. They never tried again.”
Ben Hillock Jan 17, 2025 ▶ 1:28 OpenAI o1 isn’t a chat model (and that’s the point)

Appearances (1)

EpisodeDateSpeaking time
OpenAI o1 isn’t a chat model (and that’s the point) Jan 17, 2025 13m
Made with StarZero

Turn any episode into a week of clips.

This entire site, over 200 episodes transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.