Everything Ben Hillock said on any show that made the record, most notable first. Each card names its show and opens the statement there.
Hylak: Do not micromanage OpenAI o1's step-by-step reasoning when prompting
“What I found that worked really well was not telling the model how to think about it. If that makes any sense, like almost more like you're actually interfacing with a person, which is like scary in its own right. But no, if you're working with a like a cowork…”
Hylak: Users willing to wait five minutes for AI will wait an hour
“I think that it's like the number of tasks that you're willing to wait, you know, like 3:05 minutes for it. It's like probably pretty similar to the number of tasks you're willing to wait like an hour for, which is interesting.”
Hylak: Claude.com yields better coding results than Cursor, which loops frequently
“I actually generally get better results out of Claude. It's in, you know, just Claude.com versus Cursor a lot of times which is interesting, like, I love being able to apply code, et cetera, But as far as just like, a lot of times I find it getting stuck in so…”
Hylak: OpenAI o1 struggles to match personal tone and writing styles
“I think that I've had a very hard time getting it to actually write stuff. I know that I've heard of people using it for writing where it's like processing diffs, more like providing critiques or feedback, but at least for myself, I haven't found a good way to…”
Hylak: Hidden reasoning tokens create an information asymmetry between OpenAI and developers
“I think that what makes a one even trickier than other models is that there is actually an asymmetric miss to how well open AI understands the model and how well we, for example, the fact that like reasoning tokens are hidden, right? So there's all this stuff …”
Hylak: OpenAI o1 is OpenAI's most capable yet hardest model to use
“We're finding that like, oh, one is the most capable model. I think that opening eye has made. And it's also, I think the hardest to use.”
Hylak: AI development is returning to multi-step chaining of specialized models
“It feels like we're, like, going back Into a time where having these separate steps with almost like varying levels of intelligence becomes increasingly important, like this idea of chaining.”
Hylak: Non-engineers form fixed negative mental models after one failed AI test
“Most people try something once and if it doesn't work, they like They have this like very fixed mental model. They never tried again.”