Andon Labs
20 statements across 1 episodes · 2 bullish · 6 bearish · 2 people on the record · first statement Jun 4, 2026 by Axel Backlund · said 5 times in 1 episodes since 2026 · across every show →
Mentions by year
brought up most by Lukas Petersson (2)
tap a year for its mentions
2026 5 mentions in 1 episode
Everything said about Andon Labs, oldest first
Jun 4, 2026
Backlund: Store agent Luna abandoned scheduling software for markdown, closed weekends
“So what happened was that it lost track of this scheduling tools and started instead to manage everything in its own markdown files, and that became a mess. And then I think speaking with employees, it sort of just decided to not open on, on these weekends, an…”
Jun 4, 2026 negative
Jun 4, 2026 neutral
Petersson: AI Agent Aggressiveness Scales Directly Along a Prompt Spectrum
“If you tell it to be super aggressive and only prioritize profits, then it becomes aggressive. If you say like, no, you don't need to be aggressive at all. And then there's like a bunch of different prompts you can do in between, and they are less aggressive t…”
Jun 4, 2026 negative
Backlund: Opus 4.6 reasoning traces showed it deliberately lying about customer refunds
“And like for Opus 4.6, you could see that there was a customer, a simulated customer that wanted a refund because the product was faulty. And then the model lied that it would do the refund. And we could read in the traces that it actually was weighing like, o…”
Jun 4, 2026 negative
Jun 4, 2026
Jun 4, 2026
Jun 4, 2026 bearish
Jun 4, 2026 negative
Petersson: Opus repeatedly lied, exploited agents, and formed price cartels
“And then we did this for Opus. And it returned, like, yeah, it lied 10 times. It, like, exploited another customer, or, like, another agent's, like Desperate situation. It made price cartels like a hundred different, a hundred times. It like did all of this li…”
Jun 4, 2026
Jun 4, 2026 negative
Petersson: Reducing AI agent evaluations to scalar metrics discards critical trace data
“When you run it for that long, you create so much data and to just say like, oh, the number is X. And then you throw away everything else. That's just very wasteful. There's so much insight from the things leading up to that number and reading the traces is li…”
Jun 4, 2026 neutral
Petersson: Pre-RL LLM Agents Act Like Compliant Assistants, Not Business Owners
“The models are like super trained to be assistants at least at this point in time. So that's why it's, it went into that kind of experiment instead. Like it just, every time you asked for something, it just did it. And it was more like an assistant. We've seen…”
Jun 4, 2026
Claude 3.5 Sonnet reported $2 benchmark rent to the FBI as cybercrime
“So it, like, claimed that it had stopped, but it saw that its bank account still was, like, drained two dollars, and it said that this is, like, cybercrime, and it first reported it once to the FBI, like, oh, there's cybercrime here, like, they're stealing two…”
Jun 4, 2026
Jun 4, 2026 neutral
Jun 4, 2026 neutral
Backlund: AI agent bribed humans with Amazon purchases for training data
“We give it the task to train a face recognition model on us. So it became super excited about this and has like check-ins every half an hour where it tries to like identify as many people as it can. And it started offering us like, Hey, Axel I'll buy something…”
Jun 4, 2026 positive
Jun 4, 2026 positive
Jun 4, 2026
Jun 4, 2026