Nov 11, 2024 · 58m · latent-space
Agents @ Work: Dust.tt — with Stanislas Polu
gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions
Former OpenAI reasoning researcher and Dust co-founder Stanislas Polu shares inside perspectives on OpenAI's scaling culture and details the engineering and product architectures needed to deploy reliable enterprise AI agents.
How this conversation actually went
Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. The hosts hold 25.1% of the talking time here. How this is scored →
speaking balance: gold is the hosts, purple is the guest (3 minute bins)
Stan directly dismisses Swyx's claim that open-sourcing invites commoditization and cloning, calling it a complete fantasy compared to execution velocity.
Hardest push from the hosts ▶ 57:17 Swyx pushes product-over-platform counterthesisSwyx directly challenges Dust's broad horizontal positioning by arguing founders should build specialized vertical products before ever trying to build platforms.
Biggest teaching moment ▶ 45:08 Explaining structural chunking vs generic ETLStan breaks down why standard connectors like Airbyte fail for LLMs due to structural loss in rich documents like Notion databases, demonstrating the need for bespoke infrastructure.
The host holds their own ▶ 42:39 Swyx breaks down Berkeley Function Calling leaderboardSwyx demonstrates deep domain expertise by citing the newly released Berkeley benchmark data, validating model tiers and specific performance metrics.
the scores for every segment, with the reasoning behind each
| Chapter | Topic | The hosts as informed peer | Guest teaching | Guest disagreement | The hosts pushing back | Why |
|---|---|---|---|---|---|---|
| Stanislas Polu's Background and Journey into Artificial Intelligence | 4 | 3 | 1 | 1 | Swyx sets the biographical context of Stripe and early OpenAI culture while Stan recounts his journey from Stanford and early robotics to theorem proving. The conversation is collaborative and exploratory without direct friction. | |
| Early OpenAI Days and Formal Mathematics Reasoning Research | 5 | 6 | 1 | 2 | Stan educates the hosts on how formal mathematics systems verify proofs instantaneous via type systems while search tactics require computation. Swyx engages with technical familiarity around complexity limits and competitive math benchmarks. | |
| Compute Governance, Scaling Thesis, and Ilya Sutskever's Vision | 5 | 6 | 1 | 1 | Stan explains the internal dynamics of compute allocation at OpenAI as a management tool and sheds light on Ilya Sutskever's scaling philosophy. The hosts probe into the historical timeline of the scaling laws and compute prioritization. | |
| Leadership at OpenAI and the Historical Anthropic Split | 4 | 4 | 1 | 2 | Swyx prompts Stan on the executive leadership dynamics and the split that formed Anthropic. Stan provides first-hand perspective while noting he was not in the immediate executive weeds, maintaining a measured tone. | |
| Founding Dust, Open-Source Strategy, and the XP1 Browser Extension | 6 | 5 | 4 | 5 | Swyx challenges Stan on the downsides of open-sourcing Dust, asserting that open source exposes them to cloning without business benefit. Stan pushes back firmly, calling the cloning fear a fantasy and defending transparent developer velocity. | |
| Dust's Enterprise Agent Thesis: API Integration vs Browser Automation | 6 | 6 | 5 | 3 | Stan firmly rejects the browser RPA thesis advocated by Adept and David Luan, arguing that internal enterprise automation belongs on native APIs. Swyx explicitly points out the direct contradiction with David Luan's philosophy. | |
| Function Calling Mechanics, Hierarchical Meta-Agents, and Product Usability | 5 | 5 | 2 | 2 | Stan explains the pragmatic value of deterministic single-step agents over brittle autonomous systems, elaborating on hierarchical meta-agents. The hosts contribute concepts around dependency graphs and agent orchestration protocols. | |
| Real-World Agent Evaluation, Feedback Loops, and Model Benchmarks | 7 | 5 | 2 | 3 | Swyx shows deep knowledge of the Berkeley function calling leaderboard, citing rankings of GPT-4 Turbo, 4o, and Salesforce xLAM. Stan explains why raw benchmark evaluations matter less than daily active enterprise adoption in practice. | |
| Engineering Dust's Infrastructure: Custom Connectors, Temporal, and Rust | 7 | 6 | 2 | 3 | Swyx leverages his insider experience with Airbyte and Temporal to drill into data ingestion and workflow engines. Stan explains why generic ETL connectors fail at semantic context chunking, justifying custom Rust connectors. | |
| Billion-Dollar Lean Companies and Horizontal AI Agent Strategy | 6 | 4 | 3 | 5 | Swyx challenges Stan's horizontal company-wide agent thesis by pitching his own 'product over platform' framework. Stan defends the horizontal model, arguing that enterprise workflows have long-tail variations requiring general tooling. |