Ian Fischer is the co-founder and CEO of Poetiq, discussing benchmark performance on ARC-AGI-2.
Assertion Not checkable as stated
Poetiq achieves faster, cheaper recursive self-improvement than existing methods
“The core insight that we had is that we could do recursive self-improvement far faster and cheaper than all of the other ways that people had been proposing to do this.”
Assertion Not checkable as stated
Poetiq's agentic harness outperforms new base models without code changes
“With poetic what we end up giving you is a you know, people are calling these things harnesses now, but you know, or agentic system or whatever you want to call it, that sits on top of one or more language models, and it just performs better than them. And whe…”
Disclosure
Poetiq's meta-system generates reasoning systems for problems GPT-5 cannot reliably solve
“And so the core technology that we've developed at Poetic is recursive self-improvement. So we have a recursively self-improving system, which we call the Poetic meta system. The output of that system is systems that solve hard problems where a hard problem is…”
Insight
AI is replacing human engineers for dataset understanding and failure-mode detection
“Historically in machine learning, you always, you know, it's like the rule was you have to know your data set really well. But now we're kind of outsourcing that to the AI itself, where the AI is the, it's the AI's job to understand the dataset and figure out …”
Assertion Supported
Poetiq scored 55% on Humanity's Last Exam, outperforming Claude Opus 4.6
“AI hasn't passed it yet, but we got to 55%, which is almost two percentage points higher than the previous state of the art. Which came out just last week from Anthropic with Claude Opus 4.6. They got 53.1%, and we got 55% on it.”
Assertion Not checkable as stated
Poetiq's Humanity's Last Exam optimization run cost less than $100,000
“We didn't publish any cost for this, but I can say that the optimization costs us less than a hundred K, yeah.”