Jul 6, 2018 · 1h 13m · y-combinator
Fermat's Library Cofounders João Batalha and Luís Batalha · Y Combinator
gold bands on the timeline = statements, start to end. Hover to read, click to jump. CC turns on captions
In this Y Combinator interview, Fermat's Library co-founders João and Luís Batalha discuss building open tools for scientific paper annotation, reforming academic peer review and open access, and expanding public engagement with technical research.
How this conversation actually went
Every chapter scored 0–10 on four independent dynamics. Hover any point for the reasoning behind the score. How this is scored →
speaking balance: gold is the partners, purple is the guest (3 minute bins)
Luís and João question whether an open platform like arXiv could ever survive as a commercial for-profit company with linear submission growth, gently countering standard venture assumptions.
Hardest push from the partners ▶ 1:04:55 Host pushes back on non-venture business viabilityThe host firmly rejects the notion that companies lacking exponential growth cannot raise capital or become sustainable businesses.
Biggest teaching moment ▶ 23:06 Simpson's paradox wage breakdownJoão systematically demonstrates how median wages dropped across every individual educational cohort while rising in aggregate, clearly educating the host on Simpson's paradox.
The partners hold their own ▶ 16:57 Host articulates duplicate failure cost in academic biologyThe host demonstrates deep analytical grasp by detailing his Cambridge friend's concurrent duplicate failures to substantiate systemic flaws in academic publishing incentives.
the scores for every segment, with the reasoning behind each
| Chapter | Topic | The partners as informed peer | Guest teaching | Guest disagreement | The partners pushing back | Why |
|---|---|---|---|---|---|---|
| Founders' Backgrounds and Initial Content Growth | 3 | 5 | 0 | 0 | The host asks open-ended questions about Fermat's Library's inception and early traction. The guests educate the host on open science, the Polymath project, and how Terence Tao solved the Erdos discrepancy via a blog comment. | |
| arXiv Discoverability and the Fermat Chrome Extension | 4 | 4 | 0 | 1 | The host draws on a Peter Attia podcast discussing data scrubbing on arXiv to ask about paper discoverability. The guests explain how their Chrome extension integrates annotations and references directly onto preprint drafts. | |
| Evaluating Paper Quality and Potential Rating Systems | 4 | 3 | 0 | 1 | The host presses on how quality can be evaluated without traditional peer review, proposing concepts like forking. The guests discuss user survey feedback and potential multi-factor rating mechanisms. | |
| Negative Results, P-Value Hacking, and Publication Incentives | 5 | 4 | 0 | 1 | The host demonstrates domain familiarity by highlighting academic disincentives around negative results, sharing an anecdote of a Cambridge PhD friend duplicating failed experiments. The guests explain p-value hacking and statistical significance verification. | |
| Twitter Questions: Unexpected Papers and Statistical Paradoxes | 3 | 6 | 0 | 0 | The host reads Twitter questions regarding unusual papers. João delivers an educational explanation of sleep mortality studies, the hot hand fallacy in basketball, and Simpson's paradox with wage distribution data. | |
| Historic Science Papers versus Modern Writing Length | 4 | 4 | 0 | 1 | The guests describe historic physics breakthroughs published in single-page formats. The host introduces a literary comparison to David Foster Wallace to probe why modern papers have lengthened. | |
| Twitter Questions: Elements of Impactful Science Writing | 3 | 5 | 0 | 0 | The guests detail how Freeman Dyson's synthesis paper on quantum electrodynamics had more impact than discovery papers due to clear writing. The host concurs that clarity of communication is the decisive factor. | |
| Metrics of Scientific Value and Mass-Appeal Content | 4 | 4 | 0 | 1 | The host challenges the guests on balancing high-impact accessible content like Charlie Munger transcripts against dense technical science. The guests agree that citation counts are an imperfect proxy for true value. | |
| Speed of Publishing versus Rigorous Peer Review | 3 | 4 | 0 | 0 | The host asks about the tension between fast open publishing and rigorous peer review in fields like machine learning. João reflects on the extensive research required to annotate complex papers. | |
| Annotating Books, Educational Textbooks, and Retaining Knowledge | 5 | 4 | 0 | 1 | The host shares personal techniques from his English major background, detailing an index card note-taking system and critique of Kindle highlights. The guests share classical mechanics textbook examples and the annotated Portuguese epic poem Os Lusiadas. | |
| Value of External Explainer Annotations and Managing Side Projects | 4 | 4 | 0 | 0 | João recounts consulting Vitalik Buterin on the Ethereum whitepaper to show why external annotators frequently write clearer explanations than authors. The host commends turning intellectual hobbies into structured side projects. | |
| Public Enthusiasm for Science and Long-Term Platform Growth | 4 | 5 | 0 | 0 | The guests discuss a 14-year-old Russian student producing a novel math proof inspired by Fermat annotations and emphasize the viral demand for bite-sized scientific knowledge on Twitter. The host adds the example of Dan Carlin's multi-hour history podcasts succeeding against industry assumptions. | |
| Business Sustainability, Open Access Movement, and Global Paywalls | 5 | 5 | 0 | 1 | The host asks about business sustainability and pushes back on the assumption that non-venture growth curves cannot achieve profitability. The guests contrast frictionless access on MIT Wi-Fi with global research paywall barriers. |