why aren't all 877 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Disclosure
Slack says Amp eliminated mandatory pre-merge code reviews since day one
“Mandatory code review before it gets to main. Yeah, it's dead. It's been dead ever since we started working on AMP.”
Disclosure
GLM-5.2 autonomously wrote and guided production GPU kernels for Baseten's inference engine
“Some of the GPU kernels that were on GLM-Five-two within our inference engine is written by GLM-Five-two. And the trace and the kernels were guided by GLM-Five-two as the driver.”
Disclosure
Chen: OpenAI's three-year goal is models conducting end-to-end research
“When we look at our kind of three-year roadmap, right the end goal that we want to reach is one where You know, the models are just doing end-to-end research, and I think a part of that problem is just being able to have the model come up with good taste.”
Disclosure
Ball: Amp core team has 8 people, ships 15 times daily without code reviews
“I think we're around eight people now on the AMP core team, and we still don't do formal code reviews. We still push to main. We still ship 15 times every day.”
Disclosure
McPartlon: Chai's research team largely lacks formal biology backgrounds
“Like the whole research team at CHI except for me and Kevin, really like we're the only people with quote bio background. Even still, like we're pretty far removed. So I think like we try to like look at every problem as a core ML problem.”
Disclosure
Biderman: Engram trains models to decide what to memorize vs keep in notes
“The way to work on it is to train models both, to train models to manage it themselves, and that's an active area for us. Have the model know, like, without any explicit supervision signal to determine this kind of stuff I can pull from my brain, and that kind…”
Disclosure
Chen: OpenAI favors unifying modalities in as few architectures as possible
“For a research lab, I think there are a lot of advantages for it to being under one. So you just have to maintain one infrastructure stack, for instance. I think the cost to, like, maintaining and scaling many infrastructure stacks at once I think that's somet…”
Disclosure
Nadella says Azure networking team requested AI tokens over headcount for operations
“So they built this agentic system. They even have a character for it. It's called Miles and it sort of does all this stuff. Right. They started sort of screaming for more tokens and so on, and so they were saying, look, we don't need headcount, we need tokens …”
Disclosure
Azhnyuk: The Fourth Law starting construction on two semiconductor plants
“And we're about to start construction of two semiconductor plants to make sensors for thermal cameras.”
Disclosure
Parakhin: Liquid is the only genuinely competitive non-transformer architecture Shopify found
“That's why we at Shopify, when we tried multiple, and we constantly try multiple models, multiple companies, we found that for small, particularly with low latency applications, when you have low latency and or if you need longer context lengths, Liquid was th…”
Disclosure
Lopopolo: OpenAI Frontier team operates with post-merge or zero human code review
“You know, we, we've moved beyond even the humans reviewing the code as well. Most of the human review is post merge at this point, but it's not even reviewed.”
Disclosure
Lopopolo: Symphony discards failed PRs entirely to regenerate from scratch
“In Symfony, there's this like rework state where once the PR is proposed and it's escalated to the human for review, it should be a cheap review, right? It is either mergeable or it is not. And if it's not, you move it to rework. The Elixir service will comple…”
Disclosure
Lample: Mistral plans a phased approach to full-duplex audio models
“Ultimately what we want to do is to be this, Full duplex model, but we are not going to start this, start there directly. I think it's some approach that people are doing, but. Just to confirm, full duplex means it can speak while I'm speaking, or? Okay. Yeah,…”
Disclosure
Rieseberg: Anthropic builds all prototype candidates quickly instead of writing memos
“We internally at Anthropica are now probably much closer to the point where, like, don't even write a memo. Just, like, build, like, let's build all the candidates very quickly. Let's just build all of them and then pick the best ones.”
Disclosure
Anthropic deeply worries about AI automating entry-level junior jobs
“At Anthropic as a group of people, we're deeply worried about the impact that the tools are going to have on the labor market, especially for like junior employees that, because I think it's only honest to say that when we talk about automating a lot away, a l…”
Disclosure
Pydantic built a VIP issue scorer after closing an OpenAI founder's ticket
“Basically this started off because one of the OpenAI co-founders created an issue on Pydantic. And we just closed it and said it was wrong. And so we have this that, like, injects itself and tries to summarize someone and it gives them a, like, brutal score of…”
Disclosure
Eskildsen: Offered to return capital if Turbopuffer lacked PMF by year-end
“I don't think I've said this publicly before, but I just called Locky and was like, well, Locky, like, if this doesn't have PMF by the end of the year, like, we'll just like return all the money to you. But it's just like, I don't really, Justine and I don't w…”
Disclosure
Cursor removed web file editing to force users to delegate to agents
“And we actually felt that, that in some ways by restricting and limiting what you could do there, people would naturally leave more to the agent. And fall into this new pattern of delegating, which we thought was really valuable. And so there's currently no wa…”
Disclosure
Becker: Stopped Investing in Personal Software Engineering Skills Due to AI
“Intentionally not investing in engineering skills, because the areas are getting so good, maybe that's the wrong decision.”
Disclosure
White: FutureHouse bypassed custom foundation models to focus on scientific agents
“For example, at future house, we took the opinion that scientific agents are the future, and that allows us to skip a lot of steps because a lot of other people were like, we need to build a foundation model for X. Yeah. And we just skipped all that.”
Disclosure
DeepMind Abandoned AlphaProof to Run Gemini End-to-End for IMO Math
“We wanted to try to, like, use, actually use Gemini as an end-to-end model. Basically, no, no second system with alpha proof. No second system. In, text out.”
Disclosure
Yi Tay Fixes ML Bugs Automatically Using AI Without Reading Error Traces
“I think AI coding has started to become the point where I run a job, I get a bug. I almost don't look at the bug. I paste it into, like, anti-gravity, and, like, I throw it, that will fix the bug for me. And then I relaunched the job. And, like, beyond, like, …”
Disclosure
Reggio: Tens of thousands of ROI-negative SMB customers almost threatened Brex
“It ended up being a huge burden for the business almost existential for us to have those tens of thousands of customers that all were ROI negative.”
Disclosure
Reggio: Brex engineering interview revamp makes agentic coding mandatory
“We adapted our interview loop to be more AI sort of agentic coding native. So instead of we had like a coding and a system design question that we basically have revamped into a project where we'll give you like a brief, Before you come on site and then like a…”
Disclosure
Brex forced all existing engineers to pass its new AI interview
“As soon as we had the interview ready to ship, we started, we said everybody in engineering, including all the managers are going to have to go through this interview. And so we re-interviewed everybody internally.”
Disclosure
Nina Lopatina: Contextual AI shifts from MCP to direct API calls
“And I think for us, for me personally, like in my dynamic agent configs, I'm moving more toward API calls. And something a little bit more once I kind of maybe been able to prototype with an MCP server and figure out how I'm going to use this I think then you …”
Disclosure
Fioca: Has not hand-written a single line of code in months
“I haven't written a single line of code by hand in months, because I know what I can trust it to do.”
Disclosure
Ubl: Vercel's Composite Models Are Faster Than Agentic Loops
“Basically what we do is we have this like composite model architecture. We run the frontier model and then we run the fine tune model after to fix its errors. That doesn't perform better than an agentic loop, but it's orders of magnitude faster, right?”
Disclosure
General Intuition is betting maximally on video-to-action transfer over physics simulation
“And so I think for us, it is more so about making a maximal bet on video transfer and interacting with things that are difficult to simulate. And the steerability is also really interesting with text than it is on betting against simulation or something like t…”
Disclosure
Sands: Stripe passed on Tecton, co-building Cronon with Airbnb for latency
“We evaluated Tekton. This was a couple of years ago now. At the time, we couldn't wrap our heads around using it on the charge path, just from like a latency and reliability perspective. Like we've got to be operating at six nines. We got to be like, Decisioni…”
Disclosure
Bakouch primarily uses closed AI models for deep research out of convenience
“Yeah, I think even I, to be honest, I'm mainly using a closed model because it's easier, basically. It's like the same interface when I'm like, for example, researching very deep on papers.”
Disclosure
Ries: Answer.AI will never need more than 12 employees in total
“I can remember where I was when Jeremy told me, by the way, he doesn't think the company will ever need more than 12 employees. And I was like, oh sure, you mean like this month, you know, but it's like, no, over its entire life.”
Disclosure
Ball: The Sourcegraph Amp team does not use formal evals
“I think we don't have any set evals. We don't. And this was controversial up until a week ago, I think, when I think Boris from
Or two weeks ago from Anthropix that they don't have evals for the coding agent too.
But we don't, and we haven't had them.”
Disclosure
OpenCode exactly replicated Claude Code by dumping its system prompts and schemas.
“We look at cloud code. We dump all their system prompts. We dump all their tool descriptions. We dump all their tool schemas. We re-implement the tools. And when you're using an anthropic model, we basically have the exact same implementation.”
Disclosure
Mohan: Codeium uses in-house models for autocomplete due to poor frontier FIM
“The things like autocomplete and supercomplete that run on every keystroke are entirely, like, our own models, and by the way, that is still because properties like FIM, fill in the middle capabilities are still quite bad with the current model.”
Disclosure
Windsurf aims to shift coding workflows from 80% agent to 99% agent
“And today, timelines are 80 to 90% agent, 10 to 20% human. But we're trying to build towards a future that gets the 99% agent and one percent human. We only want to ask the user for final approval.”
Disclosure
Modular completely eliminates and replaces NVIDIA's CUDA stack
“In the case of Modular, we go literally, like, we only work at that level, because we get rid of all of CUDA, right? And so we've replaced the entire stack, and so we only do that.”
Disclosure
Kilpatrick: Google's strategy is building Gemini as one single unified model
“Like we're here to make one model and like that model is Gemini.
And like, I think you do need to just trust this point, like to make the capabilities work in some cases, like you do need to have these forks that like go off and make that capability and harde…”
Disclosure
Anthropic avoided building an IDE to target future AI model scaling
“If we want some product that has like very broad a product market fit today, we would build, you know, a cursor or a Windsurf or something like this. Like these are awesome products that so many people use every day. I use them. That's not the product that we …”
Disclosure
Oleve repurposes LaunchDarkly feature flags to load balance LLM endpoints
“What we do is we use the stage roll, like we use LaunchDarky flags to route between different Azure, Azure OpenAI endpoints setting up the percentages ourselves. And so now like every new request is a random, like a new like user effectively. And so it comes i…”
Disclosure
Colvin: Logfire migrated from ClickHouse to Timescale before choosing DataFusion
“We don't use ClickHouse. We started building our database with ClickHouse, moved off ClickHouse onto Timescale, which is a Postgres extension to do analytical databases. And then moved off Timescale onto Data Fusion, and we're basically now building, it's Data…”
Disclosure
Bryk: Exa applies OpenAI's o1 variable compute paradigm to web search
“One way of thinking about what we built is like O-one for search because, Oh, one is all about like, you know, some questions require more compute than others, and we'll put as much compute into the question as we need to solve it. So similarly with our search…”
Disclosure
Teo: Singapore will ban synthetic depictions of candidates during elections
“It is, in a way, a very specific set of requirements that we are putting in place, which is that in an election setting, we should only be shown saying what we actually said, or doing what we actually did, and anything else would be an assault. On factual accu…”
Disclosure
Goyal: Under 5% of Braintrust production customers use open source
“Among customers running in production, it's less than five percent.”
Disclosure
Godement: OpenAI plans API knobs for developers to adjust safety thresholds
“And so I think the direction where we'll go here is that, basically, there will always be, like, you know, a set of behavior that will, you know, just, like, forbid, frankly, because they're illegal against our terms of services, but then there will be, like, …”
Disclosure
Cosine refuses to publish SWE-bench trajectories to prevent competitor model distillation
“For the moment, as a closed source company, like fighting for an edge, we've decided not to publish that information for that exact reason. I don't want someone basically taking my tragedies and then taking a model that's suing them in GA and just distilling i…”
Disclosure
Howard: Answer.ai operates with no managers and zero corporate hierarchy
“We don't have any managers. We don't have any hierarchy from that point of view. So, for example, I'm not a manager, which means I don't get to tell people what to do or how to do it or when to do it.”
Disclosure
Howard: Answer.ai aims to build thousands of products with 12 people
“We want to create thousands of Commercially successful products at Answer.ai. And we want to do that with like, 12 people.”