Hays: SpaceX is up 13% in five days, partly from SemiAnalysis report
“SpaceX is up 13% in the past five days, almost certainly in part due to semi-analysis.”
Baker: Dylan Patel's company spends 30% of labor cost on AI tokens
“Well our, your, our friend Dylan Patel at his company, so he's an ASI maxi, but he's at 30%.”
Patel: Nvidia B200 and B300 server pricing increased prior to production
“I think you've already seen server pricing go up like B 200, B 300 that you're buying now versus last year. The price is significantly higher. We've seen next generation hardware receive price increases before they've even started production. Like, here was th…”
Eliahu-Ontiveros: Gas turbine orders will peak this year
“We think gas turbine orders are going to pick this year.”
Burazin: CPUs Will Become the Next Critical Bottleneck for AI Agents
“You will get to the point, and Dylan Patel was at the conference talking about, from Semi-Analysis that talks usually about GPUs, was also talking about how CPUs will now be a bottleneck because it will be the constraint. You won't be able to grow, or we won't…”
Patel: SemiAnalysis spends $7M on Claude Code vs $25M payroll
“Across a firm, we're spending seven million dollars a year now on Claude code at the current rate versus our salary expense being in the neighborhood of twenty-five million dollars.”
Patel: SemiAnalysis AI spend could exceed total payroll by year-end
“And if this trajectory continues, then, you know, we'll spend more than a hundred percent by the end of the year”
Patel: SemiAnalysis built SEM chip material mapping app using Claude tokens
“One person on the team, they've been able to spend with a couple thousand dollars of Claude tokens, they've been able to create this application that is GPU accelerated, runs on a server that we have at CoreWeave, and anytime we send it an image, it's able to …”
Patel: AI can currently automate roughly 3% of BLS tasks
“The BLS has this entire Bureau of Labor Statistics has this entire, like, set of, like, 2000 tasks, And so he did that with AI, which ones can be done by AI, which ones cannot and grading them across a rubric, you know, about three percent are doable now with …”
Patel: Solo economist's AI project would have taken 200 economists a year
“And he's like, dude, this would have taken the team of 200 economists a year.”
Analyst mapped entire US power grid using $6,000 daily AI tokens
“He was spending like 6000 dollars a day. It was an insane amount, but he scraped every single power plant in the U S every single transmission line above a certain voltage. And created this entire mapping of the entire US grid, as well as a lot of demand sourc…”
Patel: Anthropic could double Opus pricing and users would still pay
“They could double their pricing on Opus, and I would continue to pay, and I bet most users would continue to pay. I bet that wouldn't solve their humongous capacity problem that they have.”
Patel: Next-Generation AI Racks Require 120 FPGAs per Rack
“There's a project we did on FPGAs, and it turns out there's a 120 FPGAs per, per next generation rack.”
Coogan: SemiAnalysis Is Launching an Investment Fund and Credit Ratings
“They're also launching a fund allegedly according to the information to invest in both semiconductor stocks and startups, and they're already doing angel investing. And then also they're going to do credit ratings.”
Parr: Niche B2B Research Playbooks in Tech Booms Mint Billionaires
“What you are explaining right now by what Dylan Patella does, it's this exact same B to B publication in a technology boom owning a very small niche and just kicking ass and doing the same playbook of research media events. And it has created many billionaires…”
Tiwari: Blackwell GPUs achieve 90 to 100x inference efficiency over Hopper
“And so for the hoppers, the H 100 or H 200 series of GPUs into the Blackwells there was a claim made that it could be 30 times more efficient. And I think the data from, you know, some analysis showed that it was 90 to a hundred times more efficient in terms o…”
O'Laughlin: Being 5% more accurate on EPS models never drives good investment decisions
“Is your model, you know, being five percent more accurate, really going to ever make a good investment decision or not? No, never, not once. Like no one's saying, oh yeah, my estimate is always one cent more tighter than everyone else. That's why I'm good at s…”
O'Laughlin: Traditional sell-side equity research is a broken business model on its last legs
“Sell side as a concept is very broken. If you're talking about waves and things that are changing sell side in a lot of ways is this hereditary child of like, let's say, 30 or 40 years of banking. Where you had you know, a company go public, so you needed some…”
O'Laughlin: Compiled a PhD-level chip cycle dataset in one day using AI
“I mean, this is like too much information to gather. It's like a lifetime of work. It's like a PhD project. I did it in a day.”
O'Laughlin: Memory supply will not catch up with demand for two years
“And then boom, you're just looking at the supply demand and you're like, yeah, this is not gonna catch up for like two years.”
SemiAnalysis reports Claude Code generates four percent of all public commits
“There were some other stuff from like semi-analysis that four percent of all public commits are made by quad code.”
Patel: SemiAnalysis advised a client on buying and restarting a coal plant
“Have clients would like, had a client buy a coal plant. And we were advising them on the transaction based on, they just like showed up and they're like, yeah, we want to buy power assets.”
Patel: Hedge funds make up 40% to 50% of SemiAnalysis's business
“Roughly like half our business is, or 40% of our business is like hedge funds”
Patel: SemiAnalysis's 2nd highest revenue product is impossible without AI
“My business model, like, this is the second highest revenue product for us, would not have been possible if it weren't for AI, like, vibe coding, like, you know, being able to dig through permits and regulatory filings, being able to run image recognition on s…”
Patel: Hyperscaler CapEx will reach $450B-$500B next year
“And my number is closer to, like, it's like four 5500 and that's based on, like, you know, all the research we do on, like, data centers and, like, tracking each individual data center in the supply chains, right?”
Coogan: SemiAnalysis clocks ChatGPT at 71% of AI query share
“So semi-analysis, clocks, chat GPTs, share of queries, it's 71%. Meta is in second at 12%, and that's with billions of users on meta products, and so OpenAI is certainly running away with consumer.”
Coogan: AMD's software is too broken to train AI models
“Even though they're hitting the benchmarks on the actual, like, computing power, like, price per flop, the software is so broken, there's so many bugs that you cannot train models on it.”
Patel: SemiAnalysis tracks all 1,500 global semiconductor fabs
“We track all 1500 fabs in the world. For your purposes, only 50 of them matter, but like, you know, all 1500 fabs around the world.”
Patel: Reasoning models will see humongous performance gains within a year
“And so this, the performance improvements we'll get out of these models is, is humongous, right? In, in the coming, you know, six months to a year in certain benchmarks where you have functional verifiers.”
Prakash: SemiAnalysis inference report erred by assuming unpriced input tokens
“I think there were some errors in that analysis. In particular we were trying to decode it, and one of the things we noticed is that it assumed that input tokens weren't being priced. So I think that may have been an error in the model.”
Patel: SemiAnalysis reported GPT-4 mixture of experts architecture in January
“Just being clear, I talked about mixture of experts in January, it's just people didn't really notice it.”
Patel: AI inference will deploy more GPUs than training by 2024
“LLM inference will be bigger than training, or multimodal, whatever, blah, blah, blah inference will be bigger than training, you know, probably next year, in fact at least in terms of GPUs deployed,”
Patel: AI chip bottlenecks lie in interconnects and memory, not matrix multiplication
“Multiplying tensors is kind of, you know, anyone can, there's a lot of people who've made good matrix multiply units, right? But it's about, like, getting good utilization out of those, and interfacing with the memory, and interfacing with other chips really e…”