People, every show

Eugene Cheah

CEO & Co-founder, Featherless.ai. On 2 shows, 6 appearances. The Shows tab opens the full record on each.

founderexecutivescientistengineer@picocreator ↗LinkedIn ↗featherless.ai ↗

Eugene Cheah co-founded and leads Featherless.ai, a serverless inference platform designed to run and serve tens of thousands of open-weight models at scale. He is best known for his research and core contributions to RWKV, an attention-free linear RNN architecture developed under the Linux Foundation.

2shows
6appearances
52statements
17resolved
13supported
3contradicted
76%fully supported
9said about them ↓

Everything Eugene Cheah said on any show that made the record, most notable first. Each card names its show and opens the statement there.

Cheah: Open Models Now Match Claude Sonnet and GPT-4o Mini
“So, and the, this growing collection of open models includes some of the best models that are already on par or surpass, let's say, Plot Sonnet or even GPT-A for Mini.”
Eugene Cheah Jul 1, 2026 ▶ 5:19 Featherless: $3.6M Revenue Running 6,700 Open Source AI Models — Eugene Cheah
Cheah: Hybrid SSM-transformer models outperform pure baselines of both
“None of us understand why a hybrid with a state-based model, the RWA state space, and transformer performs better than the baseline of both. It's like when you train one, you expect, and then you replace, you expect the same results. That's our pitch. That's o…”
Eugene Cheah Dec 24, 2024 ▶ 28:15 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
Cheah: Non-positional attention architectures remain stable beyond trained context
“One key advantage of this alternate attention mechanic that is not based on token position is that the model don't suddenly become crazy when you go past the eight K training context or a million context. It is actually still stable. It's still, it's able to r…”
Eugene Cheah Dec 24, 2024 ▶ 41:28 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
LATENT SPACE Prediction Didn’t hold up
Cheah: Standard Transformers Will Never Scale to Ten Million Tokens
“I think what was quick, I think it was rather quick after I concluded that transformer as it is will not scale to ten million tokens.”
Eugene Cheah Aug 31, 2023 ▶ 20:25 RWKV: Reinventing RNNs for the Transformer Era
Foreign Language Data Degrades English Benchmark Scores on Small LLMs
“Adding in a foreign data set is actually a loss, because once you're below a certain param count, so we're talking about the seven important, right? The more you add that's more in line with your evals, the more it will degrade, and they just exclude it.”
Eugene Cheah Aug 31, 2023 ▶ 25:13 RWKV: Reinventing RNNs for the Transformer Era
LATENT SPACE Assertion Supported
RWKV Matches GPT-NeoX Performance at Equal Parameter and Data Scales
“RWKV is a modern recursive neural network with transformer-like level of LMM performance, which can be trained in a transformer mode. And this part has already been benchmarked against GPT-NeoX in the paper, And it has similar training performance compared to …”
Eugene Cheah Aug 31, 2023 ▶ 31:59 RWKV: Reinventing RNNs for the Transformer Era
LATENT SPACE Assertion Supported
RWKV Architecture Is Proven to Scale to Any Parameter Size
“What we have already proven is that it can be scaled and trained by a transformer. How I do so, we'll cover later. And this can be scaled to as many parameters as we want.”
Eugene Cheah Aug 31, 2023 ▶ 37:52 RWKV: Reinventing RNNs for the Transformer Era
Cheah: RWKV Achieves Linear Scaling With No Trade-Offs in Reasoning
“So, so this is like literally us saying, there's no trade-offs. Yeah, you don't lose out in that process.”
Eugene Cheah Aug 31, 2023 ▶ 1:07:18 RWKV: Reinventing RNNs for the Transformer Era
TOP FOUNDERS Assertion Not checkable as stated
Cheah: RWKV Architecture Could Reduce Inference Costs by 1,000x
“Like this new AI architecture has the potential of reducing inference costs by over a thousand X.”
Eugene Cheah Jul 1, 2026 ▶ 5:58 Featherless: $3.6M Revenue Running 6,700 Open Source AI Models — Eugene Cheah
TOP FOUNDERS Prediction Not checkable as stated
Cheah: Global AI Market Will Segment into Domestic Sovereign Models
“So they are going to the direction of highly tailored sovereign AI models for the domestic market. And we actually see this happening more and more. So for Cohear, they will service the Canadian market. For the US market is going to be served by OpenAI Entropi…”
Eugene Cheah Jul 1, 2026 ▶ 15:08 Featherless: $3.6M Revenue Running 6,700 Open Source AI Models — Eugene Cheah
LATENT SPACE Disclosure
Cheah: RWKV organization has less compute than a single Google researcher
“So our entire organization has less compute than a single researcher in Google.”
Eugene Cheah Dec 24, 2024 ▶ 24:32 2024 in Post-Transformer Architectures: State Space Models, RWKV [Latent Space LIVE! @ NeurIPS 2024]
LATENT SPACE Assertion Contradicted
Cheah: Llama 3.1 405B is first frontier model using pipeline parallelism
“This is the first major model that of this cell class size, right? They're saying, hey, we are doing pipeline parallelism.”
Eugene Cheah Jul 29, 2024 ▶ 19:21 [LLM Paper Club] Llama 3.1 Paper: The Llama Family of Models
LATENT SPACE Disclosure
RWKV Raven Dataset Scrubs Out 'As an AI' Refusal Boilerplate
“Typically GPT for all, but then we scrub it for and remove all the, as a large model.”
Eugene Cheah Aug 31, 2023 ▶ 39:46 RWKV: Reinventing RNNs for the Transformer Era
LATENT SPACE Disclosure
RWKV Prioritizes User Feedback Over Benchmark Evals for Dataset Additions
“The reason why we add things to the data set was never about improving evals. It's about directly in response to user feedback.”
Eugene Cheah Aug 31, 2023 ▶ 43:11 RWKV: Reinventing RNNs for the Transformer Era
LATENT SPACE Assertion Supported
RWKV Uses Trie Tokenizer Without Space Delimiters for CJK Languages
“Instead of using like this token pairs well with this and should be paired with that we just made it a trial list. So So basically, try the data structure. Yeah. So we just find the longest matching string in that matching string that we have trained inside ou…”
Eugene Cheah Aug 31, 2023 ▶ 48:13 RWKV: Reinventing RNNs for the Transformer Era
LATENT SPACE Assertion Supported
RWKV Trains in Parallel Across GPUs, Unlike Traditional RNNs
“And in practice, once you start cascading there, you just saturate the GPU, and that's how it starts being paralysable trained. You no longer need to train in slices like traditional RNNs.”
Eugene Cheah Aug 31, 2023 ▶ 58:45 RWKV: Reinventing RNNs for the Transformer Era
The Token Shortage Crisis Only Applies to AGI, Not Small Models
“I would say if we are aiming for AGI, there is a token crisis, but if we are aiming for useful small models, I don't think there is a token crisis.”
Eugene Cheah Aug 31, 2023 ▶ 1:27:16 RWKV: Reinventing RNNs for the Transformer Era
Cheah: AI Engineers Do Not Need ML Math to Build Products
“Frankly, for an AI engineer, you don't need it. You, your main thing that you needed to do was to, frankly, just play around with ChatGPT, or all the alternatives, be aware of the alternatives, because be very mercenary, swap out to Cloudia if it's better for …”
Eugene Cheah Aug 31, 2023 ▶ 1:33:57 RWKV: Reinventing RNNs for the Transformer Era
Cheah: Pre-Transformer Academic Neural Network Research Is No Longer Relevant
“Frankly, almost everything that is, that matters, Ah, was basically in the past four years. Like, there were a lot of things that fit in academics that were before that, and you know, and they were mostly dealing with models that were under a billion parameter…”
Eugene Cheah Aug 31, 2023 ▶ 1:37:51 RWKV: Reinventing RNNs for the Transformer Era
Cheah: A Human Personality and Memories Can Fit on Two SSDs
“No offense to myself, I don't think my personality and my memories is more than this. We could, even if I can exit, I could store this in two SSDs. Two hard drives.”
Eugene Cheah Aug 31, 2023 ▶ 1:54:37 RWKV: Reinventing RNNs for the Transformer Era
TOP FOUNDERS Assertion Not checkable as stated
Cheah: Long-Tail Fine-Tuned Models Drive 50% of Featherless Workload
“You see, most providers, they only provide, let's say, less than a hundred models. That covers 50% of our inference work. It's the bottom 50% where they run all these interesting fine-tuned models that people came on board for.”
Eugene Cheah Jul 1, 2026 ▶ 6:34 Featherless: $3.6M Revenue Running 6,700 Open Source AI Models — Eugene Cheah
TOP FOUNDERS Disclosure
Featherless Targets Startups Burning $100K Monthly on OpenAI and Anthropic
“We also realized that there is a lot of money on the table right now where you can go after the startups that, hey, I just built my entire startup or SMB on OpenAI or Entropic, and I'm burning a 100,000 dollars a month. And I do not know what I was doing. And …”
Eugene Cheah Jul 1, 2026 ▶ 10:40 Featherless: $3.6M Revenue Running 6,700 Open Source AI Models — Eugene Cheah
TOP FOUNDERS Assertion Not checkable as stated
Featherless Won Multiple Contracts by Exclusively Hosting the StepFun Model
“The step one model is a particularly popular model for us that easily ship several contracts for us on this model alone. And no one else is.”
Eugene Cheah Jul 1, 2026 ▶ 13:05 Featherless: $3.6M Revenue Running 6,700 Open Source AI Models — Eugene Cheah
TOP FOUNDERS Assertion Not checkable as stated
Cheah: Featherless AI's Largest Customer Pays $1M to $2M Annually
“So currently the biggest will be around one to two million dollars a year, which may sound extremely large, but when you actually peel behind the layers, it only comes to around like five, six of the largest servers you see in the market.”
Eugene Cheah Jul 1, 2026 ▶ 16:12 Featherless: $3.6M Revenue Running 6,700 Open Source AI Models — Eugene Cheah

Show 24statements(28 left)

The other half of the tape: Eugene Cheah's own voice is left out of every number here. Other people bring the name up 9 times in 4 episodes across the shows. every mention, with the transcript →

Who brings them up most Sarah Chieng 5Jesse Hu 1Alessio Fanelli 1

Every mention by year

tap a year for its mentions
00428320242025episodesmentions
02320242025episodes it came up in
001.51.53320242025episodesmentions per episode

Latent Space 9

2025 1 mention in 1 episode
2024 8 mentions in 3 episodes 3 per episode

One line per show, most statements first. The link opens Eugene's full record on that show: the calibration, argument clarity, speaking style and every statement made there.

ShowRole thereEpsStatementsRecord
LATENT SPACELEDGER CEO & Co-founder, Featherless.ai 5 43 76% 13/17 full record on Latent Space →
TOP FOUNDERSLEDGER CEO & Co-founder, Featherless.ai 1 9 full record on Top Founders →
Made with StarZero

Turn any episode into a week of clips.

This entire site, thousands of episodes across every show transcribed, diarized, checked and made playable, runs on the StarZero media pipeline. Drop in your own episode and the podcast clipper finds the moments worth sharing, cuts them, captions them, and reframes them for every feed.