Mixtral, every mention
24 scenes across 8 shows · ← back to Mixtral
Latent Space 14
the a16z Podcast 5
the MAD Podcast 3
No Priors 2
All-In 1
BG2 Pod 1
the Official SaaStr Podcast 1
TBPN 1
every year every show
Latent Space 14
the a16z Podcast 5
the MAD Podcast 3
No Priors 2
BG2 Pod 1
the Official SaaStr Podcast 1
All-In 1
TBPN 1
Verbatim, from the transcripts: passages where Mixtral comes up on Latent Space, the a16z Podcast, the MAD Podcast, No Priors, BG2 Pod
Gavin Baker: SpaceX Might Be the Greatest Company of All Time
- ▶ 22:10 Gavin Baker I think where that ends up is, um, I saw today that the EU had restricted exports of mixed straws.
Inside xAI: Building Grok Imagine in 3 Months, Videogen vs World Models, and Video Agents— Ethan He
State of LLMs 2026: RLVR, GRPO, Inference Scaling — Sebastian Raschka
- ▶ 15:13 Sebastian Raschka And then, uh, Mixtrol had a big, um, MOE, I think it was 2024.
Why RL Won — Kyle Corbitt, OpenPipe (acq. CoreWeave)
- ▶ 8:01 Shawn Wang Mistral and mixed trial.
Building Jamba 3B: the tiny Hybrid Transformer State Space Reasoning Model - Barak Lenz, CTO of AI21
- ▶ 4:42 unnamed speaker Uh, the way, when I, when I, when you came up with it last year, I said that basically it dethroned mixed trial. 2 times in the scene
DeepSeek, Reasoning Models, and the Future of LLMs
- ▶ 9:07 Guido Appenzeller Compared to a mixed trial, which is eight, right?
DeepSeek V3, SGLang, and the state of Open Model Inference in 2025 (Quantization, MoEs, Pricing)
- ▶ 13:22 unnamed speaker Basically, like, like this time last year, Mick Strao was sort of kicking off a bit of an MOE trend with, uh, you know, eight by seven B, eight by 22 B.
AI Semiconductor Landscape feat. Dylan Patel | BG2 w/ Bill Gurley & Brad Gerstner · Bg2 Pod
- ▶ 55:25 Dylan Patel When Mistral released their Mixtral model, which was like very revolutionary sort of late last year, um, because it was such a level of performance that didn't exist in the open source, um, that it drove pricing down so fast, right?
The State of AI Startups in 2024 [LS Live @ NeurIPS]
- ▶ 3:33 Pranav Reddy Uh, the Mistral folks had just launched the Mixtral model right before the beginning of NeurIps.
Making AI Work: Fine-Tuning, Inference, Memory | Sharon Zhou, CEO, Lamini
- ▶ 27:33 Sharon Zhou I think there's been a trend of mixture of experts, GPT-IV, uh, Mixtral are all a mixture of experts.
The 10,000x Yolo Researcher Metagame — with Yi Tay of Reka
- ▶ 1:49:04 unnamed speaker Um, uh, so like, you know, it's, it, I don't know if you have any commentary on, like, uh, Mixtral, DeepSeq, Snowflake, Quen, uh, all these, um, proliferation of, uh, MOEs, MOE models that seem to all be sparse upcycle, because,
No Priors Ep. 63 | With Sarah Guo and Elad Gil
How Shopify Implements AI Across Sales and Product with the Head of AI at Shopify
- ▶ 15:57 Mike Tamir What if you want to do something like a seventy billion Lama or minstrel or mixed or any of these, those are going to be very hard to get on an ad hoc basis.
2024 will be the year of ENTERPRISE AI | Florian Douetteau, CEO of Dataiku
- ▶ 20:59 Florian Douetteau Essentially something cheaper, like, uh, nixtral self-hosted or nixtral self-hosted when moving to production.
E169: Elon sues OpenAI, Apple's decline, TikTok ban, Bitcoin $100K?, Science corner: Microplastics
- ▶ 46:58 Sandeep 'Sunny' Madra Right now, what we've done, given all the demand, is we've kind of limited it to Llama-II, Seventy-B, and Mixtral.
A Brief History of the Open Source AI Hacker - with Ben Firshman of Replicate
- ▶ 1:07:59 unnamed speaker Yeah, and then on the competitive piece, um, there was a price war on Mixtrall last year, last year, this last December. 3 times in the scene
No Priors Ep. 52 | With Pinecone CEO Edo Liberty
- ▶ 5:05 Edo Liberty We load that into Pinecone and saw what happens when you augment GPT, 3.5 and four and Lama and Mixtral and models from Cohere and Anthropic.
Building an open AI company - with Ce and Vipul of Together AI
- ▶ 42:54 Vipul Ved Prakash We know that this, you know, system B is better for, uh, Mixtrol, and system C is going to be better for Stripe Tine or Mamba.
The Four Wars of the AI Stack - Dec 2023 Recap
- ▶ 2:44 unnamed speaker I think over the last maybe like four or five months, everybody's so focused on, uh, fine tuning Lama two and like a DPO to improve these models, max trial and all these things.
- ▶ 53:25 unnamed speaker Who's like the, ah, and maybe like, the Mixtrall inference words are like another example.
The Origin and Future of RLHF: the secret ingredient for ChatGPT - with Nathan Lambert
- ▶ 1:30:27 Nathan Lambert We could swap between Llama-II and Mixtrawl and kind of see, like, does RLHF work
Safety in Numbers: Keeping AI Open
- ▶ 0:58 unnamed speaker That, naturally, they're calling Mixtral.
- ▶ 18:13 Arthur Mensch So mixed trial is as similar performance to GPT, 3.5.
- ▶ 20:57 Arthur Mensch I saw Mixtral on the stuffed parrot as well. 2 times in the scene