Comfyanonymous, creator of ComfyUI, explains on Latent Space why he avoided Gradio when building ComfyUI, contrasting it with Automatic1111's architecture.
Opinion
AMD GPUs perform horribly on Windows for generative AI tasks
“AMD works horribly on Windows. Like, on Linux, it works fine. It's lower than the price equivalent NVIDIA GPU. But it works, like, you can use it, generate images, everything works. On Linux, on Windows, You might have a hard time”
Assertion Not checkable as stated
Automatic1111's inefficient SDXL implementation drove ComfyUI's viral user adoption
“The big, one point zero release happened, and wow, Confu UI was the only way a lot of people could actually run it on their computers, because it just, like, automatic was so, like, inefficient and bad that most people couldn't act, like, it just wouldn't work…”
Insight
Comfy: Training on all internet images yields ugly outputs
“You can create a very good model that doesn't generate nice images, because most images on the internet are ugly, so if you, if that's like, if you just, oh, I have the best model that can, like, it's super smart, I create it on all the, like, I turn it on jus…”
Opinion
Comfyanonymous considers Genmo's Mochi the best open video generation model
“There's, yeah, there's actually a few of them, but the one I've implemented in Comfy is Mochi, because that, that seems to be The best one so far.”
What-if
Comfy: Extending Automatic1111 Was Harder Than Building ComfyUI From Scratch
“It would have been harder to implement that in the auto interface than to create my own interface, so that's when I decided to create my own.”
Assertion Supported
MultiDiffusion researchers published area conditioning a month after ComfyUI
“And then a month later, there was a paper that came out called multi diffusion, which was the same thing, but yeah, that's .”