| | I’ve been working on native ComfyUI support for Tencent’s HunyuanImage 3.0, the 80B mixture-of-experts image model (13B active per step). It isn’t a wrapper around Tencent’s pipeline: it uses the normal KSampler, the normal VAE Decode and ComfyUI’s own memory management, which streams the experts from system RAM so the model fits on one consumer GPU. What’s in the gallery (all Instruct-Distil, 8 steps):
These are picked from a bigger run: 60 prompts × 2 formats, 60 edits and 14 styles, one seed each, no rerolls. Most of the edits worked; a few didn’t (snow that barely shows, a logo it wouldn’t remove, a “make it night” that stayed day). The text-to-image prompts come from popular prompt posts on X. What you get
Speed (Instruct-Distil, about 1 megapixel):
It also runs with only 16 GB or 12 GB of VRAM (~30 s and ~32 s per image on a 4090 limited to that). What you need
Links
Happy to answer questions. If something breaks, open an issue on GitHub with the traceback. submitted by /u/LatentSpacer |
In this article, you will learn how synchronous and asynchronous execution patterns differ architecturally, and…
Training a single LLM agent jointly across diverse interactive environments has attracted increasing attention as…
Off-the-shelf AI assistants answer individual questions well, but they fall short on a different axis:…
Never pay full price. Bag yourself some Prime Day tech deals on our favorite WIRED-tested…
Science has increasingly used artificial intelligence (AI) as a kind of microscope—sorting data, analyzing images…
During VFX Week last week, we released seven open-weight capabilities for LTX-2.5, covering high-res editing,…