Sparse attention for H3 minimax, enjoy up to 2.5x speed up.

Sparse attention for H3 minimax, enjoy up to 2.5x speed up.

Added to my node pack, sparse attention SLA node for H3 Minimax. speed increase of up to 2.5x.

enjoy.

you can use it with whatever turbo you like, doesn’t actually require the SLA lora.
if you oom, add comfykitch attention before it, they work together. you’ll get an additional 5-10% speedup.

https://github.com/PlagueKind/ComfyUI-PlagueKind-Nodes

credit to pl0x for designing it and allowing me to be the host.

EDIT: make sure you’re on a new pytorch version and CU130.
add the node after your lora loader for now. I haven’t tested other positioning.
additional note: Blackwell will see the biggest gain, but other cards still get a big boost.

If you’re doing lower res short videos, adjust min seq accordingly if see no speedup or messages about blocks not being sparse.

Confirmed: make sure this is the LAST thing in the chain, connected DIRECTLY to the guider and scheduler.

submitted by /u/Plague_Kind
[link] [comments]