NVIDIA Sol-Attn for ComfyUI / Triton kernel on SM89 - SM121, with zero-copy MiniMax H3 nodes: memory-efficient attention, scheduled tau with graph preview, and feed-forward chunking. Measured ×ばつ vs SageAttention and −37% MLP peak VRAM on H3
-
Updated
Aug 13, 2026 - Python