Skip to content

Navigation Menu

Sign in
Sign up
#

minimax-h3

Here are 206 public repositories matching this topic...

A native .NET LLM inference engine for GGUF models. TensorSharp provides a console application, a web-based chatbot interface, and Ollama/OpenAI-compatible HTTP APIs for programmatic access. It supports Windows/MacOS/iOS/Linux with full GPU capability

  • Updated Sep 6, 2026
  • C#

NVIDIA Sol-Attn for ComfyUI / Triton kernel on SM89 - SM121, with zero-copy MiniMax H3 nodes: memory-efficient attention, scheduled tau with graph preview, and feed-forward chunking. Measured ×ばつ vs SageAttention and −37% MLP peak VRAM on H3

  • Updated Aug 13, 2026
  • Python

Add this topic to your repo

To associate your repository with the minimax-h3 topic, visit your repo's landing page and select "manage topics."

Learn more

AltStyle によって変換されたページ (->オリジナル) /