#
mlu
Here are 3 public repositories matching this topic...
Custom BANG C operators and a two-MLU tensor-parallel Qwen AWQ inference runtime.
-
Updated
Sep 11, 2026 - Python
Run MiniMax H3 video and audio generation on Cambricon MLU with BF16 CPU offload, Flash Attention, and profiling tools.
-
Updated
Aug 31, 2026 - Python
Add this topic to your repo
To associate your repository with the mlu topic, visit your repo's landing page and select "manage topics."