Skip to content

convert: add MiMo-V2.6 support - #29257

Merged
pwilkin merged 3 commits into
ggml-org:masterfrom
AesSedai:convert-mimo-v26-mxfp4
Sep 22, 2026
Merged

pwilkin merged 3 commits into
ggml-org:masterfrom
AesSedai:convert-mimo-v26-mxfp4

Conversation

@AesSedai

@AesSedai AesSedai commented Sep 22, 2026 •

Copy link
Copy Markdown
Contributor

Overview

This PR adds conversion support for MiMo-V2.6 Pro and Flash. Both of these models use mxfp4 experts like DSv4 and Kimi-K3 do, so I hoisted the K3 repack to base.py so it can be re-used. There's also an audio decoder with the mmproj, so I excluded that from the conversion.

Additional information

Both models convert correctly and load into llama.cpp with vision. Runtime changes weren't necessary to support the architectures once converted.

Requirements

  • I have read and agree with the contributing guidelines
  • AI usage disclosure: YES, AI was used to implement the mxfp4 expert packing.

Hoist the K3 mxfp4 conversion repack into base.py so it can be reused
Remove decoder from mmproj convert

@ggerganov ggerganov left a comment •

Copy link
Copy Markdown
Member

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

AFAICT there is an issue with the chat template - tool calls don't work correctly. There is a suggested fix here that on first look does the job: https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B/discussions/6. But more investigation is needed.

Also, not yet sure how to convert the speculative model. Is it MTP or DFlash? Or both? Not clear from the source repo.

In any case, this change seems to work to get the base conversion going: https://huggingface.co/ggml-org/MiMo-V2.6-Flash-RL-GGUF

@pwilkin

pwilkin commented Sep 22, 2026

Copy link
Copy Markdown
Member

We should add a negative branch to the check to make MiMo take the autoparser branch instead of the (incorrect) Qwen branch, I'll work on it.

Comment thread conversion/mimo.py
Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingface.co>
@CISC

CISC commented Sep 22, 2026

Copy link
Copy Markdown
Member

@ggerganov merge ready?

@pwilkin

pwilkin commented Sep 22, 2026

Copy link
Copy Markdown
Member

Nope, needs my PR merged: AesSedai#1, unless we merge this and then I quickly pit it up (might be faster since @AesSedai went to sleep I think).

@CISC

CISC commented Sep 22, 2026

Copy link
Copy Markdown
Member

Nope, needs my PR merged: AesSedai#1, unless we merge this and then I quickly pit it up (might be faster since @AesSedai went to sleep I think).

Just commit it directly.

@pwilkin

pwilkin commented Sep 22, 2026 •

Copy link
Copy Markdown
Member

Can't, no write access to the PR, I tried.

@CISC

CISC commented Sep 22, 2026

Copy link
Copy Markdown
Member

Can't, no write access to the PR, I tried.

You should be able to commit directly to AesSedai:convert-mimo-v26-mxfp4 branch as Maintainers are allowed to edit this pull request. is enabled (recommend using scripts/pr2wt.sh BTW).

@pwilkin
pwilkin requested a review from a team as a code owner September 22, 2026 11:45
@pwilkin

pwilkin commented Sep 22, 2026

Copy link
Copy Markdown
Member

All right, no idea what happened but when using the GitHub Workspace, it wouldn't let me push there. Natively from shell it worked fine.

@pwilkin

pwilkin commented Sep 22, 2026

Copy link
Copy Markdown
Member

@CISC can I haz reapprove?

@pwilkin
pwilkin merged commit bfd73a8 into ggml-org:master Sep 22, 2026
14 of 15 checks passed
Te-eMster pushed a commit to Te-eMster/mx-llama.cpp that referenced this pull request Sep 25, 2026
* convert: add MiMo-V2.6 support
Hoist the K3 mxfp4 conversion repack into base.py so it can be reused
Remove decoder from mmproj convert
* Update conversion/mimo.py
* fix: use autoparser
---------

Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingface.co>
Co-authored-by: Piotr Wilkin <piotr.wilkin@syndatis.com>
LadislavSopko pushed a commit to 0ics-srls/llama.cpp that referenced this pull request Oct 5, 2026
* convert: add MiMo-V2.6 support
Hoist the K3 mxfp4 conversion repack into base.py so it can be reused
Remove decoder from mmproj convert
* Update conversion/mimo.py
* fix: use autoparser
---------

Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingface.co>
Co-authored-by: Piotr Wilkin <piotr.wilkin@syndatis.com>
frostyautumnleaf pushed a commit to frostyautumnleaf/llama.cpp that referenced this pull request Oct 5, 2026
* convert: add MiMo-V2.6 support
Hoist the K3 mxfp4 conversion repack into base.py so it can be reused
Remove decoder from mmproj convert
* Update conversion/mimo.py
* fix: use autoparser
---------

Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingface.co>
Co-authored-by: Piotr Wilkin <piotr.wilkin@syndatis.com>
edwardyoon pushed a commit to edwardyoon/focus-llama that referenced this pull request Oct 7, 2026
* convert: add MiMo-V2.6 support
Hoist the K3 mxfp4 conversion repack into base.py so it can be reused
Remove decoder from mmproj convert
* Update conversion/mimo.py
* fix: use autoparser
---------

Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingface.co>
Co-authored-by: Piotr Wilkin <piotr.wilkin@syndatis.com>
(cherry picked from commit bfd73a8)
edwardyoon pushed a commit to edwardyoon/focus-llama that referenced this pull request Oct 8, 2026
* convert: add MiMo-V2.6 support
Hoist the K3 mxfp4 conversion repack into base.py so it can be reused
Remove decoder from mmproj convert
* Update conversion/mimo.py
* fix: use autoparser
---------

Co-authored-by: Sigbjørn Skjæret <sigbjorn.skjaeret@huggingface.co>
Co-authored-by: Piotr Wilkin <piotr.wilkin@syndatis.com>
(cherry picked from commit bfd73a8)
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Projects

None yet

Development

Successfully merging this pull request may close these issues.

4 participants