ggml-org/MiMo-V2.6-Flash-RL-GGUF

🤗 Hugging Face 来源image-text-to-textmit582 GBGGUF✓ 11 个校验和今天更新
一条命令提交

在你的模型文件夹旁边运行它。它会制作种子、将文件与 Hugging Face 比对,然后提交。你只需开始做种,并粘贴你账户中的密钥。它只读取你的文件,绝不修改。如果愿意,可以先阅读脚本。

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo ggml-org/MiMo-V2.6-Flash-RL-GGUF ./model-folder
需要做种者 →

MiMo-V2.6-Flash-RL

Run with https://llama.app

llama serve -hf ggml-org/MiMo-V2.6-Flash-RL-GGUF

Source models

Notes

  • The MXFP4 output keeps the routed experts at their native MXFP4 precision.
  • The Q2_K output keeps the expert down projections at MXFP4, and quantizes the gate/up projections to Q2_K.
  • Includes MTP sidecars (Q4_0 and Q8_0) for speculative decoding (--mtp).
  • Includes a DFlash drafter sidecar (BF16 and Q8_0) for speculative decoding, converted from the dflash/ subdirectory of the source repo.
  • Includes a Q8_0 mmproj for the vision and audio encoders.
  • Currently, the Q2 models do not use an imatrix calibration due to lack of one.

[!IMPORTANT] This model is automatically converted using https://github.com/ggml-org/convert