ProCreations/grug-35b-qat-q4-gguf

🤗 Hugging Face 来源apache-2.022 GBGGUF✓ 2 个校验和今天更新
一条命令提交

在你的模型文件夹旁边运行它。它会制作种子、将文件与 Hugging Face 比对,然后提交。你只需开始做种,并粘贴你账户中的密钥。它只读取你的文件,绝不修改。如果愿意,可以先阅读脚本。

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo ProCreations/grug-35b-qat-q4-gguf ./model-folder
需要做种者 →

grug-35b-qat-q4-gguf

35b MoE brother in QAT four-bit rock. brain feel rounding rock during training so Q4 squish hurt less.

recipe: expert-freeze QAT on grug-35b-v2 - ALL text linear (expert include) fake-quant int4 g32 in forward, gradient only flow to attention/DeltaNet/shared path (1.4B trainable; expert too heavy for one cave GPU). ~1.7M grug token, Adafactor lr 2e-6. release = 25% QAT + 75% original anchor (grug family standard). fresh bf16 export, ONE Q4_K_M squish.

rocks

file what
grug-35b-qat-Q4_K_M.gguf the QAT rock (~21 GB)
mmproj-grug-35b-v2-f16.gguf eye rock (vision)

if rock act broken

single-token spam = context-shift corruption, not rock. recent llama.cpp + -c 16384+ for agent frontends. see main gguf card for full troubleshoot.

how run

llama-server -m grug-35b-qat-Q4_K_M.gguf --mmproj mmproj-grug-35b-v2-f16.gguf \
  -c 16384 --temp 0.6 --top-p 0.95 --top-k 20

ordinary rocks: grug-35b-v2-gguf. 27b QAT brother: grug-27b-qat-q4-gguf. grug made by ProCreations.