ProCreations/grug-27b-nvfp4

🤗 Hugging Face 来源apache-2.016.7B 参数20 GBsafetensors✓ 2 个校验和今天更新
一条命令提交

在你的模型文件夹旁边运行它。它会制作种子、将文件与 Hugging Face 比对,然后提交。你只需开始做种,并粘贴你账户中的密钥。它只读取你的文件,绝不修改。如果愿意,可以先阅读脚本。

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo ProCreations/grug-27b-nvfp4 ./model-folder
需要做种者 →

grug-27b-nvfp4

grug brain in NVIDIA four-bit float rock. Blackwell GPU eat this format raw = big speed, small memory, quality mostly keep.

  • source: grug-27b (v2.1)
  • scheme: NVFP4 (W4A4) via llm-compressor one-shot
  • calibration: 512 samples of grug's own training distribution (agent sessions + grug reasoning) so scales match real usage
  • kept high precision: lm_head, vision tower, router gates

how run

vllm serve ProCreations/grug-27b-nvfp4 --max-model-len 32768 \
  --reasoning-parser deepseek_r1

best on Blackwell (native FP4). works on Hopper too (weight-only benefit). grug think dense inside <think> (arrives in message.reasoning), answer normal english. sampling temp 0.6-1.0, top_p 0.95, top_k 20.

family: grug-27b | grug-27b-mtp | gguf | qat-q4

grug made by ProCreations.