ProCreations/grug-v1.1-qwen-3.8-27b-awq-int4

🤗 Hugging Face 来源image-text-to-textapache-2.05.8B 参数103 GBsafetensors✓ 7 个校验和今天更新
一条命令提交

在你的模型文件夹旁边运行它。它会制作种子、将文件与 Hugging Face 比对,然后提交。你只需开始做种,并粘贴你账户中的密钥。它只读取你的文件,绝不修改。如果愿意,可以先阅读脚本。

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo ProCreations/grug-v1.1-qwen-3.8-27b-awq-int4 ./model-folder
需要做种者 →

Grug v1.1 Qwen3.8 27B — AWQ INT4

vLLM-oriented quantization of ProCreations/grug-v1.1-qwen-3.8-27b, pinned to revision 3ab073b4bb06dc8a83e819a33485469f32b945ba. Activation-aware W4A16 asymmetric INT4 with group size 128, calibrated on 128 Ultrachat samples at 1024 tokens. The vision tower, embeddings, LM head, GatedDeltaNet a/b gates, and MTP namespace remain BF16.

The trained MTP head is not included in this repository. Use the corresponding -mtp-... repository for speculative decoding.

Serve

vllm serve ProCreations/grug-v1.1-qwen-3.8-27b-awq-int4 --max-model-len 32768 \
  --reasoning-parser qwen3 --enable-auto-tool-choice --tool-call-parser qwen3_coder

The exact build script and structural report are included under reproduce/ and quantization_report.json.