aufklarer/Qwen3-4B-Instruct-2507-MLX-4bit

🤗 Hugging Face 来源text-generationapache-2.04B 参数8.0 GBsafetensors✓ 2 个校验和今天更新
一条命令提交

在你的模型文件夹旁边运行它。它会制作种子、将文件与 Hugging Face 比对,然后提交。你只需开始做种,并粘贴你账户中的密钥。它只读取你的文件,绝不修改。如果愿意,可以先阅读脚本。

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo aufklarer/Qwen3-4B-Instruct-2507-MLX-4bit ./model-folder
需要做种者 →

Qwen3-4B-Instruct-2507 — MLX int4

First-party MLX export of Qwen/Qwen3-4B-Instruct-2507, quantized to int4 (group size 64) for on-device chat on Apple Silicon. Built by our own pipeline (speech-models/export_mlx.py, via mlx_lm.convert).

Runs in the runner voice companion through a hand-written MLX dense runtime (soniqo/speech-swift → Qwen3Chat/Qwen3DenseModel), not a generic loader — the forward pass is numerically parity-verified against mlx_lm (identical next-token logits).

Params 4B (dense) · 36 layers · 32 q / 8 kv heads · head_dim 128
Quantization int4, group size 64 (~4.5 bits/weight, 2.28 GB)
Context 262144

Attribution & license