bombman/Qwen3.6-35B-A3B-4bit-Native

🤗 Hugging Face 来源text-generationapache-2.034.7B 参数激活 3B68 GBsafetensors✓ 3 个校验和今天更新
已有模型文件?提交模型种子

如果你有完整的模型文件并有权分享,请把示例文件夹路径替换为你的文件路径,再运行这条命令。它会校验文件、制作种子,并将磁力链接和校验和提交给 Pirate Face。请让种子客户端持续做种,方便其他人从节点下载。Pirate Face 不接收模型文件。你可以从账户页面获取社区密钥。也可以先阅读脚本。

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo bombman/Qwen3.6-35B-A3B-4bit-Native ./model-folder
需要做种者 →

Qwen3.6-35B-A3B-4bit-Native This repository provides the 4-bit (NF4) quantized weights for the Qwen3.6-35B-A3B Mixture-of-Experts model. These weights were generated using the bitsandbytes library with double quantization enabled to ensure maximum precision at a reduced memory footprint.

Model Details

Base Model: Qwen3.6-35B-A3B Quantization: 4-bit NormalFloat (NF4) Framework: Hugging Face Transformers Total Parameters: ~35B Expert Architecture: 256 Experts per Layer

Key Features

Native Compatibility: Designed to work seamlessly with the transformers library without additional conversion layers. Memory Efficiency: Optimized to fit within ~20GB of memory (VRAM/RAM combined), making it accessible for mid-range hardware environments. Precision: Uses Double Quantization to minimize perplexity degradation compared to the original BF16 weights.