Hemmingway-1 — AutoRound W4G128
4-bit, group-size-128 AutoRound quantization of Altworld/Hemmingway-1, a 27B-parameter fine-tune of Qwen3.8-27B. This repository is an independent quantization, not the original model release. See the original model card for training, intended use, and benchmark information.
Quantization: AutoRound 0.15.1, bits=4, group_size=128, packing_format=auto_round:auto_awq, seqlen=1024, nsamples=64, iters=50. Certain layers remain in higher precision, as specified in quantization_config.json.
The included six-shard weight index and auxiliary model_extra_tensors.safetensors are needed together. Approximate download size: 18.65 GB (17.37 GiB). This is a quantization of an existing fine-tuned model, not a new fine-tuning run. Inference compatibility depends on a Transformers/AutoRound stack supporting this model architecture and quantization format; no inference benchmark or quality claim is made here.