Hemmingway-1 — AutoRound W4G128

4-bit, group-size-128 AutoRound quantization of Altworld/Hemmingway-1, a 27B-parameter fine-tune of Qwen3.8-27B. This repository is an independent quantization, not the original model release. See the original model card for training, intended use, and benchmark information.

Quantization: AutoRound 0.15.1, bits=4, group_size=128, packing_format=auto_round:auto_awq, seqlen=1024, nsamples=64, iters=50. Certain layers remain in higher precision, as specified in quantization_config.json.

The included six-shard weight index and auxiliary model_extra_tensors.safetensors are needed together. Approximate download size: 18.65 GB (17.37 GiB). This is a quantization of an existing fine-tuned model, not a new fine-tuning run. Inference compatibility depends on a Transformers/AutoRound stack supporting this model architecture and quantization format; no inference benchmark or quality claim is made here.

Downloads last month
186
Safetensors
Model size
6B params
Tensor type
I32
·
BF16
·
F16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for groxaxo/Hemmingway-1-AutoRound-W4G128

Base model

Qwen/Qwen3.8-27B
Quantized
(42)
this model