- Q4F : Q4_K feed-forawrd (Q5_1 for ffn_down due to shape constraints)
- Q8A : Q8_0 attention, Q8_0 output, Q8_0 embeds
- Q8SH : Q8_0 shared experts
Readable speeds on a 24GiB GPU + 64GB RAM
Run it next to your model folder. It makes the torrent, checks your files against Hugging Face, and submits it. You just start seeding and paste your key from your account. It only reads your files and never changes them. Read the script first if you like.
curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo Beinsezii/GLM-4.5-Air-Q4F-Q8A-Q8SH-GGUF ./model-folderReadable speeds on a 24GiB GPU + 64GB RAM