mondk/Msh-Tiny-14M-GGUF

🤗 Hugging Face 来源text-generationapache-2.050 MBGGUF✓ 3 个校验和今天更新
已有模型文件?提交模型种子

如果你有完整的模型文件并有权分享,请把示例文件夹路径替换为你的文件路径,再运行这条命令。它会校验文件、制作种子,并将磁力链接和校验和提交给 Pirate Face。请让种子客户端持续做种,方便其他人从节点下载。Pirate Face 不接收模型文件。你可以从账户页面获取社区密钥。也可以先阅读脚本。

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo mondk/Msh-Tiny-14M-GGUF ./model-folder
需要做种者 →

hi guys, im lazy to write, so this was written by claude, ty.

msh-tiny (GGUF)

A tiny (~14M parameter) GPT-2-style chat model, trained completely from scratch — no pretrained base model, custom BPE tokenizer trained from zero, custom PyTorch transformer architecture. This repo contains GGUF builds for use with llama.cpp, Ollama, and LM Studio.

The .safetensors source model is at mondk/Safetensors.msh-tiny.

Files

File Quant Size
model-f16.gguf F16 (full precision) 28.3 MB
model-q4km.gguf Q4_K_M 11.6 MB
model-q2k.gguf Q2_K 9.64 MB

Limitations

This model was trained from random initialization on a modest amount of data with limited compute — it is a small educational project, not a production-quality assistant. Expect it to follow the chat format reliably but produce limited/inconsistent knowledge and occasional incoherent answers.

Prompt format

<|user|>
{your message}
<|assistant|>

The model was trained to stop generating at <|end|>.

Usage

Ollama

FROM ./model-f16.gguf
ollama create msh-tiny -f Modelfile
ollama run msh-tiny

LM Studio: drop the .gguf file into your models folder and load it directly.

llama.cpp

./llama-cli -m model-f16.gguf -p "<|user|>\nhi\n<|assistant|>\n"

Training data

Combining 3 well-known open instruction/chat datasets plus a small hand-written set of everyday chit-chat (greetings, thanks, small talk):

  • tatsu-lab/alpaca
  • teknium/OpenHermes-2.5
  • HuggingFaceH4/no_robots