Qwen3 Embedding 4B Q4_K_M GGUF
This repository contains the exact immutable local-model asset used by NovelAide.
Provenance
- Base model:
Qwen/Qwen3-Embedding-4B - Asset source:
Qwen/Qwen3-Embedding-4B-GGUF - Pinned source revision:
f4602530db1d980e16da9d7d3a70294cf5c190be - Runtime:
llama-cpp - NovelAide asset ID:
qwen3-embedding-4b - NovelAide CDN prefix:
qwen3-embedding-4b-gguf
The base model is Qwen/Qwen3-Embedding-4B. NovelAide consumes the official
Qwen/Qwen3-Embedding-4B-GGUF Q4_K_M artifact at revision
f4602530db1d980e16da9d7d3a70294cf5c190be.
Files
| File | Size | SHA-256 |
|---|---|---|
Qwen3-Embedding-4B-Q4_K_M.gguf |
2381.04 MiB | 2b0cf8f17b4c723c27303015383c27ec4bf2d8314bb677d05e920dd70bb0f16b |
Usage
Use the GGUF files with a compatible llama.cpp runtime and the ONNX files with a compatible ONNX / Transformers.js runtime. NovelAide pins the exact files and checksums shown above; do not substitute similarly named quantizations.
License and attribution
The model asset follows the upstream apache-2.0 license.
Review the linked base model and asset source model cards for their complete
terms, limitations, and attribution requirements. NovelAide is not affiliated
with or endorsed by the upstream model authors.