Qwen3 Embedding 0.6B Q8_0 GGUF
This repository contains the exact immutable local-model asset used by NovelAide.
Provenance
- Base model:
Qwen/Qwen3-Embedding-0.6B - Asset source:
Qwen/Qwen3-Embedding-0.6B-GGUF - Pinned source revision:
370f27d7550e0def9b39c1f16d3fbaa13aa67728 - Runtime:
llama-cpp - NovelAide asset ID:
qwen3-embedding-0.6b - NovelAide CDN prefix:
qwen3-embedding-0.6b-gguf
The base model is Qwen/Qwen3-Embedding-0.6B. NovelAide consumes the official
Qwen/Qwen3-Embedding-0.6B-GGUF Q8_0 artifact at revision
370f27d7550e0def9b39c1f16d3fbaa13aa67728.
Files
| File | Size | SHA-256 |
|---|---|---|
Qwen3-Embedding-0.6B-Q8_0.gguf |
609.54 MiB | 06507c7b42688469c4e7298b0a1e16deff06caf291cf0a5b278c308249c3e439 |
Usage
Use the GGUF files with a compatible llama.cpp runtime and the ONNX files with a compatible ONNX / Transformers.js runtime. NovelAide pins the exact files and checksums shown above; do not substitute similarly named quantizations.
License and attribution
The model asset follows the upstream apache-2.0 license.
Review the linked base model and asset source model cards for their complete
terms, limitations, and attribution requirements. NovelAide is not affiliated
with or endorsed by the upstream model authors.