meshllm/DeepSeek-V4-Flash-0731-UD-Q4_K_XL-layers

🤗 Hugging Face sourcetext-generationmit155 GBGGUF✓ 47 checksumsupdated today
Submit in one command

Run it next to your model folder. It makes the torrent, checks your files against Hugging Face, and submits it. You just start seeding and paste your key from your account. It only reads your files and never changes them. Read the script first if you like.

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo meshllm/DeepSeek-V4-Flash-0731-UD-Q4_K_XL-layers ./model-folder
Needs a seeder →

DeepSeek-V4-Flash-0731-UD-Q4_K_XL

Distributed GGUF inference package for Mesh LLM

GGUF layer package for running DeepSeek-V4-Flash-0731-UD-Q4_K_XL across a local Mesh LLM cluster.

This package is derived from unsloth/DeepSeek-V4-Flash-0731-GGUF and keeps the original GGUF distribution split into per-layer artifacts for distributed inference.

Highlights

Run locally Pool multiple machines OpenAI-compatible Package variant
Private inference on your hardware Split layers across peers Serve /v1/chat/completions locally UD-Q4_K_XL layer package

Model Overview

Property Value
Source model unsloth/DeepSeek-V4-Flash-0731-GGUF
Model id unsloth/DeepSeek-V4-Flash-0731-GGUF:UD-Q4_K_XL
Family DeepSeek
Parameter scale not recorded
Quantization UD-Q4_K_XL
Layer count 43
Activation width not recorded
Package size 0 B
Source file UD-Q4_K_XL/DeepSeek-V4-Flash-0731-UD-Q4_K_XL-00001-of-00005.gguf
Package repo meshllm/DeepSeek-V4-Flash-0731-UD-Q4_K_XL-layers
License mit from unsloth/DeepSeek-V4-Flash-0731-GGUF

Recommended Use

  • Local and private inference with Mesh LLM.
  • Multi-machine serving when the full GGUF is too large for one host.
  • OpenAI-compatible chat/completions workflows through Mesh LLM's local API.

For upstream architecture details, chat template guidance, sampling recommendations, license terms, and benchmark notes, see the source model card: unsloth/DeepSeek-V4-Flash-0731-GGUF.

Quickstart

# Run this on each machine that should contribute memory/compute.
mesh-llm serve --model "meshllm/DeepSeek-V4-Flash-0731-UD-Q4_K_XL-layers" --split
# Check the mesh and discover the OpenAI-compatible model name.
curl -s http://localhost:3131/api/status
curl -s http://localhost:3131/v1/models
# Send an OpenAI-compatible chat request.
curl -s http://localhost:3131/v1/chat/completions \
  -H "Content-Type: application/json" \
  -d '{
    "model": "unsloth/DeepSeek-V4-Flash-0731-GGUF:UD-Q4_K_XL",
    "messages": [{"role": "user", "content": "Write a tiny hello-world function in Rust."}],
    "max_tokens": 128
  }'

Package Variant

Property Value
Format gguf
Canonical source ref unsloth/DeepSeek-V4-Flash-0731-GGUF@fbbb5b93fb787c21338159b0af3318bb3f4d9768/UD-Q4_K_XL/DeepSeek-V4-Flash-0731-UD-Q4_K_XL-00001-of-00005.gguf
Source revision fbbb5b93fb787c21338159b0af3318bb3f4d9768
Source SHA-256 d13ce8f90855547bdaebe7312f531a1f2c4f822178d3103951f27fe884395cfa
Skippy ABI not recorded
Package manifest SHA-256 63806aff514458320bbaf0a9d5eb4a431d26a506f857f6b709c32b5a697b81e7

What Is Included

Artifact Path Contents SHA-256
Manifest model-package.json Package schema, source identity, checksums 63806aff514458320bbaf0a9d5eb4a431d26a506f857f6b709c32b5a697b81e7

Validation

Generated by the Mesh LLM HF Jobs splitter from mesh-llm ref 3744f53d301866c7fe59b9638a7260476a2c77b6. Each artifact is checksummed as it is written, uploaded to this repository, and removed from the job workspace before the next artifact is produced.

skippy-model-package write-package "/hf-cache/UD-Q4_K_XL/DeepSeek-V4-Flash-0731-UD-Q4_K_XL-00001-of-00005.gguf" --out-dir "/tmp/meshllm-layer-job-meshllm_DeepSeek-V4-Flash-0731-UD-Q4_K_XL-layers-1/package"

Links