chewwt/dm_qwen4b_noise_emulator

🤗 Hugging Face sourcemit1 MBsafetensors✓ 1 checksumupdated today
Submit in one command

Run it next to your model folder. It makes the torrent, checks your files against Hugging Face, and submits it. You just start seeding and paste your key from your account. It only reads your files and never changes them. Read the script first if you like.

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo chewwt/dm_qwen4b_noise_emulator ./model-folder
Needs a seeder →

dm_qwen4b_noise_emulator

Laplacian kernel regression noise model for std_math prediction in data mixture optimization.

Predicts per-configuration std_math (across seeds) given data mixture proportions, used as a heteroscedastic noise model in Bayesian optimization.

Architecture

  • Kernel: Laplacian K(x, x') = exp(-γ · ||x - x'||₁), γ = 0.5
  • Support points: 100 training configs
  • Input features (6): [if_prop1, math_prop1, code_prop1, if_prop2, math_prop2, code_prop2] (values in [0, 1])
  • Output: predicted std_math (scalar)

Usage

import torch
from huggingface_hub import hf_hub_download
from safetensors.torch import load_file


class KernelRegressionModel(torch.nn.Module):
    def __init__(self, dual_coef, X_fit, gamma=0.1):
        super().__init__()
        self.gamma = gamma
        self.register_buffer("dual_coef", dual_coef)
        self.register_buffer("X_fit", X_fit)

    def forward(self, x):
        dist = torch.cdist(x, self.X_fit, p=1)
        K = torch.exp(-self.gamma * dist)
        return K @ self.dual_coef


path = hf_hub_download("chewwt/dm_qwen4b_noise_emulator", "noise_model.safetensors")
tensors = load_file(path)
model = KernelRegressionModel(tensors["dual_coef"], tensors["X_fit"], gamma=0.5)
model.eval()

# x: (batch, 6) float64 tensor, features in [0, 1]
x = torch.tensor([[0.1, 0.1, 0.1, 0.1, 0.1, 0.1]], dtype=torch.float64)
with torch.no_grad():
    sigma = model(x)   # predicted std_math