HYU-NLP-EVAL/qwen3-4b-rar-medicine-onlinerubrics-seed11-step-033

🤗 Hugging Face sourcetext-generationapache-2.04B params8.0 GBsafetensors✓ 4 checksumsupdated today
Submit in one command

Run it next to your model folder. It makes the torrent, checks your files against Hugging Face, and submits it. You just start seeding and paste your key from your account. It only reads your files and never changes them. Read the script first if you like.

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo HYU-NLP-EVAL/qwen3-4b-rar-medicine-onlinerubrics-seed11-step-033 ./model-folder
Needs a seeder →

OnlineRubrics RaR-Medicine: step 33, seed 11

Intermediate policy from dynamic OnlineRubrics-Every GRPO training, distinct from static-rubric GRPO. Base model: Qwen/Qwen3-4B-Instruct-2507; thinking disabled. This checkpoint is a historical policy state used by the Phase-1 audit. No downstream medical capability or safety claim is made. Research use only; not validated for clinical decision-making.

Root files are the veRL-exported Hugging Face inference model (BF16). original_checkpoint/ preserves the exact original FSDP parameter checkpoint and tokenizer/configuration files. Optimizer state, training data, responses, rubrics, infrastructure configuration, and credentials are not included. The original is retained because export precision/serialization differs.

Base model revision: cdbee75f17c01a7cc42f958dc650907174af0554 Original actor tree SHA256: 4997eb6f885259e9abd7b8926a4c87a93ce9f968b5b58e642e97de5633c5f9c2