JusperLee/AudioSep-hive

🤗 Hugging Face sourceaudio-to-audioapache-2.03.6 GBother✓ 2 checksumsupdated today
Submit in one command

Run it next to your model folder. It makes the torrent, checks your files against Hugging Face, and submits it. You just start seeding and paste your key from your account. It only reads your files and never changes them. Read the script first if you like.

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo JusperLee/AudioSep-hive ./model-folder
Needs a seeder →

AudioSep-hive

Model Description

AudioSep-hive is a data-efficient, query-based universal sound separation model trained on the Hive dataset. By leveraging the high-quality, semantically consistent Hive dataset, this model achieves competitive separation accuracy and perceptual quality comparable to state-of-the-art models (such as SAM-Audio) while utilizing only a fraction (~0.2%) of the training data volume.

This model is developed by Shanda AI Research Tokyo and is introduced in the paper: A Semantically Consistent Dataset for Data-Efficient Query-Based Universal Sound Separation.

Model Details

  • Model Type:​ Query-Based Universal Sound Separation
  • Language(s):​ English (for text queries)
  • License:​ Apache 2.0 (Please update if different)
  • Trained on:​ JusperLee/Hive (2,442 hours of raw audio, 19.6M mixtures)
  • Paper:​ arXiv:2601.22599
  • Code Repository:​ GitHub - JusperLee/Hive

Uses

The model is intended for universal sound separation tasks, allowing users to extract specific sounds from complex audio mixtures using multimodal prompts (e.g., text descriptions or audio queries).