JusperLee/FlowSep-hive

🤗 Hugging Face 来源audio-to-audioapache-2.06.3 GBother✓ 2 个校验和今天更新
一条命令提交

在你的模型文件夹旁边运行它。它会制作种子、将文件与 Hugging Face 比对,然后提交。你只需开始做种,并粘贴你账户中的密钥。它只读取你的文件,绝不修改。如果愿意,可以先阅读脚本。

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo JusperLee/FlowSep-hive ./model-folder
需要做种者 →

FlowSep-hive

Model Description

FlowSep-hive is a data-efficient, query-based universal sound separation model trained on the Hive dataset. By leveraging the high-quality, semantically consistent Hive dataset, this model achieves competitive separation accuracy and perceptual quality comparable to state-of-the-art models (such as SAM-Audio) while utilizing only a fraction (~0.2%) of the training data volume.

This model is developed by Shanda AI Research Tokyo and is introduced in the paper: A Semantically Consistent Dataset for Data-Efficient Query-Based Universal Sound Separation.

Model Details

  • Model Type:​ Query-Based Universal Sound Separation
  • Language(s):​ English (for text queries)
  • License:​ Apache 2.0 (Please update if different)
  • Trained on:​ JusperLee/Hive (2,442 hours of raw audio, 19.6M mixtures)
  • Paper:​ arXiv:2601.22599
  • Code Repository:​ GitHub - JusperLee/Hive

Uses

The model is intended for universal sound separation tasks, allowing users to extract specific sounds from complex audio mixtures using multimodal prompts (e.g., text descriptions or audio queries).