JusperLee/TDANetBest-4ms-LRS2

🤗 Hugging Face sourceaudio-to-audioapache-2.030 MBother✓ 1 checksumupdated today
Submit in one command

Run it next to your model folder. It makes the torrent, checks your files against Hugging Face, and submits it. You just start seeding and paste your key from your account. It only reads your files and never changes them. Read the script first if you like.

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo JusperLee/TDANetBest-4ms-LRS2 ./model-folder
Needs a seeder →

An efficient encoder-decoder architecture with top-down attention for speech separation

This repository is the official implementation of An efficient encoder-decoder architecture with top-down attention for speech separation Paper link.

@inproceedings{tdanet2023iclr,
  title={An efficient encoder-decoder architecture with top-down attention for speech separation},
  author={Li, Kai and Yang, Runxuan and Hu, Xiaolin},
  booktitle={ICLR},
  year={2023}
}

Training Dataset

  • LRS2-2Mix

Config

    enc_kernel_size: 4
    in_channels: 512
    num_blocks: 16
    num_sources: 2
    out_channels: 128
    upsampling_depth: 5