Rongjiehuang

Rongjiehuang on Hugging Face Daily Papers: 4 papers, 1 in the top 3 of their day, 78 upvotes.

  1. WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling 42 upvotes, #3 of 2024-08-30
  2. UniAudio: An Audio Foundation Model Toward Universal Audio Generation 20 upvotes, #7 of 2023-10-06
  3. Make-An-Audio 2: Temporal-Enhanced Text-to-Audio Generation 4 upvotes, #9 of 2023-05-31
  4. CLAPSpeech: Learning Prosody from Text Context with Contrastive Language-Audio Pre-training 4 upvotes, #6 of 2023-05-19

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.