xiaobin zhuang
xiaobin zhuang on Hugging Face Daily Papers: 5 papers, 2 in the top 3 of their day, 191 upvotes.
- Sounding that Object: Interactive Object-Aware Image to Audio Generation 1 upvotes, #48 of 2025-06-05
- MagiCodec: Simple Masked Gaussian-Injected Codec for High-Fidelity Reconstruction and Generation 2 upvotes, #51 of 2025-06-03
- AudioTrust: Benchmarking the Multifaceted Trustworthiness of Audio Large Language Models 17 upvotes, #15 of 2025-05-26
- Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model 119 upvotes, #1 of 2025-04-14
- Seed-TTS: A Family of High-Quality Versatile Speech Generation Models 24 upvotes, #1 of 2024-06-05
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.