Sony

Sony on Hugging Face Daily Papers: 9 papers, 0 in the top 3 of their day, 0 paper of the day.

  1. Spectral Prior for Reducing Exposure Bias in Diffusion Models 6 upvotes, #15 of 2026-07-27
  2. Woosh: A Sound Effects Foundation Model 12 upvotes, #21 of 2026-04-03
  3. Echoes Over Time: Unlocking Length Generalization in Video-to-Audio Generation Models 1 upvotes, #25 of 2026-02-27
  4. SoundReactor: Frame-level Online Video-to-Audio Generation 2 upvotes, #28 of 2025-10-06
  5. VIRTUE: Visual-Interactive Text-Image Universal Embedder 6 upvotes, #31 of 2025-10-03
  6. Improving Inference-Time Optimisation for Vocal Effects Style Transfer with a Gaussian Prior 0 upvotes, #20 of 2025-05-19
  7. DiffVox: A Differentiable Model for Capturing and Analysing Professional Effects Distributions 2 upvotes, #21 of 2025-04-23
  8. Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis 16 upvotes, #5 of 2024-12-23
  9. SoundCTM: Uniting Score-based and Consistency Models for Text-to-Sound Generation 7 upvotes, #11 of 2024-05-30

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.