zhang yuechen

zhang yuechen on Hugging Face Daily Papers: 14 papers, 4 in the top 3 of their day, 551 upvotes.

  1. RoPE-Aware Bit Allocation for KV-Cache Quantization 8 upvotes, #16 of 2026-06-25
  2. UnityShots: Memory-Driven Multi-Shot Audio-Video Generation with Boundary-Aware Gating 26 upvotes, #9 of 2026-06-25
  3. Utonia: Toward One Encoder for All Point Clouds 161 upvotes, #1 of 2026-03-04
  4. UnityVideo: Unified Multi-Modal Multi-Task Learning for Enhancing World-Aware Video Generation 16 upvotes, #16 of 2025-12-09
  5. UniVA: Universal Video Agent towards Open-Source Next-Generation Video Generalist 36 upvotes, #5 of 2025-11-14
  6. DreamOmni2: Multimodal Instruction-based Editing and Generation 72 upvotes, #3 of 2025-10-10
  7. Training-Free Efficient Video Generation via Dynamic Token Carving 21 upvotes, #16 of 2025-05-23
  8. Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey 28 upvotes, #6 of 2025-03-18
  9. Magic Mirror: ID-Preserved Video Generation in Video Diffusion Transformers 14 upvotes, #9 of 2025-01-08
  10. Lyra: An Efficient and Speech-Centric Framework for Omni-Cognition 43 upvotes, #4 of 2024-12-13
  11. ControlNeXt: Powerful and Efficient Control for Image and Video Generation 48 upvotes, #2 of 2024-08-13
  12. Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models 33 upvotes, #2 of 2024-03-28
  13. Make-Your-Video: Customized Video Generation Using Textual and Structural Guidance 6 upvotes, #7 of 2023-06-02
  14. Real-World Image Variation by Aligning Diffusion Inversion Chain 5 upvotes, #4 of 2023-05-31

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.