Weiming Ren

Weiming Ren on Hugging Face Daily Papers: 15 papers, 2 in the top 3 of their day, 564 upvotes.

  1. Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation 8 upvotes, #15 of 2026-04-28
  2. VecGlypher: Unified Vector Glyph Generation with Language Models 11 upvotes, #13 of 2026-02-26
  3. HiStream: Efficient High-Resolution Video Generation via Redundancy-Eliminated Streaming 20 upvotes, #8 of 2025-12-25
  4. OneStory: Coherent Multi-Shot Video Generation with Adaptive Memory 43 upvotes, #4 of 2025-12-10
  5. Scaling Zero-Shot Reference-to-Video Generation 28 upvotes, #6 of 2025-12-09
  6. TUNA: Taming Unified Visual Representations for Native Unified Multimodal Models 60 upvotes, #5 of 2025-12-02
  7. Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning 48 upvotes, #4 of 2025-05-23
  8. VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation 14 upvotes, #15 of 2025-05-21
  9. Vamba: Understanding Hour-Long Videos with Hybrid Mamba-Transformers 17 upvotes, #8 of 2025-03-17
  10. VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by Video Spatiotemporal Augmentation 24 upvotes, #5 of 2024-12-03
  11. OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision 42 upvotes, #2 of 2024-11-12
  12. MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark 35 upvotes, #1 of 2024-06-04
  13. AnyV2V: A Plug-and-Play Framework For Any Video-to-Video Editing Tasks 17 upvotes, #4 of 2024-03-22
  14. StructLM: Towards Building Generalist Models for Structured Knowledge Grounding 27 upvotes, #7 of 2024-02-27
  15. ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation 27 upvotes, #6 of 2024-02-08

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.