SII-Yibin Wang

SII-Yibin Wang on Hugging Face Daily Papers: 19 papers, 5 in the top 3 of their day, 809 upvotes.

  1. WorldReward: Reward Modeling for Camera-Conditioned World Models 26 upvotes, #17 of 2026-09-04
  2. Light-WAM: Efficient World Action Models with State-Fusion Action Decoding 10 upvotes, #25 of 2026-06-09
  3. LoMo: Local Modality Substitution for Deeper Vision-Language Fusion 23 upvotes, #17 of 2026-05-29
  4. Project Imaging-X: A Survey of 1000+ Open-Access Medical Imaging Datasets for Foundation Model Development 67 upvotes, #6 of 2026-04-01
  5. From Sparse to Dense: Multi-View GRPO for Flow Models via Augmented Condition Space 14 upvotes, #12 of 2026-03-16
  6. DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing 78 upvotes, #3 of 2026-02-13
  7. Unified Personalized Reward Model for Vision Generation 19 upvotes, #14 of 2026-02-04
  8. UniReason 1.0: A Unified Reasoning Framework for World Knowledge Aligned Image Generation and Editing 75 upvotes, #6 of 2026-02-03
  9. MeepleLM: A Virtual Playtester Simulating Diverse Subjective Experiences 11 upvotes, #12 of 2026-01-26
  10. EtCon: Edit-then-Consolidate for Reliable Knowledge Editing 7 upvotes, #13 of 2025-12-11
  11. UniREditBench: A Unified Reasoning-based Image Editing Benchmark 36 upvotes, #5 of 2025-11-04
  12. UniGenBench++: A Unified Semantic Evaluation Benchmark for Text-to-Image Generation 66 upvotes, #4 of 2025-10-22
  13. G^2RPO: Granular GRPO for Precise Reward in Flow Models 5 upvotes, #26 of 2025-10-09
  14. Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding 49 upvotes, #3 of 2025-10-09
  15. Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning 85 upvotes, #2 of 2025-08-29
  16. InMind: Evaluating LLMs in Capturing and Applying Individual Human Reasoning Styles 2 upvotes, #14 of 2025-08-25
  17. GeometryZero: Improving Geometry Solving for LLM with Group Contrastive Policy Optimization 3 upvotes, #37 of 2025-06-10
  18. Unified Multimodal Chain-of-Thought Reward Model through Reinforcement Fine-Tuning 87 upvotes, #2 of 2025-05-07
  19. Unified Reward Model for Multimodal Understanding and Generation 105 upvotes, #2 of 2025-03-10

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.