Shengyuan Ding

Shengyuan Ding on Hugging Face Daily Papers: 13 papers, 4 in the top 3 of their day, 627 upvotes.

  1. Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games 46 upvotes, #2 of 2026-06-18
  2. WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation 45 upvotes, #9 of 2026-05-15
  3. Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale 125 upvotes, #1 of 2026-03-27
  4. Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning 10 upvotes, #25 of 2026-03-18
  5. Visual-ERM: Reward Modeling for Visual Equivalence 21 upvotes, #8 of 2026-03-16
  6. Trust Your Critic: Robust Reward Modeling and Reinforcement Learning for Faithful Image Editing and Generation 23 upvotes, #10 of 2026-03-13
  7. DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing 78 upvotes, #3 of 2026-02-13
  8. ARM-Thinker: Reinforcing Multimodal Generative Reward Models with Agentic Tool Use and Visual Reasoning 45 upvotes, #5 of 2025-12-05
  9. SPARK: Synergistic Policy And Reward Co-Evolving Framework 16 upvotes, #21 of 2025-09-29
  10. MM-IFEngine: Towards Multimodal Instruction Following 31 upvotes, #6 of 2025-04-11
  11. Creation-MMBench: Assessing Context-Aware Creative Intelligence in MLLM 41 upvotes, #4 of 2025-03-19
  12. OmniAlign-V: Towards Enhanced Alignment of MLLMs with Human Preference 67 upvotes, #1 of 2025-02-26
  13. InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model 39 upvotes, #6 of 2025-01-22

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.