Shengyuan Ding
Shengyuan Ding on Hugging Face Daily Papers: 13 papers, 4 in the top 3 of their day, 627 upvotes.
- Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games 46 upvotes, #2 of 2026-06-18
- WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation 45 upvotes, #9 of 2026-05-15
- Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale 125 upvotes, #1 of 2026-03-27
- Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning 10 upvotes, #25 of 2026-03-18
- Visual-ERM: Reward Modeling for Visual Equivalence 21 upvotes, #8 of 2026-03-16
- Trust Your Critic: Robust Reward Modeling and Reinforcement Learning for Faithful Image Editing and Generation 23 upvotes, #10 of 2026-03-13
- DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing 78 upvotes, #3 of 2026-02-13
- ARM-Thinker: Reinforcing Multimodal Generative Reward Models with Agentic Tool Use and Visual Reasoning 45 upvotes, #5 of 2025-12-05
- SPARK: Synergistic Policy And Reward Co-Evolving Framework 16 upvotes, #21 of 2025-09-29
- MM-IFEngine: Towards Multimodal Instruction Following 31 upvotes, #6 of 2025-04-11
- Creation-MMBench: Assessing Context-Aware Creative Intelligence in MLLM 41 upvotes, #4 of 2025-03-19
- OmniAlign-V: Towards Enhanced Alignment of MLLMs with Human Preference 67 upvotes, #1 of 2025-02-26
- InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model 39 upvotes, #6 of 2025-01-22
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.