Zehui Chen
Zehui Chen on Hugging Face Daily Papers: 18 papers, 7 in the top 3 of their day, 776 upvotes.
- UniCorn: Towards Self-Improving Unified Multimodal Models through Self-Generated Supervision 44 upvotes, #4 of 2026-01-07
- DualVLA: Building a Generalizable Embodied Agent via Partial Decoupling of Reasoning and Action 21 upvotes, #12 of 2025-12-01
- Critique-RL: Training Language Models for Critiquing through Two-Stage Reinforcement Learning 18 upvotes, #18 of 2025-10-29
- Agentic Jigsaw Interaction Learning for Enhancing Visual Perception and Reasoning in Vision-Language Models 9 upvotes, #24 of 2025-10-03
- AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning 55 upvotes, #3 of 2025-09-11
- UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning 112 upvotes, #2 of 2025-09-03
- FutureX: An Advanced Live Benchmark for LLM Agents in Future Prediction 61 upvotes, #3 of 2025-08-21
- CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios 10 upvotes, #16 of 2025-06-18
- VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning 10 upvotes, #29 of 2025-05-29
- VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning 43 upvotes, #4 of 2025-04-11
- ViDoRAG: Visual Document Retrieval-Augmented Generation via Dynamic Iterative Reasoning Agents 18 upvotes, #7 of 2025-03-03
- Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training 84 upvotes, #1 of 2025-01-22
- ToolHop: A Query-Driven Benchmark for Evaluating Large Language Models in Multi-Hop Tool Use 9 upvotes, #14 of 2025-01-07
- MMSearch: Benchmarking the Potential of Large Models as Multi-modal Search Engines 33 upvotes, #3 of 2024-09-20
- MindSearch: Mimicking Human Minds Elicits Deep AI Searcher 37 upvotes, #6 of 2024-07-30
- ShareGPT4Video: Improving Video Understanding and Generation with Better Captions 61 upvotes, #1 of 2024-06-07
- InternLM2 Technical Report 22 upvotes, #3 of 2024-03-27
- Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models 11 upvotes, #6 of 2024-03-20
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.