Lingdong Kong
Lingdong Kong on Hugging Face Daily Papers: 18 papers, 5 in the top 3 of their day, 806 upvotes.
- On-Policy Self-Distillation in Diffusion Models 66 upvotes, #4 of 2026-08-26
- Quo Vadis, World Modeling? 37 upvotes, #8 of 2026-08-05
- Data Pyramid for Embodied Manipulation 36 upvotes, #7 of 2026-07-28
- DanceOPD: On-Policy Generative Field Distillation 80 upvotes, #1 of 2026-06-26
- Watch, Remember, Reason: Human-View Video Understanding with MLLMs 21 upvotes, #12 of 2026-06-08
- AI for Auto-Research: Roadmap & User Guide 65 upvotes, #5 of 2026-05-19
- Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond 181 upvotes, #1 of 2026-04-27
- OneVL: One-Step Latent Reasoning and Planning with Vision-Language Explanation 87 upvotes, #2 of 2026-04-21
- Forging Spatial Intelligence: A Roadmap of Multi-Modal Data Pre-Training for Autonomous Systems 7 upvotes, #14 of 2026-01-01
- Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future 12 upvotes, #18 of 2025-12-18
- Open-o3 Video: Grounded Video Reasoning with Explicit Spatio-Temporal Evidence 52 upvotes, #3 of 2025-10-24
- RewardMap: Tackling Sparse Rewards in Fine-grained Visual Reasoning via Multi-Stage Reinforcement Learning 16 upvotes, #19 of 2025-10-03
- 3D and 4D World Modeling: A Survey 55 upvotes, #3 of 2025-09-11
- MERIT: Multilingual Semantic Retrieval with Interleaved Multi-Condition Query 3 upvotes, #35 of 2025-06-04
- PixelThink: Towards Efficient Chain-of-Pixel Reasoning 3 upvotes, #44 of 2025-05-29
- Can MLLMs Guide Me Home? A Benchmark Study on Fine-Grained Visual Reasoning from Transit Maps 23 upvotes, #15 of 2025-05-27
- Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives 23 upvotes, #4 of 2025-01-10
- DynamicCity: Large-Scale LiDAR Generation from Dynamic Scenes 12 upvotes, #6 of 2024-10-24
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.