Lidong Bing

Lidong Bing on Hugging Face Daily Papers: 34 papers, 16 in the top 3 of their day, 1,732 upvotes.

  1. MOOSE-Star: Unlocking Tractable Training for Scientific Discovery by Breaking the Complexity Barrier 86 upvotes, #1 of 2026-03-06
  2. LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling 150 upvotes, #2 of 2025-12-02
  3. OpenMMReasoner: Pushing the Frontiers for Multimodal Reasoning with an Open and General Recipe 88 upvotes, #3 of 2025-11-24
  4. MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling 151 upvotes, #1 of 2025-11-18
  5. UniME-V2: MLLM-as-a-Judge for Universal Multimodal Embedding Learning 11 upvotes, #18 of 2025-10-16
  6. First Try Matters: Revisiting the Role of Reflection in Reasoning Models 24 upvotes, #16 of 2025-10-10
  7. Multi-Agent Tool-Integrated Policy Optimization 29 upvotes, #8 of 2025-10-09
  8. MMR1: Enhancing Multimodal Reasoning with Variance-Aware Sampling and Open Resources 98 upvotes, #2 of 2025-09-26
  9. MiroMind-M1: An Open-Source Advancement in Mathematical Reasoning via Context-Aware Multi-Stage Policy Optimization 116 upvotes, #2 of 2025-07-22
  10. Evolving Prompts In-Context: An Open-ended, Self-replicating Perspective 19 upvotes, #7 of 2025-07-01
  11. MOOSE-Chem2: Exploring LLM Limits in Fine-Grained Scientific Hypothesis Discovery via Hierarchical Search 24 upvotes, #14 of 2025-05-27
  12. MOOSE-Chem3: Toward Experiment-Guided Hypothesis Ranking via Simulated Experimental Feedback 30 upvotes, #10 of 2025-05-26
  13. 100 Days After DeepSeek-R1: A Survey on Replication Studies and More Directions for Reasoning Language Models 29 upvotes, #5 of 2025-05-01
  14. Analyzing LLMs' Knowledge Boundary Cognition Across Languages Through the Lens of Internal Representations 17 upvotes, #6 of 2025-04-21
  15. Babel: Open Multilingual Large Language Models Serving Over 90% of Global Speakers 58 upvotes, #1 of 2025-03-06
  16. FINEREASON: Evaluating and Improving LLMs' Deliberate Reasoning through Reflective Puzzle Solving 24 upvotes, #8 of 2025-02-28
  17. Finding the Sweet Spot: Preference Data Construction for Scaling Preference Optimization 6 upvotes, #15 of 2025-02-26
  18. LongPO: Long Context Self-Evolution of Large Language Models through Short-to-Long Preference Optimization 25 upvotes, #9 of 2025-02-20
  19. VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding 75 upvotes, #3 of 2025-01-23
  20. VideoRefer Suite: Advancing Spatial-Temporal Object Understanding with Video LLM 40 upvotes, #4 of 2025-01-03
  21. 2.5 Years in Class: A Multimodal Textbook for Vision-Language Pretraining 91 upvotes, #1 of 2025-01-03
  22. M-Longdoc: A Benchmark For Multimodal Super-Long Document Understanding And A Retrieval-Aware Tuning Framework 42 upvotes, #2 of 2024-11-12
  23. Breaking the Memory Barrier: Near Infinite Batch Size Scaling for Contrastive Loss 82 upvotes, #1 of 2024-10-25
  24. Stabilize the Latent Space for Image Autoregressive Modeling: A Unified Perspective 6 upvotes, #15 of 2024-10-17
  25. The Curse of Multi-Modalities: Evaluating Hallucinations of Large Multimodal Models across Language, Visual, and Audio 29 upvotes, #3 of 2024-10-17
  26. SeaLLMs 3: Open Foundation and Chat Multilingual Large Language Models for Southeast Asian Languages 52 upvotes, #3 of 2024-07-30
  27. Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models 8 upvotes, #10 of 2024-06-27
  28. VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs 28 upvotes, #8 of 2024-06-13
  29. LLM-R2: A Large Language Model Enhanced Rule-based Rewrite System for Boosting Query Efficiency 9 upvotes, #5 of 2024-04-22
  30. Contrastive Chain-of-Thought Prompting 35 upvotes, #3 of 2023-11-17
  31. CLEX: Continuous Length Extrapolation for Large Language Models 10 upvotes, #10 of 2023-10-26
  32. INSTRUCTEVAL: Towards Holistic Evaluation of Instruction-Tuned Large Language Models 5 upvotes, #8 of 2023-06-09
  33. Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding 20 upvotes, #2 of 2023-06-06
  34. Is GPT-4 a Good Data Analyst? 6 upvotes, #1 of 2023-05-25

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.