Ning Ding

Ning Ding on Hugging Face Daily Papers: 20 papers, 12 in the top 3 of their day, 1,478 upvotes.

  1. SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning 73 upvotes, #3 of 2025-09-12
  2. A Survey of Reinforcement Learning for Large Reasoning Models 156 upvotes, #1 of 2025-09-11
  3. Towards a Unified View of Large Language Model Post-Training 67 upvotes, #3 of 2025-09-05
  4. Intern-S1: A Scientific Multimodal Foundation Model 242 upvotes, #1 of 2025-08-22
  5. From AI for Science to Agentic Science: A Survey on Autonomous Scientific Discovery 31 upvotes, #7 of 2025-08-21
  6. SSRL: Self-Search Reinforcement Learning 88 upvotes, #2 of 2025-08-18
  7. RLPR: Extrapolating RLVR to General Domains without Verifiers 31 upvotes, #5 of 2025-06-24
  8. MiniCPM4: Ultra-Efficient LLMs on End Devices 78 upvotes, #3 of 2025-06-10
  9. The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models 114 upvotes, #1 of 2025-05-29
  10. TTRL: Test-Time Reinforcement Learning 96 upvotes, #2 of 2025-04-23
  11. UltraIF: Advancing Instruction Following from the Wild 20 upvotes, #9 of 2025-02-07
  12. Process Reinforcement through Implicit Rewards 53 upvotes, #3 of 2025-02-04
  13. MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding 19 upvotes, #6 of 2025-01-31
  14. Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization 35 upvotes, #1 of 2024-12-25
  15. How to Synthesize Text Data without Model Collapse? 46 upvotes, #4 of 2024-12-20
  16. Free Process Rewards without Process Labels 26 upvotes, #4 of 2024-12-04
  17. MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies 14 upvotes, #5 of 2024-04-10
  18. Advancing LLM Reasoning Generalists with Preference Trees 36 upvotes, #2 of 2024-04-03
  19. KoLA: Carefully Benchmarking World Knowledge of Large Language Models 20 upvotes, #6 of 2023-06-16
  20. Enhancing Chat Language Models by Scaling High-quality Instructional Conversations 8 upvotes, #2 of 2023-05-24

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.