Ning Ding
Ning Ding on Hugging Face Daily Papers: 20 papers, 12 in the top 3 of their day, 1,478 upvotes.
- SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning 73 upvotes, #3 of 2025-09-12
- A Survey of Reinforcement Learning for Large Reasoning Models 156 upvotes, #1 of 2025-09-11
- Towards a Unified View of Large Language Model Post-Training 67 upvotes, #3 of 2025-09-05
- Intern-S1: A Scientific Multimodal Foundation Model 242 upvotes, #1 of 2025-08-22
- From AI for Science to Agentic Science: A Survey on Autonomous Scientific Discovery 31 upvotes, #7 of 2025-08-21
- SSRL: Self-Search Reinforcement Learning 88 upvotes, #2 of 2025-08-18
- RLPR: Extrapolating RLVR to General Domains without Verifiers 31 upvotes, #5 of 2025-06-24
- MiniCPM4: Ultra-Efficient LLMs on End Devices 78 upvotes, #3 of 2025-06-10
- The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models 114 upvotes, #1 of 2025-05-29
- TTRL: Test-Time Reinforcement Learning 96 upvotes, #2 of 2025-04-23
- UltraIF: Advancing Instruction Following from the Wild 20 upvotes, #9 of 2025-02-07
- Process Reinforcement through Implicit Rewards 53 upvotes, #3 of 2025-02-04
- MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding 19 upvotes, #6 of 2025-01-31
- Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization 35 upvotes, #1 of 2024-12-25
- How to Synthesize Text Data without Model Collapse? 46 upvotes, #4 of 2024-12-20
- Free Process Rewards without Process Labels 26 upvotes, #4 of 2024-12-04
- MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies 14 upvotes, #5 of 2024-04-10
- Advancing LLM Reasoning Generalists with Preference Trees 36 upvotes, #2 of 2024-04-03
- KoLA: Carefully Benchmarking World Knowledge of Large Language Models 20 upvotes, #6 of 2023-06-16
- Enhancing Chat Language Models by Scaling High-quality Instructional Conversations 8 upvotes, #2 of 2023-05-24
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.