Hanze Dong
Hanze Dong on Hugging Face Daily Papers: 10 papers, 5 in the top 3 of their day, 411 upvotes.
- Fractured Chain-of-Thought Reasoning 21 upvotes, #14 of 2025-05-20
- Beyond 'Aha!': Toward Systematic Meta-Abilities Alignment in Large Reasoning Models 113 upvotes, #1 of 2025-05-16
- Scalable Chain of Thoughts via Elastic Reasoning 23 upvotes, #5 of 2025-05-09
- Optimizing Chain-of-Thought Reasoners via Gradient Variance Minimization in Rejection Sampling and RL 22 upvotes, #9 of 2025-05-06
- BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation 21 upvotes, #7 of 2025-02-07
- Reward-Guided Speculative Decoding for Efficient LLM Reasoning 34 upvotes, #2 of 2025-02-03
- Offline Reinforcement Learning for LLM Multi-Step Reasoning 33 upvotes, #2 of 2024-12-23
- MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs 13 upvotes, #13 of 2024-10-08
- ThinK: Thinner Key Cache by Query-Driven Pruning 28 upvotes, #3 of 2024-07-31
- RLHF Workflow: From Reward Modeling to Online RLHF 54 upvotes, #2 of 2024-05-14
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.