Hanbin Wang
Hanbin Wang on Hugging Face Daily Papers: 2 papers, 2 in the top 3 of their day, 107 upvotes.
- Process Reinforcement through Implicit Rewards 53 upvotes, #3 of 2025-02-04
- Advancing LLM Reasoning Generalists with Preference Trees 36 upvotes, #2 of 2024-04-03
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.