Haitao Mi
Haitao Mi on Hugging Face Daily Papers: 10 papers, 4 in the top 3 of their day, 410 upvotes.
- Self-Rewarding Vision-Language Model via Reasoning Decomposition 77 upvotes, #2 of 2025-08-28
- MPS-Prover: Advancing Stepwise Theorem Proving by Multi-Perspective Search and Data Curation 7 upvotes, #9 of 2025-05-19
- Expanding RL with Verifiable Rewards Across Diverse Domains 17 upvotes, #11 of 2025-04-01
- Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs 51 upvotes, #2 of 2025-01-31
- Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning 9 upvotes, #15 of 2024-10-11
- HDFlow: Enhancing LLM Complex Problem-Solving with Hybrid Thinking and Dynamic Workflows 8 upvotes, #8 of 2024-09-30
- LiteSearch: Efficacious Tree Search for LLM 34 upvotes, #4 of 2024-07-02
- Scaling Synthetic Data Creation with 1,000,000,000 Personas 79 upvotes, #1 of 2024-07-01
- Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing 44 upvotes, #1 of 2024-04-19
- Stabilizing RLHF through Advantage Model and Selective Rehearsal 10 upvotes, #6 of 2023-09-20
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.