Zexiong Ma
Zexiong Ma on Hugging Face Daily Papers: 4 papers, 2 in the top 3 of their day, 111 upvotes.
- Tool-integrated Reinforcement Learning for Repo Deep Search 18 upvotes, #5 of 2025-08-06
- SoRFT: Issue Resolving with Subtask-oriented Reinforced Fine-Tuning 9 upvotes, #15 of 2025-02-28
- Make Your LLM Fully Utilize the Context 45 upvotes, #3 of 2024-04-26
- Learning From Mistakes Makes LLM Better Reasoner 29 upvotes, #1 of 2023-11-01
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.