zongzefang

zongzefang on Hugging Face Daily Papers: 2 papers, 1 in the top 3 of their day, 68 upvotes.

  1. AT^2PO: Agentic Turn-based Policy Optimization via Tree Search 26 upvotes, #7 of 2026-01-09
  2. Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models 35 upvotes, #3 of 2025-01-17

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.