Yunxuan Li
Yunxuan Li on Hugging Face Daily Papers: 5 papers, 2 in the top 3 of their day, 118 upvotes.
- Beyond Markovian: Reflective Exploration via Bayes-Adaptive RL for LLM Reasoning 6 upvotes, #40 of 2025-05-28
- Conditioned Language Policy: A General Framework for Steerable Multi-Objective Finetuning 7 upvotes, #16 of 2024-07-23
- Improve Mathematical Reasoning in Language Models by Automated Process Supervision 17 upvotes, #6 of 2024-06-12
- Gemini: A Family of Highly Capable Multimodal Models 50 upvotes, #2 of 2023-12-20
- Enable Language Models to Implicitly Learn Self-Improvement From Data 24 upvotes, #2 of 2023-10-03
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.