mz.w
mz.w on Hugging Face Daily Papers: 8 papers, 2 in the top 3 of their day, 212 upvotes.
- From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space 28 upvotes, #6 of 2026-04-16
- Bottom-up Policy Optimization: Your Language Model Policy Secretly Contains Internal Policies 60 upvotes, #3 of 2025-12-24
- Think on your Feet: Adaptive Thinking via Reinforcement Learning for Social Agents 17 upvotes, #12 of 2025-05-06
- DEMO: Reframing Dialogue Interaction with Fine-grained Element Modeling 7 upvotes, #15 of 2024-12-09
- The Imperative of Conversation Analysis in the Era of LLMs: A Survey of Tasks, Techniques, and Trends 10 upvotes, #8 of 2024-09-27
- MMEvol: Empowering Multimodal Large Language Models with Evol-Instruct 42 upvotes, #2 of 2024-09-10
- Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA 13 upvotes, #10 of 2024-06-26
- YAYI 2: Multilingual Open-Source Large Language Models 14 upvotes, #5 of 2023-12-26
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.