mz.w

mz.w on Hugging Face Daily Papers: 8 papers, 2 in the top 3 of their day, 212 upvotes.

  1. From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space 28 upvotes, #6 of 2026-04-16
  2. Bottom-up Policy Optimization: Your Language Model Policy Secretly Contains Internal Policies 60 upvotes, #3 of 2025-12-24
  3. Think on your Feet: Adaptive Thinking via Reinforcement Learning for Social Agents 17 upvotes, #12 of 2025-05-06
  4. DEMO: Reframing Dialogue Interaction with Fine-grained Element Modeling 7 upvotes, #15 of 2024-12-09
  5. The Imperative of Conversation Analysis in the Era of LLMs: A Survey of Tasks, Techniques, and Trends 10 upvotes, #8 of 2024-09-27
  6. MMEvol: Empowering Multimodal Large Language Models with Evol-Instruct 42 upvotes, #2 of 2024-09-10
  7. Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA 13 upvotes, #10 of 2024-06-26
  8. YAYI 2: Multilingual Open-Source Large Language Models 14 upvotes, #5 of 2023-12-26

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.