Bryan Hooi

Bryan Hooi on Hugging Face Daily Papers: 11 papers, 6 in the top 3 of their day, 743 upvotes.

  1. Rewarding the Rare: Uniqueness-Aware RL for Creative Problem Solving in LLMs 140 upvotes, #3 of 2026-01-16
  2. Collaborative Multi-Agent Test-Time Reinforcement Learning for Reasoning 82 upvotes, #4 of 2026-01-16
  3. MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning Research 10 upvotes, #36 of 2025-05-27
  4. GuardReasoner-VL: Safeguarding VLMs via Reinforced Reasoning 50 upvotes, #3 of 2025-05-19
  5. Beyond 'Aha!': Toward Systematic Meta-Abilities Alignment in Large Reasoning Models 113 upvotes, #1 of 2025-05-16
  6. FlowReasoner: Reinforcing Query-Level Meta-Agents 46 upvotes, #3 of 2025-04-22
  7. JudgeLRM: Large Reasoning Models as a Judge 55 upvotes, #2 of 2025-04-02
  8. Efficient Inference for Large Reasoning Models: A Survey 45 upvotes, #5 of 2025-04-01
  9. ReLearn: Unlearning via Learning for Large Language Models 28 upvotes, #4 of 2025-02-18
  10. GuardReasoner: Towards Reasoning-based LLM Safeguards 79 upvotes, #1 of 2025-01-31
  11. LongRecipe: Recipe for Efficient Long Context Generalization in Large Languge Models 37 upvotes, #4 of 2024-09-04

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.