Bryan Hooi
Bryan Hooi on Hugging Face Daily Papers: 11 papers, 6 in the top 3 of their day, 743 upvotes.
- Rewarding the Rare: Uniqueness-Aware RL for Creative Problem Solving in LLMs 140 upvotes, #3 of 2026-01-16
- Collaborative Multi-Agent Test-Time Reinforcement Learning for Reasoning 82 upvotes, #4 of 2026-01-16
- MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning Research 10 upvotes, #36 of 2025-05-27
- GuardReasoner-VL: Safeguarding VLMs via Reinforced Reasoning 50 upvotes, #3 of 2025-05-19
- Beyond 'Aha!': Toward Systematic Meta-Abilities Alignment in Large Reasoning Models 113 upvotes, #1 of 2025-05-16
- FlowReasoner: Reinforcing Query-Level Meta-Agents 46 upvotes, #3 of 2025-04-22
- JudgeLRM: Large Reasoning Models as a Judge 55 upvotes, #2 of 2025-04-02
- Efficient Inference for Large Reasoning Models: A Survey 45 upvotes, #5 of 2025-04-01
- ReLearn: Unlearning via Learning for Large Language Models 28 upvotes, #4 of 2025-02-18
- GuardReasoner: Towards Reasoning-based LLM Safeguards 79 upvotes, #1 of 2025-01-31
- LongRecipe: Recipe for Efficient Long Context Generalization in Large Languge Models 37 upvotes, #4 of 2024-09-04
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.