Arman Cohan
Arman Cohan on Hugging Face Daily Papers: 14 papers, 3 in the top 3 of their day, 388 upvotes.
- RbtAct: Rebuttal as Supervision for Actionable Review Feedback Generation 6 upvotes, #17 of 2026-03-12
- References Improve LLM Alignment in Non-Verifiable Domains 2 upvotes, #21 of 2026-02-20
- Patient-Similarity Cohort Reasoning in Clinical Text-to-SQL 6 upvotes, #30 of 2026-01-16
- Z1: Efficient Test-time Scaling with Code 25 upvotes, #9 of 2025-04-02
- PHYSICS: Benchmarking Foundation Models on University-Level Physics Problem Solving 16 upvotes, #12 of 2025-03-31
- MCTS-RAG: Enhancing Retrieval-Augmented Generation with Monte Carlo Tree Search 9 upvotes, #12 of 2025-03-27
- Survey on Evaluation of LLM-based Agents 78 upvotes, #2 of 2025-03-21
- TESS 2: A Large-Scale Generalist Diffusion Language Model 5 upvotes, #21 of 2025-02-20
- MMVU: Measuring Expert-Level Multi-Discipline Video Understanding 79 upvotes, #2 of 2025-01-22
- ChemAgent: Self-updating Library in Large Language Models Improves Chemical Reasoning 8 upvotes, #11 of 2025-01-14
- FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions 8 upvotes, #9 of 2024-03-25
- OLMo: Accelerating the Science of Language Models 86 upvotes, #1 of 2024-02-02
- ML-Bench: Large Language Models Leverage Open-source Libraries for Machine Learning Tasks 10 upvotes, #6 of 2023-11-17
- Struc-Bench: Are Large Language Models Really Good at Generating Complex Structured Data? 11 upvotes, #10 of 2023-09-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.