Arman Cohan

Arman Cohan on Hugging Face Daily Papers: 14 papers, 3 in the top 3 of their day, 388 upvotes.

  1. RbtAct: Rebuttal as Supervision for Actionable Review Feedback Generation 6 upvotes, #17 of 2026-03-12
  2. References Improve LLM Alignment in Non-Verifiable Domains 2 upvotes, #21 of 2026-02-20
  3. Patient-Similarity Cohort Reasoning in Clinical Text-to-SQL 6 upvotes, #30 of 2026-01-16
  4. Z1: Efficient Test-time Scaling with Code 25 upvotes, #9 of 2025-04-02
  5. PHYSICS: Benchmarking Foundation Models on University-Level Physics Problem Solving 16 upvotes, #12 of 2025-03-31
  6. MCTS-RAG: Enhancing Retrieval-Augmented Generation with Monte Carlo Tree Search 9 upvotes, #12 of 2025-03-27
  7. Survey on Evaluation of LLM-based Agents 78 upvotes, #2 of 2025-03-21
  8. TESS 2: A Large-Scale Generalist Diffusion Language Model 5 upvotes, #21 of 2025-02-20
  9. MMVU: Measuring Expert-Level Multi-Discipline Video Understanding 79 upvotes, #2 of 2025-01-22
  10. ChemAgent: Self-updating Library in Large Language Models Improves Chemical Reasoning 8 upvotes, #11 of 2025-01-14
  11. FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions 8 upvotes, #9 of 2024-03-25
  12. OLMo: Accelerating the Science of Language Models 86 upvotes, #1 of 2024-02-02
  13. ML-Bench: Large Language Models Leverage Open-source Libraries for Machine Learning Tasks 10 upvotes, #6 of 2023-11-17
  14. Struc-Bench: Are Large Language Models Really Good at Generating Complex Structured Data? 11 upvotes, #10 of 2023-09-19

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.