Lichang Chen

Lichang Chen on Hugging Face Daily Papers: 8 papers, 2 in the top 3 of their day, 181 upvotes.

  1. Learning to Reason via Mixture-of-Thought for Logical Reasoning 17 upvotes, #13 of 2025-05-22
  2. Self-rewarding correction for mathematical reasoning 75 upvotes, #1 of 2025-02-28
  3. RRM: Robust Reward Model Training Mitigates Reward Hacking 3 upvotes, #17 of 2024-09-25
  4. ODIN: Disentangled Reward Mitigates Hacking in RLHF 14 upvotes, #9 of 2024-02-13
  5. HallusionBench: You See What You Think? Or You Think What You See? An Image-Context Reasoning Benchmark Challenging for GPT-4V(ision), LLaVA-1.5, and Other Multi-modality Models 27 upvotes, #2 of 2023-10-24
  6. Virtual Prompt Injection for Instruction-Tuned Large Language Models 8 upvotes, #11 of 2023-08-01
  7. AlpaGasus: Training A Better Alpaca with Fewer Data 24 upvotes, #4 of 2023-07-18
  8. InstructZero: Efficient Instruction Optimization for Black-Box Large Language Models 5 upvotes, #5 of 2023-06-06

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.