Junkang Wu

Junkang Wu on Hugging Face Daily Papers: 5 papers, 1 in the top 3 of their day, 182 upvotes.

  1. Quantile Advantage Estimation for Entropy-Safe Reasoning 113 upvotes, #3 of 2025-09-29
  2. Robust Preference Optimization via Dynamic Target Margins 2 upvotes, #42 of 2025-06-10
  3. Aligning Multimodal LLM with Human Preference: A Survey 21 upvotes, #8 of 2025-03-19
  4. RePO: ReLU-based Preference Optimization 1 upvotes, #46 of 2025-03-11
  5. MM-RLHF: The Next Step Forward in Multimodal LLM Alignment 30 upvotes, #6 of 2025-02-17

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.