Junkang Wu
Junkang Wu on Hugging Face Daily Papers: 5 papers, 1 in the top 3 of their day, 182 upvotes.
- Quantile Advantage Estimation for Entropy-Safe Reasoning 113 upvotes, #3 of 2025-09-29
- Robust Preference Optimization via Dynamic Target Margins 2 upvotes, #42 of 2025-06-10
- Aligning Multimodal LLM with Human Preference: A Survey 21 upvotes, #8 of 2025-03-19
- RePO: ReLU-based Preference Optimization 1 upvotes, #46 of 2025-03-11
- MM-RLHF: The Next Step Forward in Multimodal LLM Alignment 30 upvotes, #6 of 2025-02-17
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.