Ximing Lu
Ximing Lu on Hugging Face Daily Papers: 7 papers, 2 in the top 3 of their day, 315 upvotes.
- GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization 191 upvotes, #1 of 2026-01-09
- StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements 9 upvotes, #9 of 2024-08-30
- WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models 7 upvotes, #13 of 2024-06-27
- Localized Symbolic Knowledge Distillation for Visual Commonsense Models 3 upvotes, #8 of 2023-12-11
- The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learning 31 upvotes, #3 of 2023-12-05
- Tailoring Self-Rationalizers with Multi-Reward Distillation 5 upvotes, #13 of 2023-11-07
- The Generative AI Paradox: "What It Can Create, It May Not Understand" 18 upvotes, #5 of 2023-11-02
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.