Ximing Lu

Ximing Lu on Hugging Face Daily Papers: 7 papers, 2 in the top 3 of their day, 315 upvotes.

  1. GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization 191 upvotes, #1 of 2026-01-09
  2. StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements 9 upvotes, #9 of 2024-08-30
  3. WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models 7 upvotes, #13 of 2024-06-27
  4. Localized Symbolic Knowledge Distillation for Visual Commonsense Models 3 upvotes, #8 of 2023-12-11
  5. The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learning 31 upvotes, #3 of 2023-12-05
  6. Tailoring Self-Rationalizers with Multi-Reward Distillation 5 upvotes, #13 of 2023-11-07
  7. The Generative AI Paradox: "What It Can Create, It May Not Understand" 18 upvotes, #5 of 2023-11-02

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.