Chao Du

Chao Du on Hugging Face Daily Papers: 11 papers, 3 in the top 3 of their day, 255 upvotes.

  1. Optimizing Anytime Reasoning via Budget Relative Policy Optimization 34 upvotes, #3 of 2025-05-21
  2. UFO2: The Desktop AgentOS 27 upvotes, #7 of 2025-04-22
  3. Understanding R1-Zero-Like Training: A Critical Perspective 36 upvotes, #7 of 2025-04-03
  4. Error Analyses of Auto-Regressive Video Diffusion Models: A Unified Framework 5 upvotes, #20 of 2025-03-18
  5. Sample-Efficient Alignment for LLMs 10 upvotes, #6 of 2024-11-06
  6. Improving Long-Text Alignment for Text-to-Image Diffusion Models 13 upvotes, #8 of 2024-10-17
  7. Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates 6 upvotes, #20 of 2024-10-11
  8. Bootstrapping Language Models with DPO Implicit Rewards 34 upvotes, #3 of 2024-06-19
  9. Weak-to-Strong Jailbreaking on Large Language Models 16 upvotes, #8 of 2024-01-31
  10. LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition 34 upvotes, #1 of 2023-07-26
  11. Efficient Diffusion Policies for Offline Reinforcement Learning 2 upvotes, #5 of 2023-06-01

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.