Chao Du
Chao Du on Hugging Face Daily Papers: 11 papers, 3 in the top 3 of their day, 255 upvotes.
- Optimizing Anytime Reasoning via Budget Relative Policy Optimization 34 upvotes, #3 of 2025-05-21
- UFO2: The Desktop AgentOS 27 upvotes, #7 of 2025-04-22
- Understanding R1-Zero-Like Training: A Critical Perspective 36 upvotes, #7 of 2025-04-03
- Error Analyses of Auto-Regressive Video Diffusion Models: A Unified Framework 5 upvotes, #20 of 2025-03-18
- Sample-Efficient Alignment for LLMs 10 upvotes, #6 of 2024-11-06
- Improving Long-Text Alignment for Text-to-Image Diffusion Models 13 upvotes, #8 of 2024-10-17
- Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates 6 upvotes, #20 of 2024-10-11
- Bootstrapping Language Models with DPO Implicit Rewards 34 upvotes, #3 of 2024-06-19
- Weak-to-Strong Jailbreaking on Large Language Models 16 upvotes, #8 of 2024-01-31
- LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition 34 upvotes, #1 of 2023-07-26
- Efficient Diffusion Policies for Offline Reinforcement Learning 2 upvotes, #5 of 2023-06-01
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.