Xie
Xie on Hugging Face Daily Papers: 9 papers, 3 in the top 3 of their day, 510 upvotes.
- RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents 78 upvotes, #5 of 2026-09-21
- Qwen-AgentWorld: Language World Models for General Agents 144 upvotes, #1 of 2026-06-24
- Claw-Eval: Toward Trustworthy Evaluation of Autonomous Agents 114 upvotes, #2 of 2026-04-08
- OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows 70 upvotes, #2 of 2025-11-03
- Attention as a Compass: Efficient Exploration for Process-Supervised RL in Reasoning Models 12 upvotes, #24 of 2025-10-01
- Teaching Language Models to Critique via Reinforcement Learning 22 upvotes, #8 of 2025-02-12
- VLRewardBench: A Challenging Benchmark for Vision-Language Generative Reward Models 10 upvotes, #8 of 2024-11-27
- Jailbreaking as a Reward Misspecification Problem 12 upvotes, #9 of 2024-06-24
- Silkie: Preference Distillation for Large Visual Language Models 10 upvotes, #9 of 2023-12-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.