jingyi Yang

jingyi Yang on Hugging Face Daily Papers: 6 papers, 0 in the top 3 of their day, 111 upvotes.

  1. WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation 45 upvotes, #9 of 2026-05-15
  2. ToolCUA: Towards Optimal GUI-Tool Path Orchestration for Computer Use Agents 27 upvotes, #14 of 2026-05-13
  3. DARE: Diffusion Large Language Models Alignment and Reinforcement Executor 21 upvotes, #16 of 2026-04-08
  4. Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents 16 upvotes, #9 of 2025-10-06
  5. Taming Masked Diffusion Language Models via Consistency Trajectory Reinforcement Learning with Fewer Decoding Step 7 upvotes, #45 of 2025-09-30
  6. RiOSWorld: Benchmarking the Risk of Multimodal Compter-Use Agents 1 upvotes, #48 of 2025-06-05

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.