Lei Wang

Lei Wang on Hugging Face Daily Papers: 14 papers, 8 in the top 3 of their day, 751 upvotes.

  1. Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents 302 upvotes, #1 of 2026-07-31
  2. EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments 137 upvotes, #2 of 2026-06-12
  3. MARS: Enabling Autoregressive Models Multi-Token Generation 38 upvotes, #3 of 2026-04-09
  4. MiroEval: Benchmarking Multimodal Deep Research Agents in Process and Outcome 68 upvotes, #3 of 2026-04-02
  5. From Perception to Action: An Interactive Benchmark for Vision Reasoning 22 upvotes, #5 of 2026-02-25
  6. Self-Rewarding Sequential Monte Carlo for Masked Diffusion Language Models 5 upvotes, #38 of 2026-02-05
  7. DeepResearchEval: An Automated Framework for Deep Research Task Construction and Agentic Evaluation 121 upvotes, #1 of 2026-01-15
  8. Truth in the Few: High-Value Data Selection for Efficient Multi-Modal Reasoning 36 upvotes, #3 of 2025-06-09
  9. Scalable Chain of Thoughts via Elastic Reasoning 23 upvotes, #5 of 2025-05-09
  10. A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce 14 upvotes, #12 of 2025-04-16
  11. What, How, Where, and How Well? A Survey on Test-Time Scaling in Large Language Models 49 upvotes, #4 of 2025-04-01
  12. MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs 13 upvotes, #13 of 2024-10-08
  13. ThinK: Thinner Key Cache by Query-Driven Pruning 28 upvotes, #3 of 2024-07-31
  14. Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language Models 3 upvotes, #1 of 2023-05-09

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.