Lei Wang
Lei Wang on Hugging Face Daily Papers: 14 papers, 8 in the top 3 of their day, 751 upvotes.
- Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents 302 upvotes, #1 of 2026-07-31
- EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments 137 upvotes, #2 of 2026-06-12
- MARS: Enabling Autoregressive Models Multi-Token Generation 38 upvotes, #3 of 2026-04-09
- MiroEval: Benchmarking Multimodal Deep Research Agents in Process and Outcome 68 upvotes, #3 of 2026-04-02
- From Perception to Action: An Interactive Benchmark for Vision Reasoning 22 upvotes, #5 of 2026-02-25
- Self-Rewarding Sequential Monte Carlo for Masked Diffusion Language Models 5 upvotes, #38 of 2026-02-05
- DeepResearchEval: An Automated Framework for Deep Research Task Construction and Agentic Evaluation 121 upvotes, #1 of 2026-01-15
- Truth in the Few: High-Value Data Selection for Efficient Multi-Modal Reasoning 36 upvotes, #3 of 2025-06-09
- Scalable Chain of Thoughts via Elastic Reasoning 23 upvotes, #5 of 2025-05-09
- A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce 14 upvotes, #12 of 2025-04-16
- What, How, Where, and How Well? A Survey on Test-Time Scaling in Large Language Models 49 upvotes, #4 of 2025-04-01
- MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs 13 upvotes, #13 of 2024-10-08
- ThinK: Thinner Key Cache by Query-Driven Pruning 28 upvotes, #3 of 2024-07-31
- Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language Models 3 upvotes, #1 of 2023-05-09
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.