Junyang Lin

Junyang Lin on Hugging Face Daily Papers: 35 papers, 24 in the top 3 of their day, 2,898 upvotes.

  1. Mobile-Agent-v3.5: Multi-platform Fundamental GUI Agents 46 upvotes, #2 of 2026-02-20
  2. SWE-Universe: Scale Real-World Verifiable Environments to Millions 59 upvotes, #7 of 2026-02-03
  3. DeepPlanning: Benchmarking Long-Horizon Agentic Planning with Verifiable Constraints 24 upvotes, #8 of 2026-01-27
  4. Qwen3-TTS Technical Report 54 upvotes, #5 of 2026-01-23
  5. Qwen3-Omni Technical Report 121 upvotes, #1 of 2025-09-23
  6. RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback 13 upvotes, #9 of 2025-07-23
  7. Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning 144 upvotes, #1 of 2025-06-03
  8. Parallel Scaling Law for Language Models 72 upvotes, #2 of 2025-05-16
  9. WorldPM: Scaling Human Preference Modeling 33 upvotes, #5 of 2025-05-16
  10. Qwen2.5-Omni Technical Report 113 upvotes, #1 of 2025-03-27
  11. START: Self-taught Reasoner with Tools 87 upvotes, #1 of 2025-03-07
  12. Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think 26 upvotes, #7 of 2025-02-28
  13. Qwen2.5-VL Technical Report 146 upvotes, #1 of 2025-02-20
  14. Qwen2.5-1M Technical Report 51 upvotes, #1 of 2025-01-28
  15. RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques 29 upvotes, #3 of 2025-01-27
  16. Demons in the Detail: On Implementing Load Balancing Loss for Training Specialized Mixture-of-Expert Models 61 upvotes, #3 of 2025-01-22
  17. The Lessons of Developing Process Reward Models in Mathematical Reasoning 83 upvotes, #1 of 2025-01-14
  18. Enabling Scalable Oversight via Self-Evolving Critic 66 upvotes, #1 of 2025-01-13
  19. CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings 44 upvotes, #3 of 2025-01-03
  20. Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey 49 upvotes, #3 of 2024-12-30
  21. Qwen2.5 Technical Report 328 upvotes, #1 of 2024-12-20
  22. Evaluating and Aligning CodeLLMs on Human Preference 47 upvotes, #2 of 2024-12-11
  23. ProcessBench: Identifying Process Errors in Mathematical Reasoning 61 upvotes, #2 of 2024-12-10
  24. Aligning Large Language Models via Self-Steering Optimization 18 upvotes, #3 of 2024-10-23
  25. Rethinking Data Selection at Scale: Random Selection is Almost All You Need 14 upvotes, #11 of 2024-10-15
  26. A Spark of Vision-Language Intelligence: 2-Dimensional Autoregressive Transformer for Efficient Finegrained Image Generation 13 upvotes, #3 of 2024-10-09
  27. Qwen2.5-Coder Technical Report 111 upvotes, #1 of 2024-09-19
  28. Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution 63 upvotes, #2 of 2024-09-19
  29. Synthesizing Text-to-SQL Data from Weak and Strong LLMs 6 upvotes, #11 of 2024-08-07
  30. OpenDevin: An Open Platform for AI Software Developers as Generalist Agents 62 upvotes, #1 of 2024-07-25
  31. Qwen2-Audio Technical Report 34 upvotes, #2 of 2024-07-17
  32. Qwen2 Technical Report 142 upvotes, #1 of 2024-07-16
  33. An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models 23 upvotes, #4 of 2024-03-12
  34. Routing to the Expert: Efficient Reward-guided Ensemble of Large Language Models 13 upvotes, #5 of 2023-11-16
  35. Qwen Technical Report 39 upvotes, #4 of 2023-09-29

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.