Junyang Lin
Junyang Lin on Hugging Face Daily Papers: 35 papers, 24 in the top 3 of their day, 2,898 upvotes.
- Mobile-Agent-v3.5: Multi-platform Fundamental GUI Agents 46 upvotes, #2 of 2026-02-20
- SWE-Universe: Scale Real-World Verifiable Environments to Millions 59 upvotes, #7 of 2026-02-03
- DeepPlanning: Benchmarking Long-Horizon Agentic Planning with Verifiable Constraints 24 upvotes, #8 of 2026-01-27
- Qwen3-TTS Technical Report 54 upvotes, #5 of 2026-01-23
- Qwen3-Omni Technical Report 121 upvotes, #1 of 2025-09-23
- RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback 13 upvotes, #9 of 2025-07-23
- Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning 144 upvotes, #1 of 2025-06-03
- Parallel Scaling Law for Language Models 72 upvotes, #2 of 2025-05-16
- WorldPM: Scaling Human Preference Modeling 33 upvotes, #5 of 2025-05-16
- Qwen2.5-Omni Technical Report 113 upvotes, #1 of 2025-03-27
- START: Self-taught Reasoner with Tools 87 upvotes, #1 of 2025-03-07
- Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think 26 upvotes, #7 of 2025-02-28
- Qwen2.5-VL Technical Report 146 upvotes, #1 of 2025-02-20
- Qwen2.5-1M Technical Report 51 upvotes, #1 of 2025-01-28
- RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques 29 upvotes, #3 of 2025-01-27
- Demons in the Detail: On Implementing Load Balancing Loss for Training Specialized Mixture-of-Expert Models 61 upvotes, #3 of 2025-01-22
- The Lessons of Developing Process Reward Models in Mathematical Reasoning 83 upvotes, #1 of 2025-01-14
- Enabling Scalable Oversight via Self-Evolving Critic 66 upvotes, #1 of 2025-01-13
- CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings 44 upvotes, #3 of 2025-01-03
- Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey 49 upvotes, #3 of 2024-12-30
- Qwen2.5 Technical Report 328 upvotes, #1 of 2024-12-20
- Evaluating and Aligning CodeLLMs on Human Preference 47 upvotes, #2 of 2024-12-11
- ProcessBench: Identifying Process Errors in Mathematical Reasoning 61 upvotes, #2 of 2024-12-10
- Aligning Large Language Models via Self-Steering Optimization 18 upvotes, #3 of 2024-10-23
- Rethinking Data Selection at Scale: Random Selection is Almost All You Need 14 upvotes, #11 of 2024-10-15
- A Spark of Vision-Language Intelligence: 2-Dimensional Autoregressive Transformer for Efficient Finegrained Image Generation 13 upvotes, #3 of 2024-10-09
- Qwen2.5-Coder Technical Report 111 upvotes, #1 of 2024-09-19
- Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution 63 upvotes, #2 of 2024-09-19
- Synthesizing Text-to-SQL Data from Weak and Strong LLMs 6 upvotes, #11 of 2024-08-07
- OpenDevin: An Open Platform for AI Software Developers as Generalist Agents 62 upvotes, #1 of 2024-07-25
- Qwen2-Audio Technical Report 34 upvotes, #2 of 2024-07-17
- Qwen2 Technical Report 142 upvotes, #1 of 2024-07-16
- An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models 23 upvotes, #4 of 2024-03-12
- Routing to the Expert: Efficient Reward-guided Ensemble of Large Language Models 13 upvotes, #5 of 2023-11-16
- Qwen Technical Report 39 upvotes, #4 of 2023-09-29
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.