Zhou
Zhou on Hugging Face Daily Papers: 36 papers, 21 in the top 3 of their day, 2,454 upvotes.
- Qwen3-Omni Technical Report 121 upvotes, #1 of 2025-09-23
- ReSum: Unlocking Long-Horizon Search Intelligence via Context Summarization 69 upvotes, #4 of 2025-09-17
- WebWeaver: Structuring Web-Scale Evidence with Dynamic Outlines for Open-Ended Deep Research 98 upvotes, #2 of 2025-09-17
- WebShaper: Agentically Data Synthesizing via Information-Seeking Formalization 44 upvotes, #5 of 2025-07-22
- WebDancer: Towards Autonomous Information Seeking Agency 18 upvotes, #16 of 2025-05-29
- QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning 83 upvotes, #2 of 2025-05-26
- Qwen3 Technical Report 152 upvotes, #1 of 2025-05-19
- WorldPM: Scaling Human Preference Modeling 33 upvotes, #5 of 2025-05-16
- Wan: Open and Advanced Large-Scale Video Generative Models 44 upvotes, #3 of 2025-03-27
- Demons in the Detail: On Implementing Load Balancing Loss for Training Specialized Mixture-of-Expert Models 61 upvotes, #3 of 2025-01-22
- The Lessons of Developing Process Reward Models in Mathematical Reasoning 83 upvotes, #1 of 2025-01-14
- Qwen2.5 Technical Report 328 upvotes, #1 of 2024-12-20
- ProcessBench: Identifying Process Errors in Mathematical Reasoning 61 upvotes, #2 of 2024-12-10
- A Simple and Provable Scaling Law for the Test-Time Compute of Large Language Models 4 upvotes, #22 of 2024-12-03
- In-Context LoRA for Diffusion Transformers 10 upvotes, #9 of 2024-11-04
- Aligning Large Language Models via Self-Steering Optimization 18 upvotes, #3 of 2024-10-23
- ACE: All-round Creator and Editor Following Instructions via Diffusion Transformer 10 upvotes, #6 of 2024-10-02
- Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution 63 upvotes, #2 of 2024-09-19
- Qwen2.5-Coder Technical Report 111 upvotes, #1 of 2024-09-19
- mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding 22 upvotes, #6 of 2024-09-06
- mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models 27 upvotes, #3 of 2024-08-12
- Very Large-Scale Multi-Agent Simulation in AgentScope 29 upvotes, #3 of 2024-07-26
- Qwen2-Audio Technical Report 34 upvotes, #2 of 2024-07-17
- Qwen2 Technical Report 142 upvotes, #1 of 2024-07-16
- Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models 14 upvotes, #13 of 2024-06-21
- mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding 24 upvotes, #1 of 2024-03-20
- AgentScope: A Flexible yet Robust Multi-Agent Platform 13 upvotes, #8 of 2024-02-23
- EE-Tuning: An Economical yet Scalable Solution for Tuning Early-Exit Large Language Models 4 upvotes, #11 of 2024-02-02
- Large Language Models are Superpositions of All Characters: Attaining Arbitrary Role-play via Self-Alignment 36 upvotes, #2 of 2024-01-24
- EE-LLM: Large-Scale Training and Inference of Early-Exit Large Language Models with 3D Parallelism 7 upvotes, #7 of 2023-12-11
- DreamVideo: Composing Your Dream Videos with Customized Subject and Motion 10 upvotes, #9 of 2023-12-08
- Routing to the Expert: Efficient Reward-guided Ensemble of Large Language Models 13 upvotes, #5 of 2023-11-16
- Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models 9 upvotes, #10 of 2023-11-15
- mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration 22 upvotes, #2 of 2023-11-09
- I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models 34 upvotes, #1 of 2023-11-08
- ModelScope-Agent: Building Your Customizable Agent System with Open-source Large Language Models 22 upvotes, #3 of 2023-09-06
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.