Xipeng Qiu
Xipeng Qiu on Hugging Face Daily Papers: 25 papers, 12 in the top 3 of their day, 1,281 upvotes.
- AI Can Learn Scientific Taste 397 upvotes, #1 of 2026-03-17
- MOVA: Towards Scalable and Synchronized Video-Audio Generation 151 upvotes, #4 of 2026-02-10
- FutureOmni: Evaluating Future Forecasting from Omni-Modal Context for Multimodal LLMs 34 upvotes, #7 of 2026-01-21
- ABC-Bench: Benchmarking Agentic Backend Coding in Real-World Development 63 upvotes, #1 of 2026-01-20
- AstroReason-Bench: Evaluating Unified Agentic Planning across Heterogeneous Space Planning Problems 4 upvotes, #20 of 2026-01-19
- MOSS Transcribe Diarize: Accurate Transcription with Speaker Diarization 52 upvotes, #3 of 2026-01-07
- Beyond Real: Imaginary Extension of Rotary Position Embeddings for Long-Context LLMs 55 upvotes, #2 of 2025-12-09
- VisuoThink: Empowering LVLM Reasoning with Multimodal Tree Search 11 upvotes, #16 of 2025-04-15
- World Modeling Makes a Better Planner: Dual Preference Optimization for Embodied Task Planning 45 upvotes, #4 of 2025-03-14
- YuE: Scaling Open Foundation Models for Long-Form Music Generation 57 upvotes, #3 of 2025-03-12
- DuoDecoding: Hardware-aware Heterogeneous Speculative Decoding with Dynamic Multi-Sequence Drafting 10 upvotes, #13 of 2025-03-04
- Thus Spake Long-Context Large Language Model 66 upvotes, #2 of 2025-02-25
- Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities? 16 upvotes, #13 of 2025-02-19
- VideoRoPE: What Makes for Good Video Rotary Position Embedding? 60 upvotes, #3 of 2025-02-10
- AgentGym: Evolving Large Language Model-based Agents across Diverse Environments 13 upvotes, #8 of 2024-06-07
- InternLM2 Technical Report 22 upvotes, #3 of 2024-03-27
- Training-Free Long-Context Scaling of Large Language Models 23 upvotes, #6 of 2024-02-28
- AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling 45 upvotes, #3 of 2024-02-20
- InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning 19 upvotes, #2 of 2024-02-12
- MouSi: Poly-Visual-Expert Vision-Language Models 9 upvotes, #13 of 2024-01-31
- Secrets of RLHF in Large Language Models Part II: Reward Modeling 26 upvotes, #6 of 2024-01-12
- Alignment for Honesty 13 upvotes, #5 of 2023-12-13
- Secrets of RLHF in Large Language Models Part I: PPO 30 upvotes, #2 of 2023-07-12
- Full Parameter Fine-tuning for Large Language Models with Limited Resources 31 upvotes, #2 of 2023-06-19
- SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities 5 upvotes, #5 of 2023-05-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.