Zhiyuan Liu
Zhiyuan Liu on Hugging Face Daily Papers: 32 papers, 7 in the top 3 of their day, 954 upvotes.
- MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction 68 upvotes, #6 of 2026-05-07
- The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models 114 upvotes, #1 of 2025-05-29
- Search and Refine During Think: Autonomous Retrieval-Augmented Reasoning of LLMs 5 upvotes, #45 of 2025-05-28
- Towards Unified Latent Space for 3D Molecular Latent Diffusion Modeling 6 upvotes, #36 of 2025-03-21
- Cost-Optimal Grouped-Query Attention for Long-Context LLMs 5 upvotes, #16 of 2025-03-13
- Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models 28 upvotes, #6 of 2025-01-13
- ACDiT: Interpolating Autoregressive Conditional Modeling and Diffusion Transformer 30 upvotes, #4 of 2024-12-11
- Densing Law of LLMs 14 upvotes, #14 of 2024-12-06
- Free Process Rewards without Process Labels 26 upvotes, #4 of 2024-12-04
- Sparsing Law: Towards Large Language Models with Greater Activation Sparsity 10 upvotes, #15 of 2024-11-05
- LLMtimesMapReduce: Simplified Long-Sequence Processing using Large Language Models 36 upvotes, #2 of 2024-10-16
- VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents 21 upvotes, #10 of 2024-10-15
- Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System 7 upvotes, #18 of 2024-10-11
- Stuffed Mamba: State Collapse and State Capacity of RNN-Based Long-Context Modeling 2 upvotes, #47 of 2024-10-10
- Configurable Foundation Models: Building LLMs from a Modular Perspective 26 upvotes, #2 of 2024-09-09
- From MOOC to MAIC: Reshaping Online Teaching and Learning through LLM-driven Agents 24 upvotes, #4 of 2024-09-06
- MiniCPM-V: A GPT-4V Level MLLM on Your Phone 69 upvotes, #1 of 2024-08-06
- Internet of Agents: Weaving a Web of Heterogeneous Agents for Collaborative Intelligence 23 upvotes, #4 of 2024-07-10
- Simulating Classroom Education with LLM-Empowered Agents 27 upvotes, #4 of 2024-06-28
- Beyond the Turn-Based Game: Enabling Real-Time Conversations with Duplex Models 14 upvotes, #12 of 2024-06-25
- LEGENT: Open Platform for Embodied Agents 17 upvotes, #4 of 2024-04-30
- MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies 14 upvotes, #5 of 2024-04-10
- Advancing LLM Reasoning Generalists with Preference Trees 36 upvotes, #2 of 2024-04-03
- LLaVA-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images 13 upvotes, #6 of 2024-03-19
- BurstAttention: An Efficient Distributed Attention Framework for Extremely Long Sequences 19 upvotes, #7 of 2024-03-15
- Ouroboros: Speculative Decoding with Large Model Enhanced Drafting 7 upvotes, #11 of 2024-02-22
- OneBit: Towards Extremely Low-bit Large Language Models 24 upvotes, #5 of 2024-02-20
- ProAgent: From Robotic Process Automation to Agentic Process Automation 8 upvotes, #13 of 2023-11-21
- ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs 102 upvotes, #1 of 2023-08-01
- Exploring Format Consistency for Instruction Tuning 8 upvotes, #6 of 2023-07-31
- KoLA: Carefully Benchmarking World Knowledge of Large Language Models 20 upvotes, #6 of 2023-06-16
- Enhancing Chat Language Models by Scaling High-quality Instructional Conversations 8 upvotes, #2 of 2023-05-24
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.