Zhiyuan Liu

Zhiyuan Liu on Hugging Face Daily Papers: 32 papers, 7 in the top 3 of their day, 954 upvotes.

  1. MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction 68 upvotes, #6 of 2026-05-07
  2. The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models 114 upvotes, #1 of 2025-05-29
  3. Search and Refine During Think: Autonomous Retrieval-Augmented Reasoning of LLMs 5 upvotes, #45 of 2025-05-28
  4. Towards Unified Latent Space for 3D Molecular Latent Diffusion Modeling 6 upvotes, #36 of 2025-03-21
  5. Cost-Optimal Grouped-Query Attention for Long-Context LLMs 5 upvotes, #16 of 2025-03-13
  6. Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models 28 upvotes, #6 of 2025-01-13
  7. ACDiT: Interpolating Autoregressive Conditional Modeling and Diffusion Transformer 30 upvotes, #4 of 2024-12-11
  8. Densing Law of LLMs 14 upvotes, #14 of 2024-12-06
  9. Free Process Rewards without Process Labels 26 upvotes, #4 of 2024-12-04
  10. Sparsing Law: Towards Large Language Models with Greater Activation Sparsity 10 upvotes, #15 of 2024-11-05
  11. LLMtimesMapReduce: Simplified Long-Sequence Processing using Large Language Models 36 upvotes, #2 of 2024-10-16
  12. VisRAG: Vision-based Retrieval-augmented Generation on Multi-modality Documents 21 upvotes, #10 of 2024-10-15
  13. Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System 7 upvotes, #18 of 2024-10-11
  14. Stuffed Mamba: State Collapse and State Capacity of RNN-Based Long-Context Modeling 2 upvotes, #47 of 2024-10-10
  15. Configurable Foundation Models: Building LLMs from a Modular Perspective 26 upvotes, #2 of 2024-09-09
  16. From MOOC to MAIC: Reshaping Online Teaching and Learning through LLM-driven Agents 24 upvotes, #4 of 2024-09-06
  17. MiniCPM-V: A GPT-4V Level MLLM on Your Phone 69 upvotes, #1 of 2024-08-06
  18. Internet of Agents: Weaving a Web of Heterogeneous Agents for Collaborative Intelligence 23 upvotes, #4 of 2024-07-10
  19. Simulating Classroom Education with LLM-Empowered Agents 27 upvotes, #4 of 2024-06-28
  20. Beyond the Turn-Based Game: Enabling Real-Time Conversations with Duplex Models 14 upvotes, #12 of 2024-06-25
  21. LEGENT: Open Platform for Embodied Agents 17 upvotes, #4 of 2024-04-30
  22. MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies 14 upvotes, #5 of 2024-04-10
  23. Advancing LLM Reasoning Generalists with Preference Trees 36 upvotes, #2 of 2024-04-03
  24. LLaVA-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images 13 upvotes, #6 of 2024-03-19
  25. BurstAttention: An Efficient Distributed Attention Framework for Extremely Long Sequences 19 upvotes, #7 of 2024-03-15
  26. Ouroboros: Speculative Decoding with Large Model Enhanced Drafting 7 upvotes, #11 of 2024-02-22
  27. OneBit: Towards Extremely Low-bit Large Language Models 24 upvotes, #5 of 2024-02-20
  28. ProAgent: From Robotic Process Automation to Agentic Process Automation 8 upvotes, #13 of 2023-11-21
  29. ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs 102 upvotes, #1 of 2023-08-01
  30. Exploring Format Consistency for Instruction Tuning 8 upvotes, #6 of 2023-07-31
  31. KoLA: Carefully Benchmarking World Knowledge of Large Language Models 20 upvotes, #6 of 2023-06-16
  32. Enhancing Chat Language Models by Scaling High-quality Instructional Conversations 8 upvotes, #2 of 2023-05-24

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.