Juanzi Li

Juanzi Li on Hugging Face Daily Papers: 21 papers, 6 in the top 3 of their day, 922 upvotes.

  1. IndexCache: Accelerating Sparse Attention via Cross-Layer Index Reuse 51 upvotes, #3 of 2026-03-13
  2. GLM-5: from Vibe Coding to Agentic Engineering 94 upvotes, #1 of 2026-02-18
  3. Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis 26 upvotes, #8 of 2025-06-05
  4. Hard Negative Contrastive Learning for Fine-Grained Geometric Understanding in Large Multimodal Models 11 upvotes, #34 of 2025-05-27
  5. AdaptThink: Reasoning Models Can Learn When to Think 72 upvotes, #2 of 2025-05-20
  6. ReaRAG: Knowledge-guided Reasoning Enhances Factuality of Large Reasoning Models with Iterative Retrieval Augmented Generation 25 upvotes, #6 of 2025-03-28
  7. Shifting Long-Context LLMs Research from Input to Output 19 upvotes, #14 of 2025-03-14
  8. LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models 23 upvotes, #9 of 2025-02-21
  9. Pairwise RM: Perform Best-of-N Sampling with Knockout Tournament 18 upvotes, #8 of 2025-01-23
  10. LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks 31 upvotes, #5 of 2024-12-20
  11. Pre-training Distillation for Large Language Models: A Design Space Exploration 15 upvotes, #11 of 2024-10-22
  12. RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style 23 upvotes, #9 of 2024-10-22
  13. From MOOC to MAIC: Reshaping Online Teaching and Learning through LLM-driven Agents 24 upvotes, #4 of 2024-09-06
  14. LongCite: Enabling LLMs to Generate Fine-grained Citations in Long-context QA 41 upvotes, #3 of 2024-09-05
  15. CogVLM2: Visual Language Models for Image and Video Understanding 55 upvotes, #2 of 2024-08-30
  16. LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs 60 upvotes, #1 of 2024-08-14
  17. LLMAEL: Large Language Models are Good Context Augmenters for Entity Linking 2 upvotes, #16 of 2024-07-09
  18. SeaKR: Self-aware Knowledge Retrieval for Adaptive Retrieval Augmented Generation 26 upvotes, #5 of 2024-06-28
  19. Simulating Classroom Education with LLM-Empowered Agents 27 upvotes, #4 of 2024-06-28
  20. Aligning Teacher with Student Preferences for Tailored Training Data Generation 22 upvotes, #6 of 2024-06-28
  21. ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools 26 upvotes, #6 of 2024-06-19

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.