Caiming Xiong

Caiming Xiong on Hugging Face Daily Papers: 24 papers, 9 in the top 3 of their day, 928 upvotes.

  1. Least-Loaded Expert Parallelism: Load Balancing An Imbalanced Mixture-of-Experts 5 upvotes, #27 of 2026-01-27
  2. Fractured Chain-of-Thought Reasoning 21 upvotes, #14 of 2025-05-20
  3. Scaling Computer-Use Grounding via User Interface Decomposition and Synthesis 44 upvotes, #6 of 2025-05-20
  4. Beyond 'Aha!': Toward Systematic Meta-Abilities Alignment in Large Reasoning Models 113 upvotes, #1 of 2025-05-16
  5. BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset 80 upvotes, #1 of 2025-05-15
  6. Scalable Chain of Thoughts via Elastic Reasoning 23 upvotes, #5 of 2025-05-09
  7. BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation 21 upvotes, #7 of 2025-02-07
  8. Reward-Guided Speculative Decoding for Efficient LLM Reasoning 34 upvotes, #2 of 2025-02-03
  9. Demystifying Domain-adaptive Post-training for Financial LLMs 10 upvotes, #13 of 2025-01-13
  10. AgentTrek: Agent Trajectory Synthesis via Guiding Replay with Web Tutorials 24 upvotes, #6 of 2024-12-13
  11. Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction 43 upvotes, #4 of 2024-12-06
  12. MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs 13 upvotes, #13 of 2024-10-08
  13. ThinK: Thinner Key Cache by Query-Driven Pruning 28 upvotes, #3 of 2024-07-31
  14. Summary of a Haystack: A Challenge to Long-Context LLMs and RAG Systems 73 upvotes, #1 of 2024-07-03
  15. RLHF Workflow: From Reward Modeling to Online RLHF 54 upvotes, #2 of 2024-05-14
  16. AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning 17 upvotes, #10 of 2024-02-26
  17. TrustLLM: Trustworthiness in Large Language Models 69 upvotes, #1 of 2024-01-12
  18. Unlocking Anticipatory Text Generation: A Constrained Approach for Faithful Decoding with Large Language Models 3 upvotes, #13 of 2023-12-12
  19. Diffusion Model Alignment Using Direct Preference Optimization 48 upvotes, #4 of 2023-11-23
  20. Lemur: Harmonizing Natural Language and Code for Language Agents 31 upvotes, #3 of 2023-10-13
  21. XGen-7B Technical Report 8 upvotes, #11 of 2023-09-08
  22. BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents 20 upvotes, #4 of 2023-08-14
  23. Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization 21 upvotes, #1 of 2023-08-07
  24. DialogStudio: Towards Richest and Most Diverse Unified Dataset Collection for Conversational AI 13 upvotes, #6 of 2023-07-20

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.