Caiming Xiong
Caiming Xiong on Hugging Face Daily Papers: 24 papers, 9 in the top 3 of their day, 928 upvotes.
- Least-Loaded Expert Parallelism: Load Balancing An Imbalanced Mixture-of-Experts 5 upvotes, #27 of 2026-01-27
- Fractured Chain-of-Thought Reasoning 21 upvotes, #14 of 2025-05-20
- Scaling Computer-Use Grounding via User Interface Decomposition and Synthesis 44 upvotes, #6 of 2025-05-20
- Beyond 'Aha!': Toward Systematic Meta-Abilities Alignment in Large Reasoning Models 113 upvotes, #1 of 2025-05-16
- BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset 80 upvotes, #1 of 2025-05-15
- Scalable Chain of Thoughts via Elastic Reasoning 23 upvotes, #5 of 2025-05-09
- BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation 21 upvotes, #7 of 2025-02-07
- Reward-Guided Speculative Decoding for Efficient LLM Reasoning 34 upvotes, #2 of 2025-02-03
- Demystifying Domain-adaptive Post-training for Financial LLMs 10 upvotes, #13 of 2025-01-13
- AgentTrek: Agent Trajectory Synthesis via Guiding Replay with Web Tutorials 24 upvotes, #6 of 2024-12-13
- Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction 43 upvotes, #4 of 2024-12-06
- MathHay: An Automated Benchmark for Long-Context Mathematical Reasoning in LLMs 13 upvotes, #13 of 2024-10-08
- ThinK: Thinner Key Cache by Query-Driven Pruning 28 upvotes, #3 of 2024-07-31
- Summary of a Haystack: A Challenge to Long-Context LLMs and RAG Systems 73 upvotes, #1 of 2024-07-03
- RLHF Workflow: From Reward Modeling to Online RLHF 54 upvotes, #2 of 2024-05-14
- AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning 17 upvotes, #10 of 2024-02-26
- TrustLLM: Trustworthiness in Large Language Models 69 upvotes, #1 of 2024-01-12
- Unlocking Anticipatory Text Generation: A Constrained Approach for Faithful Decoding with Large Language Models 3 upvotes, #13 of 2023-12-12
- Diffusion Model Alignment Using Direct Preference Optimization 48 upvotes, #4 of 2023-11-23
- Lemur: Harmonizing Natural Language and Code for Language Agents 31 upvotes, #3 of 2023-10-13
- XGen-7B Technical Report 8 upvotes, #11 of 2023-09-08
- BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents 20 upvotes, #4 of 2023-08-14
- Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization 21 upvotes, #1 of 2023-08-07
- DialogStudio: Towards Richest and Most Diverse Unified Dataset Collection for Conversational AI 13 upvotes, #6 of 2023-07-20
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.