Zhang
Zhang on Hugging Face Daily Papers: 7 papers, 3 in the top 3 of their day, 477 upvotes.
- Chain-of-Model Learning for Language Model 107 upvotes, #1 of 2025-05-20
- MMInference: Accelerating Pre-filling for Long-Context VLMs via Modality-Aware Permutation Sparse Attention 8 upvotes, #8 of 2025-04-29
- SCBench: A KV Cache-Centric Analysis of Long-Context Methods 8 upvotes, #13 of 2024-12-16
- RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval 28 upvotes, #3 of 2024-09-17
- MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention 22 upvotes, #4 of 2024-07-03
- Parrot: Efficient Serving of LLM-based Applications with Semantic Variable 3 upvotes, #10 of 2024-05-31
- Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone 222 upvotes, #1 of 2024-04-23
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.