Yushi Bai
Yushi Bai on Hugging Face Daily Papers: 22 papers, 7 in the top 3 of their day, 1,059 upvotes.
- IndexCache: Accelerating Sparse Attention via Cross-Layer Index Reuse 51 upvotes, #3 of 2026-03-13
- GLM-5: from Vibe Coding to Agentic Engineering 94 upvotes, #1 of 2026-02-18
- Glyph: Scaling Context Windows via Visual-Text Compression 61 upvotes, #2 of 2025-10-21
- DeepPrune: Parallel Scaling without Inter-trace Redundancy 23 upvotes, #17 of 2025-10-10
- SIRI: Scaling Iterative Reinforcement Learning with Interleaved Compression 12 upvotes, #34 of 2025-09-30
- CHARM: Control-point-based 3D Anime Hairstyle Auto-Regressive Modeling 15 upvotes, #12 of 2025-09-26
- GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models 144 upvotes, #1 of 2025-08-11
- LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning 51 upvotes, #3 of 2025-06-24
- SuperWriter: Reflection-Driven Long-Form Generation with Large Language Models 32 upvotes, #5 of 2025-06-05
- Hard Negative Contrastive Learning for Fine-Grained Geometric Understanding in Large Multimodal Models 11 upvotes, #34 of 2025-05-27
- An LMM for Efficient Video Understanding via Reinforced Compression of Video Cubes 10 upvotes, #17 of 2025-04-22
- Shifting Long-Context LLMs Research from Input to Output 19 upvotes, #14 of 2025-03-14
- LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models 23 upvotes, #9 of 2025-02-21
- LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks 31 upvotes, #5 of 2024-12-20
- AlphaTablets: A Generic Plane Representation for 3D Planar Reconstruction from Monocular Videos 5 upvotes, #18 of 2024-12-02
- Pre-training Distillation for Large Language Models: A Design Space Exploration 15 upvotes, #11 of 2024-10-22
- LongCite: Enabling LLMs to Generate Fine-grained Citations in Long-context QA 41 upvotes, #3 of 2024-09-05
- LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs 60 upvotes, #1 of 2024-08-14
- ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools 26 upvotes, #6 of 2024-06-19
- CogCoM: Train Large Vision-Language Models Diving into Details through Chain of Manipulations 9 upvotes, #9 of 2024-02-07
- LongAlign: A Recipe for Long Context Alignment of Large Language Models 25 upvotes, #4 of 2024-02-01
- KoLA: Carefully Benchmarking World Knowledge of Large Language Models 20 upvotes, #6 of 2023-06-16
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.