Jungang Li
Jungang Li on Hugging Face Daily Papers: 13 papers, 2 in the top 3 of their day, 1,031 upvotes.
- VTR-Bench: A Systematic Benchmark for Evaluating Visual Text Rendering in Video Generation 11 upvotes, #56 of 2026-10-02
- Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation 695 upvotes, #1 of 2026-09-15
- Vidu S1: A Real-Time Interactive Video Generation Model 138 upvotes, #1 of 2026-07-10
- CM-EVS: Sparse Panoramic RGB-D-Pose Data for Complete Scene Coverage 11 upvotes, #20 of 2026-05-18
- Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization 12 upvotes, #22 of 2026-04-14
- AndroTMem: From Interaction Trajectories to Anchored Memory in Long-Horizon GUI Agents 26 upvotes, #12 of 2026-03-20
- Temporal Gains, Spatial Costs: Revisiting Video Fine-Tuning in Multimodal Large Language Models 20 upvotes, #13 of 2026-03-19
- BPDQ: Bit-Plane Decomposition Quantization on a Variable Grid for Large Language Models 9 upvotes, #13 of 2026-02-16
- OmniSIFT: Modality-Asymmetric Token Compression for Efficient Omni-modal Large Language Models 46 upvotes, #5 of 2026-02-05
- JavisGPT: A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation 18 upvotes, #9 of 2026-01-01
- MOSS-ChatV: Reinforcement Learning with Process Reasoning Reward for Video Temporal Reasoning 4 upvotes, #27 of 2025-09-26
- Mind the Third Eye! Benchmarking Privacy Awareness in MLLM-powered Smartphone Agents 11 upvotes, #12 of 2025-08-28
- Unveiling Instruction-Specific Neurons & Experts: An Analytical Framework for LLM's Instruction-Following Capabilities 2 upvotes, #50 of 2025-05-29
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.