Jintao Zhang

Jintao Zhang on Hugging Face Daily Papers: 25 papers, 12 in the top 3 of their day, 1,753 upvotes.

  1. Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation 695 upvotes, #1 of 2026-09-15
  2. Vidu S1: A Real-Time Interactive Video Generation Model 138 upvotes, #1 of 2026-07-10
  3. TurboServe: Serving Streaming Video Generation Efficiently and Economically 34 upvotes, #2 of 2026-07-02
  4. KernelBench-X: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernels 7 upvotes, #23 of 2026-05-08
  5. Speculative Decoding for Autoregressive Video Generation 10 upvotes, #16 of 2026-04-22
  6. 6Bit-Diffusion: Inference-Time Mixed-Precision Quantization for Video Diffusion Models 10 upvotes, #13 of 2026-03-26
  7. HybridStitch: Pixel and Timestep Level Model Stitching for Diffusion Acceleration 10 upvotes, #15 of 2026-03-16
  8. SVG-EAR: Parameter-Free Linear Compensation for Sparse Video Generation via Error-aware Routing 15 upvotes, #10 of 2026-03-12
  9. Flash-KMeans: Fast and Memory-Efficient Exact K-Means 78 upvotes, #3 of 2026-03-12
  10. SageBwd: A Trainable Low-bit Attention 17 upvotes, #10 of 2026-03-06
  11. SpargeAttention2: Trainable Sparse Attention via Hybrid Top-k+Top-p Masking and Distillation Fine-Tuning 43 upvotes, #3 of 2026-02-20
  12. SLA2: Sparse-Linear Attention with Learnable Routing and QAT 52 upvotes, #1 of 2026-02-19
  13. Quant VideoGen: Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization 33 upvotes, #8 of 2026-02-05
  14. Residual Context Diffusion Language Models 31 upvotes, #10 of 2026-02-05
  15. TurboDiffusion: Accelerating Video Diffusion Models by 100-200 Times 88 upvotes, #1 of 2025-12-25
  16. Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency 8 upvotes, #30 of 2025-10-10
  17. SLA: Beyond Sparsity in Diffusion Transformers via Fine-Tunable Sparse-Linear Attention 109 upvotes, #1 of 2025-09-30
  18. SageAttention2++: A More Efficient Implementation of SageAttention2 41 upvotes, #8 of 2025-05-29
  19. Sparse VideoGen2: Accelerate Video Generation with Sparse Attention via Semantic-Aware Permutation 38 upvotes, #12 of 2025-05-28
  20. SageAttention3: Microscaling FP4 Attention for Inference and An Exploration of 8-Bit Training 58 upvotes, #2 of 2025-05-21
  21. SAGE: A Framework of Precise Retrieval for RAG 5 upvotes, #21 of 2025-03-10
  22. Identifying Sensitive Weights via Post-quantization Integral 7 upvotes, #15 of 2025-03-07
  23. SpargeAttn: Accurate Sparse Attention Accelerating Any Model Inference 50 upvotes, #3 of 2025-02-26
  24. SageAttention2 Technical Report: Accurate 4 Bit Attention for Plug-and-play Inference Acceleration 47 upvotes, #1 of 2024-11-21
  25. SageAttention: Accurate 8-Bit Attention for Plug-and-play Inference Acceleration 44 upvotes, #2 of 2024-10-04

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.