Di Zhang

Di Zhang on Hugging Face Daily Papers: 25 papers, 10 in the top 3 of their day, 1,289 upvotes.

  1. Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA 334 upvotes, #2 of 2026-08-11
  2. LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget 197 upvotes, #1 of 2026-07-17
  3. MinT: Managed Infrastructure for Training and Serving Millions of LLMs 216 upvotes, #1 of 2026-05-14
  4. δ-mem: Efficient Online Memory for Large Language Models 119 upvotes, #3 of 2026-05-13
  5. Golden Goose: A Simple Trick to Synthesize Unlimited RLVR Tasks from Unverifiable Internet Text 89 upvotes, #2 of 2026-02-02
  6. AgentDevel: Reframing Self-Evolving LLM Agents as Release Engineering 1 upvotes, #28 of 2026-01-09
  7. Error-Free Linear Attention is a Free Lunch: Exact Solution from Continuous-Time Dynamics 39 upvotes, #9 of 2025-12-16
  8. NVIDIA Nemotron Nano V2 VL 25 upvotes, #5 of 2025-11-07
  9. Chem-R: Learning to Reason as a Chemist 51 upvotes, #7 of 2025-10-22
  10. CMPhysBench: A Benchmark for Evaluating Large Language Models in Condensed Matter Physics 46 upvotes, #3 of 2025-08-27
  11. IAG: Input-aware Backdoor Attack on VLMs for Visual Grounding 7 upvotes, #14 of 2025-08-14
  12. Mol-R1: Towards Explicit Long-CoT Reasoning in Molecule Discovery 37 upvotes, #4 of 2025-08-14
  13. Control-R: Towards controllable test-time scaling 3 upvotes, #35 of 2025-06-04
  14. MOOSE-Chem3: Toward Experiment-Guided Hypothesis Ranking via Simulated Experimental Feedback 30 upvotes, #10 of 2025-05-26
  15. GameFactory: Creating New Games with Generative Interactive Videos 60 upvotes, #1 of 2025-01-21
  16. ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning 15 upvotes, #8 of 2025-01-13
  17. StyleMaster: Stylize Your Video with Artistic Generation and Translation 18 upvotes, #6 of 2024-12-12
  18. SynCamMaster: Synchronizing Multi-Camera Video Generation from Diverse Viewpoints 49 upvotes, #1 of 2024-12-12
  19. 3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation 18 upvotes, #9 of 2024-12-11
  20. Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning 28 upvotes, #1 of 2024-11-29
  21. MolReFlect: Towards In-Context Fine-grained Alignments between Molecules and Texts 3 upvotes, #17 of 2024-11-27
  22. LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning 11 upvotes, #14 of 2024-10-08
  23. Seeing and Understanding: Bridging Vision with Chemical Knowledge Via ChemVLM 19 upvotes, #6 of 2024-08-14
  24. Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B 15 upvotes, #8 of 2024-06-12
  25. ChemLLM: A Chemical Large Language Model 33 upvotes, #3 of 2024-02-13

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.