Di Zhang
Di Zhang on Hugging Face Daily Papers: 25 papers, 10 in the top 3 of their day, 1,289 upvotes.
- Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA 334 upvotes, #2 of 2026-08-11
- LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget 197 upvotes, #1 of 2026-07-17
- MinT: Managed Infrastructure for Training and Serving Millions of LLMs 216 upvotes, #1 of 2026-05-14
- δ-mem: Efficient Online Memory for Large Language Models 119 upvotes, #3 of 2026-05-13
- Golden Goose: A Simple Trick to Synthesize Unlimited RLVR Tasks from Unverifiable Internet Text 89 upvotes, #2 of 2026-02-02
- AgentDevel: Reframing Self-Evolving LLM Agents as Release Engineering 1 upvotes, #28 of 2026-01-09
- Error-Free Linear Attention is a Free Lunch: Exact Solution from Continuous-Time Dynamics 39 upvotes, #9 of 2025-12-16
- NVIDIA Nemotron Nano V2 VL 25 upvotes, #5 of 2025-11-07
- Chem-R: Learning to Reason as a Chemist 51 upvotes, #7 of 2025-10-22
- CMPhysBench: A Benchmark for Evaluating Large Language Models in Condensed Matter Physics 46 upvotes, #3 of 2025-08-27
- IAG: Input-aware Backdoor Attack on VLMs for Visual Grounding 7 upvotes, #14 of 2025-08-14
- Mol-R1: Towards Explicit Long-CoT Reasoning in Molecule Discovery 37 upvotes, #4 of 2025-08-14
- Control-R: Towards controllable test-time scaling 3 upvotes, #35 of 2025-06-04
- MOOSE-Chem3: Toward Experiment-Guided Hypothesis Ranking via Simulated Experimental Feedback 30 upvotes, #10 of 2025-05-26
- GameFactory: Creating New Games with Generative Interactive Videos 60 upvotes, #1 of 2025-01-21
- ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning 15 upvotes, #8 of 2025-01-13
- StyleMaster: Stylize Your Video with Artistic Generation and Translation 18 upvotes, #6 of 2024-12-12
- SynCamMaster: Synchronizing Multi-Camera Video Generation from Diverse Viewpoints 49 upvotes, #1 of 2024-12-12
- 3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation 18 upvotes, #9 of 2024-12-11
- Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning 28 upvotes, #1 of 2024-11-29
- MolReFlect: Towards In-Context Fine-grained Alignments between Molecules and Texts 3 upvotes, #17 of 2024-11-27
- LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning 11 upvotes, #14 of 2024-10-08
- Seeing and Understanding: Bridging Vision with Chemical Knowledge Via ChemVLM 19 upvotes, #6 of 2024-08-14
- Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B 15 upvotes, #8 of 2024-06-12
- ChemLLM: A Chemical Large Language Model 33 upvotes, #3 of 2024-02-13
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.