Zhengzhong Tu

Zhengzhong Tu on Hugging Face Daily Papers: 14 papers, 3 in the top 3 of their day, 319 upvotes.

  1. Digital Twin AI: Opportunities and Challenges from Large Language Models to World Models 17 upvotes, #11 of 2026-01-07
  2. MMHU: A Massive-Scale Multimodal Benchmark for Human Behavior Understanding 23 upvotes, #5 of 2025-07-17
  3. 4KAgent: Agentic Any Image to 4K Super-Resolution 81 upvotes, #1 of 2025-07-10
  4. Demystifying the Visual Quality Paradox in Multimodal Large Language Models 4 upvotes, #30 of 2025-06-24
  5. SAFEFLOW: A Principled Protocol for Trustworthy and Transactional Autonomous Agent Systems 6 upvotes, #27 of 2025-06-10
  6. MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning 20 upvotes, #12 of 2025-06-02
  7. DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models 25 upvotes, #8 of 2025-06-02
  8. VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction 4 upvotes, #53 of 2025-05-28
  9. Can Large Vision Language Models Read Maps Like a Human? 9 upvotes, #14 of 2025-03-24
  10. On the Trustworthiness of Generative Foundation Models: Guideline, Assessment, and Perspective 44 upvotes, #2 of 2025-02-20
  11. Edit Away and My Face Will not Stay: Personal Biometric Defense against Malicious Generative Editing 2 upvotes, #22 of 2024-11-28
  12. Bigger is not Always Better: Scaling Properties of Latent Diffusion Models 17 upvotes, #5 of 2024-04-03
  13. TIP: Text-Driven Image Processing with Semantic and Restoration Instructions 6 upvotes, #12 of 2023-12-20
  14. Conditional Diffusion Distillation 19 upvotes, #3 of 2023-10-03

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.