tyfeld

tyfeld on Hugging Face Daily Papers: 15 papers, 4 in the top 3 of their day, 540 upvotes.

  1. PerceptionDLM: Parallel Region Perception with Multimodal Diffusion Language Models 64 upvotes, #2 of 2026-06-22
  2. LoomVideo: Unifying Multimodal Inputs into Video Generation and Editing 24 upvotes, #8 of 2026-06-05
  3. Towards Customized Multimodal Role-Play 10 upvotes, #30 of 2026-05-26
  4. SAMTok: Representing Any Mask with Two Words 41 upvotes, #9 of 2026-01-23
  5. MMaDA-Parallel: Multimodal Large Diffusion Language Models for Thinking-Aware Editing and Generation 64 upvotes, #6 of 2025-11-18
  6. Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs 35 upvotes, #9 of 2025-10-22
  7. Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models 53 upvotes, #3 of 2025-09-09
  8. VMoBA: Mixture-of-Block Attention for Video Diffusion Models 31 upvotes, #3 of 2025-07-01
  9. Co-Evolving LLM Coder and Unit Tester via Reinforcement Learning 22 upvotes, #15 of 2025-06-04
  10. MMaDA: Multimodal Large Diffusion Language Models 83 upvotes, #2 of 2025-05-22
  11. Training-free Diffusion Acceleration with Bottleneck Sampling 12 upvotes, #18 of 2025-03-25
  12. Diffusion-Sharpening: Fine-tuning Diffusion Models with Denoising Trajectory Sharpening 15 upvotes, #11 of 2025-02-18
  13. HermesFlow: Seamlessly Closing the Gap in Multimodal Understanding and Generation 16 upvotes, #9 of 2025-02-18
  14. VideoTetris: Towards Compositional Text-to-Video Generation 18 upvotes, #6 of 2024-06-07
  15. RealCompo: Dynamic Equilibrium between Realism and Compositionality Improves Text-to-Image Diffusion Models 9 upvotes, #13 of 2024-02-21

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.