tyfeld
tyfeld on Hugging Face Daily Papers: 15 papers, 4 in the top 3 of their day, 540 upvotes.
- PerceptionDLM: Parallel Region Perception with Multimodal Diffusion Language Models 64 upvotes, #2 of 2026-06-22
- LoomVideo: Unifying Multimodal Inputs into Video Generation and Editing 24 upvotes, #8 of 2026-06-05
- Towards Customized Multimodal Role-Play 10 upvotes, #30 of 2026-05-26
- SAMTok: Representing Any Mask with Two Words 41 upvotes, #9 of 2026-01-23
- MMaDA-Parallel: Multimodal Large Diffusion Language Models for Thinking-Aware Editing and Generation 64 upvotes, #6 of 2025-11-18
- Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs 35 upvotes, #9 of 2025-10-22
- Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models 53 upvotes, #3 of 2025-09-09
- VMoBA: Mixture-of-Block Attention for Video Diffusion Models 31 upvotes, #3 of 2025-07-01
- Co-Evolving LLM Coder and Unit Tester via Reinforcement Learning 22 upvotes, #15 of 2025-06-04
- MMaDA: Multimodal Large Diffusion Language Models 83 upvotes, #2 of 2025-05-22
- Training-free Diffusion Acceleration with Bottleneck Sampling 12 upvotes, #18 of 2025-03-25
- Diffusion-Sharpening: Fine-tuning Diffusion Models with Denoising Trajectory Sharpening 15 upvotes, #11 of 2025-02-18
- HermesFlow: Seamlessly Closing the Gap in Multimodal Understanding and Generation 16 upvotes, #9 of 2025-02-18
- VideoTetris: Towards Compositional Text-to-Video Generation 18 upvotes, #6 of 2024-06-07
- RealCompo: Dynamic Equilibrium between Realism and Compositionality Improves Text-to-Image Diffusion Models 9 upvotes, #13 of 2024-02-21
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.