Trung Bui

Trung Bui on Hugging Face Daily Papers: 21 papers, 1 in the top 3 of their day, 340 upvotes.

  1. FlowTool: Controlling Tool Parameter in Image Retouching via Flow Matching 33 upvotes, #24 of 2026-09-29
  2. DISCO: Distributed Long Context Scaling with Grounding-Reasoning Disaggregation 3 upvotes, #78 of 2026-09-29
  3. A Survey on LLM-based Conversational User Simulation 7 upvotes, #9 of 2026-04-30
  4. SketchVLM: Vision language models can annotate images to explain thoughts and guide users 11 upvotes, #10 of 2026-04-28
  5. PageGuide: Browser extension to assist users in navigating a webpage and locating information 6 upvotes, #17 of 2026-04-28
  6. ViT-AdaLA: Adapting Vision Transformers with Linear Attention 2 upvotes, #33 of 2026-03-18
  7. Can Large Language Models Keep Up? Benchmarking Online Adaptation to Continual Knowledge Streams 17 upvotes, #9 of 2026-03-12
  8. Agentic Planning with Reasoning for Image Styling via Offline RL 3 upvotes, #30 of 2026-03-10
  9. InfinityStory: Unlimited Video Generation with World Consistency and Character-Aware Shot Transitions 8 upvotes, #13 of 2026-03-05
  10. StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos 7 upvotes, #26 of 2025-12-02
  11. MS4UI: A Dataset for Multi-modal Summarization of User Interface Instructional Videos 3 upvotes, #28 of 2025-06-17
  12. Understanding Generative AI Capabilities in Everyday Image Editing Tasks 23 upvotes, #12 of 2025-05-23
  13. CORG: Generating Answers from Complex, Interrelated Contexts 8 upvotes, #6 of 2025-05-05
  14. YoChameleon: Personalized Vision and Language Generation 11 upvotes, #9 of 2025-04-30
  15. NoLiMa: Long-Context Evaluation Beyond Literal Matching 14 upvotes, #14 of 2025-02-13
  16. Toward Robust Hyper-Detailed Image Captioning: A Multiagent Approach and Dual Evaluation Metrics for Factuality and Coverage 13 upvotes, #6 of 2024-12-23
  17. GUI Agents: A Survey 22 upvotes, #5 of 2024-12-19
  18. SlimLM: An Efficient Small Language Model for On-Device Document Assistance 12 upvotes, #7 of 2024-11-19
  19. DynaSaur: Large Language Agents Beyond Predefined Actions 13 upvotes, #13 of 2024-11-05
  20. Taipan: Efficient and Expressive State Space Language Models with Selective Attention 14 upvotes, #10 of 2024-10-25
  21. LRM: Large Reconstruction Model for Single Image to 3D 50 upvotes, #1 of 2023-11-09

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.