inclusionAI
inclusionAI on Hugging Face Daily Papers: 26 papers, 8 in the top 3 of their day, 3 paper of the day.
- RULER: Instance-aware Rubric Rewards for SVG Generation 101 upvotes, #2 of 2026-09-23
- LLaDA-UI: Bringing Block-wise Diffusion to Vision-Language GUI Agents 16 upvotes, #17 of 2026-09-15
- Realtime-Venus: A full-duplex interaction system with asynchronous delegation 3 upvotes, #30 of 2026-09-15
- LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes 233 upvotes, #3 of 2026-09-04
- Unlocking the Potential of Image Editing via Concept Scaling and Dense Supervision 51 upvotes, #4 of 2026-08-25
- Open-AoE: An Open Egocentric Manipulation Dataset and Toolchain for Embodied Learning 68 upvotes, #6 of 2026-07-21
- SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning 15 upvotes, #8 of 2026-06-29
- Rethinking Shrinkage Bias in LLM FP4 Pretraining: Geometric Origin, Systemic Impact, and UFP4 Recipe 9 upvotes, #18 of 2026-06-19
- RODS: Reward-Driven Online Data Synthesis for Multi-Turn Tool-Use Agents 4 upvotes, #24 of 2026-06-18
- Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale 80 upvotes, #6 of 2026-06-16
- MemDreamer: Decoupling Perception and Reasoning for Long Video Understanding via Hierarchical Graph Memory and Agentic Retrieval Mechanism 38 upvotes, #10 of 2026-06-10
- DR-Venus: Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data 49 upvotes, #3 of 2026-04-23
- LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model 232 upvotes, #1 of 2026-04-23
- Zooming without Zooming: Region-to-Image Distillation for Fine-Grained Multimodal Perception 58 upvotes, #4 of 2026-02-16
- UI-Venus-1.5 Technical Report 149 upvotes, #3 of 2026-02-11
- LLaDA2.1: Speeding Up Text Diffusion via Token Editing 66 upvotes, #9 of 2026-02-10
- VenusBench-GD: A Comprehensive Multi-Platform GUI Benchmark for Diverse Grounding Tasks 8 upvotes, #21 of 2025-12-19
- TwinFlow: Realizing One-step Generation on Large Models with Self-adversarial Flows 69 upvotes, #1 of 2025-12-08
- Every Activation Boosted: Scaling General Reasoner to 1 Trillion Open Language Foundation 81 upvotes, #1 of 2025-11-04
- Ming-Flash-Omni: A Sparse, Unified Architecture for Multimodal Perception and Generation 34 upvotes, #8 of 2025-10-30
- FunReason-MT Technical Report: Overcoming the Complexity Barrier in Multi-Turn Function Calling 5 upvotes, #29 of 2025-10-29
- ARGenSeg: Image Segmentation with Autoregressive Image Generation Model 8 upvotes, #15 of 2025-10-24
- Every Step Evolves: Scaling Reinforcement Learning for Trillion-Scale Thinking Model 61 upvotes, #6 of 2025-10-22
- Don't Just Fine-tune the Agent, Tune the Environment 26 upvotes, #12 of 2025-10-14
- Ming-UniVision: Joint Image Understanding and Generation with a Unified Continuous Tokenizer 69 upvotes, #2 of 2025-10-09
- MultiEdit: Advancing Instruction-based Image Editing on Diverse and Challenging Tasks 11 upvotes, #10 of 2025-09-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.