ByteDance Seed

ByteDance Seed on Hugging Face Daily Papers: 116 papers, 30 in the top 3 of their day, 7 paper of the day.

  1. Aligning One-Step Generative Models with Reward-Weighted Transport Distillation 2 upvotes, #75 of 2026-10-01
  2. Periodic Weak Spots: Phase Sensitivity from Chunked KV-Cache Compression 100 upvotes, #12 of 2026-09-30
  3. Block Sparse Attention with Log-Linear Complexity 30 upvotes, #5 of 2026-09-28
  4. Towards Full Pipeline FP8 Reinforcement Learning for LLMs 17 upvotes, #18 of 2026-09-22
  5. Paint-Anything: Unified Any-Color Control for Image Generation and Editing 57 upvotes, #6 of 2026-09-21
  6. Aspire: Can Models Self-Evolve from Vague Goals? 228 upvotes, #3 of 2026-09-03
  7. S3Gym: Can LLMs Turn Self-Testing and Self-Judging into Self-Improvement? 39 upvotes, #8 of 2026-09-03
  8. HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? 264 upvotes, #2 of 2026-09-03
  9. SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers 99 upvotes, #3 of 2026-09-02
  10. GenFirst: Generation Before Reconstruction for Stable End-to-End Latent Generative Modeling 67 upvotes, #4 of 2026-09-01
  11. Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling 116 upvotes, #2 of 2026-09-01
  12. On-Policy Self-Distillation in Diffusion Models 66 upvotes, #4 of 2026-08-26
  13. StartupBench: Benchmarking General-Purpose Agents on Market-Validated End-to-End Workflows 9 upvotes, #21 of 2026-08-19
  14. Scaling Domain Data Repetition in LLM Pretraining 15 upvotes, #14 of 2026-08-17
  15. Modular TTT: Rethinking Test-Time Training as Composable Modules 8 upvotes, #22 of 2026-08-10
  16. GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? 46 upvotes, #5 of 2026-08-07
  17. Scaling Properties of Text Conditioning in Visual Generation 38 upvotes, #7 of 2026-08-03
  18. FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry 50 upvotes, #8 of 2026-07-21
  19. Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation 83 upvotes, #3 of 2026-07-15
  20. OpenCoF: Learning to Reason Through Video Generation 28 upvotes, #7 of 2026-07-10
  21. UP: Unbounded Positive Asymmetric Optimization for Breaking the Exploration-Stability Dilemma 9 upvotes, #15 of 2026-07-10
  22. EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments 17 upvotes, #20 of 2026-07-07
  23. Morphing into Hybrid Attention Models 47 upvotes, #4 of 2026-07-03
  24. Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity 28 upvotes, #5 of 2026-07-02
  25. Translation as a Bridging Action: Transferring Manipulation Skills from Humans to Robots 39 upvotes, #4 of 2026-06-29
  26. DanceOPD: On-Policy Generative Field Distillation 80 upvotes, #1 of 2026-06-26
  27. Improved Large Language Diffusion Models 43 upvotes, #5 of 2026-06-25
  28. World Value Models for Robotic Manipulation 7 upvotes, #16 of 2026-06-24
  29. Dynamic Linear Attention 5 upvotes, #27 of 2026-06-10
  30. Workflow-GYM: Towards Long-Horizon Evaluation of Computer-use Agentic tasks in Real-World Professional Fields 21 upvotes, #15 of 2026-06-10
  31. Representation Forcing for Bottleneck-Free Unified Multimodal Models 59 upvotes, #4 of 2026-06-01
  32. Task-Focused Memorization for Multimodal Agents 38 upvotes, #10 of 2026-06-01
  33. Negligible in Size, Significant in Effect: On Scale Vectors in Large Language Models 20 upvotes, #15 of 2026-05-27
  34. Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context 85 upvotes, #4 of 2026-05-14
  35. Continuous Latent Diffusion Language Model 75 upvotes, #3 of 2026-05-08
  36. Video Generation with Predictive Latents 24 upvotes, #5 of 2026-05-06
  37. End-to-End Autoregressive Image Generation with 1D Semantic Tokenizer 11 upvotes, #10 of 2026-05-04
  38. Leveraging Verifier-Based Reinforcement Learning in Image Editing 27 upvotes, #9 of 2026-05-01
  39. Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence 80 upvotes, #3 of 2026-04-21
  40. LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories 12 upvotes, #10 of 2026-04-17
  41. Seedance 2.0: Advancing Video Generation for World Complexity 151 upvotes, #1 of 2026-04-16
  42. Continuous Adversarial Flow Models 8 upvotes, #25 of 2026-04-14
  43. UniGRPO: Unified Policy Optimization for Reasoning-Driven Visual Generation 35 upvotes, #8 of 2026-03-25
  44. SIMART: Decomposing Monolithic Meshes into Sim-ready Articulated Assets via MLLM 40 upvotes, #6 of 2026-03-25
  45. Mixture-of-Depths Attention 77 upvotes, #7 of 2026-03-17
  46. Understanding by Reconstruction: Reversing the Software Development Process for LLM Pretraining 8 upvotes, #21 of 2026-03-13
  47. Learn Hard Problems During RL with Reference Guided Fine-tuning 12 upvotes, #16 of 2026-03-03
  48. CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel Generation 80 upvotes, #2 of 2026-03-02
  49. World Guidance: World Modeling in Condition Space for Action Generation 14 upvotes, #10 of 2026-02-26
  50. When to Memorize and When to Stop: Gated Recurrent Memory for Long-Context Reasoning 28 upvotes, #7 of 2026-02-12
  51. VideoWorld 2: Learning Transferable Knowledge from Real-world Videos 14 upvotes, #22 of 2026-02-11
  52. Late-to-Early Training: LET LLMs Learn Earlier, So Faster and Better 7 upvotes, #30 of 2026-02-06
  53. Retrieval-Infused Reasoning Sandbox: A Benchmark for Decoupling Retrieval and Reasoning Capabilities 19 upvotes, #15 of 2026-02-06
  54. BABE: Biology Arena BEnchmark 10 upvotes, #27 of 2026-02-06
  55. Protein Autoregressive Modeling via Multiscale Structure Generation 3 upvotes, #40 of 2026-02-05
  56. SPARKLING: Balancing Signal Preservation and Symmetry Breaking for Width-Progressive Learning 44 upvotes, #9 of 2026-02-03
  57. ConceptMoE: Adaptive Token-to-Concept Compression for Implicit Compute Allocation 42 upvotes, #7 of 2026-01-30
  58. Visual Generation Unlocks Human-Like Reasoning through Multimodal World Models 25 upvotes, #5 of 2026-01-28
  59. Post-LayerNorm Is Back: Stable, ExpressivE, and Deep 23 upvotes, #7 of 2026-01-28
  60. Stable-DiffCoder: Pushing the Frontier of Code Diffusion Large Language Model 53 upvotes, #7 of 2026-01-23
  61. Rethinking Video Generation Model for the Embodied World 42 upvotes, #4 of 2026-01-22
  62. VLingNav: Embodied Navigation with Adaptive Reasoning and Visual-Assisted Linguistic Memory 7 upvotes, #16 of 2026-01-14
  63. Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space 54 upvotes, #2 of 2026-01-02
  64. GR-Dexter Technical Report 21 upvotes, #8 of 2026-01-01
  65. Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss 93 upvotes, #1 of 2025-12-30
  66. LLM Swiss Round: Aggregating Multi-Benchmark Performance via Competitive Swiss-System Dynamics 2 upvotes, #16 of 2025-12-25
  67. SpatialTree: How Spatial Abilities Branch Out in MLLMs 42 upvotes, #5 of 2025-12-24
  68. Seed-Prover 1.5: Mastering Undergraduate-Level Theorem Proving via Learning from Experience 48 upvotes, #5 of 2025-12-22
  69. Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model 38 upvotes, #5 of 2025-12-19
  70. End-to-End Training for Autoregressive Video Diffusion via Self-Resampling 14 upvotes, #16 of 2025-12-18
  71. UniUGP: Unifying Understanding, Generation, and Planing For End-to-end Autonomous Driving 10 upvotes, #11 of 2025-12-11
  72. DAComp: Benchmarking Data Agents across the Full Data Intelligence Lifecycle 147 upvotes, #2 of 2025-12-05
  73. Revisiting the Necessity of Lengthy Chain-of-Thought in Vision-centric Reasoning Generalization 6 upvotes, #26 of 2025-12-03
  74. GR-RL: Going Dexterous and Precise for Long-Horizon Robotic Manipulation 23 upvotes, #13 of 2025-12-02
  75. Adversarial Flow Models 20 upvotes, #14 of 2025-12-01
  76. Virtual Width Networks 34 upvotes, #3 of 2025-11-17
  77. DiscoX: Benchmarking Discourse-Level Translation task in Expert Domains 4 upvotes, #14 of 2025-11-17
  78. Depth Anything 3: Recovering the Visual Space from Any Views 77 upvotes, #2 of 2025-11-14
  79. WMPO: World Model-based Policy Optimization for Vision-Language-Action Models 15 upvotes, #6 of 2025-11-13
  80. Lumine: An Open Recipe for Building Generalist Agents in 3D Open Worlds 172 upvotes, #1 of 2025-11-13
  81. Visual Spatial Tuning 46 upvotes, #2 of 2025-11-10
  82. Learning Vision-Driven Reactive Soccer Skills for Humanoid Robots 3 upvotes, #13 of 2025-11-07
  83. MME-CC: A Challenging Multi-Modal Evaluation Benchmark of Cognitive Capacity 7 upvotes, #9 of 2025-11-06
  84. When Visualizing is the First Step to Reasoning: MIRA, a Benchmark for Visual Chain-of-Thought 53 upvotes, #3 of 2025-11-05
  85. INT v.s. FP: A Comprehensive Study of Fine-Grained Low-bit Quantization Formats 66 upvotes, #3 of 2025-11-03
  86. Scaling Latent Reasoning via Looped Language Models 201 upvotes, #1 of 2025-10-30
  87. Parallel Loop Transformer for Efficient Test-Time Computation Scaling 14 upvotes, #14 of 2025-10-30
  88. From Spatial to Actions: Grounding Vision-Language-Action Model in Spatial Foundation Priors 25 upvotes, #7 of 2025-10-29
  89. Seed3D 1.0: From Images to High-Fidelity Simulation-Ready 3D Assets 17 upvotes, #9 of 2025-10-24
  90. Any-Depth Alignment: Unlocking Innate Safety Alignment of LLMs to Any-Depth 2 upvotes, #32 of 2025-10-22
  91. Beyond Correctness: Evaluating Subjective Writing Preferences Across Cultures 10 upvotes, #23 of 2025-10-17
  92. Trace Anything: Representing Any Video in 4D via Trajectory Fields 30 upvotes, #7 of 2025-10-16
  93. Generative Universal Verifier as Multimodal Meta-Reasoner 24 upvotes, #12 of 2025-10-16
  94. Scaling Long-Horizon LLM Agent via Context-Folding 3 upvotes, #32 of 2025-10-15
  95. Memory Retrieval and Consolidation in Large Language Models through Function Tokens 7 upvotes, #33 of 2025-10-10
  96. Artificial Hippocampus Networks for Efficient Long-Context Modeling 26 upvotes, #10 of 2025-10-09
  97. Heptapod: Language Modeling on Visual Signals 3 upvotes, #30 of 2025-10-09
  98. Self-Forcing++: Towards Minute-Scale High-Quality Video Generation 86 upvotes, #2 of 2025-10-03
  99. Knapsack RL: Unlocking Exploration of LLMs via Optimizing Budget Allocation 43 upvotes, #5 of 2025-10-02
  100. Inverse IFEval: Can LLMs Unlearn Stubborn Training Conventions to Follow Real Instructions? 54 upvotes, #4 of 2025-09-05

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.