ByteDance Seed
ByteDance Seed on Hugging Face Daily Papers: 116 papers, 30 in the top 3 of their day, 7 paper of the day.
- Aligning One-Step Generative Models with Reward-Weighted Transport Distillation 2 upvotes, #75 of 2026-10-01
- Periodic Weak Spots: Phase Sensitivity from Chunked KV-Cache Compression 100 upvotes, #12 of 2026-09-30
- Block Sparse Attention with Log-Linear Complexity 30 upvotes, #5 of 2026-09-28
- Towards Full Pipeline FP8 Reinforcement Learning for LLMs 17 upvotes, #18 of 2026-09-22
- Paint-Anything: Unified Any-Color Control for Image Generation and Editing 57 upvotes, #6 of 2026-09-21
- Aspire: Can Models Self-Evolve from Vague Goals? 228 upvotes, #3 of 2026-09-03
- S3Gym: Can LLMs Turn Self-Testing and Self-Judging into Self-Improvement? 39 upvotes, #8 of 2026-09-03
- HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? 264 upvotes, #2 of 2026-09-03
- SMELT: Scaling Laws for Compute-Matched MoE Looped Transformers 99 upvotes, #3 of 2026-09-02
- GenFirst: Generation Before Reconstruction for Stable End-to-End Latent Generative Modeling 67 upvotes, #4 of 2026-09-01
- Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling 116 upvotes, #2 of 2026-09-01
- On-Policy Self-Distillation in Diffusion Models 66 upvotes, #4 of 2026-08-26
- StartupBench: Benchmarking General-Purpose Agents on Market-Validated End-to-End Workflows 9 upvotes, #21 of 2026-08-19
- Scaling Domain Data Repetition in LLM Pretraining 15 upvotes, #14 of 2026-08-17
- Modular TTT: Rethinking Test-Time Training as Composable Modules 8 upvotes, #22 of 2026-08-10
- GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? 46 upvotes, #5 of 2026-08-07
- Scaling Properties of Text Conditioning in Visual Generation 38 upvotes, #7 of 2026-08-03
- FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry 50 upvotes, #8 of 2026-07-21
- Read It Back: Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation 83 upvotes, #3 of 2026-07-15
- OpenCoF: Learning to Reason Through Video Generation 28 upvotes, #7 of 2026-07-10
- UP: Unbounded Positive Asymmetric Optimization for Breaking the Exploration-Stability Dilemma 9 upvotes, #15 of 2026-07-10
- EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments 17 upvotes, #20 of 2026-07-07
- Morphing into Hybrid Attention Models 47 upvotes, #4 of 2026-07-03
- Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity 28 upvotes, #5 of 2026-07-02
- Translation as a Bridging Action: Transferring Manipulation Skills from Humans to Robots 39 upvotes, #4 of 2026-06-29
- DanceOPD: On-Policy Generative Field Distillation 80 upvotes, #1 of 2026-06-26
- Improved Large Language Diffusion Models 43 upvotes, #5 of 2026-06-25
- World Value Models for Robotic Manipulation 7 upvotes, #16 of 2026-06-24
- Dynamic Linear Attention 5 upvotes, #27 of 2026-06-10
- Workflow-GYM: Towards Long-Horizon Evaluation of Computer-use Agentic tasks in Real-World Professional Fields 21 upvotes, #15 of 2026-06-10
- Representation Forcing for Bottleneck-Free Unified Multimodal Models 59 upvotes, #4 of 2026-06-01
- Task-Focused Memorization for Multimodal Agents 38 upvotes, #10 of 2026-06-01
- Negligible in Size, Significant in Effect: On Scale Vectors in Large Language Models 20 upvotes, #15 of 2026-05-27
- Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context 85 upvotes, #4 of 2026-05-14
- Continuous Latent Diffusion Language Model 75 upvotes, #3 of 2026-05-08
- Video Generation with Predictive Latents 24 upvotes, #5 of 2026-05-06
- End-to-End Autoregressive Image Generation with 1D Semantic Tokenizer 11 upvotes, #10 of 2026-05-04
- Leveraging Verifier-Based Reinforcement Learning in Image Editing 27 upvotes, #9 of 2026-05-01
- Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence 80 upvotes, #3 of 2026-04-21
- LeapAlign: Post-Training Flow Matching Models at Any Generation Step by Building Two-Step Trajectories 12 upvotes, #10 of 2026-04-17
- Seedance 2.0: Advancing Video Generation for World Complexity 151 upvotes, #1 of 2026-04-16
- Continuous Adversarial Flow Models 8 upvotes, #25 of 2026-04-14
- UniGRPO: Unified Policy Optimization for Reasoning-Driven Visual Generation 35 upvotes, #8 of 2026-03-25
- SIMART: Decomposing Monolithic Meshes into Sim-ready Articulated Assets via MLLM 40 upvotes, #6 of 2026-03-25
- Mixture-of-Depths Attention 77 upvotes, #7 of 2026-03-17
- Understanding by Reconstruction: Reversing the Software Development Process for LLM Pretraining 8 upvotes, #21 of 2026-03-13
- Learn Hard Problems During RL with Reference Guided Fine-tuning 12 upvotes, #16 of 2026-03-03
- CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel Generation 80 upvotes, #2 of 2026-03-02
- World Guidance: World Modeling in Condition Space for Action Generation 14 upvotes, #10 of 2026-02-26
- When to Memorize and When to Stop: Gated Recurrent Memory for Long-Context Reasoning 28 upvotes, #7 of 2026-02-12
- VideoWorld 2: Learning Transferable Knowledge from Real-world Videos 14 upvotes, #22 of 2026-02-11
- Late-to-Early Training: LET LLMs Learn Earlier, So Faster and Better 7 upvotes, #30 of 2026-02-06
- Retrieval-Infused Reasoning Sandbox: A Benchmark for Decoupling Retrieval and Reasoning Capabilities 19 upvotes, #15 of 2026-02-06
- BABE: Biology Arena BEnchmark 10 upvotes, #27 of 2026-02-06
- Protein Autoregressive Modeling via Multiscale Structure Generation 3 upvotes, #40 of 2026-02-05
- SPARKLING: Balancing Signal Preservation and Symmetry Breaking for Width-Progressive Learning 44 upvotes, #9 of 2026-02-03
- ConceptMoE: Adaptive Token-to-Concept Compression for Implicit Compute Allocation 42 upvotes, #7 of 2026-01-30
- Visual Generation Unlocks Human-Like Reasoning through Multimodal World Models 25 upvotes, #5 of 2026-01-28
- Post-LayerNorm Is Back: Stable, ExpressivE, and Deep 23 upvotes, #7 of 2026-01-28
- Stable-DiffCoder: Pushing the Frontier of Code Diffusion Large Language Model 53 upvotes, #7 of 2026-01-23
- Rethinking Video Generation Model for the Embodied World 42 upvotes, #4 of 2026-01-22
- VLingNav: Embodied Navigation with Adaptive Reasoning and Visual-Assisted Linguistic Memory 7 upvotes, #16 of 2026-01-14
- Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space 54 upvotes, #2 of 2026-01-02
- GR-Dexter Technical Report 21 upvotes, #8 of 2026-01-01
- Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss 93 upvotes, #1 of 2025-12-30
- LLM Swiss Round: Aggregating Multi-Benchmark Performance via Competitive Swiss-System Dynamics 2 upvotes, #16 of 2025-12-25
- SpatialTree: How Spatial Abilities Branch Out in MLLMs 42 upvotes, #5 of 2025-12-24
- Seed-Prover 1.5: Mastering Undergraduate-Level Theorem Proving via Learning from Experience 48 upvotes, #5 of 2025-12-22
- Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model 38 upvotes, #5 of 2025-12-19
- End-to-End Training for Autoregressive Video Diffusion via Self-Resampling 14 upvotes, #16 of 2025-12-18
- UniUGP: Unifying Understanding, Generation, and Planing For End-to-end Autonomous Driving 10 upvotes, #11 of 2025-12-11
- DAComp: Benchmarking Data Agents across the Full Data Intelligence Lifecycle 147 upvotes, #2 of 2025-12-05
- Revisiting the Necessity of Lengthy Chain-of-Thought in Vision-centric Reasoning Generalization 6 upvotes, #26 of 2025-12-03
- GR-RL: Going Dexterous and Precise for Long-Horizon Robotic Manipulation 23 upvotes, #13 of 2025-12-02
- Adversarial Flow Models 20 upvotes, #14 of 2025-12-01
- Virtual Width Networks 34 upvotes, #3 of 2025-11-17
- DiscoX: Benchmarking Discourse-Level Translation task in Expert Domains 4 upvotes, #14 of 2025-11-17
- Depth Anything 3: Recovering the Visual Space from Any Views 77 upvotes, #2 of 2025-11-14
- WMPO: World Model-based Policy Optimization for Vision-Language-Action Models 15 upvotes, #6 of 2025-11-13
- Lumine: An Open Recipe for Building Generalist Agents in 3D Open Worlds 172 upvotes, #1 of 2025-11-13
- Visual Spatial Tuning 46 upvotes, #2 of 2025-11-10
- Learning Vision-Driven Reactive Soccer Skills for Humanoid Robots 3 upvotes, #13 of 2025-11-07
- MME-CC: A Challenging Multi-Modal Evaluation Benchmark of Cognitive Capacity 7 upvotes, #9 of 2025-11-06
- When Visualizing is the First Step to Reasoning: MIRA, a Benchmark for Visual Chain-of-Thought 53 upvotes, #3 of 2025-11-05
- INT v.s. FP: A Comprehensive Study of Fine-Grained Low-bit Quantization Formats 66 upvotes, #3 of 2025-11-03
- Scaling Latent Reasoning via Looped Language Models 201 upvotes, #1 of 2025-10-30
- Parallel Loop Transformer for Efficient Test-Time Computation Scaling 14 upvotes, #14 of 2025-10-30
- From Spatial to Actions: Grounding Vision-Language-Action Model in Spatial Foundation Priors 25 upvotes, #7 of 2025-10-29
- Seed3D 1.0: From Images to High-Fidelity Simulation-Ready 3D Assets 17 upvotes, #9 of 2025-10-24
- Any-Depth Alignment: Unlocking Innate Safety Alignment of LLMs to Any-Depth 2 upvotes, #32 of 2025-10-22
- Beyond Correctness: Evaluating Subjective Writing Preferences Across Cultures 10 upvotes, #23 of 2025-10-17
- Trace Anything: Representing Any Video in 4D via Trajectory Fields 30 upvotes, #7 of 2025-10-16
- Generative Universal Verifier as Multimodal Meta-Reasoner 24 upvotes, #12 of 2025-10-16
- Scaling Long-Horizon LLM Agent via Context-Folding 3 upvotes, #32 of 2025-10-15
- Memory Retrieval and Consolidation in Large Language Models through Function Tokens 7 upvotes, #33 of 2025-10-10
- Artificial Hippocampus Networks for Efficient Long-Context Modeling 26 upvotes, #10 of 2025-10-09
- Heptapod: Language Modeling on Visual Signals 3 upvotes, #30 of 2025-10-09
- Self-Forcing++: Towards Minute-Scale High-Quality Video Generation 86 upvotes, #2 of 2025-10-03
- Knapsack RL: Unlocking Exploration of LLMs via Optimizing Budget Allocation 43 upvotes, #5 of 2025-10-02
- Inverse IFEval: Can LLMs Unlearn Stubborn Training Conventions to Follow Real Instructions? 54 upvotes, #4 of 2025-09-05
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.