KAIST AI

KAIST AI on Hugging Face Daily Papers: 79 papers, 8 in the top 3 of their day, 2 paper of the day.

  1. Learning What to Recall: Adaptive Multi-Cue Episodic Memory for World Models 11 upvotes, #55 of 2026-10-02
  2. World Observer: Joint Actor-Observer Generation for Persistent World Modeling 76 upvotes, #10 of 2026-10-02
  3. Imagine3D-LLM: Teaching MLLMs to Imagine 3D Scenes Before Answering 59 upvotes, #14 of 2026-10-01
  4. Overcoming Scaling Limits in On-Policy Self-Distillation for LLM Reasoning 7 upvotes, #49 of 2026-10-01
  5. Learning Beyond What You Sample: Off-Policy-Aware Cross-Model Trajectory Exchange for RLVR 62 upvotes, #19 of 2026-09-30
  6. Surprising Success, Repeated Failure: Entropy-Guided Credit Assignment for Exploration in LLM Reasoning 45 upvotes, #14 of 2026-09-29
  7. Knowing When Thinking Is Not Enough: Teaching Small Reasoning Models to Reason Beyond Their Parametric Knowledge 40 upvotes, #17 of 2026-09-29
  8. EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents 39 upvotes, #9 of 2026-09-17
  9. PLC-DPO: Posterior Label Correction in Noisy and Ambiguous Preference Optimization 27 upvotes, #10 of 2026-09-14
  10. SpatialBlock: Enhancing Spatial Intelligence in LVLMs via Synthetic Block-Stacking Problem 137 upvotes, #3 of 2026-09-11
  11. Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation 94 upvotes, #4 of 2026-09-09
  12. Language Models Can Control Their Own Attention 68 upvotes, #7 of 2026-09-03
  13. Super Library Agent: Joint Generation and Maintenance of Multiple Applications Beyond the Single Codebase 27 upvotes, #13 of 2026-09-01
  14. J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data 43 upvotes, #6 of 2026-08-31
  15. PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents 9 upvotes, #18 of 2026-08-21
  16. MBA: Multimodal Benchmark and Agents for Real-World Business Ideation 6 upvotes, #16 of 2026-08-13
  17. Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory 49 upvotes, #7 of 2026-08-11
  18. ReDesign: Recovering Editable Design Structures from Images via Agentic Decomposition 65 upvotes, #4 of 2026-07-29
  19. See like a Robot: Robot-Centric Pointmaps for Vision-Language-Action Models 7 upvotes, #21 of 2026-07-20
  20. 3D HAMSTER: Bridging Planning and Control in Hierarchical Vision Language Action Models through 3D Trajectory Guidance 7 upvotes, #25 of 2026-07-08
  21. LLM-as-a-Tutor: Policy-Aware Prompt Adaptation for Non-Verifiable RL 32 upvotes, #8 of 2026-07-08
  22. PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents 8 upvotes, #27 of 2026-06-30
  23. MVTrack4Gen: Multi-View Point Tracking as Geometric Supervision for 4D Video Generation 35 upvotes, #7 of 2026-06-25
  24. Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views 5 upvotes, #33 of 2026-06-23
  25. TRIAGE: Dialectical Reasoning for Explainable Risk Prediction on Irregularly Sampled Medical Time Series with LLMs 30 upvotes, #8 of 2026-06-17
  26. Who Should Lead Decoding Now? Tracking Reliable Trajectories for Ensembling Masked Diffusion Language Models 33 upvotes, #7 of 2026-06-16
  27. Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization 33 upvotes, #11 of 2026-06-10
  28. TIDE: Proactive Multi-Problem Discovery via Template-Guided Iteration 44 upvotes, #3 of 2026-06-05
  29. Trust Region Q Adjoint Matching 3 upvotes, #34 of 2026-06-05
  30. Mitigating Perceptual Judgment Bias in Multimodal LLM-as-a-Judge via Perceptual Perturbation and Reward Modeling 5 upvotes, #32 of 2026-06-03
  31. OmniRetrieval: Unified Retrieval across Heterogeneous Knowledge Sources 76 upvotes, #3 of 2026-05-29
  32. Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases 7 upvotes, #48 of 2026-05-29
  33. Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents 38 upvotes, #10 of 2026-05-28
  34. Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction 41 upvotes, #5 of 2026-05-27
  35. HINT-SD: Targeted Hindsight Self-Distillation for Long-Horizon Agents 11 upvotes, #18 of 2026-05-25
  36. WorldKV: Efficient World Memory with World Retrieval and Compression 41 upvotes, #10 of 2026-05-22
  37. FlowLong: Inference-time Long Video Generation via Manifold-constrained Tweedie Matching 29 upvotes, #14 of 2026-05-22
  38. It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs 30 upvotes, #10 of 2026-05-21
  39. Nudging Beyond the Comfort Zone: Efficient Strategy-Guided Exploration for RLVR 33 upvotes, #11 of 2026-05-18
  40. PREPING: Building Agent Memory without Tasks 28 upvotes, #13 of 2026-05-15
  41. MEME: Multi-entity & Evolving Memory Evaluation 7 upvotes, #36 of 2026-05-13
  42. CollabVR: Collaborative Video Reasoning with Vision-Language and Video Generation Models 68 upvotes, #3 of 2026-05-12
  43. Towards Autonomous Mechanistic Reasoning in Virtual Cells 6 upvotes, #20 of 2026-04-17
  44. Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents 29 upvotes, #5 of 2026-04-16
  45. Sommelier: Scalable Open Multi-turn Audio Pre-processing for Full-duplex Speech Language Models 35 upvotes, #5 of 2026-03-30
  46. Representation Alignment for Just Image Transformers is not Easier than You Think 13 upvotes, #11 of 2026-03-27
  47. T-MAP: Red-Teaming LLM Agents with Trajectory-aware Evolutionary Search 36 upvotes, #5 of 2026-03-26
  48. DA-Flow: Degradation-Aware Optical Flow Estimation with Diffusion Models 50 upvotes, #5 of 2026-03-25
  49. SNAP: Speaker Nulling for Artifact Projection in Speech Deepfake Detection 3 upvotes, #31 of 2026-03-24
  50. RoboAlign: Learning Test-Time Reasoning for Language-Action Alignment in Vision-Language-Action Models 23 upvotes, #14 of 2026-03-24
  51. SpatialBoost: Enhancing Visual Representation through Language-Guided Reasoning 45 upvotes, #7 of 2026-03-24
  52. Repurposing Geometric Foundation Models for Multi-view Diffusion 45 upvotes, #7 of 2026-03-24
  53. ECG-Reasoning-Benchmark: A Benchmark for Evaluating Clinical Reasoning Capabilities in ECG Interpretation 1 upvotes, #42 of 2026-03-18
  54. MA-EgoQA: Question Answering over Egocentric Videos from Multiple Embodied Agents 28 upvotes, #6 of 2026-03-12
  55. Can Large Language Models Keep Up? Benchmarking Online Adaptation to Continual Knowledge Streams 17 upvotes, #9 of 2026-03-12
  56. MolHIT: Advancing Molecular-Graph Generation with Hierarchical Discrete Diffusion Models 54 upvotes, #1 of 2026-02-26
  57. RoboCurate: Harnessing Diversity with Action-Verified Neural Trajectory for Robot Learning 10 upvotes, #11 of 2026-02-24
  58. THINKSAFE: Self-Generated Safety Alignment for Reasoning Models 38 upvotes, #5 of 2026-02-02
  59. Lost in the Noise: How Reasoning Models Fail with Contextual Distractors 29 upvotes, #6 of 2026-01-13
  60. InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion 95 upvotes, #2 of 2025-12-29
  61. Infinite-Homography as Robust Conditioning for Camera-Controlled Video Generation 27 upvotes, #6 of 2025-12-23
  62. Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure 28 upvotes, #7 of 2025-12-17
  63. Directional Textual Inversion for Personalized Text-to-Image Generation 2 upvotes, #34 of 2025-12-16
  64. EgoX: Egocentric Video Generation from a Single Exocentric Video 106 upvotes, #1 of 2025-12-15
  65. Aligned but Stereotypical? The Hidden Influence of System Prompts on Social Bias in LVLM-Based Text-to-Image Models 7 upvotes, #26 of 2025-12-05
  66. Instruction-Guided Lesion Segmentation for Chest X-rays with Automatically Generated Large-Scale Dataset 25 upvotes, #5 of 2025-11-20
  67. KLASS: KL-Guided Fast Inference in Masked Diffusion Models 35 upvotes, #3 of 2025-11-12
  68. When to Ensemble: Identifying Token-Level Points for Stable and Fast LLM Ensembling 32 upvotes, #7 of 2025-10-21
  69. Temporal Alignment Guidance: On-Manifold Sampling in Diffusion Models 29 upvotes, #11 of 2025-10-15
  70. Multimodal Prompt Optimization: Why Not Leverage Multiple Modalities for MLLMs 46 upvotes, #5 of 2025-10-13
  71. Meta-Awareness Enhances Reasoning Models: Self-Alignment Reinforcement Learning 54 upvotes, #7 of 2025-10-10
  72. Verifier-free Test-Time Sampling for Vision Language Action Models 2 upvotes, #34 of 2025-10-08
  73. No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping 37 upvotes, #9 of 2025-09-29
  74. ReviewScore: Misinformed Peer Review Detection with Large Language Models 62 upvotes, #7 of 2025-09-29
  75. PatientSim: A Persona-Driven Simulator for Realistic Doctor-Patient Interactions 11 upvotes, #29 of 2025-05-30
  76. CXReasonBench: A Benchmark for Evaluating Structured Diagnostic Reasoning in Chest X-rays 6 upvotes, #42 of 2025-05-30
  77. Lunguage: A Benchmark for Structured and Sequential Chest X-ray Interpretation 4 upvotes, #49 of 2025-05-30
  78. SphereDiff: Tuning-free Omnidirectional Panoramic Image and Video Generation via Spherical Latent Representation 27 upvotes, #7 of 2025-04-22
  79. Towards Predicting Temporal Changes in a Patient's Chest X-ray Images based on Electronic Health Records 3 upvotes, #13 of 2024-09-17

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.