KAIST AI
KAIST AI on Hugging Face Daily Papers: 79 papers, 8 in the top 3 of their day, 2 paper of the day.
- Learning What to Recall: Adaptive Multi-Cue Episodic Memory for World Models 11 upvotes, #55 of 2026-10-02
- World Observer: Joint Actor-Observer Generation for Persistent World Modeling 76 upvotes, #10 of 2026-10-02
- Imagine3D-LLM: Teaching MLLMs to Imagine 3D Scenes Before Answering 59 upvotes, #14 of 2026-10-01
- Overcoming Scaling Limits in On-Policy Self-Distillation for LLM Reasoning 7 upvotes, #49 of 2026-10-01
- Learning Beyond What You Sample: Off-Policy-Aware Cross-Model Trajectory Exchange for RLVR 62 upvotes, #19 of 2026-09-30
- Surprising Success, Repeated Failure: Entropy-Guided Credit Assignment for Exploration in LLM Reasoning 45 upvotes, #14 of 2026-09-29
- Knowing When Thinking Is Not Enough: Teaching Small Reasoning Models to Reason Beyond Their Parametric Knowledge 40 upvotes, #17 of 2026-09-29
- EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents 39 upvotes, #9 of 2026-09-17
- PLC-DPO: Posterior Label Correction in Noisy and Ambiguous Preference Optimization 27 upvotes, #10 of 2026-09-14
- SpatialBlock: Enhancing Spatial Intelligence in LVLMs via Synthetic Block-Stacking Problem 137 upvotes, #3 of 2026-09-11
- Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation 94 upvotes, #4 of 2026-09-09
- Language Models Can Control Their Own Attention 68 upvotes, #7 of 2026-09-03
- Super Library Agent: Joint Generation and Maintenance of Multiple Applications Beyond the Single Codebase 27 upvotes, #13 of 2026-09-01
- J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data 43 upvotes, #6 of 2026-08-31
- PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents 9 upvotes, #18 of 2026-08-21
- MBA: Multimodal Benchmark and Agents for Real-World Business Ideation 6 upvotes, #16 of 2026-08-13
- Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory 49 upvotes, #7 of 2026-08-11
- ReDesign: Recovering Editable Design Structures from Images via Agentic Decomposition 65 upvotes, #4 of 2026-07-29
- See like a Robot: Robot-Centric Pointmaps for Vision-Language-Action Models 7 upvotes, #21 of 2026-07-20
- 3D HAMSTER: Bridging Planning and Control in Hierarchical Vision Language Action Models through 3D Trajectory Guidance 7 upvotes, #25 of 2026-07-08
- LLM-as-a-Tutor: Policy-Aware Prompt Adaptation for Non-Verifiable RL 32 upvotes, #8 of 2026-07-08
- PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents 8 upvotes, #27 of 2026-06-30
- MVTrack4Gen: Multi-View Point Tracking as Geometric Supervision for 4D Video Generation 35 upvotes, #7 of 2026-06-25
- Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views 5 upvotes, #33 of 2026-06-23
- TRIAGE: Dialectical Reasoning for Explainable Risk Prediction on Irregularly Sampled Medical Time Series with LLMs 30 upvotes, #8 of 2026-06-17
- Who Should Lead Decoding Now? Tracking Reliable Trajectories for Ensembling Masked Diffusion Language Models 33 upvotes, #7 of 2026-06-16
- Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization 33 upvotes, #11 of 2026-06-10
- TIDE: Proactive Multi-Problem Discovery via Template-Guided Iteration 44 upvotes, #3 of 2026-06-05
- Trust Region Q Adjoint Matching 3 upvotes, #34 of 2026-06-05
- Mitigating Perceptual Judgment Bias in Multimodal LLM-as-a-Judge via Perceptual Perturbation and Reward Modeling 5 upvotes, #32 of 2026-06-03
- OmniRetrieval: Unified Retrieval across Heterogeneous Knowledge Sources 76 upvotes, #3 of 2026-05-29
- Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases 7 upvotes, #48 of 2026-05-29
- Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents 38 upvotes, #10 of 2026-05-28
- Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction 41 upvotes, #5 of 2026-05-27
- HINT-SD: Targeted Hindsight Self-Distillation for Long-Horizon Agents 11 upvotes, #18 of 2026-05-25
- WorldKV: Efficient World Memory with World Retrieval and Compression 41 upvotes, #10 of 2026-05-22
- FlowLong: Inference-time Long Video Generation via Manifold-constrained Tweedie Matching 29 upvotes, #14 of 2026-05-22
- It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs 30 upvotes, #10 of 2026-05-21
- Nudging Beyond the Comfort Zone: Efficient Strategy-Guided Exploration for RLVR 33 upvotes, #11 of 2026-05-18
- PREPING: Building Agent Memory without Tasks 28 upvotes, #13 of 2026-05-15
- MEME: Multi-entity & Evolving Memory Evaluation 7 upvotes, #36 of 2026-05-13
- CollabVR: Collaborative Video Reasoning with Vision-Language and Video Generation Models 68 upvotes, #3 of 2026-05-12
- Towards Autonomous Mechanistic Reasoning in Virtual Cells 6 upvotes, #20 of 2026-04-17
- Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents 29 upvotes, #5 of 2026-04-16
- Sommelier: Scalable Open Multi-turn Audio Pre-processing for Full-duplex Speech Language Models 35 upvotes, #5 of 2026-03-30
- Representation Alignment for Just Image Transformers is not Easier than You Think 13 upvotes, #11 of 2026-03-27
- T-MAP: Red-Teaming LLM Agents with Trajectory-aware Evolutionary Search 36 upvotes, #5 of 2026-03-26
- DA-Flow: Degradation-Aware Optical Flow Estimation with Diffusion Models 50 upvotes, #5 of 2026-03-25
- SNAP: Speaker Nulling for Artifact Projection in Speech Deepfake Detection 3 upvotes, #31 of 2026-03-24
- RoboAlign: Learning Test-Time Reasoning for Language-Action Alignment in Vision-Language-Action Models 23 upvotes, #14 of 2026-03-24
- SpatialBoost: Enhancing Visual Representation through Language-Guided Reasoning 45 upvotes, #7 of 2026-03-24
- Repurposing Geometric Foundation Models for Multi-view Diffusion 45 upvotes, #7 of 2026-03-24
- ECG-Reasoning-Benchmark: A Benchmark for Evaluating Clinical Reasoning Capabilities in ECG Interpretation 1 upvotes, #42 of 2026-03-18
- MA-EgoQA: Question Answering over Egocentric Videos from Multiple Embodied Agents 28 upvotes, #6 of 2026-03-12
- Can Large Language Models Keep Up? Benchmarking Online Adaptation to Continual Knowledge Streams 17 upvotes, #9 of 2026-03-12
- MolHIT: Advancing Molecular-Graph Generation with Hierarchical Discrete Diffusion Models 54 upvotes, #1 of 2026-02-26
- RoboCurate: Harnessing Diversity with Action-Verified Neural Trajectory for Robot Learning 10 upvotes, #11 of 2026-02-24
- THINKSAFE: Self-Generated Safety Alignment for Reasoning Models 38 upvotes, #5 of 2026-02-02
- Lost in the Noise: How Reasoning Models Fail with Contextual Distractors 29 upvotes, #6 of 2026-01-13
- InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion 95 upvotes, #2 of 2025-12-29
- Infinite-Homography as Robust Conditioning for Camera-Controlled Video Generation 27 upvotes, #6 of 2025-12-23
- Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure 28 upvotes, #7 of 2025-12-17
- Directional Textual Inversion for Personalized Text-to-Image Generation 2 upvotes, #34 of 2025-12-16
- EgoX: Egocentric Video Generation from a Single Exocentric Video 106 upvotes, #1 of 2025-12-15
- Aligned but Stereotypical? The Hidden Influence of System Prompts on Social Bias in LVLM-Based Text-to-Image Models 7 upvotes, #26 of 2025-12-05
- Instruction-Guided Lesion Segmentation for Chest X-rays with Automatically Generated Large-Scale Dataset 25 upvotes, #5 of 2025-11-20
- KLASS: KL-Guided Fast Inference in Masked Diffusion Models 35 upvotes, #3 of 2025-11-12
- When to Ensemble: Identifying Token-Level Points for Stable and Fast LLM Ensembling 32 upvotes, #7 of 2025-10-21
- Temporal Alignment Guidance: On-Manifold Sampling in Diffusion Models 29 upvotes, #11 of 2025-10-15
- Multimodal Prompt Optimization: Why Not Leverage Multiple Modalities for MLLMs 46 upvotes, #5 of 2025-10-13
- Meta-Awareness Enhances Reasoning Models: Self-Alignment Reinforcement Learning 54 upvotes, #7 of 2025-10-10
- Verifier-free Test-Time Sampling for Vision Language Action Models 2 upvotes, #34 of 2025-10-08
- No Prompt Left Behind: Exploiting Zero-Variance Prompts in LLM Reinforcement Learning via Entropy-Guided Advantage Shaping 37 upvotes, #9 of 2025-09-29
- ReviewScore: Misinformed Peer Review Detection with Large Language Models 62 upvotes, #7 of 2025-09-29
- PatientSim: A Persona-Driven Simulator for Realistic Doctor-Patient Interactions 11 upvotes, #29 of 2025-05-30
- CXReasonBench: A Benchmark for Evaluating Structured Diagnostic Reasoning in Chest X-rays 6 upvotes, #42 of 2025-05-30
- Lunguage: A Benchmark for Structured and Sequential Chest X-ray Interpretation 4 upvotes, #49 of 2025-05-30
- SphereDiff: Tuning-free Omnidirectional Panoramic Image and Video Generation via Spherical Latent Representation 27 upvotes, #7 of 2025-04-22
- Towards Predicting Temporal Changes in a Patient's Chest X-ray Images based on Electronic Health Records 3 upvotes, #13 of 2024-09-17
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.