University of Illinois at Urbana-Champaign

University of Illinois at Urbana-Champaign on Hugging Face Daily Papers: 52 papers, 6 in the top 3 of their day, 3 paper of the day.

  1. MotorMind: Scaffolding General Vision Language Models for Zero-Shot Robot Manipulation 30 upvotes, #1 of 2026-10-05
  2. InterEvolve: Test-Time Evolution of Reward Programs for Humanoid Loco-Manipulation 55 upvotes, #21 of 2026-10-02
  3. Hierarchical Continuous Diffusion Language Models 80 upvotes, #8 of 2026-10-02
  4. ANTMAN: Adaptive Need Tracking for Multi-Agent Navigation in Large Information Spaces 35 upvotes, #31 of 2026-09-30
  5. LeRF: Learning Reference Coordinate Frames for Perspective Taking Reasoning 2 upvotes, #93 of 2026-09-29
  6. Feature Recovery for Object Understanding After Irreversible Fire Damage 1 upvotes, #24 of 2026-09-14
  7. Competence-Gated Pooling of Language Models and Priors for Event Forecasting 2 upvotes, #22 of 2026-09-14
  8. VGI-BENCH: Probing Visual Intelligence in Video Generation Models 177 upvotes, #2 of 2026-08-27
  9. CLEAR: Continuous Latent Adapter Routing for Utility-Preserving LLM Safety Alignment 5 upvotes, #18 of 2026-08-24
  10. LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers 108 upvotes, #3 of 2026-08-14
  11. AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses 111 upvotes, #3 of 2026-08-13
  12. AgentDebugX: An Open-Source Toolkit for Failure Observability, Attribution, and Recovery in LLM Agents 24 upvotes, #9 of 2026-07-22
  13. SVR-R1: Bootstrapping Multi-modal Reasoning with Self-verification in Reinforcement Learning 5 upvotes, #22 of 2026-07-20
  14. Trimming the Long-Tail of Visual World Modeling Evaluation 42 upvotes, #7 of 2026-06-30
  15. GBC: Gradient-Based Connections for Optimizing Multi-Agent Systems 13 upvotes, #10 of 2026-06-29
  16. PlanBench-XL: Evaluating Long-Horizon Planning of LLM Tool-Use Agents in Large-Scale Tool Ecosystems 95 upvotes, #1 of 2026-06-23
  17. Building Social World Models with Large Language Models 1 upvotes, #39 of 2026-06-11
  18. BenSyc: Benchmarking Conversational Sycophancy and Human Alignment in LLMs for Bengali Contexts 1 upvotes, #44 of 2026-06-10
  19. AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints 40 upvotes, #4 of 2026-06-05
  20. CausaLab: A Scalable Environment for Interactive Causal Discovery Toward AI Scientists 18 upvotes, #20 of 2026-05-29
  21. Advancing Creative Physical Intelligence in Large Multimodal Models 19 upvotes, #21 of 2026-05-28
  22. Spreadsheet-RL: Advancing Large Language Model Agents on Realistic Spreadsheet Tasks via Reinforcement Learning 35 upvotes, #12 of 2026-05-22
  23. RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably 0 upvotes, #53 of 2026-05-20
  24. GRASP: Learning to Ground Social Reasoning in Multi-Person Non-Verbal Interactions 3 upvotes, #41 of 2026-05-19
  25. RouteProfile: Elucidating the Design Space of LLM Profiles for Routing 30 upvotes, #12 of 2026-05-15
  26. Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance 4 upvotes, #45 of 2026-05-15
  27. Retrieval is Cheap, Show Me the Code: Executable Multi-Hop Reasoning for Retrieval-Augmented Generation 9 upvotes, #20 of 2026-05-14
  28. The Many Faces of On-Policy Distillation: Pitfalls, Mechanisms, and Fixes 5 upvotes, #41 of 2026-05-13
  29. Who Prices Cognitive Labor in the Age of Agents? Compute-Anchored Wages 3 upvotes, #46 of 2026-05-11
  30. CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposing 21 upvotes, #10 of 2026-05-07
  31. Agentic AI Systems Should Be Designed as Marginal Token Allocators 4 upvotes, #18 of 2026-05-05
  32. LangFlow: Continuous Diffusion Rivals Discrete in Language Modeling 14 upvotes, #13 of 2026-04-16
  33. Narrative-Driven Paper-to-Slide Generation via ArcDeck 7 upvotes, #22 of 2026-04-16
  34. STRIDE: When to Speak Meets Sequence Denoising for Streaming Video Understanding 11 upvotes, #22 of 2026-03-31
  35. HandX: Scaling Bimanual Motion and Interaction Generation 12 upvotes, #21 of 2026-03-31
  36. Sparking Scientific Creativity via LLM-Driven Interdisciplinary Inspiration 2 upvotes, #33 of 2026-03-18
  37. Learning Humanoid End-Effector Control for Open-Vocabulary Visual Loco-Manipulation 26 upvotes, #4 of 2026-02-19
  38. Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs 10 upvotes, #27 of 2026-02-11
  39. CodeCircuit: Toward Inferring LLM-Generated Code Correctness via Attribution Graphs 6 upvotes, #32 of 2026-02-10
  40. Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning 39 upvotes, #13 of 2026-02-03
  41. Agentic Reasoning for Large Language Models 182 upvotes, #1 of 2026-01-22
  42. When Reasoning Meets Its Laws 54 upvotes, #4 of 2025-12-22
  43. BEAVER: An Efficient Deterministic LLM Verifier 29 upvotes, #6 of 2025-12-12
  44. Visual Sync: Multi-Camera Synchronization via Cross-View Object Motion 3 upvotes, #32 of 2025-12-03
  45. VisPlay: Self-Evolving Vision-Language Models from Images 41 upvotes, #4 of 2025-11-20
  46. Multi-Agent Evolve: LLM Self-Improve through Co-evolution 8 upvotes, #18 of 2025-10-28
  47. SCas4D: Structural Cascaded Optimization for Boosting Persistent 4D Novel View Synthesis 2 upvotes, #44 of 2025-10-17
  48. ERA: Transforming VLMs into Embodied Agents via Embodied Prior Learning and Online Reinforcement Learning 25 upvotes, #12 of 2025-10-15
  49. How to Teach Large Multimodal Models New Skills 2 upvotes, #37 of 2025-10-13
  50. GTAlign: Game-Theoretic Alignment of LLM Assistants for Mutual Welfare 2 upvotes, #37 of 2025-10-13
  51. GRACE: Generative Representation Learning via Contrastive Policy Optimization 9 upvotes, #17 of 2025-10-08
  52. Where LLM Agents Fail and How They can Learn From Failures 11 upvotes, #36 of 2025-09-30

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.