Shanghai Jiao Tong University

Shanghai Jiao Tong University on Hugging Face Daily Papers: 63 papers, 12 in the top 3 of their day, 4 paper of the day.

  1. EVOKE: Eliciting World Knowledge in Agents for Transferable Decision-Making 69 upvotes, #12 of 2026-10-01
  2. APM-Bench: Benchmarking Cross-session Persistent Memory for Egocentric Streaming Video Assistants 46 upvotes, #25 of 2026-09-30
  3. GeoVerse: World-Consistent Novel View Synthesis in Geometric Latent Space 8 upvotes, #51 of 2026-09-29
  4. Fewer Tokens, More Self-Teaching: On-Policy Self-Distillation for Extreme Visual Token Reduction 8 upvotes, #51 of 2026-09-29
  5. Schrödinger's Code Repository: Have LLMs Learned SWE-bench or Memorized It? 31 upvotes, #7 of 2026-09-24
  6. Streaming Video Editing with Easy Adaptation 11 upvotes, #23 of 2026-09-22
  7. HyQuant: Hybrid-Precision Quantization for LLM Attention 28 upvotes, #16 of 2026-09-11
  8. EarlyEval: Cheaper Agent Evaluation via Early Outcome Prediction 118 upvotes, #5 of 2026-09-03
  9. SafeAtlas-VL: Beyond Binary Multimodal Safety with Large-Scale Data and Guard Models 7 upvotes, #26 of 2026-09-01
  10. Weaving Visual Narratives: Agentic Image Bundle Composition Beyond Atomic Visual Matching 8 upvotes, #22 of 2026-09-01
  11. UrbanGround: From Local Perception to Spatial Agency in a Real-Scale City 106 upvotes, #3 of 2026-08-28
  12. ParaTempo: Efficient Parallel Reasoning via Temporal Confidence 38 upvotes, #3 of 2026-08-24
  13. Repo0: Design-Driven Zero-to-All Code Generation 20 upvotes, #8 of 2026-08-21
  14. SkillGate: Training In-Policy Skill Selection in Long-Horizon Agents 7 upvotes, #15 of 2026-08-20
  15. SkillForge: Self-Distilling Agents for Project-Specific Issue Resolution 11 upvotes, #17 of 2026-08-19
  16. TRACE-Bench: Decomposing and Diagnosing Multi-Reference Image Generation 13 upvotes, #19 of 2026-08-18
  17. H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models 17 upvotes, #16 of 2026-08-14
  18. ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment 65 upvotes, #2 of 2026-08-06
  19. Know Before Fix: QA-Driven Repository Knowledge Acquisition for Software Issue Resolution 22 upvotes, #6 of 2026-07-15
  20. Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation 39 upvotes, #4 of 2026-07-10
  21. Are We Ready For An Agent-Native Memory System? 123 upvotes, #1 of 2026-06-25
  22. ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? 19 upvotes, #11 of 2026-06-19
  23. ViT-Up: Faithful Feature Upsampling for Vision Transformers 9 upvotes, #17 of 2026-06-18
  24. LLM Agents Can See Code Repositories 20 upvotes, #14 of 2026-06-15
  25. Revisiting Articulated Parts Perception in Robot Manipulation 2 upvotes, #35 of 2026-06-12
  26. Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning 30 upvotes, #9 of 2026-06-11
  27. SWE-Explore: Benchmarking How Coding Agents Explore Repositories 112 upvotes, #2 of 2026-06-09
  28. AHA-WAM:Asynchronous Horizon-Adaptive World-Action Modeling with Observation-Guided Context Routing 14 upvotes, #18 of 2026-06-09
  29. PBSD: Privileged Bayesian Self-Distillation for Long-Horizon Credit Assignment 1 upvotes, #44 of 2026-06-09
  30. Domino: Decoupling Causal Modeling from Autoregressive Drafting in Speculative Decoding 143 upvotes, #3 of 2026-06-02
  31. GradSentry: Gradient Spectral Entropy for Backdoor Sample Filtering in Large Language Model Fine-Tuning 12 upvotes, #32 of 2026-05-28
  32. OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents 6 upvotes, #43 of 2026-05-28
  33. Anticipate and Learn: Unleashing Idle-Time Compute in Proactive Agents 16 upvotes, #22 of 2026-05-26
  34. Semantic Generative Tuning for Unified Multimodal Models 10 upvotes, #21 of 2026-05-20
  35. Towards Self-Evolving Agentic Literature Retrieval 3 upvotes, #34 of 2026-05-14
  36. ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration 110 upvotes, #1 of 2026-05-06
  37. River-LLM: Large Language Model Seamless Exit Based on KV Share 6 upvotes, #23 of 2026-04-21
  38. Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanization 30 upvotes, #6 of 2026-04-15
  39. Externalization in LLM Agents: A Unified Review of Memory, Skills, Protocols and Harness Engineering 50 upvotes, #9 of 2026-04-10
  40. Automating Database-Native Function Code Synthesis with LLMs 17 upvotes, #21 of 2026-04-10
  41. From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation 6 upvotes, #29 of 2026-03-18
  42. EmbTracker: Traceable Black-box Watermarking for Federated Language Models 2 upvotes, #37 of 2026-03-13
  43. RoboPocket: Improve Robot Policies Instantly with Your Phone 31 upvotes, #5 of 2026-03-06
  44. AgentConductor: Topology Evolution for Multi-Agent Competition-Level Code Generation 2 upvotes, #29 of 2026-03-04
  45. ProtegoFed: Backdoor-Free Federated Instruction Tuning with Interspersed Poisoned Data 1 upvotes, #36 of 2026-03-03
  46. Grounding and Enhancing Informativeness and Utility in Dataset Distillation 19 upvotes, #15 of 2026-02-06
  47. CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding 91 upvotes, #1 of 2026-02-04
  48. MMFineReason: Closing the Multimodal Reasoning Gap via Open Data-Centric Methods 57 upvotes, #5 of 2026-01-30
  49. Innovator-VL: A Multimodal Large Language Model for Scientific Discovery 76 upvotes, #3 of 2026-01-29
  50. Scientific Image Synthesis: Benchmarking, Methodologies, and Downstream Utility 41 upvotes, #4 of 2026-01-27
  51. Can LLMs Clean Up Your Mess? A Survey of Application-Ready Data Preparation with LLMs 181 upvotes, #1 of 2026-01-27
  52. AgentEHR: Advancing Autonomous Clinical Decision-Making via Retrospective Summarization 5 upvotes, #18 of 2026-01-22
  53. Toward Ultra-Long-Horizon Agentic Science: Cognitive Accumulation for Machine Learning Engineering 36 upvotes, #7 of 2026-01-16
  54. GlimpRouter: Efficient Collaborative Inference by Glimpsing One Token of Thoughts 27 upvotes, #7 of 2026-01-13
  55. SpeContext: Enabling Efficient Long-context Reasoning with Speculative Context Sparsity in LLMs 14 upvotes, #18 of 2025-12-02
  56. LoopTool: Closing the Data-Training Loop for Robust LLM Tool Calls 15 upvotes, #6 of 2025-11-13
  57. EHR-R1: A Reasoning-Enhanced Foundational Language Model for Electronic Health Record Analysis 9 upvotes, #14 of 2025-10-31
  58. Evolving Diagnostic Agents in a Virtual Clinical Environment 10 upvotes, #19 of 2025-10-30
  59. RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling 11 upvotes, #12 of 2025-10-27
  60. AI for Service: Proactive Assistance with AI Glasses 71 upvotes, #4 of 2025-10-17
  61. Efficient Multi-modal Large Language Models via Progressive Consistency Distillation 38 upvotes, #3 of 2025-10-06
  62. PARROT: A Benchmark for Evaluating LLMs in Cross-System SQL Translation 4 upvotes, #56 of 2025-09-30
  63. Shifting AI Efficiency From Model-Centric to Data-Centric Compression 141 upvotes, #2 of 2025-05-27

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.