Shanghai Jiao Tong University
Shanghai Jiao Tong University on Hugging Face Daily Papers: 63 papers, 12 in the top 3 of their day, 4 paper of the day.
- EVOKE: Eliciting World Knowledge in Agents for Transferable Decision-Making 69 upvotes, #12 of 2026-10-01
- APM-Bench: Benchmarking Cross-session Persistent Memory for Egocentric Streaming Video Assistants 46 upvotes, #25 of 2026-09-30
- GeoVerse: World-Consistent Novel View Synthesis in Geometric Latent Space 8 upvotes, #51 of 2026-09-29
- Fewer Tokens, More Self-Teaching: On-Policy Self-Distillation for Extreme Visual Token Reduction 8 upvotes, #51 of 2026-09-29
- Schrödinger's Code Repository: Have LLMs Learned SWE-bench or Memorized It? 31 upvotes, #7 of 2026-09-24
- Streaming Video Editing with Easy Adaptation 11 upvotes, #23 of 2026-09-22
- HyQuant: Hybrid-Precision Quantization for LLM Attention 28 upvotes, #16 of 2026-09-11
- EarlyEval: Cheaper Agent Evaluation via Early Outcome Prediction 118 upvotes, #5 of 2026-09-03
- SafeAtlas-VL: Beyond Binary Multimodal Safety with Large-Scale Data and Guard Models 7 upvotes, #26 of 2026-09-01
- Weaving Visual Narratives: Agentic Image Bundle Composition Beyond Atomic Visual Matching 8 upvotes, #22 of 2026-09-01
- UrbanGround: From Local Perception to Spatial Agency in a Real-Scale City 106 upvotes, #3 of 2026-08-28
- ParaTempo: Efficient Parallel Reasoning via Temporal Confidence 38 upvotes, #3 of 2026-08-24
- Repo0: Design-Driven Zero-to-All Code Generation 20 upvotes, #8 of 2026-08-21
- SkillGate: Training In-Policy Skill Selection in Long-Horizon Agents 7 upvotes, #15 of 2026-08-20
- SkillForge: Self-Distilling Agents for Project-Specific Issue Resolution 11 upvotes, #17 of 2026-08-19
- TRACE-Bench: Decomposing and Diagnosing Multi-Reference Image Generation 13 upvotes, #19 of 2026-08-18
- H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models 17 upvotes, #16 of 2026-08-14
- ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment 65 upvotes, #2 of 2026-08-06
- Know Before Fix: QA-Driven Repository Knowledge Acquisition for Software Issue Resolution 22 upvotes, #6 of 2026-07-15
- Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation 39 upvotes, #4 of 2026-07-10
- Are We Ready For An Agent-Native Memory System? 123 upvotes, #1 of 2026-06-25
- ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? 19 upvotes, #11 of 2026-06-19
- ViT-Up: Faithful Feature Upsampling for Vision Transformers 9 upvotes, #17 of 2026-06-18
- LLM Agents Can See Code Repositories 20 upvotes, #14 of 2026-06-15
- Revisiting Articulated Parts Perception in Robot Manipulation 2 upvotes, #35 of 2026-06-12
- Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning 30 upvotes, #9 of 2026-06-11
- SWE-Explore: Benchmarking How Coding Agents Explore Repositories 112 upvotes, #2 of 2026-06-09
- AHA-WAM:Asynchronous Horizon-Adaptive World-Action Modeling with Observation-Guided Context Routing 14 upvotes, #18 of 2026-06-09
- PBSD: Privileged Bayesian Self-Distillation for Long-Horizon Credit Assignment 1 upvotes, #44 of 2026-06-09
- Domino: Decoupling Causal Modeling from Autoregressive Drafting in Speculative Decoding 143 upvotes, #3 of 2026-06-02
- GradSentry: Gradient Spectral Entropy for Backdoor Sample Filtering in Large Language Model Fine-Tuning 12 upvotes, #32 of 2026-05-28
- OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents 6 upvotes, #43 of 2026-05-28
- Anticipate and Learn: Unleashing Idle-Time Compute in Proactive Agents 16 upvotes, #22 of 2026-05-26
- Semantic Generative Tuning for Unified Multimodal Models 10 upvotes, #21 of 2026-05-20
- Towards Self-Evolving Agentic Literature Retrieval 3 upvotes, #34 of 2026-05-14
- ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration 110 upvotes, #1 of 2026-05-06
- River-LLM: Large Language Model Seamless Exit Based on KV Share 6 upvotes, #23 of 2026-04-21
- Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanization 30 upvotes, #6 of 2026-04-15
- Externalization in LLM Agents: A Unified Review of Memory, Skills, Protocols and Harness Engineering 50 upvotes, #9 of 2026-04-10
- Automating Database-Native Function Code Synthesis with LLMs 17 upvotes, #21 of 2026-04-10
- From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation 6 upvotes, #29 of 2026-03-18
- EmbTracker: Traceable Black-box Watermarking for Federated Language Models 2 upvotes, #37 of 2026-03-13
- RoboPocket: Improve Robot Policies Instantly with Your Phone 31 upvotes, #5 of 2026-03-06
- AgentConductor: Topology Evolution for Multi-Agent Competition-Level Code Generation 2 upvotes, #29 of 2026-03-04
- ProtegoFed: Backdoor-Free Federated Instruction Tuning with Interspersed Poisoned Data 1 upvotes, #36 of 2026-03-03
- Grounding and Enhancing Informativeness and Utility in Dataset Distillation 19 upvotes, #15 of 2026-02-06
- CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding 91 upvotes, #1 of 2026-02-04
- MMFineReason: Closing the Multimodal Reasoning Gap via Open Data-Centric Methods 57 upvotes, #5 of 2026-01-30
- Innovator-VL: A Multimodal Large Language Model for Scientific Discovery 76 upvotes, #3 of 2026-01-29
- Scientific Image Synthesis: Benchmarking, Methodologies, and Downstream Utility 41 upvotes, #4 of 2026-01-27
- Can LLMs Clean Up Your Mess? A Survey of Application-Ready Data Preparation with LLMs 181 upvotes, #1 of 2026-01-27
- AgentEHR: Advancing Autonomous Clinical Decision-Making via Retrospective Summarization 5 upvotes, #18 of 2026-01-22
- Toward Ultra-Long-Horizon Agentic Science: Cognitive Accumulation for Machine Learning Engineering 36 upvotes, #7 of 2026-01-16
- GlimpRouter: Efficient Collaborative Inference by Glimpsing One Token of Thoughts 27 upvotes, #7 of 2026-01-13
- SpeContext: Enabling Efficient Long-context Reasoning with Speculative Context Sparsity in LLMs 14 upvotes, #18 of 2025-12-02
- LoopTool: Closing the Data-Training Loop for Robust LLM Tool Calls 15 upvotes, #6 of 2025-11-13
- EHR-R1: A Reasoning-Enhanced Foundational Language Model for Electronic Health Record Analysis 9 upvotes, #14 of 2025-10-31
- Evolving Diagnostic Agents in a Virtual Clinical Environment 10 upvotes, #19 of 2025-10-30
- RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling 11 upvotes, #12 of 2025-10-27
- AI for Service: Proactive Assistance with AI Glasses 71 upvotes, #4 of 2025-10-17
- Efficient Multi-modal Large Language Models via Progressive Consistency Distillation 38 upvotes, #3 of 2025-10-06
- PARROT: A Benchmark for Evaluating LLMs in Cross-System SQL Translation 4 upvotes, #56 of 2025-09-30
- Shifting AI Efficiency From Model-Centric to Data-Centric Compression 141 upvotes, #2 of 2025-05-27
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.