Amazon
Amazon on Hugging Face Daily Papers: 35 papers, 2 in the top 3 of their day, 1 paper of the day.
- Breaking Babel: A Self-Evolving Multi-Agent System for Long-Form Subtitle Translation 35 upvotes, #24 of 2026-10-01
- MILO: Automated Harness Discovery via Orchestrated Multi-Agent Evolution 16 upvotes, #33 of 2026-10-01
- Training LLM Judges from Language Feedback via Position-Selective Self-Distillation 5 upvotes, #53 of 2026-10-01
- Rethinking Latent Visual Reasoning: Grounding Latent Reasoning in Visual Evidence 209 upvotes, #5 of 2026-10-01
- On the Off-Policy Teacher in On-Policy Distillation 22 upvotes, #42 of 2026-09-30
- StructRL: Online Structured Reinforcement Learning for Long-Horizon Vision-Language-Action Tasks 5 upvotes, #74 of 2026-09-30
- Rufus-Air: An Open LLM Post-Training Recipe 37 upvotes, #5 of 2026-09-25
- Don't Mask the Environment: Observation Supervision Changes How Agents Explore Under RL 42 upvotes, #16 of 2026-09-18
- MInTRL: Off-policy Intervention can boost On-policy RL 14 upvotes, #18 of 2026-09-15
- Group Adaptive Clipping Policy Optimization 10 upvotes, #16 of 2026-09-07
- Training Leaves Traces: Centered Residual Signatures for Language Model Lineage Verification 18 upvotes, #8 of 2026-08-20
- Statistically Reliable LLM-Based Ranking Evaluation via Prediction-Powered Inference 3 upvotes, #39 of 2026-06-15
- An Empirical Study of Automating Agent Evaluation 3 upvotes, #34 of 2026-05-14
- Emergent Strategic Reasoning Risks in AI: A Taxonomy-Driven Evaluation Framework 5 upvotes, #16 of 2026-04-27
- Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts 17 upvotes, #9 of 2026-04-23
- A Frame is Worth One Token: Efficient Generative World Modeling with Delta Tokens 10 upvotes, #15 of 2026-04-09
- Perceptio: Perception Enhanced Vision Language Models via Spatial Token Generation 16 upvotes, #19 of 2026-03-24
- Scalable Prompt Routing via Fine-Grained Latent Task Discovery 6 upvotes, #25 of 2026-03-24
- Supervised Fine-Tuning versus Reinforcement Learning: A Study of Post-Training Methods for Large Language Models 10 upvotes, #20 of 2026-03-17
- ByteFlow: Language Modeling through Adaptive Byte Compression without a Tokenizer 2 upvotes, #35 of 2026-03-10
- DDiT: Dynamic Patch Scheduling for Efficient Diffusion Transformers 11 upvotes, #10 of 2026-02-20
- Approximation of Log-Partition Function in Policy Mirror Descent Induces Implicit Regularization for LLM Post-Training 5 upvotes, #33 of 2026-02-06
- LikeBench: Evaluating Subjective Likability in LLMs for Personalization 2 upvotes, #28 of 2025-12-18
- VIDEOP2R: Video Understanding from Perception to Reasoning 107 upvotes, #1 of 2025-11-19
- WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance 9 upvotes, #14 of 2025-11-18
- Adaptive Multi-Agent Response Refinement in Conversational Systems 39 upvotes, #2 of 2025-11-12
- Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs 6 upvotes, #20 of 2025-10-24
- Chronos-2: From Univariate to Universal Forecasting 14 upvotes, #14 of 2025-10-21
- MTSQL-R1: Towards Long-Horizon Multi-Turn Text-to-SQL via Agentic Training 2 upvotes, #31 of 2025-10-16
- The Personalization Trap: How User Memory Alters Emotional Reasoning in LLMs 6 upvotes, #35 of 2025-10-14
- Multimodal Policy Internalization for Conversational Agents 4 upvotes, #40 of 2025-10-14
- When Thoughts Meet Facts: Reusable Reasoning for Long-Context LMs 44 upvotes, #9 of 2025-10-10
- TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning 61 upvotes, #4 of 2025-10-08
- CompLLM: Compression for Long Context Q&A 4 upvotes, #27 of 2025-09-26
- Quantifying Fairness in LLMs Beyond Tokens: A Semantic and Statistical Perspective 1 upvotes, #39 of 2025-06-24
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.