Amazon

Amazon on Hugging Face Daily Papers: 35 papers, 2 in the top 3 of their day, 1 paper of the day.

  1. Breaking Babel: A Self-Evolving Multi-Agent System for Long-Form Subtitle Translation 35 upvotes, #24 of 2026-10-01
  2. MILO: Automated Harness Discovery via Orchestrated Multi-Agent Evolution 16 upvotes, #33 of 2026-10-01
  3. Training LLM Judges from Language Feedback via Position-Selective Self-Distillation 5 upvotes, #53 of 2026-10-01
  4. Rethinking Latent Visual Reasoning: Grounding Latent Reasoning in Visual Evidence 209 upvotes, #5 of 2026-10-01
  5. On the Off-Policy Teacher in On-Policy Distillation 22 upvotes, #42 of 2026-09-30
  6. StructRL: Online Structured Reinforcement Learning for Long-Horizon Vision-Language-Action Tasks 5 upvotes, #74 of 2026-09-30
  7. Rufus-Air: An Open LLM Post-Training Recipe 37 upvotes, #5 of 2026-09-25
  8. Don't Mask the Environment: Observation Supervision Changes How Agents Explore Under RL 42 upvotes, #16 of 2026-09-18
  9. MInTRL: Off-policy Intervention can boost On-policy RL 14 upvotes, #18 of 2026-09-15
  10. Group Adaptive Clipping Policy Optimization 10 upvotes, #16 of 2026-09-07
  11. Training Leaves Traces: Centered Residual Signatures for Language Model Lineage Verification 18 upvotes, #8 of 2026-08-20
  12. Statistically Reliable LLM-Based Ranking Evaluation via Prediction-Powered Inference 3 upvotes, #39 of 2026-06-15
  13. An Empirical Study of Automating Agent Evaluation 3 upvotes, #34 of 2026-05-14
  14. Emergent Strategic Reasoning Risks in AI: A Taxonomy-Driven Evaluation Framework 5 upvotes, #16 of 2026-04-27
  15. Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts 17 upvotes, #9 of 2026-04-23
  16. A Frame is Worth One Token: Efficient Generative World Modeling with Delta Tokens 10 upvotes, #15 of 2026-04-09
  17. Perceptio: Perception Enhanced Vision Language Models via Spatial Token Generation 16 upvotes, #19 of 2026-03-24
  18. Scalable Prompt Routing via Fine-Grained Latent Task Discovery 6 upvotes, #25 of 2026-03-24
  19. Supervised Fine-Tuning versus Reinforcement Learning: A Study of Post-Training Methods for Large Language Models 10 upvotes, #20 of 2026-03-17
  20. ByteFlow: Language Modeling through Adaptive Byte Compression without a Tokenizer 2 upvotes, #35 of 2026-03-10
  21. DDiT: Dynamic Patch Scheduling for Efficient Diffusion Transformers 11 upvotes, #10 of 2026-02-20
  22. Approximation of Log-Partition Function in Policy Mirror Descent Induces Implicit Regularization for LLM Post-Training 5 upvotes, #33 of 2026-02-06
  23. LikeBench: Evaluating Subjective Likability in LLMs for Personalization 2 upvotes, #28 of 2025-12-18
  24. VIDEOP2R: Video Understanding from Perception to Reasoning 107 upvotes, #1 of 2025-11-19
  25. WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance 9 upvotes, #14 of 2025-11-18
  26. Adaptive Multi-Agent Response Refinement in Conversational Systems 39 upvotes, #2 of 2025-11-12
  27. Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs 6 upvotes, #20 of 2025-10-24
  28. Chronos-2: From Univariate to Universal Forecasting 14 upvotes, #14 of 2025-10-21
  29. MTSQL-R1: Towards Long-Horizon Multi-Turn Text-to-SQL via Agentic Training 2 upvotes, #31 of 2025-10-16
  30. The Personalization Trap: How User Memory Alters Emotional Reasoning in LLMs 6 upvotes, #35 of 2025-10-14
  31. Multimodal Policy Internalization for Conversational Agents 4 upvotes, #40 of 2025-10-14
  32. When Thoughts Meet Facts: Reusable Reasoning for Long-Context LMs 44 upvotes, #9 of 2025-10-10
  33. TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning 61 upvotes, #4 of 2025-10-08
  34. CompLLM: Compression for Long Context Q&A 4 upvotes, #27 of 2025-09-26
  35. Quantifying Fairness in LLMs Beyond Tokens: A Semantic and Statistical Perspective 1 upvotes, #39 of 2025-06-24

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.