Daily Papers of 2025-07-28
- Deep Researcher with Test-Time Diffusion 53 upvotes, #1 of 2025-07-28
- The Geometry of LLM Quantization: GPTQ as Babai's Nearest Plane Algorithm 37 upvotes, #2 of 2025-07-28
- MMBench-GUI: Hierarchical Multi-Platform Evaluation Framework for GUI Agents 28 upvotes, #3 of 2025-07-28
- When Tokens Talk Too Much: A Survey of Multimodal Long-Context Token Compression across Images, Videos, and Audios 24 upvotes, #4 of 2025-07-28
- GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning 20 upvotes, #5 of 2025-07-28
- CLEAR: Error Analysis via LLM-as-a-Judge Made Easy 17 upvotes, #6 of 2025-07-28
- PRIX: Learning to Plan from Raw Pixels for End-to-End Autonomous Driving 5 upvotes, #7 of 2025-07-28
- Specification Self-Correction: Mitigating In-Context Reward Hacking Through Test-Time Refinement 5 upvotes, #7 of 2025-07-28
- Chat with AI: The Surprising Turn of Real-time Video Communication from Human to AI 4 upvotes, #9 of 2025-07-28
- Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report 4 upvotes, #9 of 2025-07-28
- AFRDA: Attentive Feature Refinement for Domain Adaptive Semantic Segmentation 1 upvotes, #11 of 2025-07-28
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.