Daily Papers of 2024-10-24
- MIA-DPO: Multi-Image Augmented Direct Preference Optimization For Large Vision-Language Models 34 upvotes, #1 of 2024-10-24
- LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding 18 upvotes, #2 of 2024-10-24
- WorldSimBench: Towards Video Generation Models as World Simulators 16 upvotes, #3 of 2024-10-24
- Scalable Ranked Preference Optimization for Text-to-Image Generation 14 upvotes, #4 of 2024-10-24
- Scaling Diffusion Language Models via Adaptation from Autoregressive Models 13 upvotes, #5 of 2024-10-24
- DynamicCity: Large-Scale LiDAR Generation from Dynamic Scenes 12 upvotes, #6 of 2024-10-24
- M-RewardBench: Evaluating Reward Models in Multilingual Settings 10 upvotes, #7 of 2024-10-24
- Lightweight Neural App Control 8 upvotes, #8 of 2024-10-24
- MedINST: Meta Dataset of Biomedical Instructions 6 upvotes, #9 of 2024-10-24
- ARKit LabelMaker: A New Scale for Indoor 3D Scene Understanding 6 upvotes, #9 of 2024-10-24
- TP-Eval: Tap Multimodal LLMs' Potential in Evaluation by Customizing Prompts 6 upvotes, #9 of 2024-10-24
- Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance 1 upvotes, #12 of 2024-10-24
- LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias 1 upvotes, #12 of 2024-10-24
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.