Daily Papers of 2024-07-09
- MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation? 49 upvotes, #1 of 2024-07-09
- LLaMAX: Scaling Linguistic Horizons of LLM by Enhancing Translation Capabilities Beyond 100 Languages 32 upvotes, #2 of 2024-07-09
- Learning Action and Reasoning-Centric Image Editing from Videos and Simulations 24 upvotes, #3 of 2024-07-09
- Associative Recurrent Memory Transformer 24 upvotes, #3 of 2024-07-09
- ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation 19 upvotes, #5 of 2024-07-09
- Evaluating Language Model Context Windows: A "Working Memory" Test and Inference-time Correction 14 upvotes, #6 of 2024-07-09
- Compositional Video Generation as Flow Equalization 12 upvotes, #7 of 2024-07-09
- UltraEdit: Instruction-based Fine-Grained Image Editing at Scale 8 upvotes, #8 of 2024-07-09
- InverseCoder: Unleashing the Power of Instruction-Tuned Code LLMs with Inverse-Instruct 8 upvotes, #8 of 2024-07-09
- PAS: Data-Efficient Plug-and-Play Prompt Augmentation System 8 upvotes, #8 of 2024-07-09
- Tailor3D: Customized 3D Assets Editing and Generation with Dual-Side Images 8 upvotes, #8 of 2024-07-09
- Multi-Object Hallucination in Vision-Language Models 7 upvotes, #12 of 2024-07-09
- Training Task Experts through Retrieval Based Distillation 6 upvotes, #13 of 2024-07-09
- Understanding Visual Feature Reliance through the Lens of Complexity 4 upvotes, #14 of 2024-07-09
- PartCraft: Crafting Creative Objects by Parts 3 upvotes, #15 of 2024-07-09
- LLMAEL: Large Language Models are Good Context Augmenters for Entity Linking 2 upvotes, #16 of 2024-07-09
- ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models 1 upvotes, #17 of 2024-07-09
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.