Daily Papers of 2024-07-09

  1. MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation? 49 upvotes, #1 of 2024-07-09
  2. LLaMAX: Scaling Linguistic Horizons of LLM by Enhancing Translation Capabilities Beyond 100 Languages 32 upvotes, #2 of 2024-07-09
  3. Learning Action and Reasoning-Centric Image Editing from Videos and Simulations 24 upvotes, #3 of 2024-07-09
  4. Associative Recurrent Memory Transformer 24 upvotes, #3 of 2024-07-09
  5. ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation 19 upvotes, #5 of 2024-07-09
  6. Evaluating Language Model Context Windows: A "Working Memory" Test and Inference-time Correction 14 upvotes, #6 of 2024-07-09
  7. Compositional Video Generation as Flow Equalization 12 upvotes, #7 of 2024-07-09
  8. UltraEdit: Instruction-based Fine-Grained Image Editing at Scale 8 upvotes, #8 of 2024-07-09
  9. InverseCoder: Unleashing the Power of Instruction-Tuned Code LLMs with Inverse-Instruct 8 upvotes, #8 of 2024-07-09
  10. PAS: Data-Efficient Plug-and-Play Prompt Augmentation System 8 upvotes, #8 of 2024-07-09
  11. Tailor3D: Customized 3D Assets Editing and Generation with Dual-Side Images 8 upvotes, #8 of 2024-07-09
  12. Multi-Object Hallucination in Vision-Language Models 7 upvotes, #12 of 2024-07-09
  13. Training Task Experts through Retrieval Based Distillation 6 upvotes, #13 of 2024-07-09
  14. Understanding Visual Feature Reliance through the Lens of Complexity 4 upvotes, #14 of 2024-07-09
  15. PartCraft: Crafting Creative Objects by Parts 3 upvotes, #15 of 2024-07-09
  16. LLMAEL: Large Language Models are Good Context Augmenters for Entity Linking 2 upvotes, #16 of 2024-07-09
  17. ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models 1 upvotes, #17 of 2024-07-09

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.