Daily Papers of 2024-12-31

  1. Explanatory Instructions: Towards Unified Vision Tasks Understanding and Zero-shot Generalization 64 upvotes, #1 of 2024-12-31
  2. On the Compositional Generalization of Multimodal LLMs for Medical Imaging 40 upvotes, #2 of 2024-12-31
  3. Bringing Objects to Life: 4D generation from 3D objects 32 upvotes, #3 of 2024-12-31
  4. Efficiently Serving LLM Reasoning Programs with Certaindex 31 upvotes, #4 of 2024-12-31
  5. Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 27 upvotes, #5 of 2024-12-31
  6. TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization 22 upvotes, #6 of 2024-12-31
  7. Edicho: Consistent Image Editing in the Wild 20 upvotes, #7 of 2024-12-31
  8. Training Software Engineering Agents and Verifiers with SWE-Gym 17 upvotes, #8 of 2024-12-31
  9. OneKE: A Dockerized Schema-Guided LLM Agent-based Knowledge Extraction System 16 upvotes, #9 of 2024-12-31
  10. PERSE: Personalized 3D Generative Avatars from A Single Portrait 15 upvotes, #10 of 2024-12-31
  11. Facilitating large language model Russian adaptation with Learned Embedding Propagation 14 upvotes, #11 of 2024-12-31
  12. Slow Perception: Let's Perceive Geometric Figures Step-by-step 12 upvotes, #12 of 2024-12-31
  13. HumanEval Pro and MBPP Pro: Evaluating Large Language Models on Self-invoking Code Generation 9 upvotes, #13 of 2024-12-31

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.