Daily Papers of 2024-03-22

  1. MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? 45 upvotes, #1 of 2024-03-22
  2. DreamReward: Text-to-3D Generation with Human Preference 31 upvotes, #2 of 2024-03-22
  3. Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference 30 upvotes, #3 of 2024-03-22
  4. AnyV2V: A Plug-and-Play Framework For Any Video-to-Video Editing Tasks 17 upvotes, #4 of 2024-03-22
  5. ReNoise: Real Image Inversion Through Iterative Noising 17 upvotes, #4 of 2024-03-22
  6. Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition 16 upvotes, #6 of 2024-03-22
  7. MyVLM: Personalizing VLMs for User-Specific Queries 13 upvotes, #7 of 2024-03-22
  8. GRM: Large Gaussian Reconstruction Model for Efficient 3D Reconstruction and Generation 13 upvotes, #7 of 2024-03-22
  9. Gaussian Frosting: Editable Complex Radiance Fields with Real-Time Rendering 10 upvotes, #9 of 2024-03-22
  10. Explorative Inbetweening of Time and Space 10 upvotes, #9 of 2024-03-22
  11. StyleCineGAN: Landscape Cinemagraph Generation using a Pre-trained StyleGAN 8 upvotes, #11 of 2024-03-22
  12. Recourse for reclamation: Chatting with generative language models 6 upvotes, #12 of 2024-03-22

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.