Daily Papers of 2024-03-22
- MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? 45 upvotes, #1 of 2024-03-22
- DreamReward: Text-to-3D Generation with Human Preference 31 upvotes, #2 of 2024-03-22
- Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference 30 upvotes, #3 of 2024-03-22
- AnyV2V: A Plug-and-Play Framework For Any Video-to-Video Editing Tasks 17 upvotes, #4 of 2024-03-22
- ReNoise: Real Image Inversion Through Iterative Noising 17 upvotes, #4 of 2024-03-22
- Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition 16 upvotes, #6 of 2024-03-22
- MyVLM: Personalizing VLMs for User-Specific Queries 13 upvotes, #7 of 2024-03-22
- GRM: Large Gaussian Reconstruction Model for Efficient 3D Reconstruction and Generation 13 upvotes, #7 of 2024-03-22
- Gaussian Frosting: Editable Complex Radiance Fields with Real-Time Rendering 10 upvotes, #9 of 2024-03-22
- Explorative Inbetweening of Time and Space 10 upvotes, #9 of 2024-03-22
- StyleCineGAN: Landscape Cinemagraph Generation using a Pre-trained StyleGAN 8 upvotes, #11 of 2024-03-22
- Recourse for reclamation: Chatting with generative language models 6 upvotes, #12 of 2024-03-22
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.