Daily Papers of 2024-03-21
- Mora: Enabling Generalist Video Generation via A Multi-Agent Framework 64 upvotes, #1 of 2024-03-21
- LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models 47 upvotes, #2 of 2024-03-21
- Evolutionary Optimization of Model Merging Recipes 41 upvotes, #3 of 2024-03-21
- SceneScript: Reconstructing Scenes With An Autoregressive Structured Language Model 30 upvotes, #4 of 2024-03-21
- When Do We Not Need Larger Vision Models? 23 upvotes, #5 of 2024-03-21
- IDAdapter: Learning Mixed Features for Tuning-Free Personalization of Text-to-Image Models 20 upvotes, #6 of 2024-03-21
- RadSplat: Radiance Field-Informed Gaussian Splatting for Robust Real-Time Rendering with 900+ FPS 17 upvotes, #7 of 2024-03-21
- HyperLLaVA: Dynamic Visual and Language Expert Tuning for Multimodal Large Language Models 16 upvotes, #8 of 2024-03-21
- RewardBench: Evaluating Reward Models for Language Modeling 14 upvotes, #9 of 2024-03-21
- ZigMa: Zigzag Mamba Diffusion Model 14 upvotes, #9 of 2024-03-21
- DepthFM: Fast Monocular Depth Estimation with Flow Matching 13 upvotes, #11 of 2024-03-21
- Magic Fixup: Streamlining Photo Editing by Watching Dynamic Videos 11 upvotes, #12 of 2024-03-21
- Be-Your-Outpainter: Mastering Video Outpainting through Input-Specific Adaptation 10 upvotes, #13 of 2024-03-21
- Reverse Training to Nurse the Reversal Curse 10 upvotes, #13 of 2024-03-21
- Towards 3D Molecule-Text Interpretation in Language Models 9 upvotes, #15 of 2024-03-21
- VSTAR: Generative Temporal Nursing for Longer Dynamic Video Synthesis 9 upvotes, #15 of 2024-03-21
- Compress3D: a Compressed Latent Space for 3D Generation from a Single Image 8 upvotes, #17 of 2024-03-21
- Evaluating Frontier Models for Dangerous Capabilities 7 upvotes, #18 of 2024-03-21
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.