Daily Papers of 2024-03-20
- mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding 24 upvotes, #1 of 2024-03-20
- LLMLingua-2: Data Distillation for Efficient and Faithful Task-Agnostic Prompt Compression 20 upvotes, #2 of 2024-03-20
- AnimateDiff-Lightning: Cross-Model Diffusion Distillation 17 upvotes, #3 of 2024-03-20
- TnT-LLM: Text Mining at Scale with Large Language Models 15 upvotes, #4 of 2024-03-20
- Vid2Robot: End-to-end Video-conditioned Policy Learning with Cross-Attention Transformers 13 upvotes, #5 of 2024-03-20
- Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models 11 upvotes, #6 of 2024-03-20
- GaussianFlow: Splatting Gaussian Dynamics for 4D Content Creation 9 upvotes, #7 of 2024-03-20
- Chart-based Reasoning: Transferring Capabilities from LLMs to VLMs 9 upvotes, #7 of 2024-03-20
- ComboVerse: Compositional 3D Assets Creation Using Spatially-Aware Diffusion Guidance 8 upvotes, #9 of 2024-03-20
- FRESCO: Spatial-Temporal Correspondence for Zero-Shot Video Translation 6 upvotes, #10 of 2024-03-20
- FouriScale: A Frequency Perspective on Training-Free High-Resolution Image Synthesis 6 upvotes, #10 of 2024-03-20
- GVGEN: Text-to-3D Generation with Volumetric Representation 4 upvotes, #12 of 2024-03-20
- TexDreamer: Towards Zero-Shot High-Fidelity 3D Human Texture Generation 3 upvotes, #13 of 2024-03-20
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.