Daily Papers of 2024-07-23

  1. NNsight and NDIF: Democratizing Access to Foundation Model Internals 32 upvotes, #1 of 2024-07-23
  2. Compact Language Models via Pruning and Knowledge Distillation 32 upvotes, #1 of 2024-07-23
  3. SlowFast-LLaVA: A Strong Training-Free Baseline for Video Large Language Models 32 upvotes, #1 of 2024-07-23
  4. Knowledge Mechanisms in Large Language Models: A Survey and Perspective 31 upvotes, #4 of 2024-07-23
  5. POGEMA: A Benchmark Platform for Cooperative Multi-Agent Navigation 18 upvotes, #5 of 2024-07-23
  6. VideoGameBunny: Towards vision assistants for video games 18 upvotes, #5 of 2024-07-23
  7. LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding 18 upvotes, #5 of 2024-07-23
  8. BoostMVSNeRFs: Boosting MVS-based NeRFs to Generalizable View Synthesis in Large-scale Scenes 15 upvotes, #8 of 2024-07-23
  9. BOND: Aligning LLMs with Best-of-N Distillation 13 upvotes, #9 of 2024-07-23
  10. Artist: Aesthetically Controllable Text-Driven Stylization without Training 10 upvotes, #10 of 2024-07-23
  11. Consent in Crisis: The Rapid Decline of the AI Data Commons 9 upvotes, #11 of 2024-07-23
  12. MusiConGen: Rhythm and Chord Control for Transformer-Based Text-to-Music Generation 9 upvotes, #11 of 2024-07-23
  13. HoloDreamer: Holistic 3D Panoramic World Generation from Text Descriptions 9 upvotes, #11 of 2024-07-23
  14. Cinemo: Consistent and Controllable Image Animation with Motion Diffusion Models 9 upvotes, #11 of 2024-07-23
  15. Discrete Flow Matching 8 upvotes, #15 of 2024-07-23
  16. AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks? 7 upvotes, #16 of 2024-07-23
  17. Conditioned Language Policy: A General Framework for Steerable Multi-Objective Finetuning 7 upvotes, #16 of 2024-07-23
  18. CGB-DM: Content and Graphic Balance Layout Generation with Transformer-based Diffusion Model 6 upvotes, #18 of 2024-07-23
  19. MIBench: Evaluating Multimodal Large Language Models over Multiple Images 6 upvotes, #18 of 2024-07-23
  20. Temporal Residual Jacobians For Rig-free Motion Transfer 5 upvotes, #20 of 2024-07-23
  21. ThermalNeRF: Thermal Radiance Fields 5 upvotes, #20 of 2024-07-23
  22. Local All-Pair Correspondence for Point Tracking 5 upvotes, #20 of 2024-07-23
  23. GET-Zero: Graph Embodiment Transformer for Zero-shot Embodiment Generalization 4 upvotes, #23 of 2024-07-23
  24. Visual Haystacks: Answering Harder Questions About Sets of Images 2 upvotes, #24 of 2024-07-23

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.