Daily Papers of 2024-06-26

  1. The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale 72 upvotes, #1 of 2024-06-26
  2. YouDream: Generating Anatomically Controllable Consistent Text-to-3D Animals 39 upvotes, #2 of 2024-06-26
  3. Unlocking Continual Learning Abilities in Language Models 27 upvotes, #3 of 2024-06-26
  4. Aligning Diffusion Models with Noise-Conditioned Perception 24 upvotes, #4 of 2024-06-26
  5. DiffusionPDE: Generative PDE-Solving Under Partial Observation 23 upvotes, #5 of 2024-06-26
  6. LongIns: A Challenging Long-context Instruction-based Exam for LLMs 18 upvotes, #6 of 2024-06-26
  7. MG-LLaVA: Towards Multi-Granularity Visual Instruction Tuning 18 upvotes, #6 of 2024-06-26
  8. MotionBooth: Motion-Aware Customized Text-to-Video Generation 17 upvotes, #8 of 2024-06-26
  9. APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets 17 upvotes, #8 of 2024-06-26
  10. Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA 13 upvotes, #10 of 2024-06-26
  11. Segment Any Text: A Universal Approach for Robust, Efficient and Adaptable Sentence Segmentation 12 upvotes, #11 of 2024-06-26
  12. On the Transformations across Reward Model, Parameter Update, and In-Context Prompt 11 upvotes, #12 of 2024-06-26
  13. FreeTraj: Tuning-Free Trajectory Control in Video Diffusion Models 10 upvotes, #13 of 2024-06-26
  14. DialSim: A Real-Time Simulator for Evaluating Long-Term Dialogue Understanding of Conversational Agents 9 upvotes, #14 of 2024-06-26
  15. Image Conductor: Precision Control for Interactive Video Synthesis 8 upvotes, #15 of 2024-06-26
  16. Grass: Compute Efficient Low-Memory LLM Training with Structured Sparse Gradients 5 upvotes, #16 of 2024-06-26
  17. Large Language Models Assume People are More Rational than We Really are 4 upvotes, #17 of 2024-06-26
  18. Multi-property Steering of Large Language Models with Dynamic Activation Composition 4 upvotes, #17 of 2024-06-26
  19. Cross-Modality Safety Alignment 3 upvotes, #19 of 2024-06-26
  20. Fast and Uncertainty-Aware SVBRDF Recovery from Multi-View Capture using Frequency Domain Analysis 3 upvotes, #19 of 2024-06-26

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.