Daniil Gavrilov

Daniil Gavrilov on Hugging Face Daily Papers: 15 papers, 6 in the top 3 of their day, 653 upvotes.

  1. Rank-Then-Act: Reward-Free Control from Frame-Order Progress 7 upvotes, #25 of 2026-07-08
  2. Unstable Features, Reproducible Subspaces: Understanding Seed Dependence in Sparse Autoencoders 8 upvotes, #28 of 2026-06-16
  3. Interpreting and Steering a Text-to-Speech Language Model with Sparse Autoencoders 12 upvotes, #19 of 2026-06-10
  4. Trust-Region Behavior Blending for On-Policy Distillation 65 upvotes, #3 of 2026-06-01
  5. Next Embedding Prediction Makes World Models Stronger 17 upvotes, #11 of 2026-03-04
  6. F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare 70 upvotes, #1 of 2026-02-09
  7. Enhancing Vision-Language Model Training with Reinforcement Learning in Synthetic Worlds for Real-World Success 34 upvotes, #7 of 2025-08-07
  8. Teach Old SAEs New Domain Tricks with Boosting 11 upvotes, #12 of 2025-07-18
  9. Train Sparse Autoencoders Efficiently by Utilizing Features Correlation 21 upvotes, #16 of 2025-05-30
  10. You Do Not Fully Utilize Transformer's Representation Capacity 33 upvotes, #8 of 2025-02-19
  11. Analyze Feature Flow to Enhance Interpretation and Steering in Language Models 53 upvotes, #1 of 2025-02-07
  12. The Differences Between Direct Alignment Algorithms are a Blur 109 upvotes, #2 of 2025-02-04
  13. Mechanistic Permutability: Match Features Across Layers 16 upvotes, #7 of 2024-10-14
  14. Learn Your Reference Model for Real Good Alignment 75 upvotes, #1 of 2024-04-16
  15. Linear Transformers with Learnable Kernel Functions are Better In-Context Models 81 upvotes, #1 of 2024-02-19

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.