yd yang

yd yang on Hugging Face Daily Papers: 11 papers, 4 in the top 3 of their day, 394 upvotes.

  1. Heterogeneous Agent Collaborative Reinforcement Learning 170 upvotes, #1 of 2026-03-05
  2. ECO: Energy-Constrained Optimization with Reinforcement Learning for Humanoid Walking 3 upvotes, #36 of 2026-02-10
  3. Your Group-Relative Advantage Is Biased 144 upvotes, #1 of 2026-01-19
  4. Re:Form -- Reducing Human Priors in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny 17 upvotes, #7 of 2025-07-24
  5. A Survey on Vision-Language-Action Models: An Action Tokenization Perspective 30 upvotes, #4 of 2025-07-03
  6. A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment 13 upvotes, #11 of 2025-04-24
  7. DexGraspVLA: A Vision-Language-Action Framework Towards General Dexterous Grasping 7 upvotes, #12 of 2025-03-03
  8. ProgressGym: Alignment with a Millennium of Moral Progress 2 upvotes, #27 of 2024-07-02
  9. In-Context Editing: Learning Knowledge from Self-Induced Distributions 14 upvotes, #13 of 2024-06-18
  10. JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models 37 upvotes, #2 of 2023-11-13
  11. Safe RLHF: Safe Reinforcement Learning from Human Feedback 28 upvotes, #3 of 2023-10-20

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.