Rafael Rafailov

Rafael Rafailov on Hugging Face Daily Papers: 8 papers, 2 in the top 3 of their day, 225 upvotes.

  1. PERSONA: A Reproducible Testbed for Pluralistic Alignment 16 upvotes, #5 of 2024-07-25
  2. MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation? 49 upvotes, #1 of 2024-07-09
  3. OpenVLA: An Open-Source Vision-Language-Action Model 28 upvotes, #5 of 2024-06-14
  4. Offline Regularised Reinforcement Learning for Large Language Models Alignment 11 upvotes, #6 of 2024-05-30
  5. Diffusion Model Alignment Using Direct Preference Optimization 48 upvotes, #4 of 2023-11-23
  6. Contrastive Prefence Learning: Learning from Human Feedback without RL 25 upvotes, #1 of 2023-10-23
  7. An Emulator for Fine-Tuning Large Language Models using Small Language Models 13 upvotes, #8 of 2023-10-20
  8. Contrastive Example-Based Control 4 upvotes, #5 of 2023-07-26

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.