Misha Khalman
Misha Khalman on Hugging Face Daily Papers: 5 papers, 2 in the top 3 of their day, 128 upvotes.
- Direct Language Model Alignment from Online AI Feedback 36 upvotes, #4 of 2024-02-08
- LiPO: Listwise Preference Optimization through Learning-to-Rank 20 upvotes, #6 of 2024-02-06
- Gemini: A Family of Highly Capable Multimodal Models 50 upvotes, #2 of 2023-12-20
- Statistical Rejection Sampling Improves Preference Optimization 15 upvotes, #4 of 2023-09-14
- SLiC-HF: Sequence Likelihood Calibration with Human Feedback 7 upvotes, #3 of 2023-05-18
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.