Peter J Liu
Peter J Liu on Hugging Face Daily Papers: 7 papers, 3 in the top 3 of their day, 103 upvotes.
- LiPO: Listwise Preference Optimization through Learning-to-Rank 20 upvotes, #6 of 2024-02-06
- Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models 29 upvotes, #2 of 2023-12-12
- Frontier Language Models are not Robust to Adversarial Arithmetic, or "What do I need to say so you agree 2+2=5? 4 upvotes, #12 of 2023-11-15
- Improving Large Language Model Fine-tuning for Solving Math Problems 6 upvotes, #10 of 2023-10-17
- Small-scale proxies for large-scale Transformer training instabilities 22 upvotes, #3 of 2023-09-26
- Statistical Rejection Sampling Improves Preference Optimization 15 upvotes, #4 of 2023-09-14
- SLiC-HF: Sequence Likelihood Calibration with Human Feedback 7 upvotes, #3 of 2023-05-18
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.