Weizhe Yuan

Weizhe Yuan on Hugging Face Daily Papers: 7 papers, 2 in the top 3 of their day, 332 upvotes.

  1. O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? 35 upvotes, #2 of 2024-11-26
  2. Thinking LLMs: General Instruction Following with Thought Generation 7 upvotes, #15 of 2024-10-15
  3. Self-Taught Evaluators 16 upvotes, #6 of 2024-08-06
  4. Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge 19 upvotes, #14 of 2024-07-30
  5. Iterative Reasoning Preference Optimization 35 upvotes, #5 of 2024-05-01
  6. Self-Rewarding Language Models 156 upvotes, #1 of 2024-01-19
  7. System-Level Natural Language Feedback 11 upvotes, #5 of 2023-06-26

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.