TongZheng

TongZheng on Hugging Face Daily Papers: 8 papers, 1 in the top 3 of their day, 254 upvotes.

  1. LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling 64 upvotes, #5 of 2026-05-11
  2. VOGUE: Guiding Exploration with Visual Uncertainty Improves Multimodal Reasoning 19 upvotes, #16 of 2025-10-03
  3. CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models 28 upvotes, #5 of 2025-09-11
  4. Parallel-R1: Towards Parallel Thinking via Reinforcement Learning 95 upvotes, #2 of 2025-09-10
  5. R1-RE: Cross-Domain Relationship Extraction with RLVR 6 upvotes, #21 of 2025-07-08
  6. Learning to Reason via Mixture-of-Thought for Logical Reasoning 17 upvotes, #13 of 2025-05-22
  7. Beyond Decoder-only: Large Language Models Can be Good Encoders for Machine Translation 5 upvotes, #29 of 2025-03-12
  8. Towards Optimal Multi-draft Speculative Decoding 4 upvotes, #21 of 2025-02-27

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.