TongZheng
TongZheng on Hugging Face Daily Papers: 8 papers, 1 in the top 3 of their day, 254 upvotes.
- LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling 64 upvotes, #5 of 2026-05-11
- VOGUE: Guiding Exploration with Visual Uncertainty Improves Multimodal Reasoning 19 upvotes, #16 of 2025-10-03
- CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models 28 upvotes, #5 of 2025-09-11
- Parallel-R1: Towards Parallel Thinking via Reinforcement Learning 95 upvotes, #2 of 2025-09-10
- R1-RE: Cross-Domain Relationship Extraction with RLVR 6 upvotes, #21 of 2025-07-08
- Learning to Reason via Mixture-of-Thought for Logical Reasoning 17 upvotes, #13 of 2025-05-22
- Beyond Decoder-only: Large Language Models Can be Good Encoders for Machine Translation 5 upvotes, #29 of 2025-03-12
- Towards Optimal Multi-draft Speculative Decoding 4 upvotes, #21 of 2025-02-27
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.