TY.Zheng
TY.Zheng on Hugging Face Daily Papers: 27 papers, 13 in the top 3 of their day, 1,347 upvotes.
- LoopCoder-v2: Only Loop Once for Efficient Test-Time Computation Scaling 203 upvotes, #2 of 2026-06-17
- InCoder-32B: Code Foundation Model for Industrial Scenarios 297 upvotes, #2 of 2026-03-18
- Search More, Think Less: Rethinking Long-Horizon Agentic Search for Efficiency and Generalization 22 upvotes, #8 of 2026-02-27
- Dynamic Large Concept Models: Latent Reasoning in an Adaptive Semantic Space 54 upvotes, #2 of 2026-01-02
- Beyond Correctness: Evaluating Subjective Writing Preferences Across Cultures 10 upvotes, #23 of 2025-10-17
- COIG-Writer: A High-Quality Dataset for Chinese Creative Writing with Thought Processes 13 upvotes, #18 of 2025-10-17
- IWR-Bench: Can LVLMs reconstruct interactive webpage from a user interaction video? 3 upvotes, #64 of 2025-09-30
- TreePO: Bridging the Gap of Policy Optimization and Efficacy and Inference Efficiency with Heuristic Tree-based Modeling 76 upvotes, #2 of 2025-08-27
- First Return, Entropy-Eliciting Explore 23 upvotes, #8 of 2025-07-10
- Does Math Reasoning Improve General LLM Capabilities? Understanding Transferability of LLM Reasoning 62 upvotes, #2 of 2025-07-02
- COIG-P: A High-Quality and Large-Scale Chinese Preference Dataset for Alignment with Human Values 41 upvotes, #5 of 2025-04-09
- YuE: Scaling Open Foundation Models for Long-Form Music Generation 57 upvotes, #3 of 2025-03-12
- SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines 92 upvotes, #3 of 2025-02-21
- Steel-LLM:From Scratch to Open Source -- A Personal Journey in Building a Chinese-Centric LLM 4 upvotes, #22 of 2025-02-11
- MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale 42 upvotes, #3 of 2024-12-09
- AutoKaggle: A Multi-Agent Framework for Autonomous Data Science Competitions 36 upvotes, #2 of 2024-10-30
- A Comparative Study on Reasoning Patterns of OpenAI's o1 Model 15 upvotes, #15 of 2024-10-18
- MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark 27 upvotes, #4 of 2024-09-05
- I-SHEEP: Self-Alignment of LLM from Scratch through an Iterative Self-Enhancement Paradigm 30 upvotes, #2 of 2024-08-16
- MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series 41 upvotes, #1 of 2024-05-30
- MuPT: A Generative Symbolic Music Pretrained Transformer 14 upvotes, #5 of 2024-04-10
- Chinese Tiny LLM: Pretraining a Chinese-Centric Large Language Model 8 upvotes, #7 of 2024-04-08
- CodeEditorBench: Evaluating Code Editing Capability of Large Language Models 14 upvotes, #7 of 2024-04-05
- ChatMusician: Understanding and Generating Music Intrinsically with LLM 57 upvotes, #1 of 2024-02-27
- StructLM: Towards Building Generalist Models for Structured Knowledge Grounding 27 upvotes, #7 of 2024-02-27
- OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement 84 upvotes, #1 of 2024-02-23
- CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark 27 upvotes, #4 of 2024-01-23
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.