Zhenting Wang

Zhenting Wang on Hugging Face Daily Papers: 7 papers, 3 in the top 3 of their day, 294 upvotes.

  1. Memex(RL): Scaling Long-Horizon LLM Agents via Indexed Experience Memory 19 upvotes, #7 of 2026-03-05
  2. EPO: Entropy-regularized Policy Optimization for LLM Agents Reinforcement Learning 124 upvotes, #2 of 2025-09-29
  3. MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers 56 upvotes, #3 of 2025-08-29
  4. DUMP: Automated Distribution-Level Curriculum Learning for RL-based LLM Post-training 19 upvotes, #10 of 2025-04-15
  5. MLLM-as-a-Judge for Image Safety without Human Labeling 23 upvotes, #8 of 2025-01-03
  6. Token-Budget-Aware LLM Reasoning 40 upvotes, #1 of 2024-12-26
  7. Uncertainty is Fragile: Manipulating Uncertainty in Large Language Models 1 upvotes, #19 of 2024-07-17

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.