Zhengxi Lu
Zhengxi Lu on Hugging Face Daily Papers: 30 papers, 9 in the top 3 of their day, 1,428 upvotes.
- HybridCUA: Learning to Orchestrate GUI and CLI for Computer-Use Agents 47 upvotes, #24 of 2026-09-30
- IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis 20 upvotes, #12 of 2026-09-25
- Reflect, Revise, Reuse: Training-Free Skill Evolution for GUI Agents 45 upvotes, #14 of 2026-09-18
- RetireOPD: Self-Retiring On-Policy Distillation for Agentic Reinforcement Learning 56 upvotes, #9 of 2026-09-18
- PaperGym: Rubric-Centered Evolution for Research-Plan Generation 39 upvotes, #7 of 2026-09-01
- TTPO: Test-Time Policy Optimization 72 upvotes, #5 of 2026-08-28
- Agent-G^2: Gaussian Guidance for Agentic Reinforcement Learning 27 upvotes, #8 of 2026-08-27
- EnvACE: Internalizing Environment Dynamics via World Rehearsal for Agentic Reinforcement Learning 43 upvotes, #6 of 2026-08-07
- AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning 94 upvotes, #1 of 2026-08-07
- Distill Where You Fail: Recovering Learning Signals of Negative RL-Groups from Adaptive Teacher Guidance 16 upvotes, #15 of 2026-08-06
- VAD: Attributing Visual Evidence for Target Reconstruction in Multimodal On-Policy Distillation 46 upvotes, #8 of 2026-08-04
- SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution 28 upvotes, #8 of 2026-07-30
- SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning 99 upvotes, #3 of 2026-07-17
- OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning 54 upvotes, #3 of 2026-06-26
- GUI-CIDER: Mid-training GUI Agents via Causal Internalization and Density-aware Exemplar Reselection 23 upvotes, #17 of 2026-05-28
- Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles 20 upvotes, #19 of 2026-05-22
- Self-Distilled Agentic Reinforcement Learning 107 upvotes, #2 of 2026-05-15
- Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning 106 upvotes, #1 of 2026-05-08
- UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Grounding 10 upvotes, #18 of 2026-04-16
- UI-Copilot: Advancing Long-Horizon GUI Automation via Tool-Integrated Policy Optimization 6 upvotes, #23 of 2026-04-16
- KnowU-Bench: Towards Interactive, Proactive, and Personalized Mobile Agent Evaluation 47 upvotes, #12 of 2026-04-10
- SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization 92 upvotes, #4 of 2026-04-03
- Code-A1: Adversarial Evolving of Code LLM and Test LLM via Reinforcement Learning 10 upvotes, #20 of 2026-03-17
- MemGUI-Bench: Benchmarking Memory of Mobile GUI Agents in Dynamic Environments 13 upvotes, #15 of 2026-02-09
- UI-S1: Advancing GUI Automation via Semi-online Reinforcement Learning 45 upvotes, #2 of 2025-09-16
- MAS-Bench: A Unified Benchmark for Shortcut-Augmented Hybrid Mobile GUI Agents 2 upvotes, #21 of 2025-09-09
- Mobile-Agent-v3: Foundamental Agents for GUI Automation 55 upvotes, #3 of 2025-08-22
- Test-Time Reinforcement Learning for GUI Grounding via Region Consistency 20 upvotes, #12 of 2025-08-13
- GUI-G^2: Gaussian Reward Modeling for GUI Grounding 122 upvotes, #1 of 2025-07-22
- UI-R1: Enhancing Action Prediction of GUI Agents by Reinforcement Learning 54 upvotes, #3 of 2025-03-28
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.