Hao Liang

Hao Liang on Hugging Face Daily Papers: 20 papers, 8 in the top 3 of their day, 1,431 upvotes.

  1. DataFlex-RL: An Evaluation Platform for RLVR Data Policies 162 upvotes, #2 of 2026-09-14
  2. DataPrep-Bench: Benchmarking LLMs as Training Data Preparators 55 upvotes, #1 of 2026-07-27
  3. K12-KGraph: A Curriculum-Aligned Knowledge Graph for Benchmarking and Training Educational LLMs 63 upvotes, #2 of 2026-07-24
  4. DataFlow-Harness: A Grounded Code-Agent Platform for Constructing Editable LLM Data Pipelines 137 upvotes, #2 of 2026-07-22
  5. LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning 46 upvotes, #8 of 2026-05-22
  6. OpenWorldLib: A Unified Codebase and Definition of Advanced World Models 200 upvotes, #2 of 2026-04-07
  7. DataFlex: A Unified Framework for Data-Centric Dynamic Training of Large Language Models 345 upvotes, #1 of 2026-04-03
  8. One-Eval: An Agentic System for Automated and Traceable LLM Evaluation 11 upvotes, #23 of 2026-03-18
  9. BrowseComp-V^3: A Visual, Vertical, and Verifiable Benchmark for Multimodal Browsing Agents 8 upvotes, #15 of 2026-02-17
  10. Research on World Models Is Not Merely Injecting World Knowledge into Specific Tasks 46 upvotes, #7 of 2026-02-04
  11. DataFlow: An LLM-Driven Framework for Unified Data Preparation and Workflow Automation in the Era of Data-Centric AI 193 upvotes, #1 of 2025-12-23
  12. VABench: A Comprehensive Benchmark for Audio-Video Generation 7 upvotes, #19 of 2025-12-18
  13. Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling 40 upvotes, #4 of 2025-12-17
  14. MorphoBench: A Benchmark with Difficulty Adaptive to Model Reasoning 19 upvotes, #13 of 2025-10-20
  15. Multimodal Reasoning for Science: Technical Report and 1st Place Solution to the ICML 2025 SeePhys Challenge 6 upvotes, #11 of 2025-09-17
  16. Learning What Reinforcement Learning Can't: Interleaved Online Fine-Tuning for Hardest Questions 3 upvotes, #37 of 2025-06-10
  17. Baichuan-Omni-1.5 Technical Report 48 upvotes, #2 of 2025-01-28
  18. Document Parsing Unveiled: Techniques, Challenges, and Prospects for Structured Information Extraction 28 upvotes, #5 of 2024-10-29
  19. Baichuan Alignment Technical Report 46 upvotes, #6 of 2024-10-22
  20. PAS: Data-Efficient Plug-and-Play Prompt Augmentation System 8 upvotes, #8 of 2024-07-09

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.