Kai-Wei Chang

Kai-Wei Chang on Hugging Face Daily Papers: 14 papers, 4 in the top 3 of their day, 411 upvotes.

  1. OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks 48 upvotes, #11 of 2026-04-10
  2. X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents 30 upvotes, #6 of 2025-04-22
  3. When To Solve, When To Verify: Compute-Optimal Problem Solving and Generative Verification for LLM Reasoning 15 upvotes, #18 of 2025-04-02
  4. OpenVLThinker: An Early Exploration to Complex Vision-Language Reasoning via Iterative Self-Improvement 20 upvotes, #9 of 2025-03-24
  5. STIV: Scalable Text and Image Conditioned Video Generation 67 upvotes, #1 of 2024-12-11
  6. LongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory 9 upvotes, #13 of 2024-10-15
  7. Data Advisor: Dynamic Data Curation for Safety Alignment of Large Language Models 3 upvotes, #22 of 2024-10-11
  8. MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding 17 upvotes, #9 of 2024-06-14
  9. MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? 45 upvotes, #1 of 2024-03-22
  10. ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models 11 upvotes, #5 of 2024-01-25
  11. TrustLLM: Trustworthiness in Large Language Models 69 upvotes, #1 of 2024-01-12
  12. Lumos: Learning Agents with Unified Data, Modular Design, and Open-Source LLMs 28 upvotes, #4 of 2023-11-13
  13. FLIRT: Feedback Loop In-context Red Teaming 14 upvotes, #3 of 2023-08-09
  14. AVIS: Autonomous Visual Information Seeking with Large Language Models 5 upvotes, #20 of 2023-06-16

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.