Hans Zhao
Hans Zhao on Hugging Face Daily Papers: 9 papers, 3 in the top 3 of their day, 201 upvotes.
- Teaching Large Language Models to Maintain Contextual Faithfulness via Synthetic Tasks and Reinforcement Learning 10 upvotes, #23 of 2025-05-26
- Multimodal Representation Alignment for Image Generation: Text-Image Interleaved Control Is Easier Than You Think 26 upvotes, #7 of 2025-02-28
- Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey 49 upvotes, #3 of 2024-12-30
- Selecting Influential Samples for Long Context Alignment via Homologous Models' Guidance and Contextual Awareness Measurement 7 upvotes, #14 of 2024-10-22
- A Spark of Vision-Language Intelligence: 2-Dimensional Autoregressive Transformer for Efficient Finegrained Image Generation 13 upvotes, #3 of 2024-10-09
- UltraEdit: Instruction-based Fine-Grained Image Editing at Scale 8 upvotes, #8 of 2024-07-09
- MMEvalPro: Calibrating Multimodal Benchmarks Towards Trustworthy and Efficient Evaluation 35 upvotes, #3 of 2024-07-02
- An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models 23 upvotes, #4 of 2024-03-12
- ML-Bench: Large Language Models Leverage Open-source Libraries for Machine Learning Tasks 10 upvotes, #6 of 2023-11-17
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.