VISIONx @ NYU
VISIONx @ NYU on Hugging Face Daily Papers: 7 papers, 1 in the top 3 of their day, 0 paper of the day.
- PaintBench: Deterministic Evaluation of Precise Visual Editing 3 upvotes, #34 of 2026-06-04
- Benchmarking Visual State Tracking in Multimodal Video Understanding 23 upvotes, #10 of 2026-06-03
- Solaris: Building a Multiplayer Video World Model in Minecraft 27 upvotes, #6 of 2026-02-26
- Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders 51 upvotes, #8 of 2026-01-23
- Benchmark Designers Should "Train on the Test Set" to Expose Exploitable Non-Visual Shortcuts 7 upvotes, #9 of 2025-11-07
- SIMS-V: Simulated Instruction-Tuning for Spatial Video Understanding 4 upvotes, #11 of 2025-11-07
- Diffusion Transformers with Representation Autoencoders 155 upvotes, #2 of 2025-10-14
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.