théo gigant
théo gigant on Hugging Face Daily Papers: 3 papers, 0 in the top 3 of their day, 57 upvotes.
- Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation 11 upvotes, #20 of 2026-05-21
- Efficient Pre-Training with Token Superposition 42 upvotes, #8 of 2026-05-13
- Summarization of Multimodal Presentations with Vision-Language Models: Study of the Effect of Modalities and Structure 3 upvotes, #26 of 2025-04-16
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.