Jie Fu
Jie Fu on Hugging Face Daily Papers: 23 papers, 6 in the top 3 of their day, 717 upvotes.
- Re:Form -- Reducing Human Priors in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny 17 upvotes, #7 of 2025-07-24
- Thinker: Learning to Think Fast and Slow 10 upvotes, #35 of 2025-05-28
- Learning from Failures in Multi-Attempt Reinforcement Learning 17 upvotes, #12 of 2025-03-10
- Generating Symbolic World Models via Test-time Scaling of Large Language Models 16 upvotes, #11 of 2025-02-10
- OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models 99 upvotes, #1 of 2024-11-08
- MIO: A Foundation Model on Multimodal Tokens 46 upvotes, #2 of 2024-09-30
- Layerwise Recurrent Router for Mixture-of-Experts 30 upvotes, #4 of 2024-08-14
- A Closer Look into Mixture-of-Experts in Large Language Models 12 upvotes, #5 of 2024-06-27
- Unlocking Continual Learning Abilities in Language Models 27 upvotes, #3 of 2024-06-26
- Efficient Continual Pre-training by Mitigating the Stability Gap 18 upvotes, #9 of 2024-06-25
- PIN: A Knowledge-Intensive Dataset for Paired and Interleaved Multimodal Documents 20 upvotes, #10 of 2024-06-21
- VCR: Visual Caption Restoration 10 upvotes, #16 of 2024-06-13
- LoGAH: Predicting 774-Million-Parameter Transformers using Graph HyperNetworks with 1/100 Parameters 10 upvotes, #11 of 2024-05-28
- Stacking Your Transformers: A Closer Look at Model Growth for Efficient LLM Pre-Training 20 upvotes, #5 of 2024-05-27
- CodeEditorBench: Evaluating Code Editing Capability of Large Language Models 14 upvotes, #7 of 2024-04-05
- StructLM: Towards Building Generalist Models for Structured Knowledge Grounding 27 upvotes, #7 of 2024-02-27
- ChatMusician: Understanding and Generating Music Intrinsically with LLM 57 upvotes, #1 of 2024-02-27
- OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement 84 upvotes, #1 of 2024-02-23
- AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling 45 upvotes, #3 of 2024-02-20
- CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark 27 upvotes, #4 of 2024-01-23
- E^2-LLM: Efficient and Extreme Length Extension of Large Language Models 26 upvotes, #4 of 2024-01-17
- MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training 6 upvotes, #7 of 2023-06-02
- Think Before You Act: Decision Transformers with Internal Working Memory 4 upvotes, #5 of 2023-05-29
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.