Junxiao Yang
Junxiao Yang on Hugging Face Daily Papers: 9 papers, 1 in the top 3 of their day, 241 upvotes.
- AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security 142 upvotes, #1 of 2026-05-29
- LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety 5 upvotes, #26 of 2026-04-15
- Loong: Synthesize Long Chain-of-Thoughts at Scale through Verifiers 22 upvotes, #8 of 2025-09-05
- Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen! 13 upvotes, #19 of 2025-05-22
- How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study 13 upvotes, #19 of 2025-05-22
- BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs 11 upvotes, #23 of 2025-05-22
- AISafetyLab: A Comprehensive Framework for AI Safety Evaluation and Improvement 5 upvotes, #17 of 2025-02-27
- Agent-SafetyBench: Evaluating the Safety of LLM Agents 8 upvotes, #14 of 2024-12-24
- Safe Unlearning: A Surprisingly Effective and Generalizable Solution to Defend Against Jailbreak Attacks 9 upvotes, #12 of 2024-07-08
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.