Akbir Khan

Akbir Khan on Hugging Face Daily Papers: 4 papers, 0 in the top 3 of their day, 46 upvotes.

  1. Alignment faking in large language models 7 upvotes, #18 of 2024-12-19
  2. BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games 17 upvotes, #5 of 2024-11-25
  3. Language Models Learn to Mislead Humans via RLHF 8 upvotes, #11 of 2024-09-20
  4. JaxMARL: Multi-Agent RL Environments in JAX 7 upvotes, #7 of 2023-11-17

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.