1. 16 Oct, 2025 6 commits
  2. 15 Oct, 2025 9 commits
  3. 14 Oct, 2025 2 commits
  4. 02 Oct, 2025 1 commit
  5. 21 Sep, 2025 1 commit
  6. 12 Sep, 2025 1 commit
  7. 08 Sep, 2025 2 commits
  8. 27 Aug, 2025 1 commit
  9. 26 Aug, 2025 1 commit
    • Janna's avatar
      Support for AIME dataset (#3248) · 5ac7cdf8
      Janna authored
      * add AIME tasks
      
      * standardize the repeats
      
      * fix task naming
      
      * aime25 only has test set
      
      * edit readme
      
      * add utils
      
      * standardize
      
      * fix case sensitivity
      
      * repeat once
      
      * lint
      
      * more linting
      
      * lint huggingface.py
      5ac7cdf8
  10. 25 Aug, 2025 1 commit
  11. 21 Aug, 2025 2 commits
  12. 13 Aug, 2025 1 commit
  13. 02 Aug, 2025 1 commit
  14. 24 Jul, 2025 2 commits
  15. 23 Jul, 2025 3 commits
  16. 18 Jul, 2025 2 commits
  17. 16 Jul, 2025 1 commit
    • Baber Abbasi's avatar
      truncate thinking tags in generations (#3145) · 51ede33c
      Baber Abbasi authored
      * feat: add postprocessing for generated text to strip stop sequences and thinking tokens
      
      * nit
      
      * fix: trim leading whitespace after stripping thinking tokens from generation
      
      * feat: add think_end_token to model_args
      
      * nit
      
      * nit
      
      * nit
      
      * add to readme
      
      * nit
      51ede33c
  18. 15 Jul, 2025 1 commit
  19. 14 Jul, 2025 1 commit
  20. 06 Jul, 2025 1 commit