7 ms·
Where did you get the top ten from? https://huggingface.co/spaces/TIGER-Lab/MMLU-Pro https://huggingface.co/spaces/TIGER-Lab/MMLU-Pro Are you discounting all
by Cicero22 1y ago
Where did you get the top ten from?
https://huggingface.co/spaces/TIGER-Lab/MMLU-Pro https://huggingface.co/spaces/TIGER-Lab/MMLU-Pro
Are you discounting all of the self reported scores?
- zwischenzug 1y agoCame here to say this. It's behind the 14b Phi-reasoning-plus (which is self-reported). I don't understand why "TIGER-LAb"-sourced scores are 'unknown' in terms of model size?