5 ms·
Outperforms GPT-3.5 by 16% with less than 1B parameters. There is even a 220M parameter version that scores well. Interesting model to watch. Referenced snippe
by txtai 4y ago
Outperforms GPT-3.5 by 16% with less than 1B parameters. There is even a 220M parameter version that scores well. Interesting model to watch.
Referenced snippet from the abstract:
With Multimodal-CoT, our model under 1 billion parameters outperforms the previous state-of-the-art LLM (GPT-3.5) by 16% (75.17%->91.68%) on the ScienceQA benchmark and even surpasses human performance.