Two identically priced AI models tested on Indonesian benchmark: 56.52% vs 26.09%
Mistral Small 3 and Llama 3.1 8B Instruct are sold at the exact same price and are equally fast. On 69 items testing normalization of abbreviated writing to standard Indonesian, one scored 56.52%, the other 26.09%.
69 itemsItems56.52%Mistral Small 326.09%Llama 3.1 8B