This tiny AI model tried to Google our exam answers
The two smallest models we have ever tested, LFM2.5 at 2.6 billion parameters and hy-mt2 at 1.8 billion, faced off on 60 Indonesian slang questions. The scores ended in a dead heat, but the two failed in opposite ways, and one of them tried calling Google mid-exam.
34/60hy-mt2 (1.8B)32/60LFM2.5 (2.6B)358xToken ratio