Measured Collection

Articles

We test how well AI actually understands Indonesian and its regional languages, map the words that trip up native speakers, and publish the numbers along with the raw data and their limits, including when the results are not what we hoped for.

2 articlesEvery article links to its raw data
FilteredOpenaiClear filter

2 articles

Empirical

GPT-6 Astra and Luna tie on Indonesian slang items

GPT-6 Astra scored 42 of 42 on the public half of our Indonesian slang benchmark, the joint best result in the panel. Counting the 18 withheld items too, it finished 59 of 60, 1 behind OpenAI's own cheaper GPT-5.6 Luna. 1 item apart on 60 is not a statistically meaningful gap, and that is the more interesting finding: the public half no longer separates the strongest models at all.

42/42Astra public59/60Astra total$0.15Astra cost