Insight

How far can AI and our own language tools be trusted?

We test our own correction engine and various outside AI models against Indonesian and regional languages, then publish the numbers as they are. On this page, we sum up each finding in a single picture. The full method and raw data are one click away on every card.

2 findingsEvery finding links to its technical version
FilteredGrammarClear filter

2 findings

Correction engineStructured Data
32.8
Unusable 32.8Adaptable 64.3
Of the 2,909 LanguageTool rules we mined

Our Closest Language Relative on LanguageTool Only Has 44 Rules

Before writing grammar rules from scratch, we first checked what could be borrowed from LanguageTool, the largest open-source grammar checker there is. It has modules for 35 languages. The world's fourth most spoken language isn't one of them.

August 21, 2026 · 2 menit bacaRead the finding →
JavaneseEmpirical
Voided, book vs speaker 6Kept in the set 34
Out of the 40 Javanese items we built

6 of 40 Javanese Test Items Voided Because We Trusted the Book Too Much

We built Javanese test items from grammar-book rules. Six of 40 items were dropped, not because AI answered wrong, but because the form the book called wrong turned out to be widely accepted by native speakers.

August 14, 2026 · 3 menit bacaRead the finding →