6 of 26 models asked a question written entirely in Gujarati, in its own script, answered it correctly. Answering in English is a failure here, not a pass — a model with no coverage cannot parse the question in the first place. Every verdict below is signed by the machine that took it.
6 of 26 models answered correctly · 20 asked, and could not
Answered correctly
Best recall-proved context first — the order in which you would try them.
A model absent from both lists was never asked in Gujarati, and that is not a claim either way. The probe set covers a fixed list of languages; a script missing from it is missing from the evidence, not from the model.