Sinhala deserves real evaluation
Most language models are measured almost entirely in English. This benchmark tests genuine knowledge and reasoning in native Sinhala, with no translation.
Build open-weight systems of 8 billion parameters or fewer that answer Sinhala multiple-choice questions across Easy, Medium and Hard levels.
Most language models are measured almost entirely in English. This benchmark tests genuine knowledge and reasoning in native Sinhala, with no translation.
Every component at inference counts towards an 8B budget; models, retrievers, rerankers, embeddings. Efficient ideas win.
Open-weight models only, no closed APIs or internet at test time. Every ranked result is reproduced.