How much do compact models really understand සිංහල?

Build open-weight systems of 8 billion parameters or fewer that answer Sinhala multiple-choice questions across Easy, Medium and Hard levels.

Next Up: -
00days
00hours
00min
00sec
why this task
Low-resource

Sinhala deserves real evaluation

Most language models are measured almost entirely in English. This benchmark tests genuine knowledge and reasoning in native Sinhala, with no translation.

≤ 8B

Small enough to run locally

Every component at inference counts towards an 8B budget; models, retrievers, rerankers, embeddings. Efficient ideas win.

Reproducible

Open and verifiable

Open-weight models only, no closed APIs or internet at test time. Every ranked result is reproduced.

how it works
  1. Register your team (1-4 members) with one corresponding member.
  2. Develop on the dev set with the public baseline.
  3. Submit test predictions on the Codabench leaderboard.
  4. Share your system & code to finalize your submission.
  5. Present your work at ICTer 2026.