A long-term home for Kokborok knowledge.

Yapiri Project expands the work of script creation into language documentation, teaching, publishing, community participation, and technology.

Why this exists

Kokborok has always been spoken. It has never had a stable digital home.

Kokborok is spoken across Tripura and beyond, but much of it lives only in memory — passed between generations without a structured written record, a shared spelling system, or a place where the language can be searched, studied, and built upon. Words get lost between generations. Regional variation goes undocumented. Learners have no single, reliable place to start.

The Yapiri Project addresses this directly: a dictionary that grows and gets verified over time, a grammar that documents how the language actually works, a sentence corpus that shows Kokborok in real use, and the Yapiri script — a writing system built specifically for Kokborok's sounds, not borrowed from another language.

This is meant to serve the Tiprasa community in two directions at once. For today's speakers, teachers, and writers, it's a growing reference — something to check a word against, cite in teaching material, or build lessons around. For the next generation, it's meant to be the foundation younger speakers and learners find already built: a dictionary that already has structure, a script that already works on modern devices, and a body of real sentences to learn from instead of starting from nothing.

The project depends on the community it serves. Speakers can contribute words, corrections, regional variants, and example sentences. Teachers and writers can help verify entries and flag what's missing. Researchers can use the structured data for study. None of this replaces the community's own knowledge — it exists to organise, preserve, and make that knowledge usable for as long as the language is spoken.

Where things stand

Real progress, honestly reported.

11,700+ words collectedUnverified, working through review
UCSUR registeredYapiri script documented in the Under-ConScript Unicode Registry
Omniglot documentedRecognised as a writing-system resource
Font, keyboard, and web toolsLive and in active development
Development roadmap

Build in layers, not disconnected features.

1

Language foundation

Dictionary entry model, grammar architecture, sentence corpus, numerals, spelling rules, and editorial standards.

2

Learning system

Stable static lessons, guided practice, progress tracking, word-based reading, and teacher-ready materials.

3

Publishing and translation

Parallel readers, educational book translation workflows, style guides, terminology management, and reviewed content releases.

4

Language technology

Search, transliteration, spell-checking, keyboard support, OCR experiments, speech datasets, and machine translation research.

Community contribution

Many people can contribute without editing the website.

Speakers, teachers, writers, and researchers can all contribute — every submission preserves who contributed it, where it came from, which variety it represents, and who reviewed it before anything is published. A structured contribution form is in development; until then, reach out directly to get involved.

W

Words

Meanings, parts of speech, example sentences, related words, regional variants, and sources.

S

Sentences

Natural speech, translations, contexts, audio, speaker metadata, and consent.

R

Review

Editorial checking for spelling, grammar, meaning, dialect, cultural accuracy, and publication readiness.

yapiriproject.com can become the public front door.

Behind it, a structured editorial database should manage dictionary entries, grammar records, corpus sentences, media, contributors, and release versions.