The Institut bouddhique's rolling library, a truck racked with manuscript bundles, Phnom Penh, 1930

AI research and products for Khmer

Khmer speakers should be able to use AI in their own language. We’re doing the research to make that possible, and sharing every model and dataset we make, free.

“Rolling Library” of the Institut bouddhique de Phnom Penh, 1930 · public domain

Our models

Read the full write-ups →

Why Khmer

More than 16 million people speak Khmer. Most AI still treats it as an afterthought.

There’s little public data to train on, and tools built for English stumble on a script where consonants stack and words run together without spaces.

We build for Khmer from the start, and publish everything, so the next person working on Khmer doesn’t have to start from zero.

Our commitments

Built for the people who read and write Khmer.

Read all our commitments →
  • You won't need an account, an API key, or a credit card to use anything we make. Code, weights, and reports are released under licenses that allow commercial use, changes, and redistribution.

    This isn't a trial period. We won't lock the work away later, even if a business reason to do so comes up. Khmer-language tools shouldn't need anyone's permission to build on.

    Read more →