Skip to content
@Maelis-Research

Maelis Research

Language intelligence for Odia and low-resource Indian languages. Odia-exclusive tokenizer, LLMs, speech & translation. Apache 2.0.

Maelis Research

Language intelligence for Odia and low-resource Indian languages.

We are a language infrastructure company building the AI layer for the Odia language (~50M speakers). Based in Dhenkanal, Odisha, India.

What We Build

  • Tokenizer research — Efficient Unicode-aware tokenization for Brahmic scripts
  • Lekhani model family — Odia-optimized LLMs (Pada, Chhanda, Kavya, Mahakavya)
  • Shruti — Odia speech recognition and synthesis
  • Khoja — Odia semantic search
  • Anuvada — Odia translation
  • Patra — Odia OCR

Principles

  1. Deep specialization — One language, done well. Not 22 languages done adequately.
  2. Open source by default — Apache 2.0 license. The ecosystem grows when the foundation is open.
  3. Built for Odisha — Government, media, education, and enterprise-grade reliability.
  4. Efficiency first — Every API call should cost less. Our tokenizer is 3x more efficient for Odia.

Repositories

Repository Description
odia-tokenizer-benchmarks Benchmarking framework for Odia tokenizer evaluation
lekhani-pada Lekhani Pada — lightweight Odia language model
odia-pretrain-dataset CC-BY-4.0 corpus for Odia language model pretraining

Connect

Popular repositories Loading

  1. .github .github Public

Repositories

Showing 1 of 1 repositories

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Top languages

Loading…

Most used topics

Loading…