03 / The project

A multilingual archive built through relations.

SHABD India is research infrastructure: a structured corpus for preserving, translating, contextualising, narrating and circulating underrepresented Indian-language literature.

Our first corpus begins with one selected short story from each of India’s twenty-two scheduled languages. Each story is progressively translated into the other twenty-one languages—creating 462 translations and 484 textual versions at completion.

01

Archive first

Stories are scholarly objects with versions, people, provenance, rights and relationships.

02

Selection with context

A source story is not “content”; its significance, edition and rights pathway must be understood.

03

Translation as record

Credits, review status, notes and changes belong to each translation version.

04

Build in public

Gaps and uncertainties remain visible rather than being disguised as completeness.

05

Access across scripts

Typography, navigation, search and reading interfaces must respect multilingual use.

06

Stewardship over extraction

Communities and rights holders shape how materials may be presented and reused.

Roadmap

From prototype to durable infrastructure.

  1. NowCorpus and record model

    Establish the language network, the Heengwala scholarly-object prototype and rights-aware metadata.

  2. NextSelection and editorial workflow

    Document source-story criteria, authority records, translation review and contribution agreements.

  3. ThenPublishing infrastructure

    Migrate verified texts, add audio, persistent identifiers, exports and multilingual search.