The plan divided into 160 numbered units across four stages; the sources the corpus is assembled from; and an honest account of what is loaded, what is queued, and what has still to be acquired.
The work divides into four stages. They are thematic rather than strictly sequential — collection continues while the lexicon is built, and the lexicon deepens while the applications are written.
Veda, Vedāṅga, Upāṅga and Upaveda; the Purāṇas, Itihāsa, Kāvya and the ancillary literature, together with their commentaries. Where a text is not already proof-read, it enters through the digitisation route.
Loading is not mere ingestion. The Bhāṣya is processed so that its Vyākaraṇa and śāstra references are linked to their targets; the related portions are brought together, so a Mantra or Brāhmaṇa passage stands beside its Pada-pāṭha and its Bhāṣyam; and the MAP analysis is produced for each mantra on demand rather than stored. Throughout, a master list is maintained that can absorb further texts and references without being restructured.
Three vocabularies, built by different means because they are different in kind.
Rūḍha — the conventional vocabulary, drawn from the kośas. Being non-compositional by definition, it must be enumerated rather than derived.
Vaidika — every available Pada-pāṭha ingested, Pada-pāṭha generated where none exists, and the unique words with their meanings extracted from the Veda Bhāṣyam. A Pada-pāṭha processor brings uniformity at load time and keeps the processed padas separately, so that the Krama and Vikṛti sandhi layers can be built upon them.
Yaugika — the Vyākaraṇa Vocabulary Engine: the Pāṇinian word-space, over a hundred million forms, each carrying its complete sūtra-by-sūtra derivation, with meanings in Telugu, Kannada, Hindi and English. Telugu leads, because the team's native expertise gives the subtlest sense.
Alongside these: Nirukta etymology, with traditional nirvacana set beside the formal Pāṇinian vyutpatti; the liṅga rules from Nāma-liṅgānuśāsana, Mahābhāṣya and Kāśikā; Krama and the eight Vikṛti forms from Jaṭā to Ghana, with rule-verified sandhi and svara at every junctura; and Varṇa Krama — akṣara-by-akṣara decomposition with svara and mātrā, built directly on Śikṣā and Prātiśākhya.
Data synthesised across the verticals to establish Vākya analysis, and the cross-references that run through Veda, Vedāṅga, Upāṅga and Upaveda connected so the corpus reads as one body rather than many.
Anuvāda — translation of the mantras into Indian languages founded solely on the Bhāṣya; where no Bhāṣya exists, on the principles of Sāyaṇācārya. Samānatā — an engine determining similarity in śabda and in artha.
On these rest the applications: Śrauta Prayoga generation with every step carrying a traceable sūtra citation; the Mīmāṃsā Nyāya engine; Sāmaveda stotra generation with viṣṭuti patterns and stobha insertion; Tarka sentence structure in the śābda-bodha style with pariṣkāra; Sāma sound analysis, audio in and notation out, with the Ūha engine that re-frames melody onto new chandas; and the Chatur-Darśana engine, which interprets one sentence in parallel through Tarka, Vyākaraṇa, Mīmāṃsā and Vedānta — the śāstrārtha assembly, computed.
The work reaches its audience: integration with general search; the Vedic perspective offered on any topic on demand, with societal reference; APIs for researchers seeking Vedic references with plain meanings; video, audio and web publication; webinars; the VVS textbooks; short-form material across thousands of topics; a daily presence; and a considered śāstra perspective on contemporary events.
The plan is deliberately granular. Each unit carries a permanent number, a definition of done, and its dependencies — so progress is a matter of record rather than impression.
Unit numbers (SVU-001 onward) never change and never encode a stage, so work can be resequenced without renumbering. The same discipline governs every identifier in the project: an identifier records what a thing is, never where it currently sits.
The work is owned by group, not unit by unit, so new work inherits a group instead of waiting to be allocated. Each group has a convener who is answerable for its units. The Tech Team is not a further group — it carries the units whose deliverable is software.
| Group | Working group | Convener | Units | Complete |
|---|---|---|---|---|
| G1 | Book Collection & Source Management | VKG | 9 | 3 |
| G2 | Data Processing, Analysis & Ingestion | Chandra Chaganti | 19 | 5 |
| G3 | Pada Kośa Review | VKG | 33 | 3 |
| G4 | Specialist Tasks — Audio Research | Natraj Kavuri | 11 | 0 |
| G5 | Mīmāṃsā / Vedavākya Categorization | Vijay Krishna | 20 | 0 |
| G6 | Testing | Gayathri Ramasubramanian | 3 | 0 |
| G7 | Review — Book Review & User Interaction | Vishwanath | 3 | 0 |
| G8 | Presentation / Multimedia (MMP) | Chandra Chaganti | 7 | 0 |
| Tech | Tech Team | Tech Team Lead | 55 | 4 |
The corpus is assembled from the Foundation's own digitised holdings, the open Sanskrit lexical corpora, and a Pāṇinian derivation engine that generates grammatical forms rather than storing them.
| Source | What it contributes | Items | Size |
|---|---|---|---|
| Veda mūla | Saṃhitā text of the four Vedas, with svara | 30 | 86 MB |
| Pada-pāṭha | Word-separated recitational text | 7 | 7 MB |
| Bhāṣya | Traditional commentary — Sāyaṇa, Bhaṭṭa Bhāskara and others | 94 | 2,324 MB |
| Vedāṅga | Śikṣā, Prātiśākhya, Kalpa and the ancillary śāstras | 63 | 233 MB |
| Prayoga | Ritual manuals and performance texts | 33 | 321 MB |
| Lexical — Cologne Digital Sanskrit Lexicon | 44 dictionaries: Monier-Williams, Apte, Amarakośa, Vācaspatyam, Śabdakalpadruma and others | 44 | — |
| Lexical — indic-dict collections | 19 further kośa builds used as independent cross-check witnesses | 19 | — |
| Grammatical engine | Pāṇinian derivation engine with the Dhātupāṭha — 2,229 dhātus, 5,160 sūtras, 128 kṛt and 181 taddhita pratyayas | 1 | — |
| Pāṇinian reference corpora | Aṣṭādhyāyī commentary and annotation sets surveyed for reuse | 4 | — |
Every division of the corpus, what is held in it, and what the seventy-two volume plan still requires. This table is generated from the corpus register itself and is rebuilt whenever the holdings change, so it reports the present position rather than an intention.
The Saṃhitā text itself, accented. Of what is held, 26 can enter the pipeline as it stands and 4 needs recognition or conversion first.
| Work | Position |
|---|---|
| 1 Poorvarchikam May 16 | Ready |
| 101DV2021 | Ready |
| 106 Aitareya Brahmana Aranyakam 2021 | Ready |
| 2 Uttararchikam May16 | Ready |
| 201_TS_2024_DV | Ready |
| 2025__Atharva_Shounaka_Samhitaa | Ready |
| 202_TB_2024_DV | Ready |
| 3 Aagneyam final | Ready |
| 301 Aarchika 2016 | Ready |
| 302 Prakruti Gaanam 1A | Ready |
| 303 Prakruti Gaanam 2A | Ready |
| 304 Chhaandogya Upanishat | Ready |
| 305 Samaveda PadaPaathah | Ready |
| 306 Ooha Ganam 1A | Ready |
| 307 Ooha Ganam 2A | Ready |
| 308 Tandya 8 Brahmanas | Ready |
| 4 Aiindram final | Ready |
| 5 Paavamanam final | Ready |
| 6 Aaranyakam final | Ready |
| Bruhadaranyaka Upanishat Moolam | Ready |
| Gopatha Brahmanam DV 2021 | Ready |
| Kaanva Samhitaa | In preparation |
| Kaanva Samhitaa PuBlisher File | In preparation |
| Maadhyandina Samhitaa | Ready |
| Maitrayaneeya Samhita Satvalekar | In preparation |
| Maitrayaneeya Samhitaa | Ready |
| Samaveda Samhitaa - Aarchikam | Ready |
| Taittireeya BrahmanaAranyakam AnukramanikaProof Reading | Ready |
| Taittireeya Samhita MantraAnukramanika Proof Reading | Ready |
The word-by-word recitational text. Of what is held, 7 can enter the pipeline as it stands and 0 needs recognition or conversion first.
| Work | Position |
|---|---|
| 207 Yajurveda PadaPaatha TSPDV2022 | Ready |
| 305 Samaveda PadaPaathah | Ready |
| 408 Atharva Padam 2025 | Ready |
| 409 Gopatha Brahmanam Padapatha 2024 | Ready |
| Rigveda_Padam 1-4 | Ready |
| Rigveda_Padam 5-8 | Ready |
| TBP DV 2012 | Ready |
Commentary — Sāyaṇa, Bhaṭṭa Bhāskara and others. Of what is held, 47 can enter the pipeline as it stands and 46 needs recognition or conversion first.
| Work | Position |
|---|---|
| 103510524-Atharv-Ved-Part-1-Bhashya-by-Shri-Ram-Sharma-Acharya | In preparation |
| 103568228-Atharv-Ved-Samhita-Part-2-Bhashya-by-Shri-Ram-Sharma-Acharya | In preparation |
| 179354273-Shatpath-Brahman-Hindi-Vigyan-Bhashya-Dwitiya-Khanda-Motilal-Shastri-Part3 | In preparation |
| AB Index | In preparation |
| Aitareya Aaranyakam | Ready |
| Aitareya Brahmanam Part 1- Omkar | Ready |
| Aitareya Brahmanam Part 2- Priyadarshini | Ready |
| Aitareya_Brahmanam_with_Sayanabhashya_Part_1_-_Kasinathsastri_Agase_1896ASS_032_ | In preparation |
| Aitareya_Brahmanam_with_Sayanabhashya_Part_2_-_Kasinathsastri_Agase_1896ASS_032_ | In preparation |
| Aitareyaranyakam_with_Sayanabhashya_-_Babasastri_Phadke_1898ASS_038_ | In preparation |
| Aranya Samhita Samaveda Sayana Bhashya Jivanand Vidya Sagar 1891 - Copy | In preparation |
| Arsheya Brahmana Bhashyam | Ready |
| Atharva Sounaka 1-5 | In preparation |
| Atharva Sounaka 10-18 | In preparation |
| Atharva Sounaka 19-20 | In preparation |
| Atharvaveda_Bhashyam (1-5) | Ready |
| Atharvaveda_Bhashyam (11-18) | Ready |
| Atharvaveda_Bhashyam (19-20) | Ready |
| Atharvaveda_Bhashyam (6-8) | Ready |
| Atharvaveda_Bhashyam (9 -10) | Ready |
| BBB Ashtakam 1 | Ready |
| BBB Ashtakam 2 | Ready |
| BBB Ashtakam 3 | Ready |
| BBB Kanda 1 | Ready |
| BBB Kanda 2 | Ready |
| BBB Kanda 3 (A5- 258) | Ready |
| BBB Kanda 5 (A5- 263) | Ready |
| BBB Kanda 6 (A5-262) | Ready |
| BBB Kanda 7 (A5-187) | Ready |
| BBB T Aaranyakam | Ready |
| Bruhadarayaka Bhashyam | Ready |
| Devatadhyaya - Samhitopanisad - Vamsa Brahmanam | Ready |
| Ganesa_Atharvasirsham_Sabhashyam_-_Vamansastri_Islampurkar_1889ASS_001_ | In preparation |
| Gopatha Meanings Old Edition | Ready |
| Praataranuvaaka Agni-Ushas-Ashwin | In preparation |
| Rigveda Sahmita Bhashyam Ashtakam 1 | Ready |
| Rigveda Sahmita Bhashyam Ashtakam 2 | Ready |
| Rigveda Sahmita Bhashyam Ashtakam 3 | Ready |
| Rigveda Sahmita Bhashyam Ashtakam 4 | Ready |
| Rigveda Sahmita Bhashyam Ashtakam 5 | Ready |
| Rigveda Sahmita Bhashyam Ashtakam 6 | Ready |
| Rigveda Sahmita Bhashyam Ashtakam 7 | Ready |
| Rigveda Sahmita Bhashyam Ashtakam 8 | Ready |
| Rudra Adhyaaya Bhashyam | Retired |
| Rudra Adhyaaya Bhashyam 2016 | Ready |
| Rudradhyaya_with_Commentaries_of_Sayana__Bhattabhaskara_1935_ASS_002_ | In preparation |
| Saayana Bhashyam TB 2.6-3.7 | Ready |
| Saayana Bhashyam TS 1 Kaanda | Ready |
| Saayana Bhashyam TS 3 Kaanda | Ready |
| Saayana Bhashyam TS 4 Kaanda | Ready |
| Samaveda Arsheyadeepa | Ready |
| Samaveda Sahmita Bhashyam | Ready |
| Samavidhana brahmanam | Ready |
| Sandhya Vandanam With Meanings with Details | Ready |
| Shadvimsha brahmanam | Ready |
| TA3 Ekagni Kaandam | In preparation |
| TAPart_1_-_Babasastri_Phadke_1898ASS_036_ | In preparation |
| TAPart_2_-_Babasastri_Phadke_1927ASS_036_ | In preparation |
| TB1.1_Part_1_-_Narayanasastri_Godbole_1934ASS_037_ | Ready |
| TB2.6_Part_2_-_Narayanasastri_Godbole_1898ASS_037_ | In preparation |
| TB3.8_Part_3_-_Narayanasastri_Godbole_1898ASS_037_ | In preparation |
| TS1.1_Part_1_-_Kasinath_Sastri_Agase_1940ASS_042_ | Ready |
| TS1.31_Part_2_-_Kasinath_Sastri_Agase_1940ASS_042_ | In preparation |
| TS1.71_Part_3_-_Kasinath_Sastri_Agase_1947ASS_042_ | In preparation |
| TS2.1_Part_4_-_Kasinath_Sastri_Agase_1946ASS_042_ | In preparation |
| TS2.51_Part_5_-_Kasinath_Sastri_Agase_1946ASS_042_ | In preparation |
| TS3.5_Part_6_-_Kasinath_Sastri_Agase_1949ASS_042_ | In preparation |
| TS5.1_Part_7_-_Kasinath_Sastri_Agase_1949ASS_042_ | In preparation |
| TS6.1_Part_8_-_Kasinath_Sastri_Agase_1951ASS_042_ | Ready |
| Taittiriyopanishat Satikaa Shaankarabhashya | In preparation |
| Tandya Brahmana Bhashyam | Ready |
| ekagni_kanda_haradatta_taittiriya | In preparation |
| shadvimsha_brahmana with bhashyam | In preparation |
| shukla_yajurveda_two_commentaries | In preparation |
| ssk-samaveda-with-commentary-of-madhva | In preparation |
| t_aranyaka_bhaskara_01 | In preparation |
| t_aranyaka_bhaskara_02 | In preparation |
| t_brahmana_bhaskara_01 | In preparation |
| t_brahmana_bhaskara_02 | In preparation |
| t_brahmana_bhaskara_03.1 | In preparation |
| t_brahmana_bhaskara_03.2 | In preparation |
| t_samhita_bhaskara_01(1.1-1.3) | In preparation |
| t_samhita_bhaskara_02(1.4-1.6) | In preparation |
| t_samhita_bhaskara_03(1.7-2.2) | In preparation |
| t_samhita_bhaskara_04(2.3-2.6) | In preparation |
| t_samhita_bhaskara_05(3.1-3.5) | In preparation |
| t_samhita_bhaskara_06(5.1-5.4) | In preparation |
| t_samhita_bhaskara_07(5.5-5.7 | In preparation |
| t_samhita_bhaskara_08(6.1-6.4) | In preparation |
| t_samhita_bhaskara_09(6.5-7.3) | In preparation |
| t_samhita_bhaskara_10(7.4-7.5) | In preparation |
The six auxiliary disciplines. Of what is held, 43 can enter the pipeline as it stands and 16 needs recognition or conversion first.
| Work | Position |
|---|---|
| 01 Atharva Veda Chaturadhyayika (224 P) | Ready |
| 02 Atharva Veda Parishitham (346 P) | Ready |
| 03 Atharva Vediya Panchapatalika (26 P) | Ready |
| 04 Atharva_Praatishakhya (DV) (14 P) | Ready |
| 05 Atharva Veda Mandukeeya Shiksha (15 P) | Ready |
| 06 Atharva Veda Bhashyam (1-5 Khanda`s) - (344 P) | Ready |
| 07 Atharvaveda_Bhashyam (6-10) Kanda`s- (359 P) | Ready |
| 08 Athrvaveda_Bhashyam (11-18) Kand`s (364 P) | Ready |
| 09 Atharvaveda_Bhashyam (19-20) Kanda`s (361 P) | Ready |
| 10 Shounaka Samhitaa Padam 408DV-A4 -772P | Ready |
| 11 Koushika Paddhati (Keshava Kruta) - 273P | Ready |
| 11 Koushika Paddhati (Keshava Kruta) - 273P Atharva Karmaani | Ready |
| 401DV-A4. (working) | Ready |
| Aashwalaayana Gruhya Sutram with Commentary | Ready |
| Aranyaka Siksha 2016 | Ready |
| Aranyaka Siksha DV | Ready |
| Atharva Books List with Page Numbers | Ready |
| Atharva Veda Pratishakhyam | Ready |
| Atharva_Praatishakya(TL) | Ready |
| Atharva_Rishi_Chandas_Devata(1-20 Kand`s) | Ready |
| Atharvaveda Chandas | Ready |
| Bharadwaja Shiksha | Ready |
| Jata Darpanam | Ready |
| Kaala Nirnaya Pattikaa | Ready |
| Lakshana Grantha Rigveda | Ready |
| Lakshana Moolam Atharva | Ready |
| Lakshana Moolam Yajurveda | Ready |
| Manduki Shiksha | In preparation |
| Naradiya Shiksha Commentary 2 | In preparation |
| Naradiya Shiksha with Bhatta Shobhakar's Shiksha Vivarana Commentary - Narad | In preparation |
| Others_Yohi -Prapti with commentary | In preparation |
| Others_yohi_prapti_shiksha | In preparation |
| Panchavidha Sutram | Ready |
| Pushpa Sutram of Samaveda_5262__Alm_24_Shlf_1_Devanagari - Sutra Paddhati | In preparation |
| Pushpasutram | Ready |
| Pushpasutram Moolam | Ready |
| Rigveda Brahmakarma samuchaya | In preparation |
| Rigveda Praatishaakhyam with Uvata Bhaashyam | Ready |
| Rigveda Pratishaakhyam - Prayaga Print | In preparation |
| Rik Pratishakhya Moolam 2023 | Ready |
| Rik Tantram | Ready |
| Saama Tantram | Ready |
| Sama Lakshana Stabakam | Ready |
| Sama Model | Ready |
| Sapta Lakshanam - Priyadarshini | Ready |
| Sapta Lakshanam 1st Edition | In preparation |
| Shiksha Yajur Moolam 2 | Ready |
| Shukla Yajurveda Pratishakhya | Ready |
| Taittireeya PraatishaakhyamDV | In preparation |
| Taittireeya Pratishaakhyam 2 Vyaakhya | In preparation |
| Taittireeya Pratishakhayam Savyakhyam | Ready |
| Taittireeya Pratishakhyam Brief En | In preparation |
| Taittiriya Praatishaakhyam 2022 With 3 Commentaries | Ready |
| Taittiriya-Pratisakhya Mahisheya | In preparation |
| Taittiriya-Pratisakhya Whitney | In preparation |
| Vyaasa Shiksha VedaTaijasa Sarvalakshana Manjari | In preparation |
| Vyasa Shiksha Vyakhya 2020 | Ready |
| Yohi Shiksha - Priyadarshini | Ready |
| sama_veda_pratishakhya | In preparation |
Ritual manuals and their sequence. Of what is held, 29 can enter the pipeline as it stands and 0 needs recognition or conversion first.
| Work | Position |
|---|---|
| 20260108_000932_bhagpur-01 | Ready |
| 20260120_140835_Agnihotra Prayoga | Ready |
| 20260122_222646_Yajusha Shraaddha Prayogah | Ready |
| 20260131_084705_Agnishtoma_1 | Ready |
| 20260131_084725_Agnishtoma_2 | Ready |
| 20260131_084732_Agnishtoma_3 | Ready |
| 20260131_085930_Agnishtoma_5 | Ready |
| 20260131_090831_Agnishtoma_41 | Ready |
| 20260131_090839_Agnishtoma_42 | Ready |
| 20260131_091027_chaturmasya1 | Ready |
| 20260131_091053_chaturmasya2 | Ready |
| 20260201_155430_Ramayana Muktaavali | Ready |
| 20260202_113118_saraswati_vidya_prarthanam | Ready |
| 20260210_152838_Sachchidananda Neeti Maala 2018 Final | Ready |
| 20260217_230159_Apara Prayoga (Bharatula) | Ready |
| 20260221_054507_RUDRA PRAPANC FINAL BOOK | Ready |
| 20260309_234910_plan1 | Ready |
| 20260312_024159_sample_nirnaya_sagar | Ready |
| 20260421_233658_154075388-Asvalayana-Srautasutra-1917-pdf_compressed | Ready |
| 20260525_012330_Oudgaatra_Agnishtotma | Ready |
| 20260526_023432_MA_Sanskrit | Ready |
| 20260617_233440_missing pages Pravargya | Ready |
| 20260617_234656_1 | Ready |
| 20260625_085924_Agnyadheeya prayoga | Ready |
| 20260625_090425_Anvarambhaneeya | Ready |
| 20260625_091149_Niroodha pashubandha prayoga | Ready |
| 20260625_092223_Chaturmaasya prayoga 2 | Ready |
| 20260625_092814_Chaturmaasya prayoga | Ready |
| 20260625_093752_Niroodha pashubandha prayoga | Ready |
Mīmāṃsā, Nyāya, Vedānta, Vyākaraṇa. Of what is held, 4 can enter the pipeline as it stands and 0 needs recognition or conversion first.
| Work | Position |
|---|---|
| 20260109_003401_Mimamsa Nyaya Prakasha NSP 1_text | Ready |
| 20260120_130641_Pratibandhakata Vada Gadhadhara Narayana Shastri Patwardhan | Ready |
| 20260309_234940_Ananda Giri Teeka | Ready |
| 20260309_235044_Vedanta Sutra Muktavali | Ready |
The kośa collections held for corroboration. Of what is held, 63 can enter the pipeline as it stands and 0 needs recognition or conversion first.
| Work | Position |
|---|---|
| Abhidhanacintamani | Ready |
| Abhidhanacintamani (Hemacandra) | Ready |
| Abhidhanacintamani - Parisista | Ready |
| Abhidhanacintamani - Siloncha | Ready |
| Abhidhanaratnamala | Ready |
| Abhidhanaratnamala (Halayudha) | Ready |
| Amarakosa (ontology build) | Ready |
| Amarakosa with Sudha commentary | Ready |
| Anekarthadhvanimanjari | Ready |
| Apte, English-Sanskrit Dictionary | Ready |
| Apte, The Practical Sanskrit-English Dictionary | Ready |
| Benfey, Sanskrit-English Dictionary | Ready |
| Boehtlingk & Roth, Sanskrit-Woerterbuch (7 Baende) | Ready |
| Boehtlingk, Sanskrit-Woerterbuch in kuerzerer Fassung | Ready |
| Bopp, Glossarium Sanscritum | Ready |
| Borooah, English-Sanskrit Dictionary | Ready |
| Burnouf, Dictionnaire classique Sanscrit-Francais | Ready |
| Cappeller, Sanskrit-English Dictionary | Ready |
| Cappeller, Sanskrit-Woerterbuch | Ready |
| Ekaksaranamamala | Ready |
| Goldstuecker, Sanskrit-English Dictionary | Ready |
| Grassmann, Woerterbuch zum Rig-Veda | Ready |
| L. R. Vaidya, Sanskrit-English Dictionary | Ready |
| Lanman, Sanskrit Reader vocabulary | Ready |
| Macdonell, A Practical Sanskrit Dictionary | Ready |
| Monier-Williams (1872 edition) | Ready |
| Monier-Williams, A Sanskrit-English Dictionary | Ready |
| Monier-Williams, English-Sanskrit Dictionary | Ready |
| Sabda-Sagara, Sanskrit-English Dictionary | Ready |
| Sabdakalpadruma (Radhakantadeva) | Ready |
| Sabdakalpadruma (StarDict build) | Ready |
| Soerensen, Index to the Names in the Mahabharata | Ready |
| Vacaspatyam (StarDict build) | Ready |
| Vacaspatyam (Taranatha Tarkavacaspati) | Ready |
| Vedic Index of Names and Subjects (Macdonell & Keith) | Ready |
| Wilson, Sanskrit-English Dictionary | Ready |
| Yates, Sanskrit-English Dictionary | Ready |
| pwkvn | Ready |
Inputs to the Pāṇinian generator. Of what is held, 5 can enter the pipeline as it stands and 0 needs recognition or conversion first.
The surrounding literature. Of what is held, 0 can enter the pipeline as it stands and 0 needs recognition or conversion first.
Everything not yet classified. Of what is held, 0 can enter the pipeline as it stands and 4 needs recognition or conversion first.
Every held artefact is classified by what stands between it and the pipeline. Nothing has been loaded to the production corpus yet — the schema is still under scholarly review — so the whole holding is queued.
| State | Meaning | Files | Size |
|---|---|---|---|
| Ready to load | Text-bearing files that need no conversion. | 156 | 569 MB |
| Awaiting assembly | Recognition already complete; output not yet assembled. | 30 | 1,300 MB |
| Awaiting triage | To be checked for a text layer before any recognition. | 34 | 1,085 MB |
| Awaiting conversion | Legacy word-processor formats. | 2 | 7 MB |
| Superseded | A better copy of the same work is held. | 5 | 10 MB |
The seventy-two volume plan names 390 distinct works. Reconciling that requirement against the holdings shows precisely what has still to be sourced.
| Position | Distinct works |
|---|---|
| Held | 19 |
| Held, pending verification | 68 |
| Not yet acquired | 303 |
| Category | Works to acquire |
|---|---|
| Darśana / Philosophy | 34 |
| Saṃhitā | 30 |
| Saṅgīta / Nāṭya | 23 |
| Purāṇa | 21 |
| Dharmasūtra / Smṛti | 19 |
| Kāvya / Alaṃkāra | 18 |
| Brāhmaṇa | 15 |
| Stotra / Nāmāvali | 14 |
| Upaniṣad | 14 |
| Gṛhyasūtra | 13 |
The holdings are strongest in Veda mūla, bhāṣya and vedāṅga — the Foundation's own scholarly territory — and thinnest in the darśana, purāṇa, kāvya and applied-śāstra divisions. Acquisition is therefore sequenced by what the volumes actually require, not by what is easiest to obtain.
Four stages, 160 units. Stage numbering is thematic; several stages run concurrently. Items marked Awaiting decision are held pending a scholarly determination and are deliberately not started, because the determination may change the work.
Infrastructure, schema, governance and the registers that everything else is tracked against.
| Unit | Task | Track | Status |
|---|---|---|---|
| SVU-001 | Approve the corpus database design and create it | Governance | Awaiting decision |
| SVU-002 | Add the dictionary tables to the corpus database | Governance | Awaiting decision |
| SVU-003 | Commission the two computing servers | Infra | Complete |
| SVU-004 | Install the databases and search services on the servers | Infra | Complete |
| SVU-005 | Constitute the scholarly review panel | Governance | Planned |
| SVU-006 | Declare the licence for each of our three own works | Governance | Awaiting decision |
| SVU-007 | Repository governance and access control | Governance | Planned |
| SVU-008 | Number every document and keep a register of them | Governance | Complete |
| SVU-009 | Register every text we hold, with its loading status | Corpus | Complete |
| SVU-010 | Match the books the 72 volumes need against what we hold | Corpus | Complete |
| SVU-011 | Generate project documents reproducibly, and reject invalid files | Tooling | Complete |
| SVU-012 | Corpus resilience and off-site replication | Risk | Planned |
| SVU-013 | Define what each group hands to the next, and when | Governance | Ready to start |
| SVU-014 | Bring in outside funding through non-profits and matching grants | Governance | Planned |
Encoding the accent correctly, then loading the text — mūla, pada-pāṭha, bhāṣya and the ancillary śāstras — with provenance intact.
| Unit | Task | Track | Status |
|---|---|---|---|
| SVU-020 | Count every special VijayaDV character in the corpus | Encoding | Complete |
| SVU-021 | Sort the VijayaDV special characters into accent, marker and sandhi | Encoding | Complete |
| SVU-022 | Decide the standard character each remaining accent mark maps to | Encoding | Awaiting decision |
| SVU-023 | Confirm that position-variant accent glyphs mean one accent | Encoding | Awaiting decision |
| SVU-024 | Convert the mūla of all four Vedas to standard characters | Encoding | Awaiting decision |
| SVU-025 | Prove every converted file converts back unchanged | Encoding | Awaiting decision |
| SVU-026 | Load the 120 files that are already machine-readable | Loading | Awaiting decision |
| SVU-027 | Assemble the volumes already put through OCR | Loading | Complete |
| SVU-028 | Check each un-OCR'd PDF for existing text before paying for OCR | Loading | Complete |
| SVU-029 | Run OCR on the books that are genuine scans | Loading | Awaiting decision |
| SVU-030 | Extract the text from PDFs that already carry it | Loading | Ready to start |
| SVU-031 | Convert the two old-format Kāṇva Saṃhitā files | Loading | Complete |
| SVU-032 | Archive the 5 PDFs whose DOCX we already hold | Loading | In progress |
| SVU-033 | Load the Kalpa Sūtra collection, split at sūtra level | Corpus | Awaiting decision |
| SVU-034 | List every work quoted in the 72 volumes | Corpus | Ready to start |
| SVU-035 | Put the 303 missing works in the order we should obtain them | Corpus | Complete |
| SVU-036 | Bring the outside digital corpora into our own format | Corpus | Planned |
| SVU-037 | Mark where each ṛk, sūtra and śloka begins and ends | Corpus | Planned |
| SVU-038 | Freeze the numbering that addresses every sentence in the corpus | Architecture | Awaiting decision |
| SVU-039 | Index every passage for meaning-based search | Retrieval | Planned |
| SVU-040 | Link ṛk to devatā, sūkta to ṛṣi, mantra to rite | Retrieval | Planned |
| SVU-041 | Load the four 2020 Bhāṣya Pilot volumes as the gold standard | MAP | Awaiting decision |
| SVU-042 | Resolve every grammatical and śāstric citation in the Bhāṣya | MAP | Planned |
| SVU-043 | Group each mantra and brāhmaṇa passage with its Padapāṭha and Bhāṣyam | MAP | Planned |
| SVU-044 | Compute meaning, analysis and presentation for a mantra on demand | MAP | Planned |
| SVU-044.1 | Spec — MAP analysis contract — what the reader is shown and from what | MAP | Ready to start |
| SVU-044.2 | Build — Compute meaning, analysis and presentation for a mantra on demand | MAP | Planned |
| SVU-046 | Publish finished material to vaakya.vedanidhi.in as it is ready | Delivery | Planned |
| SVU-048 | Reconcile filing of recently added documents | Hygiene | In progress |
| SVU-049 | Recover the text from PDFs written in legacy fonts | Loading | Ready to start |
The lexical foundation: the conventional vocabulary drawn from the kośas, the Vedic vocabulary drawn from pada-pāṭha and bhāṣya, and the compositional vocabulary generated from Pāṇini's rules.
| Unit | Task | Track | Status |
|---|---|---|---|
| SVU-050 | Load the 62 dictionaries and make their words searchable | Pada Kośa A | Awaiting decision |
| SVU-051 | Publish the census of what the dictionaries contain | Pada Kośa A | Complete |
| SVU-052 | Access and serving policy for third-party lexical sources | Pada Kośa A | Awaiting decision |
| SVU-053 | Fix which dictionary the reader is shown first | Pada Kośa A | Awaiting decision |
| SVU-054 | Flag the 106,014 words attested in only one dictionary | Pada Kośa A | Awaiting decision |
| SVU-055 | Load every Padapāṭha we hold, with its accents intact | Pada Kośa V | Awaiting decision |
| SVU-057 | Generate a Padapāṭha for texts that lack one | Pada Kośa V | Planned |
| SVU-057.1 | Spec — Rules for generating Pada-pāṭha where none is attested | Pada Kośa V | Ready to start |
| SVU-057.2 | Build — Generate a Padapāṭha for texts that lack one | Pada Kośa V | Planned |
| SVU-058 | Build the Vedic lexicon: words and meanings drawn from the Bhāṣya | Pada Kośa V | Planned |
| SVU-058.1 | Spec — Method for extracting the Vaidika lexicon from Bhāṣya | Pada Kośa V | Ready to start |
| SVU-058.2 | Build — Build the Vedic lexicon: words and meanings drawn from the Bhāṣya | Pada Kośa V | Planned |
| SVU-059 | Derive words from dhātu and pratyaya, with the sūtra chain | Pada Kośa B | Complete |
| SVU-060 | Fix how accents are printed in published forms | Pada Kośa B | Awaiting decision |
| SVU-061 | Decide whether pracaya and ekaśruti are applied | Pada Kośa B | Awaiting decision |
| SVU-062 | Measure our derived accents against the gold corpus, vowel by vowel | Pada Kośa B | Awaiting decision |
| SVU-062.1 | Spec — Accent-agreement test design and acceptance threshold | Pada Kośa B | Awaiting decision |
| SVU-062.2 | Build — Measure our derived accents against the gold corpus, vowel by vowel | Pada Kośa B | Planned |
| SVU-063 | Decide how many derived forms we generate | Pada Kośa B | Awaiting decision |
| SVU-064 | Generate the full set of forms at the agreed size | Pada Kośa B | Awaiting decision |
| SVU-065 | Build the fast word-lookup index | Pada Kośa B | Awaiting decision |
| SVU-066 | Write the ~3,400 morpheme meanings in Telugu, Kannada, Hindi and English | Pada Kośa B | Planned |
| SVU-067 | Give the nirvacana of a word, with competing etymologies | Nirukta | Awaiting decision |
| SVU-067.1 | Spec — Nirvacana model — sources, competing etymologies, presentation | Nirukta | Ready to start |
| SVU-067.2 | Build — Give the nirvacana of a word, with competing etymologies | Nirukta | Planned |
| SVU-068 | Establish the liṅga of each stem from the liṅga authorities | Vyākaraṇa | Planned |
| SVU-069 | Apply sandhi by rule, with the accent change at each junctura | Pāṭha | Planned |
| SVU-069.1 | Spec — Sandhi rule inventory and svara behaviour at each junctura | Pāṭha | Ready to start |
| SVU-069.2 | Build — Apply sandhi by rule, with the accent change at each junctura | Pāṭha | Planned |
| SVU-070 | Generate the Krama pāṭha of any Saṃhitā passage | Pāṭha | Planned |
| SVU-070.1 | Spec — Krama construction rules | Pāṭha | Ready to start |
| SVU-070.2 | Build — Generate the Krama pāṭha of any Saṃhitā passage | Pāṭha | Planned |
| SVU-071 | Generate the eight Vikṛti pāṭhas, with accents preserved | Pāṭha | Planned |
| SVU-071.1 | Spec — The eight Vikṛti forms — construction rules per form | Pāṭha | Ready to start |
| SVU-071.2 | Build — Generate the eight Vikṛti pāṭhas, with accents preserved | Pāṭha | Planned |
| SVU-072 | Varṇa Krama — already shipped, so settle reuse terms only | Pāṭha | Complete |
| SVU-073 | Look up a form and return its analyses for the hover | Pada Kośa | In progress |
| SVU-074 | Join a Veda word to its dictionary entry and its derivation | Pada Kośa V | Awaiting decision |
Cross-referencing, translation, the generative engines, and the reading application built on top of them.
| Unit | Task | Track | Status |
|---|---|---|---|
| SVU-080 | Connect cross-references across Veda, Vedāṅga, Upāṅga and Upaveda | Synthesis | Planned |
| SVU-081 | Translate mantras into Indian languages, grounded in the Bhāṣya | Synthesis | Planned |
| SVU-081.1 | Spec — Anuvāda method — grounding every rendering in Bhāṣya | Synthesis | Ready to start |
| SVU-081.2 | Build — Translate mantras into Indian languages, grounded in the Bhāṣya | Synthesis | Planned |
| SVU-082 | Find passages similar in word and in meaning | Synthesis | Planned |
| SVU-082.1 | Spec — Samānatā — what counts as similarity in śabda and in artha | Synthesis | Ready to start |
| SVU-082.2 | Build — Find passages similar in word and in meaning | Synthesis | Planned |
| SVU-083 | Analyse a sentence across all the verticals | Synthesis | Planned |
| SVU-084 | Assemble the instruction-and-answer set that trains the model | Model | Planned |
| SVU-085 | Fine-tune the base model on the Vedic corpus | Model | Planned |
| SVU-086 | Answer from retrieved passages rather than from memory | Model | Planned |
| SVU-087 | Improve answers from scholars' ratings of paired outputs | Model | Planned |
| SVU-088 | Re-measure training time on the servers we actually have | Model | Awaiting decision |
| SVU-089 | Generate a complete Śrauta rite sequence, every step cited | Engine | Planned |
| SVU-089.1 | Spec — Śrauta Prayoga sequence model and citation requirements | Engine | Ready to start |
| SVU-089.2 | Build — Generate a complete Śrauta rite sequence, every step cited | Engine | Planned |
| SVU-090 | Generate Sāmaveda stotras with viṣṭuti and stobha | Engine | Planned |
| SVU-090.1 | Spec — Stotriyā assembly, viṣṭuti patterns and stobha rules | Engine | Ready to start |
| SVU-090.2 | Build — Generate Sāmaveda stotras with viṣṭuti and stobha | Engine | Planned |
| SVU-091 | Teach the model the Vikṛti pāṭhas from rule-generated forms | Engine | Planned |
| SVU-092 | Analyse arguments in Tarka form: pañcāvayava and Navya-Nyāya | Engine | Planned |
| SVU-092.1 | Spec — Śābda-bodha and pariṣkāra representation for Tarka | Engine | Ready to start |
| SVU-092.2 | Build — Analyse arguments in Tarka form: pañcāvayava and Navya-Nyāya | Engine | Planned |
| SVU-094 | Analyse a recitation's sound: svara, modulation and volume | Audio | Planned |
| SVU-094.1 | Spec — Svara extraction and notation conventions for Sāmagāna | Audio | Ready to start |
| SVU-094.2 | Build — Analyse a recitation's sound: svara, modulation and volume | Audio | Planned |
| SVU-095 | Re-frame a melody onto a new chandas | Engine | Planned |
| SVU-095.1 | Spec — Ūha — how melody is re-framed onto new chandas | Engine | Ready to start |
| SVU-095.2 | Build — Re-frame a melody onto a new chandas | Engine | Planned |
| SVU-097 | Interpret a passage through four darśanas side by side | Engine | Planned |
| SVU-097.1 | Spec — The four lenses — scope and retrieval boundary of each | Engine | Ready to start |
| SVU-097.2 | Build — Interpret a passage through four darśanas side by side | Engine | Planned |
| SVU-098 | Show a passage in five layers, mūla to related passages | App | In progress |
| SVU-099 | Show every word's grammar on hover, by lookup not guesswork | App | In progress |
| SVU-100 | Rite generator for scholars, with its citation trail | App | Planned |
| SVU-101 | Screen that turns a Saṃhitā passage into all eight Vikṛti forms | App | Planned |
| SVU-102 | Sāma Studio: stotra generator, notation viewer, sound panel | App | Planned |
| SVU-103 | Tarka workbench, and the akṣara-by-akṣara Varṇa Krama view | App | Planned |
| SVU-104 | Stotra library by devatā and chandas, and an etymology explorer | App | Planned |
| SVU-105 | Notes connecting a passage to science, philosophy and the arts | App | Planned |
| SVU-106 | Serve the model on the second node, with caching | Infra | Planned |
| SVU-107 | Run the pilot with 25–50 faculty and scholars | Validation | Planned |
| SVU-108 | Score answers on all 2,000 Swadharma topics against scholar answers | Validation | Planned |
| SVU-109 | Check generated rites against the 2026 Agniṣṭoma performance | Validation | Planned |
| SVU-160 | Improve the speech-to-text model on Vedic recitation | Audio | In progress |
| SVU-161 | Record volunteers chanting each śākhā, with correct intonation | Audio | Ready to start |
| SVU-162 | Speak Vedic text aloud with the svara correct | Audio | Planned |
| SVU-163 | Settle the licence and access terms for the training audio | Audio | Ready to start |
| SVU-170 | Test every group's output for function and for content | Testing | Ready to start |
| SVU-171 | Sample scholarly quality: random at first, then systematic | Testing | Ready to start |
| SVU-175 | Final review before anything reaches the public | Review | Planned |
| SVU-140 | Approve the Mīmāṃsā sentence-classification vocabularies and record formats | MVVF | In progress |
| SVU-141 | Tag the Darśapūrṇamāsa pilot passage by hand | MVVF | Planned |
| SVU-142 | Tag sentences automatically from the visible marks in the text | MVVF | Planned |
| SVU-143 | Review console where a scholar confirms or overrides each automatic tag | MVVF | Planned |
| SVU-144 | Enter the first Adhyāya of the Nyāyamālā as Adhikaraṇa records | MVVF | Planned |
| SVU-145 | Enter all the Nyāyamālā Adhikaraṇas, with Bhāṭṭa and Prābhākara positions | MVVF | Planned |
| SVU-146 | Registry of rites and episodes, so one sentence can be traced through every rite | MVVF | Planned |
| SVU-147 | Three services: classify a sentence, explain its tag, walk the episode graph | MVVF | Planned |
| SVU-148 | Training set of tag decisions with the reasoning behind each | MVVF | Planned |
Publication, the researcher interface, and bringing the material to a general audience.
| Unit | Task | Track | Status |
|---|---|---|---|
| SVU-115 | Put this plan and its current status on sarvaveda.info | Site | Ready to start |
| SVU-116 | Publish pages as fixed HTML with sitemaps, so search engines read them | Site | Planned |
| SVU-117 | Let a reader's correction on the page become a proposed change | Site | Awaiting decision |
| SVU-118 | Open the site with General, Scholar and Institutional access | Release | Planned |
| SVU-119 | Machine access for researchers, documented and rate-limited | Release | Planned |
| SVU-121 | Answer any topic from the śāstra, with sourced citations | Reach | Planned |
| SVU-122 | Production route from corpus to published video, audio and web | Media | Planned |
| SVU-123 | Run regular scholarly webinars from the corpus | Media | Planned |
| SVU-124 | Generate the VVS textbook series from the validated corpus | Media | Planned |
| SVU-125 | Produce short videos at scale, each scholar-approved before release | Media | Planned |
| SVU-180 | Presentation that explains the initiative to a general audience | Media | Ready to start |
| SVU-126 | Publish daily across the channels, including on current events | Media | Planned |
| SVU-128 | Publish the consultation papers with their licence settled | Release | Awaiting decision |
| SVU-129 | Retrain each quarter on new volumes and feedback | Release | Planned |
| SVU-130 | Transcribe 1,000 hours of video and align it to the text | Corpus | Planned |
| SVU-149 | Public Mīmāṃsā pages: corpus browser, tag explorer, Adhikaraṇa reader | MVVF | Planned |
| SVU-150 | Reorganise vedavishtaram.in into seven top-level sections | Site | Ready to start |
| SVU-151 | Build the VVS catalogue straight from the Master Register | Site | Planned |