Published October 26, 2024 | Version v3

The morphologically glossed Rigveda - The Zurich annotation corpus revised and extended. Hosted by VedaWeb - Online Research Platform for Old Indic Texts.

  • 1. University of Freiburg, General Linguistics (Department of Linguistics)
  • 2. ROR icon University of Cologne
  • 3. University of Wuppertal, Digital Humanities (Department of History)
  • 4. University of Cologne, General Linguistics (Department of Linguistics)
  • 5. University of Würzburg, Chair of Comparative Philology (Department of Ancient Studies)

Description

This file contains morphological and lexicographic annotations for the Rigveda. It was created in the DFG-funded research project Vedaweb and used as source data for the linguistic research platform vedaweb.uni-koeln.de.

Prof. Dr. Paul Widmer and Dr. Salvatore Scarlata from the "Institut für Vergleichende Sprachwissenschaft" (Universität Zürich) provided the VedaWeb project a Filemaker file that was later transformed in Cologne into an Excel file. This data contained a version of the Rigveda by Prof. Dr. A. Lubotsky ("Indo-European Linguistics", Leiden University) that had been morphosytactically annotated over the course of more than 10 years at the University of Zurich. It also contained for each token, if available, a reference to an entry in Grassmann's dictionary for the Rigveda.

 

Modifications made by Jakob Halfmann and Natalie Korobzow to the data in 2020:

Disambiguation of the relevant categories, if unspecified in Zurich data, according to the Grassmann dictionary (updates from 6th edition partially included up to page 274):

  • case, gender and number for nouns, pronouns (columns G–I)
  • number, person, mood, tense and voice for verbs (columns I–M) up to line 109216
  • case, gender, number, tense and voice for participles (columns G–I, L–M) up to line 109216
  • absolutives are marked as Abs. in columns N and V
  • Inconsistencies between the original file from Zurich and the Grassmann dictionary as well as internal inconsistencies in Grassmann are noted in column AE, whenever they were noticed.
  • Zurich data was overwritten by conflicting Grassmann data in columns G–M but retained elsewhere.
  • Verb classes according to Whitney (1885) and Jamison (1983) for class 10 in column Y, differences in root spelling between Whitney and Grassmann are noted in column Z. All potential verb classes provided by Whitney are given for every occurrence of the root.
  • Local particles and verbal forms containing them are marked as LP in column AF.
  • Comparatives and superlatives are marked as such in column X and desideratives as Des. in column Y.

 

Modifications made by Anna Fischer (data transformation, technical realisation) to the data:

New structure of data table for linguistic annotations with new column titles:

  • A - "VERS_NR": renamed column (from "belege::stelleMMSSSRR")
  • B - "PADA_NR": renamed column (from "belege::pada")
  • C - "PADA_TEXT_LUBOTSKY": renamed column (from "belege::lubotskypada")
  • D - "TOKEN_NR_VERS": renamed column (from "belege::wortnummer rc")
  • E - "TOKEN_NR_PADA": renamed column (from "belege::wortnummer pada")
  • F - "FORM": renamed column (from "belege::form")
  • G - "KASUS": renamed column (from "belege::kasus")
  • H - "GENUS": renamed column (from "belege::genus")
  • I - "NUMERUS": renamed column (from "belege::numerus")
  • J - "PERSON": renamed column (from "belege::person")
  • K - "TEMPUS": moved and renamed column (from L "belege::tempus")
  • L - "PRAESENSKLASSE": created new column for present stem class for each form
  • M - "LEMMA_PRAESENSKLASSEN": created column for present stem classes of respective lemma: Moved and renamed column (from Y "formen::zusätzliche merkmale verb"), moved values "Abs." and "Inf." to column P "INFINIT", moved values "Prek." and "si-Ipv." to column N "MOOD", moved value "Des." to column Q "ABGELEITETE_KONJUGATION": moved value "se-Form" to column W "WEITERE_WERTE"
  • N - "MODUS": moved and renamed column (from K "belege::modus")
  • O - "DIATHESE": moved and renamed column (from M "belege::diathese")
  • P - "INFINIT": created new column for infinite forms "Abs.", "Inf.", "Ptz.", "ta-Ptz.", "na-Ptz."
  • Q - "ABGELEITETE_KONJUGATION": created new column for secondary conjugation "Des.", "Int.", "Kaus."
  • R - "GRADUS": created new column for degree: "Comp.", "Sup."
  • S - "LOKALPARTIKEL": moved and renamed column (from AF "LP")
  • T - "LEMMA_ZÜRICH": moved and renamed column (from AA "lemmata klassisch::lemma")
  • U - "LEMMA_ZÜRICH_LEMMATYP": moved and renamed column (from AB "lemmata klassisch::lemmatyp")
  • V - "LEMMA_ZÜRICH_BEDEUTUNG": moved and renamed column (from AC "lemmata klassisch::bedeutung")
  • W - "WEITERE_WERTE": created new column for all miscellaneous values: e.g. "Hyperchar.", "n-haltig", "se-Form"
  • X - "KOMMENTAR": created new column merging former columns Z "formen::HELPformbestimmung", AD "lemmata klassisch::HELPbedeutung" and AE "anmerkungen abweichungen"

Columns that were removed due to redundant information:

  • "formen::zusätzliche merkmale nomen": values "superlative" And "comparative" were renamed "sup." and "comp." and moved to new column R "GRADUS", all other values were moved to new column for miscellaneous W "WEITERE_WERTE"
  • "belege::belegbestimmung summe simpel": values "Ptz.", "ta-Ptz." and "na-Ptz." were moved to new new column P "INFINIT"
  • "belege::kasus bestof"
  • "belege::genus bestof"
  • "belege::numerus bestof"
  • "belege::person bestof"
  • "belege::modus bestof"
  • "belege::tempus bestof"
  • "belege::diathese bestof"
  • "belege::belegbestimmung bestof summe sophistiziert"

 

Revisions and additions made by Antje Casaretto to the data in 2023:

  • F-T: - revision and correction (wherever necessary) of all annotations (books 1-7)
  • G,H,I - disambiguation of case forms, reg. pronouns and nominal forms, if unspecified in Zurich data (books 1-7)
  • L - disambiguation of present stem classes (book 7 and book 1 up to line 21050 vers 01.125.01)
  • M - disambiguation of denominal verbs from primary verbs of the 10th class (books 1-10)
  • N - disambiguation of precative and optative forms wherever possible (books 1-7)
  • Q - new annotations for "Int." (intensives) and "Kaus." (causatives) (books 1-7)

 

Revisions and additions made by Antje Casaretto to the data in 2024 with support in data modeling and automation by Anna Fischer:

  • F-T: revision and correction (wherever necessary) of all annotations (books 8-10)
  • G,H,I: disambiguation of case and gender forms in nominal and pronominal forms, if unspecified in Zurich data (books 8-10)
  • L, M: disambiguation of present stem classes (books 1-10)
  • N: disambiguation of precative and optative forms wherever possible (books 8-10)
  • P: new annotations for "Gdv." (gerundives)
  • Q: new annotations for "Den." (denominatives) (books 1-10) and further annotations of “Kaus.” (causatives) and “Int.” (intensives) (books 8-10)
  • T: revision of lemmatization (books 1-10)
  • V: update of meanings according to revised lemmatization; minimal revision
  • W: revised annotation of ending -se (“se-Form”) (books 1-10); no systematic revision
  • X: no systematic revision
  • A-U: general revision of formal inconsistencies and typing errors (book 1-10)

 

Revisions made by Natalie Korobzov and Pascal Coenen to the data in 2024 with computational support by Anna Fischer:

  • Y - "LEMMA_GRASSMANN_ID": new column for references to Grassmann dictionary (books 1-10) and revision of Grassmann references

Other (En)

Abbreviations

Description Code TEI and platform Code vedaweb_zurich.xlsx Category TEI and platform Category vedaweb_zurich.xlsx Comment
first person 1 1. person PERSON  
second person 2 2. person PERSON  
third person 3 3. person PERSON  
irregular form - ir. - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
present stem class 1 - 1 - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
present stem class 2 - 2 - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
present stem class 3 - 3 - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
present stem class 4 - 4 - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
present stem class 5 - 5 - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
present stem class 6 - 6 - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
present stem class 7 - 7 - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
present stem class 8 - 8 - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
present stem class 9 - 9 - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
present stem class 10 - 10 - PRAESENSKLASSE present stems classes not imported in TEI or platform (due 2025)
ablative ABL Abl. case KASUS  
accusative ACC Akk. case KASUS  
active ACT Akt. voice DIATHESE  
aorist AOR Aor. tense TEMPUS  
causative CAUS Kaus. secondary conjugation ABGELEITETE_KONJUGATION  
comparative CMP comp. degree GRADUS  
conditional COND Kond. mood MODUS  
converb/absolutiva CVB Abs. non-finite INFINIT  
dative DAT Dat. case KASUS  
denominative DEN Den. secondary conjugation ABGELEITETE_KONJUGATION  
desiderative DES Des. secondary conjugation ABGELEITETE_KONJUGATION  
dual DU Du. number NUMERUS  
feminine F f. gender GENUS  
future FUT Fut. tense TEMPUS  
gerundive GDV Gdv. non-finite INFINIT  
genitive GEN Gen. case KASUS  
imperative IMP Ipv. mood MODUS  
imperative si IMP-si si-Ipv. mood MODUS  
indicative IND Ind. mood MODUS  
infinitive INF Inf. non-finite INFINIT  
injuctive INJ Inj. mood MODUS  
instrumental INS Instr. case KASUS  
intensive INT Int. secondary conjugation ABGELEITETE_KONJUGATION  
imperfect IPRF Imperf. tense TEMPUS  
locative LOC Lok. case KASUS  
local particle LP LP local particle LOKALPARTIKEL  
masculine M m. gender GENUS  
middle voice MED med. voice DIATHESE  
neuter N n. gender GENUS  
nominative NOM Nom. case KASUS  
optative OPT Opt. mood MODUS  
passive voice PASS pass. voice DIATHESE  
plural PL Pl. number NUMERUS  
past perfect PLUPRF Pluperf. tense TEMPUS  
precative PREC Prek. mood MODUS  
perfect PRF Perf. tense TEMPUS  
present PRS Präs. tense TEMPUS  
participle PTCP Ptz. non-finite INFINIT  
na participle perfective passive PTCP-na na-Ptz. non-finite INFINIT  
ta participle perfective passive PTCP-ta ta-Ptz. non-finite INFINIT  
subjunctive SBJV Konj. mood MODUS  
singular SG Sg. number NUMERUS  
superlative SUP sup. degree GRADUS  
vocative VOC Vok. case KASUS  

Files

Files (10.4 MB)

Name Size Download all
md5:cafa0415fde0a8a9232069a7de234e00
10.4 MB Download