Vedic Sanskrit has a number of linguistic features which are alien to most other Indo-European languages. Prominent examples include: phonologically, the introduction of retroflexes, which alternate with dentals, and morphologically, the formation of gerunds. Some philologists attribute such features, as well as the presence of non-Indo-European vocabulary, to a local substratum of languages encountered by Indo-Aryan peoples in Central Asia (Bactria-Marghiana) and within the Indian subcontinent during Indo-Aryan migrations, including the Dravidian languages. Scholars have claimed to identify a substantial body of loanwords in the earliest Indian texts, including evidence of Non-Indo-Aryan elements (such as -s- following -u- in Rigvedic busa). While some postulated loanwords are from Dravidian, and other forms are traceable to Munda or Proto-Burushaski, the bulk have no proven basis in any of the known families, suggesting a source in one or more lost languages. The discovery that some words taken to be loans from one of these lost sources had also been preserved in the earliest Old Avestan texts, and also in Tocharian, convinced Michael Witzel and Alexander Lubotsky that the source lay in Central Asia and could be associated with the Bactria–Margiana Archaeological Complex (BMAC). Another lost language is that of the Indus Valley civilization, which Witzel initially labelled Para-Munda, but later the Kubhā-Vipāś substrate.
Phonology Retroflex phonemes are now found throughout the Burushaski, Nuristani, Dravidian and Munda families. They are reconstructed for Proto-Burushaski, Proto-Dravidian and (to a minimal extent) for Proto-Munda, and are thus clearly an areal feature of the Indian subcontinent. They are not reconstructible for either Proto-Indo-European or Proto-Indo-Iranian, and they are also not found in Mitanni–Indo-Aryan loanwords. The acquisition of the phonological trait by early Indo-Aryan is thus unsurprising, but it does not immediately permit identification of the donor language. Since the adoption of a retroflex series does not affect poetic meter, it is impossible to say if it predates the early portions of the Rigveda or was a part of Indo-Aryan when the Rigvedic verses were being composed; however, it is certain that at the time of the redaction of the Rigveda (ca. 500 BCE), the retroflex series had become part of Sanskrit phonology. There is a clear predominance of retroflexion in the Northwest (Nuristani, Dardic, Khotanese Saka, Burushaski), involving affricates, sibilants and even vowels (in Kalasha), compared to other parts of the subcontinent. It has been suggested that this points to the regional, northwestern origin of the phenomenon in Rigvedic Sanskrit. Bertil Tikkanen is open to the idea that various syntactical developments in Indo-Aryan could have been the result of adstratum rather than the result of substrate influences. However Tikkanen states that "in view of the strictly areal implications of retroflexion and the occurrence of retroflexes in many early loanwords, it is hardly likely that Indo-Aryan retroflexion arose in a region that did not have a substratum with retroflexes." Not only the typological development of Old to Middle Indo-Aryan, but already the phonological development from Pre-Vedic to Vedic (including even the oldest attested form in the Rig-Veda) has been seen as suggestive of Dravidian influence. However, Hock argues that Dravidian should not be considered as significant, but that retroflexion is, rather, the outcome of areal features cutting across language boundaries in the Northwest of the Indian subcontinent, and extending into Central Asia.
Vocabulary In 1955 Burrow listed some 500 words in Sanskrit that he considered to be loans from non-Indo-European languages. He noted that in the earliest form of the language such words are comparatively few, but they progressively become more numerous. Though mentioning the likelihood that one source was lost Indian languages extinguished by the advance of Indo-Aryan, he concentrated on finding loans from Dravidian. Kuiper identified 383 specifically Rigvedic words as non-Indo-Aryan – roughly 4% of its vocabulary. Oberlies prefers to consider 344–358 "secure" non-Indo-European words in the Rigveda. Even if all local non-Indo-Aryan names of persons and places are subtracted from Kuiper's list, that still leaves some 211–250 "foreign" words, around 2% of the total vocabulary of the Rigveda. These loanwords cover local flora and fauna, agriculture and artisanship, terms of toilette, clothing and household. Dancing and music are particularly prominent, and there are some items of religion and beliefs. They only reflect village life, and not the intricate civilization of the Indus cities, befitting a post-Harappan time frame. In particular, Indo-Aryan words for plants stem in large part from other language families, especially from the now-lost substrate languages. Mayrhofer identified a "prefixing" language as the source of many non-Indo-European words in the Rigveda, based on recurring prefixes like ka- or ki-, that have been compared by Michael Witzel to the Munda prefix k- for designation of persons, and the plural prefix ki seen in Khasi, though he notes that in Vedic, k- also applies to items merely connected with humans and animals. Examples include:
… excerpt ends here. Continue reading the full article.
