Wednesday, 12 August 2026

CHAPTER 3 The Language Question: Tajik, Uzbek, Mughat In-Group Speech, and the Secret Lexicon From the proposed monograph: The Lyuli (Mughat) of Uzbekistan: Language, Ritual, Memory and the Search for South Asian Connections Dr. Manish Kumar C. Mishra

 CHAPTER 3

The Language Question:
Tajik, Uzbek, Mughat In-Group Speech, and the Secret Lexicon

From the proposed monograph:
The Lyuli (Mughat) of Uzbekistan: Language, Ritual, Memory and the Search for South Asian Connections

Dr. Manish Kumar C. Mishra

3.1 Introduction: why language is the decisive archive

Among all the possible forms of evidence concerning Lyuli/Mughat history, language has the greatest potential to move the discussion beyond resemblance, legend, and ethnographic stereotype. Dress can change rapidly; occupations can be shared by unrelated communities; ritual forms can diffuse across religious and regional boundaries; and ethnonyms can be imposed from outside. A structured linguistic system, by contrast, can preserve traces of historical contact and descent in ways that can be tested comparatively.

The linguistic situation of the Lyuli/Mughat is nevertheless unusually difficult. The community lives within a multilingual Central Asian environment in which Tajik and Uzbek are major languages of everyday interaction. At the same time, scholarship has repeatedly referred to a restricted internal repertoire—variously described as Mugat language, secret language, argot, or community speech. These descriptions are not automatically equivalent. The central task of this chapter is therefore conceptual before it is etymological: what exactly is the linguistic object that later chapters will compare with South Asian languages?

A recent Uzbek dissertation explicitly describes the “Mugat language” as the mutual communication language of the Lyuli and discusses its present state and preservation. Earlier comparative linguistic scholarship, however, classifies Jugi/Mugat of the Hissar Valley as Tajik-based, and describes Luli/Multani of the Fergana area as an argot with a Tajik base. The apparent contradiction is productive rather than fatal. It may reflect different meanings of the word language, different communities and regions, or different layers of the same repertoire.

3.2 Four terms that must not be confused

The first distinction is between language, ethnolect, register, and argot. A language is a relatively autonomous linguistic system with grammar and lexicon capable of ordinary communication across domains. An ethnolect is a socially or ethnically associated variety of a wider language. A register is a context-dependent mode of speech used for particular purposes. An argot is a specialised vocabulary or speech practice often associated with occupational, social, or secrecy functions.

The label “secret language” can cover more than one of these realities. A community may speak ordinary Tajik grammar while replacing selected content words with special lexical items. That would differ fundamentally from a complete inherited language whose grammar, pronouns, numerals, verbs, and syntax differ from Tajik. Between these poles are mixed systems: a Tajik grammatical frame with a large non-Tajik lexicon; an ethnolect containing distinctive phonology and morphology; or a repertoire in which ordinary speech and restricted argot alternate according to audience.

Until primary recordings and paradigms are analysed, this monograph will use the neutral expression “Mughat in-group speech” for the restricted repertoire and reserve “Mughat language” for contexts in which the source itself uses that designation.

3.3 The public linguistic ecology: Tajik and Uzbek

The Lyuli/Mughat do not exist outside the linguistic ecology of Uzbekistan and Tajikistan. Tajik, an Iranian language closely related to Persian, has historically played a central role in many Mughat communities, while Uzbek, a Turkic language, is indispensable in much of Uzbekistan’s public life. Russian may also enter the repertoire through education, administration, migration, media, and Soviet/post-Soviet experience.

This multilingualism has methodological consequences. A word found in Mughat speech may have entered from Tajik, Uzbek, Russian, Persianate religious vocabulary, or another regional argot. Even an apparently South Asian-looking word must therefore be tested against all plausible Central Asian sources before it is classified as Indic.

The distinction between language of home, language of neighbourhood, language of school, language of work, and language used when outsiders are present should be documented in future fieldwork. “What language do you speak?” is too crude a question for a community with layered repertoires.

3.4 What Encyclopaedia Iranica tells us

Gernot L. Windfuhr’s survey in Encyclopaedia Iranica provides one of the clearest comparative classifications. For Central Asia it identifies Jugi, with the endonym Mugat, in the Hissar Valley as Tajik-based and relates it to Jogi populations of Afghanistan and Jugi of northern Iran. It separately records Luli of the Fergana area with the endonym Multani and describes their argot as Tajik-based. It also distinguishes Kara-Luli, whose endonym is reported as Hindustani, and an Afghon group with an Indic-based dialect.

The importance of this classification lies in its heterogeneity. The historical umbrella of “Gypsy dialects” in Central Asia includes Tajik-based, Afghan-Persian-based, Persian-based, and Indic-based varieties. Therefore the mere presence of a community within the same social category cannot establish linguistic genealogy.

For the Mughat question, the Iranica classification makes one proposition especially important: the structural base may be Tajik even where the lexicon contains elements of other origins. This is precisely the kind of situation in which lexical stratigraphy becomes necessary.

3.5 The Uzbek dissertation and the expression “Mugat language”

A dissertation associated with the Academy of Sciences of Uzbekistan provides a valuable contemporary perspective. Its English summary states that the scientific findings concern “the present state of the Mugat language which is the mutual communication language of the lyuli and its preservation.” The dissertation also links language with customs, rituals, and cultural heritage.

This formulation should be taken seriously because it reflects recent Uzbek scholarship and may correspond closely to community usage. Yet it should not be used to override the earlier structural classification without examining the underlying data. The expression “mutual communication language” may refer to an in-group ethnolect or argot that functions socially as the community’s own language even if much of its grammar is Tajik-based.

Sociolinguistic identity and genealogical classification answer different questions. A community can legitimately call a repertoire “our language” while a historical linguist describes its grammatical matrix as Tajik. Both statements can be true at the same time.

3.6 Secret speech as a social institution

Restricted speech is not only a linguistic phenomenon; it is a social institution. Its function may include privacy in the presence of outsiders, reinforcement of group solidarity, occupational secrecy, joking, taboo management, protection of commercial information, or the marking of boundaries between insiders and non-members.

This matters for etymology. A vocabulary deliberately designed to be opaque is unusually open to lexical replacement and borrowing. Speakers may select unfamiliar words precisely because outsiders do not understand them. Consequently, the lexicon of an argot can be historically cosmopolitan even when its grammar belongs overwhelmingly to one language.

The social function of Mughat in-group speech must therefore be recorded alongside every lexical item: Who uses it? With whom? In which situations? Do children understand it? Are women and men equally fluent? Does it differ between generations? Is it used in complete sentences or only by inserting special words into Tajik or Uzbek sentences? These questions are essential to deciding whether the object is a language, ethnolect, register, or argot.

3.7 Loterāʾi, Abdoltili, and the wider Persianate argot network

The wider Persianate world contains a long history of specialised argots. Encyclopaedia Iranica’s treatment of Loterāʾi is particularly important because it compares argots of Jugi, Luli, Chistoni, Kavoli, musicians, mendicant dervishes, and other mobile or occupational groups. It defines Abdoltili as the “language of itinerants” associated with Uzbek-speaking artisans and musicians, preachers, and qalandars.

This evidence complicates any attempt to identify every non-Tajik Mughat word as ancestral. Lexical items could circulate across professional and mobile networks without the speakers sharing a single ancestry. The same source nevertheless records explicitly Indic-derived words in related Iranian and Central Asian argots, demonstrating that an Indic layer is not imaginary. The analytical problem is to distinguish inherited material from borrowed argot material.

The existence of an argot network therefore weakens simplistic etymology but strengthens the case for systematic comparison.

3.8 A small verified sample from the comparative argot literature

The Loterāʾi survey gives a useful demonstration of how mixed such vocabularies can be. In the Djougi material of northern Iran—historically compared with Jugi of Tajikistan—it lists forms classified as Indic alongside Arabic and Iranian items. Examples of Indic-classified vocabulary include forms glossed as ‘man’, ‘woman/wife’, ‘iron’, ‘water’, ‘big’, ‘goat’, ‘horse’, and ‘donkey’. The same corpus contains Arabic and Iranian words.

These examples must not be transferred automatically into modern Uzbek Lyuli speech. They belong to a documented comparative argot corpus and are introduced here only to demonstrate the method and the possibility of stratification. The later lexical chapter will include a word only when its community, locality, source, transcription, and gloss can be identified.

This rule is essential: a word attested in Djougi, Loterāʾi, Domari, Romani, or Parya is not a Lyuli/Mughat word merely because the communities have been compared historically.

3.9 The concept of lexical stratigraphy

Lexical stratigraphy treats vocabulary as layers deposited through different periods of contact. For Mughat in-group speech, at least five layers must be tested: Tajik/Persian; Uzbek and other Turkic material; regional professional or religious argot; possible Indic material; and later Russian or modern borrowings. A sixth category—unresolved—must remain available for forms whose origin cannot be established.

The central question is not how many words can be made to resemble Hindi. It is whether the putative Indic forms cluster in historically conservative semantic domains and display regular sound correspondences. Basic verbs, kinship terms, body parts, numerals, pronouns, common animals, and elementary material culture generally carry more genealogical weight than occupational code words that are easily borrowed.

The book will therefore assign both an etymological category and an evidentiary confidence level to every analysed item.

3.10 Grammar is more important than attractive word matches

Popular linguistic comparisons often begin and end with similar-sounding words. Historical linguistics requires more. If Mughat in-group speech possesses Tajik word order, Tajik verb morphology, Tajik case/prepositional structures, and Tajik pronouns, while substituting selected secret nouns, the system is fundamentally different from an Indo-Aryan language such as Parya.

Conversely, if primary material reveals non-Tajik grammatical morphology, inherited pronouns, numeral systems, verb paradigms, or productive suffixes that correspond regularly with Indo-Aryan, the historical implications would be much stronger.

For this reason, future field elicitation must collect sentences and paradigms, not merely vocabulary lists. A hundred isolated words cannot substitute for ten carefully recorded grammatical constructions.

3.11 Parya/Tajuzbeki as the positive control

Parya provides a crucial control because it represents a demonstrably Indo-Aryan language preserved in the Central Asian environment. Bholanath Tiwari’s Tajuzbeki study, together with Oransky’s Parya research, allows the book to observe what genuine Indo-Aryan structural survival looks like after prolonged contact with Tajik and Uzbek.

The comparison should therefore operate at two levels. First, lexical: do Mughat in-group forms correspond systematically with Parya and north-western Indo-Aryan vocabulary? Second, grammatical: does Mughat preserve any structures comparable with Parya that cannot be explained through Tajik?

If the answer is lexical but not grammatical, the likely history may involve borrowing or residual vocabulary after language shift. If both lexicon and grammar align systematically, a much stronger genealogical argument would become possible.

3.12 The India test: Hindi alone is not enough

A serious search for South Asian connections cannot compare Mughat only with Standard Hindi. If the Multoni/Multani clue is historically meaningful, the most relevant comparison may lie farther northwest. Punjabi, Saraiki, Sindhi, Rajasthani, Haryanvi, and related Indo-Aryan varieties must therefore be included, alongside Parya, Romani, and Domari where appropriate.

Geography matters because cognates can reveal subgrouping. A form shared broadly across Indo-Aryan proves less about a precise homeland than a distinctive cluster concentrated in one region. At the same time, Persian loans in Hindi, Punjabi, or Urdu must not be mistaken for Indic inheritance when the same word could have reached Mughat directly through Persian/Tajik.

Every proposed Indian connection must therefore pass a contact-filter test: could this word be explained more economically through Persian, Tajik, Uzbek, Arabic, or regional argot? Only after those alternatives are excluded should an Indic etymology be preferred.

3.13 Proposed lexical database for the monograph

The final book will build a transparent lexical database. Each entry should contain the original Mughat form exactly as recorded, phonetic/transliteration information, English gloss, locality, speaker or source where available, date, grammatical category, example sentence, Tajik comparison, Uzbek comparison, Persian/Dari comparison, Abdoltili or other argot parallels, Parya/Tajuzbeki comparison, South Asian comparisons, proposed etymology, and confidence level.

The database will distinguish ‘probable inherited Indic’, ‘possible Indic’, ‘Iranian’, ‘Turkic’, ‘regional argot’, ‘Russian/modern’, and ‘uncertain’. The category ‘probable inherited Indic’ will require more than phonetic resemblance: semantic fit, plausible sound correspondence, and absence of a more immediate contact-language explanation will all be required.

This format makes the argument reproducible. Readers will be able to see not only the conclusion but the evidence behind each lexical decision.

3.14 Fieldwork protocol: how the language should be recorded

If fieldwork becomes possible in Uzbekistan, the linguistic component should be designed ethically and scientifically. Participation must be voluntary and based on informed consent. Because an in-group repertoire may be intentionally restricted, researchers must not pressure speakers to disclose words they regard as private, sacred, or socially sensitive.

Elicitation should begin with sociolinguistic questions before vocabulary: languages used with parents, spouse, children, neighbours, school, market, officials, and outsiders; self-name for the internal speech; age at which it is learned; and whether younger speakers retain it. Recordings should include natural conversation where consent permits, followed by controlled elicitation of basic vocabulary and grammar.

Regional sampling is essential. Samarkand, Bukhara, Surkhandarya, Fergana, and communities in Tajikistan may not share identical repertoires. Variation itself may preserve migration history.

3.15 Language vitality and intergenerational transmission

The contemporary Uzbek dissertation’s emphasis on preservation raises another major question: is Mughat in-group speech endangered? Language vitality cannot be inferred merely from the existence of older speakers. The decisive issue is intergenerational transmission.

The study should examine whether children understand and actively use the restricted repertoire, whether vocabulary is shrinking, whether Uzbek is replacing Tajik in some regions, and whether schooling, urbanisation, migration, smartphones, and social media are changing the communicative function of secret speech.

A paradox may emerge: a repertoire originally valued because outsiders could not understand it may become less useful as social boundaries change, while simultaneously becoming more important as a symbol of cultural identity. Such a transition from functional secrecy to emblematic heritage would be sociolinguistically significant.

3.16 What the present evidence allows us to conclude

Three conclusions can already be stated with reasonable confidence. First, the Lyuli/Mughat linguistic repertoire is multilingual and cannot be described adequately by a single label. Second, authoritative comparative scholarship describes major Jugi/Mugat and Luli/Multani varieties as Tajik-based, while also situating them within a wider network of specialised argots containing vocabulary from multiple origins. Third, contemporary Uzbek scholarship recognises a socially meaningful “Mugat language” used for mutual communication and treats its preservation as part of Lyuli cultural heritage.

These propositions are compatible if linguistic structure and social identity are separated. A Tajik-based ethnolect or argot can function as an in-group language. The unresolved question is historical: how much of its distinctive vocabulary is inherited from an earlier South Asian language, how much entered through Persianate argot networks, and how much reflects later Central Asian contact?

That question cannot be answered by terminology alone. It requires the corpus.

3.17 Conclusion: from the language question to the lexical test

This chapter has established the methodological foundation for the linguistic core of the book. The phrase “Mughat language” should neither be dismissed nor accepted uncritically as proof of an autonomous genealogical language. Earlier scholarship points to a Tajik grammatical base in important Jugi/Mugat and Luli/Multani varieties; contemporary Uzbek research emphasises the community function and preservation of Mugat speech; and comparative argot scholarship demonstrates that specialised vocabularies in the Persianate world can combine Indic, Iranian, Arabic, Turkic, and other elements.

The strongest India connection will therefore not come from the existence of secrecy itself. It will come, if at all, from a demonstrable pattern inside the lexicon and grammar.

Chapter 4 will undertake that test. It will reconstruct the available Mughat/Lyuli lexical corpus source by source, separate locality-specific data, compare each form with Tajik, Persian, Uzbek, regional argots, Parya/Tajuzbeki, and relevant Indo-Aryan languages, and assign an explicit confidence level to every proposed South Asian correspondence. The aim will not be to prove an Indian origin in advance, but to discover exactly what the linguistic evidence permits us to say.

Table 3.1. Working distinction among the linguistic categories

Category

Working definition

What would demonstrate it?

Risk if misused

Language

Relatively autonomous grammar and lexicon used across domains

Paradigms, syntax, basic vocabulary, connected speech

Calling an argot a separate language

Ethnolect

Community-associated variety of a wider language

Systematic phonological/lexical/grammatical features

Ignoring social identity

Register

Context-specific speech style

Clear situational distribution

Mistaking function for genealogy

Argot / secret speech

Restricted repertoire used for opacity or group functions

Special vocabulary, switching, insider use

Treating borrowed code words as ancestral

Table 3.2. Comparative classification relevant to the Mughat question

Group / label

Reported base or type

Area

Analytical significance

Jugi; endonym Mugat

Tajik base

Hissar Valley

Core evidence for Tajik-based Mugat classification

Luli; endonym Multani

Argot, Tajik base

Fergana

Links Multani designation with Tajik-based argot

Kara-Luli; endonym Hindustani

Distinct comparative category

Fergana

Shows heterogeneity under Luli labels

Afghon

Indic-based dialect group

Central Asian/Afghan comparative context

Demonstrates genuine Indic-based varieties existed in the wider field

Parya

Indo-Aryan

Hissar/Surkhandarya zone

Positive control for inherited Indo-Aryan structure

Figure 3.1. Proposed linguistic stratigraphy

CURRENT PUBLIC REPERTOIRE
Tajik  ↔  Uzbek  ↔  Russian (context-dependent)

↓ overlapping with ↓

MUGHAT IN-GROUP SPEECH

Possible lexical layers:
Tajik/Persian
+ Uzbek/Turkic
+ Persianate/Central Asian argot (including Abdoltili networks)
+ possible Indic residue
+ later Russian/modern vocabulary
+ unresolved forms

Table 3.3. Template for the Chapter 4 lexical database

Mughat form

Gloss

Source/locality

Tajik/Persian

Uzbek

Parya/Tajuzbeki

South Asian parallels

Judgement

References cited in Chapter 3

Windfuhr, Gernot L. 2002. ‘Gypsy ii. Gypsy Dialects’. Encyclopaedia Iranica, XI/4, pp. 415–421.

Windfuhr, Gernot L. 2012. ‘Loterāʾi’. Encyclopaedia Iranica. Survey of Persianate secret languages and argots, including Jugi/Luli-related material and Abdoltili.

Marushiakova, Elena, and Vesselin Popov. 2016. Gypsies of Central Asia and the Caucasus. Cham: Palgrave Macmillan.

Oranskiĭ, I. M. 1961. Studies cited in comparative classifications of Central Asian Jugi/Mugat and related groups.

Oranskiĭ, I. M. 1983. Comparative ethnolinguistic work cited for Mugat/Jugi and Central Asian argot material, especially the corpus discussed in later scholarship.

Tiwari, Bholanath. 1970. Tajuzbeki: Soviet Sangh mein Boli Jane Vali Hindi Boli: Aitihasik aur Tulnatmak Adhyayan tatha Sankshipt Shabdkosh. Delhi: National Publishing House.

Uzbekistan Academy of Sciences-associated dissertation, English summary, 150 pp. Contemporary study describing the Mugat language as the mutual communication language of the Lyuli and discussing its preservation. ZiyoNET digital copy.

Source-critical note

This chapter does not present a fabricated Mughat word list. The few comparative lexical examples discussed in the prose derive from the Encyclopaedia Iranica Loterāʾi survey and are explicitly identified there as Djougi/comparative argot material, not automatically as modern Uzbek Mughat vocabulary. Chapter 4 should proceed only from recoverable primary or clearly attributed lexical corpora. The contemporary Uzbek dissertation is used here for its explicit statement about the social status and preservation of “Mugat language”; its detailed linguistic data must be extracted and checked before being used for word-level etymology.

India, Multan, Afghanistan, and Central Asia: Reconstructing the Lyuli/Mughat Migration Problem From the proposed monograph: The Lyuli (Mughat) of Uzbekistan: Language, Ritual, Memory and the Search for South Asian Connections Dr. Manish Kumar C. Mishra

 CHAPTER 2

India, Multan, Afghanistan, and Central Asia:
Reconstructing the Lyuli/Mughat Migration Problem

From the proposed monograph:
The Lyuli (Mughat) of Uzbekistan: Language, Ritual, Memory and the Search for South Asian Connections

Dr. Manish Kumar C. Mishra

2.1 Introduction: from origin stories to an evidence-based migration history

The question of where the Lyuli/Mughat came from is among the most frequently repeated—and least securely resolved—questions in the literature on Central Asian minority communities. South Asian origin has often been asserted through a combination of names, legends, linguistic fragments, occupational parallels, and analogies with Roma or other mobile populations. Yet these forms of evidence do not possess equal historical value. A migration history must therefore be reconstructed in layers rather than narrated as a single uninterrupted journey from “India” to Uzbekistan.

This chapter asks a narrower question: what can presently be established about the geographical and chronological stages that may connect the Mughat/Lyuli with the Indian subcontinent? It separates four evidentiary fields: medieval Persian literary traditions concerning Luri/Luli; the ethnonym Multoni/Multani and its apparent reference to Multan; direct evidence for Lyuli in Central Asia; and linguistic classifications linking some Central Asian in-group varieties with South Asian or Afghan populations.

The conclusion is deliberately provisional. The available scholarship makes a South Asian connection historically plausible and, in some formulations, highly probable. But it does not yet permit a single precise route, date, or homogeneous ancestral population to be reconstructed for all communities subsequently called Lyuli.

2.2 The oldest narrative: Bahram Gur and musicians from India

One of the most influential origin narratives appears in the Persian historical-literary tradition surrounding the Sasanian ruler Bahram Gur. Later accounts tell of musicians or entertainers brought from India and subsequently dispersed. Marushiakova and Popov review the long academic tradition that connected the Luri/Luli of these narratives with later Central Asian Lyuli, while also warning that the narrative was recorded centuries after the events it purportedly describes.

This warning is methodologically decisive. A medieval literary tradition may preserve a historical memory, but it cannot be treated as a contemporary migration register. The Bahram Gur narrative is therefore important evidence for the existence of a Persianate memory linking groups called Luri/Luli with India. It is not, by itself, proof that every modern Mughat family descends directly from the people described in that story.

Later Persian and Arabic writers repeated or elaborated the narrative. The persistence of the India–Luli association is historically significant because it shows that the connection was not invented in modern nationalist discourse. Yet repetition can also reproduce literary convention. The task is therefore to distinguish the history of the tradition from the history of the population.

2.3 Luli in Persianate memory: literature as evidence and as problem

Marushiakova and Popov note references to Luli in medieval Persian and Arabic sources and in Persian poetry. These references establish that Luli was a recognisable social-cultural designation within the Persianate world. They also demonstrate that the category carried associations with performance, mobility, marginality, and particular occupations.

However, literary representation is not equivalent to ethnographic description. Poets may employ a social label metaphorically; chroniclers may reproduce inherited stereotypes; and a name may shift meaning across centuries. Consequently, the historical method adopted here asks three questions of every reference: Does the source describe an identifiable population? Does it provide a location? Does it contain information independent of literary convention?

This distinction becomes crucial when tracing a path toward Central Asia. The fact that Luli occur in Persian literature strengthens the historical depth of the ethnonym, but it does not yet demonstrate continuity between all medieval Luli and modern Mughat.

2.4 Multoni/Multani: the most geographically specific clue

The designation Multoni or Multani deserves special attention because it is geographically more specific than the general term Lyuli. Marushiakova and Popov record its use in parts of Uzbekistan and historically in Samarkand and Surkhandarya, explaining it with reference to the city of Multan in medieval India, now Pakistan. They also describe an older social division between sedentary craft-associated Kasib/Kosib and more mobile Multoni/Multani.

Multan was a major urban, commercial, religious, and strategic centre linking the north-western subcontinent with routes toward Afghanistan, Iran, and Central Asia. For that reason, an ethnonym meaning “from Multan” would fit a historically plausible geography of transregional mobility. But plausibility is not proof.

The strongest test will be convergence. If Multoni is shown in early documents to designate the same population later called Mughat; if oral traditions independently remember Multan or neighbouring regions; and if the in-group lexicon contains systematic correspondences with Punjabi, Saraiki, Sindhi, or other languages of the north-western subcontinent, then the Multan hypothesis would gain substantial evidentiary force. Until that comparison is completed, Multoni should be classified as a strong geographical clue rather than a proven point of origin.

2.5 Afghanistan as corridor, contact zone, or source?

Any serious reconstruction of South Asia–Central Asia migration must consider Afghanistan not merely as a line on a map but as a historical contact zone. Routes linking Punjab, Sindh, Multan, Kabul, Balkh, Badakhshan, and Transoxiana carried merchants, soldiers, pilgrims, craftsmen, religious specialists, entertainers, and mobile communities in multiple directions.

The linguistic literature strengthens the importance of this zone. Encyclopaedia Iranica’s survey of so-called Gypsy dialects distinguishes several Central Asian groups and explicitly relates some to populations in Afghanistan. Jugi/Mugat in the Hissar Valley is described as Tajik-based and related to Jogi populations of Afghanistan and northern Iran; other groups show Afghan-Persian, Pashto, or Indic-based features. This does not establish that all such groups shared one ancestry. Rather, it demonstrates that Afghanistan formed part of a larger network within which mobile populations, argots, and social labels circulated.

The book will therefore avoid a simplistic arrow—India → Afghanistan → Uzbekistan—as though a single migration occurred once. Afghanistan may have functioned at different times as corridor, temporary homeland, linguistic contact zone, and source of later migration.

2.6 The first direct Central Asian evidence and the Baburnama problem

Specialist scholarship identifies the Baburnama of Zahir al-Din Muhammad Babur as containing the first direct evidence for Lyuli in Central Asia. This is potentially a major chronological anchor because Babur’s own world connected Fergana, Samarkand, Kabul, and northern India.

At the present stage, however, this book will not reproduce an exact Baburnama quotation until the passage has been checked against an authoritative edition and translation. Secondary scholarship is sufficient to establish that Marushiakova and Popov treat the Baburnama as the first direct Central Asian evidence, but a monograph of this kind must verify the exact wording, context, personal name, and location before drawing further conclusions.

This is an example of the source-critical principle that will govern the book: a secondary claim may identify where evidence exists, but the primary text should be consulted before it becomes the foundation of a historical argument.

2.7 From medieval presence to Central Asian rootedness

By the period for which the Mughat/Lyuli are more securely documented, they are not simply travellers passing through Central Asia. Scholarship describes them as an integral part of Central Asian life, living in towns and villages primarily in Uzbekistan and Tajikistan and also in neighbouring states. This long regional presence complicates any language of “foreignness.”

A community may have a migration history extending beyond Central Asia and still be deeply Central Asian by historical experience. Indeed, the very persistence of Tajik-based speech, Uzbek bilingualism, Sunni Muslim practice, local clan structures, and regionally specific occupations demonstrates adaptation across generations.

Thus the historical question is not whether Lyuli are “Indian or Uzbek/Tajik.” Such a binary is analytically misleading. The relevant question is how a population with possible South Asian historical connections became socially, linguistically, and culturally embedded in Central Asia.

2.8 Migration is rarely a single event

One of the central weaknesses in popular accounts of Lyuli origins is the assumption of a single departure date from India. Mobile communities can form through repeated movements, fission, incorporation of outsiders, occupational alliances, marriage, and language shift. A name can travel farther than a lineage, while a professional argot can circulate between communities without large-scale population movement.

The evidence therefore permits several migration models. One is a single-origin model in which an ancestral South Asian population migrated through Iran or Afghanistan and diversified in Central Asia. A second is a multiple-stream model in which different South Asian-connected populations entered Central Asia at different times and were later grouped under similar labels. A third is a convergence model in which populations of partly different origins acquired common social classifications and overlapping argot traditions within the Persianate and Central Asian environment.

At present, the multiple-stream and convergence models deserve serious consideration because specialist linguistic classifications reveal considerable heterogeneity among groups historically placed under broad “Gypsy” categories.

2.9 What language can contribute to migration history

Historical linguistics can test migration narratives more rigorously than visual resemblance or occupational analogy. If the Mughat in-group vocabulary contains an inherited Indic substrate, its geographical affinities may help identify the South Asian zone from which some ancestors came.

The key is systematic correspondence. A handful of words resembling Hindi would be weak evidence. A cluster of basic vocabulary showing regular phonological correspondences with Punjabi, Saraiki, Sindhi, Rajasthani, Haryanvi, or another Indo-Aryan group would be much stronger. Equally important would be negative evidence: if supposedly “Indian” terms prove to be Persian, Tajik, Uzbek, Abdoltili, or widely circulating argot words, the South Asian linguistic case would need to be narrowed.

The next major linguistic chapters will therefore test migration hypotheses rather than merely illustrate them.

2.10 Parya/Tajuzbeki as a comparative control

Parya is especially useful because its South Asian linguistic ancestry is demonstrable. Encyclopaedia Iranica separates Parya-type Indic-based varieties from Tajik-based Jugi/Mugat and other Central Asian argots. Bholanath Tiwari’s Tajuzbeki corpus provides an Indian scholarly documentation of the Parya language in the Tajik-Uzbek zone.

This comparison prevents circular reasoning. If a lexical method identifies Indo-Aryan inheritance clearly in Parya but finds only a small or specialised Indic layer in Mughat speech, the two communities likely experienced different linguistic histories. If, however, Mughat basic vocabulary reveals systematic correspondences with Parya or a specific north-western Indo-Aryan cluster, the possibility of deeper historical connection would deserve renewed investigation.

Parya is therefore not evidence that Lyuli are Indo-Aryan speakers; it is a control against which the strength of Lyuli linguistic evidence can be measured.

2.11 An evidence-weighted migration model

The present evidence can be arranged by confidence rather than forced into a single narrative. The long Persianate association of Luli with India is historically old but partly literary. The Multoni/Multani ethnonym is geographically suggestive and potentially important. Direct Central Asian presence is securely established by the early modern period in specialist scholarship. Long residence in the Tajik-Uzbek cultural zone is unquestionable. The exact South Asian homeland, departure date, number of migration streams, and relationship among Mughat, Parya, Roma, Dom, Jogi, and other mobile communities remain open.

This produces a cautious model:

Possible South Asian source region(s)
→ Persianate literary memory and/or movement through Iran-Afghanistan
→ differentiated migrations and contacts
→ established Central Asian communities
→ Tajik/Uzbek linguistic and cultural adaptation
→ modern Mughat/Lyuli identities.

The arrows are not equal. Some represent documented historical presence; others represent hypotheses that require linguistic or archival corroboration.

2.12 Research questions generated by the migration evidence

The migration problem can now be converted into a series of testable questions. Does the name Multoni occur in records early enough to predate modern speculation about Indian origins? Do Mughat oral histories name Multan, Punjab, Sindh, Afghanistan, or particular Central Asian staging points? Does clan nomenclature preserve geographical memory? Are there regional differences between Samarkand, Bukhara, Surkhandarya, Fergana, and Tajik communities that correspond to different migration histories? Does in-group vocabulary point toward one South Asian linguistic zone or several? Are words regarded as Indic actually inherited, or did they circulate through professional argots?

These questions require different sources. No single archive can answer them all. Historical texts, Soviet ethnography, Uzbek and Tajik scholarship, linguistic corpora, oral history, and contemporary fieldwork must be brought into dialogue.

2.13 Conclusion: what can be said before the lexical evidence is tested

The migration history of the Lyuli/Mughat cannot responsibly be reduced to a single sentence such as “they came from India.” The available evidence is both richer and more complicated.

A South Asian connection has deep roots in Persianate traditions concerning Luri/Luli and is reinforced by the geographically suggestive Multoni/Multani designation. Specialist scholarship places Mughat/Lyuli firmly within the long history of Central Asia and identifies direct evidence by the early modern period. Linguistic surveys further show that the region contained several mobile or marginal communities with different language bases—Tajik, Persian, Afghan-Persian, Pashto-influenced, and Indic—making a single homogeneous migration model unlikely.

The most defensible conclusion at this stage is therefore that the Mughat/Lyuli history probably belongs to a wider field of South Asia–Afghanistan–Iran–Central Asia mobility, but the precise ancestry, route, and chronology of the focal community remain to be demonstrated. The name Multoni may preserve an especially important geographical memory; Afghanistan was likely a crucial zone of movement and contact; and Central Asia became not merely a destination but the setting in which modern Mughat identity was historically formed.

The decisive next step is linguistic. Chapter 3 will therefore turn to the public and private linguistic repertoire of the Lyuli/Mughat, asking what is meant by “Mughat language,” “secret language,” “argot,” and Tajik-based community speech. Only after establishing those categories can the book undertake the word-by-word Indic lexical analysis that will test the migration hypotheses developed here.

Table 2.1. Migration claims and their present evidentiary status

Claim

Evidence

Assessment

Persianate Luri/Luli traditions connect the group-name with India

Medieval literary/historical tradition as reviewed in specialist scholarship

Historically significant tradition; not direct migration proof

Multoni/Multani refers to Multan

Ethnographic nomenclature and specialist interpretation

Strong clue; requires independent corroboration

Lyuli were present in Central Asia by Babur’s period

Secondary specialist identification of Baburnama evidence

Strong, pending direct primary-text verification

Afghanistan formed part of the relevant mobility zone

Comparative linguistic/ethnographic links among Jugi/Mugat and Afghan populations

Strong regional-context evidence; route specifics unresolved

All Lyuli descend from one migration from India

No adequate single-source proof

Unproven

Modern Mughat in-group speech preserves an Indic substrate

Limited Indic lexical claims in scholarship; corpus not yet fully tested

Open hypothesis

Figure 2.1. Evidence-weighted migration model

Possible South Asian source region(s)

Multan / north-western subcontinent?  [hypothesis]

Afghanistan–Iranian contact corridor  [plausible regional context]

Central Asian presence by early modern period  [strong]

Long-term Tajik–Uzbek adaptation  [established]

Modern Mughat/Lyuli communities

Note: The diagram distinguishes levels of evidence; it is not intended as a proven single migration route.

References cited in Chapter 2

Marushiakova, Elena, and Vesselin Popov. 2016. Gypsies of Central Asia and the Caucasus. Cham: Palgrave Macmillan. DOI: 10.1007/978-3-319-41057-9.

Marushiakova, Elena, and Vesselin Popov. 2015. ‘Central Asian Gypsies: Identities and Migrations’. Sprawy Narodowościowe / Nationalities Affairs, no. 47. DOI: 10.11649/sn.2015.031.

Windfuhr, Gernot L. 2002. ‘Gypsy ii. Gypsy Dialects’. Encyclopaedia Iranica. Comparative classification of Jugi/Mugat, Luli/Multani and other Central Asian varieties.

Ozierski, Przemysław. 2020. ‘Central Asian Gypsies – Lyuli: The Overview of Current Socio-Economic Problems’. Review of Nationalities, no. 3.

Tiwari, Bholanath. 1970. Tajuzbeki: Soviet Sangh mein Boli Jane Vali Hindi Boli: Aitihasik aur Tulnatmak Adhyayan tatha Sankshipt Shabdkosh. Delhi: National Publishing House.

Babur, Zahir al-Din Muhammad. Baburnama. Primary passage concerning Lyuli to be verified against an authoritative edition before quotation in the final monograph.

Source-critical note

This chapter deliberately distinguishes between primary historical evidence and claims presently available through specialist secondary scholarship. The Bahram Gur/Luri tradition is treated as a historical literary tradition rather than as a literal migration record. The Baburnama reference is not quoted because the exact passage has not yet been verified against an authoritative primary edition. The Multoni–Multan relationship is treated as a research hypothesis requiring independent linguistic and historical corroboration. These distinctions will be retained in the final book.

CHAPTER 3 The Language Question: Tajik, Uzbek, Mughat In-Group Speech, and the Secret Lexicon From the proposed monograph: The Lyuli (Mughat) of Uzbekistan: Language, Ritual, Memory and the Search for South Asian Connections Dr. Manish Kumar C. Mishra

  CHAPTER 3 The Language Question: Tajik, Uzbek, Mughat In-Group Speech, and the Secret Lexicon From the proposed monograph: The Lyuli (Mugh...

Popular Posts