ArticleslgStudy

science

Linguistic distance

Linguistic distance is a science topic covered in the lgStudy science library. This page brings together a partial reference excerpt, illustrations, worked examples, real-world applications and a short study plan, so you can understand Linguistic distance rather than just read about it. In short: Linguistic distance is the measure of how different one language (or dialect) is from another. Although they lack a uniform approach to quantifying linguistic distance between languages, linguists apply the concept to a variety of linguistic contexts, such as second-language acquisition, historical linguistics, language-based conflicts, and the effects of language differences on trade.

Key takeaways

  • Linguistic distance belongs to science; place it in that map before memorising details.
  • Learn the definition first, then one example that makes the definition concrete.
  • Connect Linguistic distance to a quantity you can measure, compute or draw — that is where exam questions come from.
  • Reproduce the core statement of Linguistic distance from memory before moving on to harder problems.

Reference excerpt

Linguistic distance is the measure of how different one language (or dialect) is from another. Although they lack a uniform approach to quantifying linguistic distance between languages, linguists apply the concept to a variety of linguistic contexts, such as second-language acquisition, historical linguistics, language-based conflicts, and the effects of language differences on trade.

Measures

Lexicostatistics The proposed measures used for linguistic distance reflect varying understandings of the term itself. One approach is based on mutual intelligibility, i.e. the ability of speakers of one language to understand the other language. With this, the higher the linguistic distance, the lower is the level of mutual intelligibility. Because cognate words play an important role in mutual intelligibility between languages, these figure prominently in such analyses. The higher the percentage of cognate (as opposed to non-cognate) words in the two languages with respect to one another, the lower is their linguistic distance. Also, the greater the degree of grammatical relatedness (i.e. the cognates mean roughly similar things) and lexical relatedness (i.e. the cognates are easily discernible as related words), the lower is the linguistic distance. As an example of this, the Hindustani word pānch is grammatically identical and lexically similar (but non-identical) to its cognate Punjabi and Persian word panj as well as to the lexically dissimilar but still grammatically identical Greek pent- and English five. As another example, the English dish and German Tisch ('table') are lexically (phonologically) similar but grammatically (semantically) dissimilar. Cognates in related languages can even be identical in form, but semantically distinct, such as caldo and largo, which mean respectively 'hot' and 'wide' in Italian but 'broth, soup' and 'long' in Spanish. Using a statistical approach (called lexicostatistics) by comparing each language's mass of words, distances can be calculated between them; in technical terms, what is calculated is the Levenshtein distance. Based on this, one study compared both Afrikaans and West Frisian with Dutch to see which was closer to Dutch. It determined that Dutch and Afrikaans (mutual distance of 20.9%) were considerably closer than Dutch and West Frisian (mutual distance of 34.2%). However, lexicostatistical methods, which are based on retentions from a common proto-language – and not innovations – are problematic due to a number of reasons, so some linguists argue they cannot be relied upon during the tracing of a phylogenetic tree (for example, highest retention rates can sometimes be found in the opposite, peripheral ends of a language family). Unusual innovativeness or conservativeness of a language can distort linguistic distance and the assumed separation date, examples being Romani language and East Baltic languages respectively. On the one hand, continued adjacency of closely related languages after their separation can make some loanwords 'invisible' (indistinguishable from cognates), therefore, from lexicostatistical point of view these languages appear less distant then they actually are (examples being Finnic and Saami languages). On the other hand, strong foreign influence of languages spreading far from their homeland can make them share fewer inherited words than they ought to (examples being Hungarian and Samoyedic languages in the East Uralic branch).

Other internal aspects Besides cognates, other aspects that are often measured are similarities of syntax and written forms. To overcome the aforementioned problems of the lexicostatistical methods, Donald Ringe, Tandy Warnow and Luay Nakhleh developed a complex phylogenetical method relying on phonological and morphological innovations in the 2000s.

Language learning A 2005 paper by economists Barry Chiswick and Paul Miller attempted to put forth a metric for linguistic distances that was based on empirical observations of how rapidly speakers of a given language gained proficiency in another one when immersed in a society that overwhelmingly communicated in the latter language. In this study, the speed of English language acquisition was studied for immigrants of various linguistic backgrounds in the United States and Canada.

See also Abstand and ausbau languages Language transfer Second-language acquisition Historical linguistics

References

Worked examples

Example 1 — a first encounter with Linguistic distance

Start with the simplest possible case. Write down what Linguistic distance claims or describes in one sentence, then invent the smallest concrete situation in which that sentence is true. In science, the smallest case is usually a single object, a single equation or a single measurement. Check that every symbol or term in your sentence has a meaning in that case.

Example 2 — changing one variable

Take the situation from Example 1 and change exactly one quantity: double it, halve it, or set it to zero. Predict what should happen to Linguistic distance before you calculate. Comparing your prediction with the result is the fastest way to find out whether you understand the idea or only the words.

Example 3 — an exam-style question

Typical questions about Linguistic distance ask you to (a) state it precisely, (b) apply it to given data, and (c) explain a limitation. Practise writing all three answers in under five minutes; the third part is what separates a full-mark answer from an average one.

Applications of Linguistic distance

In research
Linguistic distance appears in science research whenever the underlying quantities have to be modelled precisely. Papers usually cite it as a starting assumption and then explore where it breaks down.
In technology and industry
Engineering practice reuses Linguistic distance in design rules, simulations and safety margins. Knowing the idea lets you read a specification sheet and understand why the numbers look the way they do.
In the classroom
Linguistic distance is common in secondary-school and first-year university syllabi. It links to neighbouring topics Applied linguistics, Historical linguistics, Language acquisition, so understanding it makes those chapters shorter.
In everyday life
Look for Linguistic distance outside the textbook — in sport, cooking, traffic, electronics or the sky above you. An example you found yourself is remembered far longer than one you were given.
Ask Teacher Smith questions about this articleOpens your AI tutor with a question about “Linguistic distance” →

Affiliate

Preply — study more efficiently by working with a personal tutor. 50% off.

How to study Linguistic distance in 20 minutes

  1. Read the reference excerpt below once, without taking notes.
  2. Close the page and write down what Linguistic distance means in your own words.
  3. Compare your version with the excerpt and mark what you missed.
  4. Work through the three examples above with pen and paper.
  5. Explain Linguistic distance out loud to somebody else — or to Teacher Smith in the lgStudy chat.

Frequently asked questions

What is Linguistic distance in simple terms?

Linguistic distance is the measure of how different one language (or dialect) is from another. Although they lack a uniform approach to quantifying linguistic distance between languages, linguists apply the concept to a variety of linguistic contexts, such as second-language acquisition, historical…

Why does Linguistic distance matter?

Because it connects several science ideas at once: it gives you a definition you can apply, a quantity you can calculate, and a way to check whether a result is plausible.

How should I study Linguistic distance?

Read the excerpt, restate it from memory, then work through the examples and applications listed on this page. The five-step study plan above takes about twenty minutes.

What does this page cover?

It gives you a compact reference excerpt plus original lgStudy explanations, examples, applications and study material on Linguistic distance.

Tags

  • Applied linguistics
  • Historical linguistics
  • Language acquisition
  • Language comparison
  • Quantitative linguistics

Keep exploring