ArticleslgStudy

science

Russian National Corpus

Russian National Corpus is a science topic covered in the lgStudy science library. This page brings together a partial reference excerpt, illustrations, worked examples, real-world applications and a short study plan, so you can understand Russian National Corpus rather than just read about it. In short: The Russian National Corpus (Russian: Национальный корпус русского языка, lit. 'National Corpus of the Russian Language') is a corpus of the Russian language that has been partially accessible through a query interface online since April 29, 2004. It is being created by the Institute of Russian language, Russian Academy of Sciences.

Key takeaways

  • Russian National Corpus belongs to science; place it in that map before memorising details.
  • Learn the definition first, then one example that makes the definition concrete.
  • Connect Russian National Corpus to a quantity you can measure, compute or draw — that is where exam questions come from.
  • Reproduce the core statement of Russian National Corpus from memory before moving on to harder problems.

Reference excerpt

The Russian National Corpus (Russian: Национальный корпус русского языка, lit. 'National Corpus of the Russian Language') is a corpus of the Russian language that has been partially accessible through a query interface online since April 29, 2004. It is being created by the Institute of Russian language, Russian Academy of Sciences. It currently contains more than 1 billion word forms that are automatically lemmatized and POS-/grammeme-tagged, i.e. all the possible morphological analyses for each orthographic form are ascribed to it. Lemmata, POS, grammatical items, and their combinations are searchable. Additionally, 6 million word forms are in the subcorpus with manually resolved homonymy. The subcorpus with resolved morphological homonymy is also automatically accentuated. The whole corpus has a searchable tagging concerning lexical semantics (LS), including morphosemantic POS subclasses (proper noun, reflexive pronoun etc.), LS characteristics proper (thematic class, causativity, evaluation), derivation (diminutive, adverb formed from adjective etc.). The RNC includes also the following subcorpora:

a treebank of syntactical dependencies (largely based on the Igor Mel'čuk's Meaning-Text Theory) English⇔Russian, German⇒Russian, Ukrainian⇔Russian and Belorussian⇔Russian parallel corpora; a large (100+ million words) separate corpus of modern newspapers (2001–2011); a corpus of Russian poetry, where the rhyming words and poetic prosody (including meter, stanzas etc.) is additionally tagged; a corpus of Russian dialects with specific dialect grammar tagging; a multimedia corpus with searchable tagged fragments of Russian-language movies; a corpus showing the history of Russian stress an educational subcorpus reflecting school standards. All the texts have tags bearing metatextual information - the author, his/her birth date, creation date, text size, text genres (general fiction, detective story, newspaper article etc.); all these categories are browsable and searchable separately. It is possible to define a user's subcorpus to search lemmata/POS-grammeme/semantic tags combinations only within this subset.

See also General Internet Corpus of Russian

References

External links Russian National corpus

Worked examples

Example 1 — a first encounter with Russian National Corpus

Start with the simplest possible case. Write down what Russian National Corpus claims or describes in one sentence, then invent the smallest concrete situation in which that sentence is true. In science, the smallest case is usually a single object, a single equation or a single measurement. Check that every symbol or term in your sentence has a meaning in that case.

Example 2 — changing one variable

Take the situation from Example 1 and change exactly one quantity: double it, halve it, or set it to zero. Predict what should happen to Russian National Corpus before you calculate. Comparing your prediction with the result is the fastest way to find out whether you understand the idea or only the words.

Example 3 — an exam-style question

Typical questions about Russian National Corpus ask you to (a) state it precisely, (b) apply it to given data, and (c) explain a limitation. Practise writing all three answers in under five minutes; the third part is what separates a full-mark answer from an average one.

Applications of Russian National Corpus

In research
Russian National Corpus appears in science research whenever the underlying quantities have to be modelled precisely. Papers usually cite it as a starting assumption and then explore where it breaks down.
In technology and industry
Engineering practice reuses Russian National Corpus in design rules, simulations and safety margins. Knowing the idea lets you read a specification sheet and understand why the numbers look the way they do.
In the classroom
Russian National Corpus is common in secondary-school and first-year university syllabi. It links to neighbouring topics Applied linguistics, Corpora, Corpus linguistics stubs, so understanding it makes those chapters shorter.
In everyday life
Look for Russian National Corpus outside the textbook — in sport, cooking, traffic, electronics or the sky above you. An example you found yourself is remembered far longer than one you were given.

Affiliate

Preply — study more efficiently by working with a personal tutor. 50% off.

How to study Russian National Corpus in 20 minutes

  1. Read the reference excerpt below once, without taking notes.
  2. Close the page and write down what Russian National Corpus means in your own words.
  3. Compare your version with the excerpt and mark what you missed.
  4. Work through the three examples above with pen and paper.
  5. Explain Russian National Corpus out loud to somebody else — or to Teacher Smith in the lgStudy chat.

Frequently asked questions

What is Russian National Corpus in simple terms?

The Russian National Corpus (Russian: Национальный корпус русского языка, lit. 'National Corpus of the Russian Language') is a corpus of the Russian language that has been partially accessible through a query interface online since April 29, 2004. It is being created by the Institute of Russian lan…

Why does Russian National Corpus matter?

Because it connects several science ideas at once: it gives you a definition you can apply, a quantity you can calculate, and a way to check whether a result is plausible.

How should I study Russian National Corpus?

Read the excerpt, restate it from memory, then work through the examples and applications listed on this page. The five-step study plan above takes about twenty minutes.

What does this page cover?

It gives you a compact reference excerpt plus original lgStudy explanations, examples, applications and study material on Russian National Corpus.

Tags

  • Applied linguistics
  • Corpora
  • Corpus linguistics stubs
  • Library and information science stubs
  • Linguistic research
  • Russian language
  • Slavic language stubs

Keep exploring