29.07.2026

We continue to improve the functionality of the corpora within the Russian National Corpus.

In the Parallel Corpora, search results can now be sorted by context in KWIC output mode, and the Statistics output view is now available. Users can now analyze the frequency of search results by the text’s sphere of use, the language of the original, and the language variety or country.

In the Syntactic Corpus, it is now possible to select texts whose sentences have semantic structure annotation.

29.07.2026

We continue to expand the corpus functionality for teaching Russian at school. The Practice Example Generator has been updated with rules for spelling vowels in prefixes. The prefixes are divided into three groups: prefixes with unstressed vowels, such as взо-, во-, до-, за-, пере-, по-, со-, and others; homophonous and/or variable prefixes, such as пра-/про-, пре-/при-, раз-/рас- и роз-/рос-; and borrowed prefixes, such as а-, анти-, архи-, интро-, ультра-, and others. Altogether, the new rules cover more than 30 groups of words with prefixes.

You can access the generator page from the RNC for Schools section by clicking on the corresponding banner.

29.07.2026

The Main Corpus of the Russian National Corpus has grown by 36,7 million words and has reached a total size of 424 million tokens. The update includes works of 21st-century Russian prose in a wide range of literary trends and genres (novels, novellas, short stories, and essays), as well as new batches of journal publications featuring both fiction and “thick-journal” nonfiction: reviews, articles, memoirs, and other genres. Around 100 new works of children’s fiction have also been added.

At the same time, the educational and scholarly segment has been expanded: the corpus now includes additional reference and encyclopedic texts, including materials aimed at school and teenage audiences.

Drama has been updated separately: the new collection contains nearly 200 plays spanning a broad historical range and diverse theatrical forms, from the 18th century to the present day. In addition, more than 80 plays already included in the corpus have been annotated to distinguish speakers’ lines and stage directions.

The update also fills gaps related to a number of authors and texts of the 19th–20th centuries (fiction, memoirs, philosophy, and early translations).

Show all