Harvested from: LINDAT/CLARIAH-CZ repository / Language: Russian - LINDAT/CLARIAH-CZ Catalog Search Results

Start Over Language Russian Harvested from LINDAT/CLARIAH-CZ repository

51. Prague Czech-English Dependency Treebank 2.0 - Russian translation

Creator:: Novák, Michal, Nedoluzhko, Anna, and Schwarz (Khoroshkina), Anna
Publisher:: Charles University, Faculty of Mathematics and Physics, Institute of Formal and Applied Linguistics (UFAL)
Type:: text and corpus
Subject:: multilingual and coreference
Language:: English, Czech, and Russian
Description:: Prague Czech-English Dependency Treebank - Russian translation (PCEDT-R) is a project of translating a subset of Prague Czech-English Dependency Treebank 2.0 (PCEDT 2.0) to Russian and linguistically annotating the Russian translations with emphasis on coreference and cross-lingual alignment of coreferential expressions. Cross-lingual comparison of coreference means is currently the purpose that drives development of this corpus. The current version 0.5 is a preliminary version, which contains (+ denotes new features): * complete PCEDT 2.0 documents "wsj_1900"-"wsj_1949" * Czech-English word alignment of coreferential expressions annotated manually mainly on the t-layer + Russian translations of the original English sentences + automatic tokenization, part-of-speech tagging and morphological analysis for Russian + automatic word alignment between all Czech and Russian words + manual alignment between Russian and the other two languages on possessive pronouns
Rights:: CC-BY-NC-SA + LDC99T42, https://lindat.mff.cuni.cz/repository/xmlui/page/license-pcedt2, and RES

52. Progress test on Russian language KARTTU

Publisher:: The Department of Modern Languages, University of Helsinki and University of Helsinki
Format:: application/octet-stream
Type:: toolService
Language:: Russian
Description:: Progress test on language competence in Russian
Rights:: Not specified

53. Project Gutenberg

Type:: corpus
Language:: Danish, Dutch, English, Finnish, French, German, Italian, Latin, Portuguese, Russian, Spanish, Swedish, and Telugu
Description:: Possibility to download or to browse free electronic books; Angebot: Download von und Online-Zugang zu frei verfügbaren E-Books; deutschsprachige Literatur stellt nur einen Teilbereich der verfügbaren E-Books dar
Rights:: Not specified

54. Run (Russian meets Norwegian )

Publisher:: Department of Literature, Area Studies and European Languages, University of Oslo and Department of Linguistics and Nordic Studies, University of Oslo
Type:: corpus
Language:: English, Norwegian, and Russian
Description:: The RuN corpus is a parallel corpus consisting of Norwegian, Russian and English texts. The texts are aligned at the sentence level and have been tagged for grammatical information at the word level.
Rights:: Not specified

55. SpeechDat-East databases

Type:: corpus
Subject:: These databases serve as an important resource for the performance of voice driven teleservice systems in practical implementations
Language:: Czech, Hungarian, Polish, Russian, and Slovak
Description:: 5 telephone databases recorded over the PSTN. Contains interesting phonetically rich material. All orthographically transcribed. Speaker information included for gender, age, accent. Including pronunciation lexicon.
Rights:: Not specified

56. Speecon databases

Type:: corpus
Language:: Czech, Danish, Dutch, English, Finnish, French, German, Hungarian, Italian, Polish, Portuguese, Russian, Spanish, Swedish, Turkish, Chinese, Hebrew, Japanese, Korean, and Thai
Description:: 28 speech databases containing broadband recordings from 550 adults and 50 children per language. Contains interesting phonetically rich material. All orthographically transcribed. Speaker information included for gender, age, accent. Including pronunciation lexicon.
Rights:: Not specified

57. SynTagRus gapping test set

Creator:: Droganova, Kira, Ponomareva, Maria, Smurov, Ivan, and Shavrina, Tatiana
Publisher:: Charles University, Faculty of Mathematics and Physics, Institute of Formal and Applied Linguistics (UFAL)
Type:: text and corpus
Subject:: linguistic data, gapping, and ellipsis
Language:: Russian
Description:: A test set that contains manually annotated sentences with gapping. The test set was compiled from SynTagRus (v. 2015) the dependency treebank for Russian that provides comprehensive manually-corrected morphological and syntactic annotation.
Rights:: Creative Commons - Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0), http://creativecommons.org/licenses/by-nc-sa/4.0/, and PUB

58. The National Certificates corpus

Publisher:: Centre for Applied Language Studies, University of Jyväskylä
Type:: corpus
Language:: English, Finnish, French, German, Italian, Russian, Spanish, and Swedish
Description:: The NC test results, background information, speaking and writing performances in 9 foreign / second languages. A web-based data base (html files).
Rights:: Not specified

59. The Use of Machine Translation by Ukrainian War Refugees in Czechia

Creator:: Agapova, Anna and Špačková, Stanislava
Publisher:: Oxford University Press
Type:: TEXT and Spreadsheet
Subject:: machine translation, migration, Ukrainian refugees, Russo-Ukrainian war, and Czech Republic
Language:: Ukrainian and Russian
Description:: Data from a questionnaire survey conducted from 2022-08-25 to 2022-11-15 and exploring the use of machine translation by Ukrainian refugees in the Czech Republic. The presented spreadsheet contains minimally processed data exported from the two questionnaires that were created in Google Forms in the Ukrainian and the Russian language. The links to these questionnaires were distributed by three methods: direct email to particular refugees whose contact details the authors obtained while volunteering; through a non-profit organisation helping refugees (Vesna women’s education institution) and on social networks by posting links to the survey in groups associating the Ukrainian community across Czech regions and towns. Since we asked potential respondents to spread the questionnaire further, we could not prevent it from reaching Ukrainians who had arrived in Czechia previously, or received temporary protection in other countries. Due to this fact, the textual answers to the question 1.5 "Which country are you in right now?" were replaced in the dataset by numbers (1 for the Czech Republic, 2 for other countries) in order for us to be able to separate the data of respondents not located in the Czech Republic, which were irrelevant for our survey. Also, in this version of the dataset, the textual answers to the question 1.6 "How many months have you been to this country?" were replaced by numbers, so that we could separate the data of respondents who arrived in the Czech Republic in February 2022 or later from the other data (0 for those staying in Czechia before February 2022, 1 for those staying in Czechia since February 2022 or later, 2 for those staying in other countries).
Rights:: Creative Commons - Attribution 4.0 International (CC BY 4.0), http://creativecommons.org/licenses/by/4.0/, and PUB

60. TITUS Old Russian

Format:: text/html
Type:: corpus
Language:: Russian
Description:: ca. 200.000 tokens; linked with relational database; XML-encoding in progress
Rights:: http://titus.uni-frankfurt.de/texte/texte2.htm#Estart

« Previous
Next »
1
2
3
4
5
6
7
8
9
10

51. Prague Czech-English Dependency Treebank 2.0 - Russian translation

52. Progress test on Russian language KARTTU

53. Project Gutenberg

54. Run (Russian meets Norwegian )

55. SpeechDat-East databases

56. Speecon databases

57. SynTagRus gapping test set

58. The National Certificates corpus

59. The Use of Machine Translation by Ukrainian War Refugees in Czechia

60. TITUS Old Russian

Limit your search

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Search

Search Constraints

Search Results

Limit your search

Contributor

Show values starting with

Coverage

Creator

Show values starting with

Format

Language

Show values starting with

Publisher

Show values starting with

Rights

Show values starting with

Subject

Show values starting with

Type

Date

Original context has metadata only

Harvested from