Language: Swedish / Rights: http://creativecommons.org/licenses/by-nc-sa/4.0/

Start Over Language Swedish Rights http://creativecommons.org/licenses/by-nc-sa/4.0/

Creator:: Gurevych, Iryna, Habernal, Ivan, and Zayed, Omnia
Publisher:: Technische Universität Darmstadt
Type:: text and corpus
Subject:: CommonCrawl, Creative Commons, Web corpus, and Amazon Web Services
Language:: Afrikaans, Arabic, Bengali, Bulgarian, Czech, Danish, German, Modern Greek (1453-), English, Estonian, Persian, Finnish, French, Gujarati, Hebrew, Hindi, Croatian, Hungarian, Indonesian, Italian, Japanese, Korean, Latvian, Lithuanian, Malayalam, Marathi, Macedonian, Nepali (macrolanguage), Dutch, Norwegian, Polish, Portuguese, Romanian, Russian, Slovak, Slovenian, Somali, Spanish, Albanian, Swahili (macrolanguage), Swedish, Tamil, Telugu, Tagalog, Thai, Turkish, Ukrainian, Undetermined, Urdu, Vietnamese, and Chinese
Description:: A large web corpus (over 10 billion tokens) licensed under CreativeCommons license family in 50+ languages that has been extracted from CommonCrawl, the largest publicly available general Web crawl to date with about 2 billion crawled URLs.
Rights:: Creative Commons - Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0), http://creativecommons.org/licenses/by-nc-sa/4.0/, and PUB

Creator:: Jedličková, Alice, Ferjenčík, Mikuláš, and Richterová, Olga
Publisher:: Akropolis
Format:: electronic and 173 s.
Type:: model:monograph and TEXT
Subject:: Literatura (teorie), literární teorie, naratologie, deskripce (filozofie), rétorika, poetika, komunikace (sdělování), pragmatika, intermedialita, literary theory, narratology, description (philosophy), rhetoric, poetics, human communication, pragmatics, intermediality, ekfráze, 82.0, 808.543-027.21, 16-028.44, 808.5, 316.77, 81'2/'44, 7.08:316.776.33/.34, (062.534), and 11
Language:: Czech, Slovak, English, German, and Swedish
Description:: Alice Jedličková (ed.) ; překlady Mikuláš Ferjenčík a Olga Richterová., Obsahuje bibliografie a bibliografické odkazy, and Část. slovenský text, anglické resumé
Rights:: http://creativecommons.org/licenses/by-nc-sa/4.0/ and policy:public

Creator:: Stymne, Sara and Östman, Carin
Publisher:: Uppsala University
Type:: text and corpus
Subject:: literature, literary fiction, dialogue, narrative, and cited materials
Language:: Swedish
Description:: SLäNDa, the Swedish literature corpus of narrative and dialogue, is a corpus made up of eight Swedish literary novels from the late 19th and early 20th centuries, manually annotated mainly for different aspects of dialogue. The full annotation also contains other cited materials, like thoughts, signs and letters. The main motivation for including these categories as well, is to be able to identify the main narrative, which is all remaining unannotated text.
Rights:: Creative Commons - Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0), http://creativecommons.org/licenses/by-nc-sa/4.0/, and PUB

Creator:: Stymne, Sara and Östman, Carin
Publisher:: Uppsala University
Type:: text and corpus
Subject:: literature, literary fiction, dialogue, narrative, and cited materials
Language:: Swedish
Description:: SLäNDa, the Swedish literature corpus of narrative and dialogue, is a corpus made up of eight Swedish literary novels from the 19th and early 20th centuries, manually annotated mainly for different aspects of dialogue. The full annotation also contains other cited materials, like thoughts, signs and letters. The main motivation for including these categories as well, is to be able to identify the main narrative, which is all remaining unannotated text. SLäNDa version 2.0 extends version 1.0 mainly by adding more data, but also by additional quality control, and a slight modification of the annotation scheme. In addition, the data is organized into test sets with different types of speech marking: quotation marks, dashes, and no marking.
Rights:: Creative Commons - Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0), http://creativecommons.org/licenses/by-nc-sa/4.0/, and PUB

Limit your search