Skip to search
Skip to main content
Skip to first result
Search
Search Results
Creator:
Frijhoff, Willem
Type:
text , články , and edice
Subject:
Rukopisy, prvotisky, staré tisky. Vzácná a pozoruhodná díla , Komenský, Jan Amos, , Heereboort, Adriaen, , Court, Pieter de la, , komeniana , vztahy česko-nizozemské , filozofové nizozemští , korespondence osobní , edice korespondence , komeniologie , české země 1620-1740 , Nizozemí , světové dějiny 1492-1648 , dějiny literatury, jazyka a knihy , filozofie, filozofové , and zahraniční politika, mezinárodní vztahy
Language:
English and Latin
Rights:
unknown
Publisher:
Tiskarna Ceskych bratri
Type:
model:monograph and TEXT
Language:
Czech and Latin
Rights:
http://creativecommons.org/publicdomain/mark/1.0/ and policy:public
Creator:
Rosa, Rudolf
Publisher:
Charles University, Faculty of Mathematics and Physics, Institute of Formal and Applied Linguistics (UFAL)
Type:
text and corpus
Subject:
Wikipedia , text corpora , and monolingual corpus
Language:
Abkhazian , Achinese , Adyghe , Afrikaans , Akan , Tosk Albanian , Amharic , Old English (ca. 450-1100) , Arabic , Official Aramaic (700-300 BCE) , Aragonese , Egyptian Arabic , Assamese , Asturian , Atikamekw , Avaric , Aymara , South Azerbaijani , Azerbaijani , Bashkir , Bambara , Bavarian , Central Bikol , Belarusian , Bengali , Bislama , Banjar , Tibetan , Bosnian , Bishnupriya , Breton , Buginese , Bulgarian , Russia Buriat , Catalan , Min Dong Chinese , Cebuano , Czech , Chamorro , Chechen , Cherokee , Church Slavic , Chuvash , Cheyenne , Central Kurdish , Cornish , Corsican , Cree , Crimean Tatar , Kashubian , Welsh , Danish , German , Dinka , Dimli (individual language) , Dhivehi , Lower Sorbian , Dzongkha , Modern Greek (1453-) , English , Esperanto , Estonian , Basque , Ewe , Extremaduran , Faroese , Persian , Fijian , Finnish , French , Arpitan , Northern Frisian , Western Frisian , Fulah , Friulian , Gagauz , Gan Chinese , Scottish Gaelic , Irish , Galician , Gilaki , Manx , Goan Konkani , Gothic , Guarani , Gujarati , Hakka Chinese , Haitian , Hausa , Hawaiian , Serbo-Croatian , Hebrew , Herero , Fiji Hindi , Hindi , Hiri Motu , Croatian , Upper Sorbian , Hungarian , Armenian , Igbo , Ido , Inuktitut , Interlingue , Iloko , Interlingua (International Auxiliary Language Association) , Indonesian , Inupiaq , Icelandic , Italian , Jamaican Creole English , Javanese , Lojban , Japanese , Kara-Kalpak , Kabyle , Kalaallisut , Kannada , Kashmiri , Georgian , Kanuri , Kazakh , Kabardian , Kabiyè , Khmer , Kikuyu , Kinyarwanda , Kirghiz , Komi-Permyak , Komi , Kongo , Korean , Karachay-Balkar , Kölsch , Kurdish , Ladino , Lao , Latin , Latvian , Lak , Lezghian , Ligurian , Limburgan , Lingala , Lithuanian , Lombard , Northern Luri , Latgalian , Luxembourgish , Ganda , Literary Chinese , Marshallese , Maithili , Malayalam , Marathi , Moksha , Eastern Mari , Minangkabau , Macedonian , Malagasy , Maltese , Mongolian , Maori , Western Mari , Malay (macrolanguage) , Creek , Mirandese , Burmese , Erzya , Mazanderani , Min Nan Chinese , Neapolitan , Nauru , Navajo , Ndonga , Low German , Nepali (macrolanguage) , Newari , Dutch , Norwegian Nynorsk , Norwegian , Novial , Pedi , Nyanja , Occitan (post 1500) , Livvi , Oriya (macrolanguage) , Oromo , Ossetian , Pangasinan , Pampanga , Panjabi , Papiamento , Picard , Pennsylvania German , Pfaelzisch , Pitcairn-Norfolk , Pali , Piemontese , Western Panjabi , Pontic , Polish , Portuguese , Pushto , Quechua , Vlax Romani , Romansh , Romanian , Rusyn , Rundi , Macedo-Romanian , Russian , Sango , Yakut , Sanskrit , Sicilian , Scots , Samogitian , Sinhala , Slovak , Slovenian , Northern Sami , Samoan , Shona , Sindhi , Somali , Southern Sotho , Spanish , Albanian , Sardinian , Sranan Tongo , Serbian , Swati , Saterfriesisch , Sundanese , Swahili (macrolanguage) , Swedish , Silesian , Tahitian , Tamil , Tatar , Tulu , Telugu , Tama (Colombia) , Tetum , Tajik , Tagalog , Thai , Tigrinya , Tonga (Tonga Islands) , Tok Pisin , Tswana , Tsonga , Turkmen , Tumbuka , Turkish , Twi , Tuvinian , Udmurt , Uighur , Ukrainian , Urdu , Uzbek , Venetian , Venda , Veps , Vietnamese , Vlaams , Volapük , Võro , Waray (Philippines) , Walloon , Wolof , Wu Chinese , Kalmyk , Xhosa , Mingrelian , Yiddish , Yoruba , Yue Chinese , Zeeuws , Zhuang , Chinese , Zulu , and Dotyali
Description:
Wikipedia plain text data obtained from Wikipedia dumps with WikiExtractor in February 2018.
The data come from all Wikipedias for which dumps could be downloaded at [https://dumps.wikimedia.org/]. This amounts to 297 Wikipedias, usually corresponding to individual languages and identified by their ISO codes. Several special Wikipedias are included, most notably "simple" (Simple English Wikipedia) and "incubator" (tiny hatching Wikipedias in various languages).
For a list of all the Wikipedias, see [https://meta.wikimedia.org/wiki/List_of_Wikipedias].
The script which can be used to get new version of the data is included, but note that Wikipedia limits the download speed for downloading a lot of the dumps, so it takes a few days to download all of them (but one or a few can be downloaded fast).
Also, the format of the dumps changes time to time, so the script will probably eventually stop working one day.
The WikiExtractor tool [http://medialab.di.unipi.it/wiki/Wikipedia_Extractor] used to extract text from the Wikipedia dumps is not mine, I only modified it slightly to produce plaintext outputs [https://github.com/ptakopysk/wikiextractor].
Rights:
Attribution-ShareAlike 3.0 Unported (CC BY-SA 3.0) , http://creativecommons.org/licenses/by-sa/3.0/ , and PUB
Creator:
Fradelius Štiavnický, Petr and Pardubský, Matěj
Publisher:
Pardubský, Matěj
Format:
print and [4] ff ; 4°
Type:
model:monograph and TEXT
Subject:
Rosaciová, Ludmila , století 17. , poezie , natalitia , and 094
Language:
Latin
Description:
Rukověť II., s.153. and BCBT39131
Rights:
http://creativecommons.org/licenses/by-nc-sa/4.0/ and policy:public
Creator:
Včelín, Jakub z Lumenštejna and Jan Stříbrský
Publisher:
Stříbrský, Jan
Format:
print and [12] ff ; 4°
Type:
model:monograph and TEXT
Subject:
století 17. , poezie , and 094
Language:
Latin
Description:
BCBT42214
Rights:
http://creativecommons.org/publicdomain/mark/1.0/ and policy:public
Creator:
Rokyta, Jan
Type:
text , studie , prameny , and edice
Subject:
Náboženství , Dějiny Česka a Slovenska , Michal Pražský , Jan z Příbramě, , Jakoubek, , Jan, , Biskupec z Pelhřimova, Mikuláš, , Chelčický, Petr, , husitství , myšlení teologické , válka , násilí , literatura náboženská , chiliasmus , pacifismus , české země 1419-1471 , and teologie, ikonografie, zbožnost, hagiografie
Language:
Czech and Latin
Description:
Traktát Jana z Příbramě De bello z roku 1420 s. 111-113 and The Perspectives on war, truce and defend of God's truth by sword.
Rights:
unknown
Type:
text and tisky pamětní
Subject:
Liturgie. Křesťanské umění a symbolika. Duchovní život , Vlk, Miloslav, , arcibiskupové pražští , kardinálové , pohřby , Československo 1918-1992 , and jednotlivci (církevní dějiny)
Language:
Czech , Italian , and Latin
Rights:
unknown
Creator:
Bouttats, Gaspar
Publisher:
s.n.
Format:
1 mapa : 34 x 43 cm and kartografický dokument
Type:
model:map and IMAGE
Language:
Latin
Rights:
http://creativecommons.org/publicdomain/mark/1.0/ and policy:public
Creator:
Lotter, Tobias Conrad
Publisher:
Tob. Conradi Lotter
Format:
1 mapa : 50 x 57 cm and kartografický dokument
Type:
model:map and IMAGE
Language:
Latin
Description:
Mapa ručně kolorována. Obsahuje parergon
Rights:
http://creativecommons.org/publicdomain/mark/1.0/ and policy:public
Creator:
Seutter, Matthäus
Publisher:
Matth. Seutteri
Format:
1 mapa : 50 x 59 cm and kartografický dokument
Type:
model:map and IMAGE
Language:
Latin
Description:
Mapa ručně kolorována
Rights:
http://creativecommons.org/publicdomain/mark/1.0/ and policy:public