Number of results to display per page
Search Results
462. Plaintext Wikipedia dump 2018
- Creator:
- Rosa, Rudolf
- Publisher:
- Charles University, Faculty of Mathematics and Physics, Institute of Formal and Applied Linguistics (UFAL)
- Type:
- text and corpus
- Subject:
- Wikipedia, text corpora, and monolingual corpus
- Language:
- Abkhazian, Achinese, Adyghe, Afrikaans, Akan, Tosk Albanian, Amharic, Old English (ca. 450-1100), Arabic, Official Aramaic (700-300 BCE), Aragonese, Egyptian Arabic, Assamese, Asturian, Atikamekw, Avaric, Aymara, South Azerbaijani, Azerbaijani, Bashkir, Bambara, Bavarian, Central Bikol, Belarusian, Bengali, Bislama, Banjar, Tibetan, Bosnian, Bishnupriya, Breton, Buginese, Bulgarian, Russia Buriat, Catalan, Min Dong Chinese, Cebuano, Czech, Chamorro, Chechen, Cherokee, Church Slavic, Chuvash, Cheyenne, Central Kurdish, Cornish, Corsican, Cree, Crimean Tatar, Kashubian, Welsh, Danish, German, Dinka, Dimli (individual language), Dhivehi, Lower Sorbian, Dzongkha, Modern Greek (1453-), English, Esperanto, Estonian, Basque, Ewe, Extremaduran, Faroese, Persian, Fijian, Finnish, French, Arpitan, Northern Frisian, Western Frisian, Fulah, Friulian, Gagauz, Gan Chinese, Scottish Gaelic, Irish, Galician, Gilaki, Manx, Goan Konkani, Gothic, Guarani, Gujarati, Hakka Chinese, Haitian, Hausa, Hawaiian, Serbo-Croatian, Hebrew, Herero, Fiji Hindi, Hindi, Hiri Motu, Croatian, Upper Sorbian, Hungarian, Armenian, Igbo, Ido, Inuktitut, Interlingue, Iloko, Interlingua (International Auxiliary Language Association), Indonesian, Inupiaq, Icelandic, Italian, Jamaican Creole English, Javanese, Lojban, Japanese, Kara-Kalpak, Kabyle, Kalaallisut, Kannada, Kashmiri, Georgian, Kanuri, Kazakh, Kabardian, Kabiyè, Khmer, Kikuyu, Kinyarwanda, Kirghiz, Komi-Permyak, Komi, Kongo, Korean, Karachay-Balkar, Kölsch, Kurdish, Ladino, Lao, Latin, Latvian, Lak, Lezghian, Ligurian, Limburgan, Lingala, Lithuanian, Lombard, Northern Luri, Latgalian, Luxembourgish, Ganda, Literary Chinese, Marshallese, Maithili, Malayalam, Marathi, Moksha, Eastern Mari, Minangkabau, Macedonian, Malagasy, Maltese, Mongolian, Maori, Western Mari, Malay (macrolanguage), Creek, Mirandese, Burmese, Erzya, Mazanderani, Min Nan Chinese, Neapolitan, Nauru, Navajo, Ndonga, Low German, Nepali (macrolanguage), Newari, Dutch, Norwegian Nynorsk, Norwegian, Novial, Pedi, Nyanja, Occitan (post 1500), Livvi, Oriya (macrolanguage), Oromo, Ossetian, Pangasinan, Pampanga, Panjabi, Papiamento, Picard, Pennsylvania German, Pfaelzisch, Pitcairn-Norfolk, Pali, Piemontese, Western Panjabi, Pontic, Polish, Portuguese, Pushto, Quechua, Vlax Romani, Romansh, Romanian, Rusyn, Rundi, Macedo-Romanian, Russian, Sango, Yakut, Sanskrit, Sicilian, Scots, Samogitian, Sinhala, Slovak, Slovenian, Northern Sami, Samoan, Shona, Sindhi, Somali, Southern Sotho, Spanish, Albanian, Sardinian, Sranan Tongo, Serbian, Swati, Saterfriesisch, Sundanese, Swahili (macrolanguage), Swedish, Silesian, Tahitian, Tamil, Tatar, Tulu, Telugu, Tama (Colombia), Tetum, Tajik, Tagalog, Thai, Tigrinya, Tonga (Tonga Islands), Tok Pisin, Tswana, Tsonga, Turkmen, Tumbuka, Turkish, Twi, Tuvinian, Udmurt, Uighur, Ukrainian, Urdu, Uzbek, Venetian, Venda, Veps, Vietnamese, Vlaams, Volapük, Võro, Waray (Philippines), Walloon, Wolof, Wu Chinese, Kalmyk, Xhosa, Mingrelian, Yiddish, Yoruba, Yue Chinese, Zeeuws, Zhuang, Chinese, Zulu, and Dotyali
- Description:
- Wikipedia plain text data obtained from Wikipedia dumps with WikiExtractor in February 2018. The data come from all Wikipedias for which dumps could be downloaded at [https://dumps.wikimedia.org/]. This amounts to 297 Wikipedias, usually corresponding to individual languages and identified by their ISO codes. Several special Wikipedias are included, most notably "simple" (Simple English Wikipedia) and "incubator" (tiny hatching Wikipedias in various languages). For a list of all the Wikipedias, see [https://meta.wikimedia.org/wiki/List_of_Wikipedias]. The script which can be used to get new version of the data is included, but note that Wikipedia limits the download speed for downloading a lot of the dumps, so it takes a few days to download all of them (but one or a few can be downloaded fast). Also, the format of the dumps changes time to time, so the script will probably eventually stop working one day. The WikiExtractor tool [http://medialab.di.unipi.it/wiki/Wikipedia_Extractor] used to extract text from the Wikipedia dumps is not mine, I only modified it slightly to produce plaintext outputs [https://github.com/ptakopysk/wikiextractor].
- Rights:
- Attribution-ShareAlike 3.0 Unported (CC BY-SA 3.0), http://creativecommons.org/licenses/by-sa/3.0/, and PUB
463. Počátek dějin
- Creator:
- Jan Patočka
- Publisher:
- Str. 46–88. Stať.
- Type:
- Text
- Subject:
- 1975, 1979/25, 1981/6, 1981/7, 1988/28, 1988/31, 1988/32, 1988/34, 1994/7, 1996/4, 1996/7, 1998/3, 1999/8, 2, 2001/9, 2002/21, 2002/6, 2006/1, 2007/1, 2008/3, bg, cs, de, en, es, fr, fulltext, hu, it, lt, no, pl, ru, SS-3/PD-III, sv, uk, and v
- Language:
- Czech, English, Bulgarian, French, Italian, Lithuanian, Hungarian, German, Norwegian, Polish, Russian, Spanish, Swedish, and Ukrainian
- Rights:
- open access and Rights holder: Archiv Jana Patočky, z.s.
464. Podjebrad György /
- Creator:
- Šmahel, František,
- Type:
- text and monografie
- Subject:
- Dějiny Česka a Slovenska, Jiří z Poděbrad,, panovníci čeští, české země 1419-1471, and politické dějiny, politici
- Language:
- Hungarian
- Rights:
- unknown
465. Pokus o českou národní filosofii a jeho nezdar
- Creator:
- Jan Patočka
- Publisher:
- Str. 3–57. Stať. [Psáno počátkem sedmdestých let.]
- Type:
- Text
- Subject:
- 1977, 1979/1, 1985/11, 1991/1, 1992/2, 1996/7, 2006/11, AS/M, cs, de, en, fr, hu, SS-12/Češi-I, and v/??
- Language:
- Czech, English, French, Hungarian, and German
- Rights:
- open access and Rights holder: Archiv Jana Patočky, z.s.
466. Pre-historické úvahy
- Creator:
- Jan Patočka
- Publisher:
- Str. 1–45. Stať.
- Type:
- Text
- Subject:
- 1975, 1979/25, 1981/6, 1981/7, 1988/28, 1988/31, 1988/32, 1988/34, 1994/7, 1996/4, 1996/7, 1998/3, 1999/8, 2001/9, 2002/1, 2002/21, 2002/5, 2006/1, 2007/1, 2008/3, bg, cs, de, en, es, fr, fulltext, hu, it, lt, no, pl, ru, SS-3/PD-III, sv, and uk
- Language:
- Czech, English, Bulgarian, French, Italian, Lithuanian, Hungarian, German, Norwegian, Polish, Russian, Spanish, Swedish, and Ukrainian
- Rights:
- open access and Rights holder: Archiv Jana Patočky, z.s.
467. Přirozený svět a fenomenologie
- Creator:
- Jan Patočka
- Publisher:
- 1.3, 48 s. Stať. [Předloha pro slovenský překlad (v. 1967/2). Začátky textů se mírně liší.] — 2. otisk in: Fenomenologické spisy II (SS-7/Fen-II), Praha 2009, str. 202–237 (v. 2009/1).
- Type:
- Text
- Subject:
- 1967/2, 1969/8, 1970/10, 1972/1, 1972/2, 1976/7, 1980, 1988/29, 1989/16, 1991/2, 1996/7, 2003/23, 2004/10, 2009/1, cs, de, en, es, fr, hu, it, sk, SS-7/Fen-II, and stať
- Language:
- English, French, Italian, Hungarian, German, Slovak, Spanish, and Czech
- Rights:
- open access and Rights holder: Archiv Jana Patočky, z.s.
468. Přirozený svět v meditaci svého autora po třiatřiceti letech
- Creator:
- Jan Patočka
- Publisher:
- Str. 155–234. Dosl.
- Type:
- Text
- Subject:
- 1969/8, 1970, 1980/1, 1988/29, 1990/3, 1992/10, 2009/1, AS/PS-2, cs, de, fr, hu, and SS-7/Fen–II
- Language:
- Czech, French, Hungarian, and German
- Rights:
- open access and Rights holder: Archiv Jana Patočky, z.s.
469. Radvaň n. Dunajom
- Publisher:
- Vojenský zeměpisný ústav
- Format:
- map and 1 mapa : barevná ; 39 x 52 cm na listu 48 x 62 cm
- Type:
- model:map, cartographic, and IMAGE
- Subject:
- udc:913(4), Konspekt:7, udc:912, udc:913(437.6), udc:912.43, udc:(084.3), Konspekt:Geografie Evropy, reálie, cestování, Konspekt:Mapy. Atlasy. Glóby, and czenas:Radvaň nad Dunajom (Slovensko : oblast)
- Language:
- Czech and Hungarian
- Description:
- 4961, Edice dle kladu listů, and (Language) Místní názvy maďarsky
- Rights:
- http://creativecommons.org/publicdomain/mark/1.0/ and policy:public
470. Rákóczi Ferenc felségárulási perének története 1701 /
- Creator:
- Lukinich, Imre,
- Type:
- text and monografie
- Subject:
- Dějiny zemí střední Evropy, František, procesy soudní, povstání stavovská, povstání protihabsburská, panovníci sedmihradští, Maďarsko, politické dějiny, politici, and světové dějiny 1648-1789
- Language:
- Hungarian
- Rights:
- unknown