Harvested from: LINDAT/CLARIAH-CZ repository / Original context has metadata only: false

1121. On-line Dictionary of medieval latin in the Czech lands

Creator:: Ctibor, Jan and Nývlt, Pavel
Publisher:: Institute of Philosophy of the Czech Academy of Sciences
Type:: text, lexicon, and lexicalConceptualResource
Subject:: dictionary, latin, Medieval, digital humanities, lexicography, and Medieval Latin
Language:: Latin and Czech
Description:: The Dictionary of Medieval Latin in the Czech Lands registers and explains the vocabulary of Medieval Latin as used in the Czech lands since the beginnings of Latin writing in this area (from about 1000 CE) to 1500 CE, so far covering the letters A-M. For more information about the Dictionary, see the webpage of the Department of Medieval Lexicography of the Institute of Philosophy of Czech Academy of Sciences. The data uploaded present the on-line version of the dictionary (API and XML data), making it possible to put the application into operation at a localhost.
Rights:: Dictionary of Medieval Latin in the Czech Lands - digital version 2.2 License Agreement, https://lindat.mff.cuni.cz/repository/xmlui/page/license-lb, and ACA

1122. On-Request Concert in Lucerna Palace

Creator:: Aktualita
Publisher:: Národní filmový archiv
Type:: video and clip
Subject:: akce Kuratorium pro výchovu mládeže, Kuratorium pro výchovu mládeže akce, akce Koncert podle přání, sál slavnostně vyzdobený, dirigent, harmonikáři dětští, pěvci operní, posluchači koncertu, posluchači tleskající, sbor pěvecký dětský, kříž hákový, orchestr, akce Umění mládeži, Umění mládeži akce, Kuratorium, Places::Praha::Nové Město::palác Lucerna::velký sál, People::Moravec Emanuel (1893-1945), and Český zvukový týdeník Aktualita::1944/16
Language:: Czech
Description:: Segment from Český zvukový týdeník Aktualita (Czech Aktualita Sound Newsreel) issue no. 16A from 1944 was shot at a big on-request concert held in the Great Hall of Lucerna Palace on 4 April as the 3000th performance of the Art for Youth initiative of the Board of Trustees for the Education of Youth. The event was attended by Minister of Education and People´s Enlightenment and Chairman of the Board Emanuel Moravec and General Secretary of the Board František Teuner. The diverse programme included the aria "While a Mother´s Love Means Blessing" from Bedřich Smetana´s The Bartered Bride performed by Marie Budíková and Oldřich Kovář as well as performances by young accordion players and a children´s choir.
Rights:: http://creativecommons.org/licenses/by-nc-nd/4.0/, PUB, and Creative Commons - Attribution-NonCommercial-NoDerivatives 4.0 International (CC BY-NC-ND 4.0)

1123. Ondřej Sekora (painter)

Creator:: Veselý, Bohumil
Publisher:: Národní filmový archiv
Type:: video and clip
Subject:: Galerie osobností, Places::Praha::Nové Město::Školská::pavlač domu, and People::Sekora Ondřej (1899-1967)
Language:: No linguistic content
Description:: Painter Ondřej Sekora on Bohumil Veselý's balcony.
Rights:: http://creativecommons.org/licenses/by-nc-nd/4.0/, PUB, and Creative Commons - Attribution-NonCommercial-NoDerivatives 4.0 International (CC BY-NC-ND 4.0)

1124. One Year Since the Death of President Masaryk

Creator:: Aktualita
Publisher:: Národní filmový archiv
Type:: video and clip
Subject:: cvičení vojenské československé, výročí úmrtí Masaryk Tomáš Garrigue 1., Mnichovská dohoda, Places::Lány::zámecká zahrada, People::Masaryk Tomáš Garrigue (1850-1937), People::Klofáč Václav Jaroslav (1868-1942), and Československý zvukový týdeník Aktualita::1938/10
Language:: Czech
Description:: The segment of Československý zvukový týdeník Aktualita (Czechoslovak Aktualita Sound Newsreel), 1938, issue no. 10 consists of a montage of archive film material created to mark the 88th anniversary of the birth of the late President Tomáš Garrique Masaryk.
Rights:: http://creativecommons.org/licenses/by-nc-nd/4.0/, PUB, and Creative Commons - Attribution-NonCommercial-NoDerivatives 4.0 International (CC BY-NC-ND 4.0)

1125. onion

Creator:: Pomikálek, Jan
Publisher:: Masaryk University, NLP Centre
Type:: toolService and tool
Subject:: deduplication, corpus, text deduplication, n-gram deduplication, and n-gram model
Language:: English
Description:: onion (ONe Instance ONly) is a tool for removing duplicate parts from large collections of texts. The tool has been implemented in Python, licensed under New BSD License and made an open source software (available for download including the source code at http://code.google.com/p/onion/). It is being successfuly used for cleaning large textual corpora at Natural language processing centre at Faculty of informatics, Masaryk university Brno and it's industry partners. The research leading to this piece of software was published in author's Ph.D. thesis "Removing Boilerplate and Duplicate Content from Web Corpora". The deduplication algorithm is based on comparing n-grams of words of text. The author's algorithm has been shown to be more suitable for textual corpora deduplication than competing algorithms (Broder, Charikar): in addition to detection of identical or very similar (95 %) duplicates, it is able to detect even partially similar duplicates (50 %) still achieving great performace (further described in author's Ph.D. thesis). The unique deduplication capabilities and scalability of the algorithm were been demonstrated while building corpora of American Spanish, Arabic, Czech, French, Japanese, Russian, Tajik, and six Turkic languages consisting --- several TB of text documents were deduplicated resulting in corpora of 70 billions tokens altogether. and PRESEMT, Lexical Computing Ltd
Rights:: BSD 3-Clause "New" or "Revised" license, http://opensource.org/licenses/BSD-3-Clause, and PUB

1126. Open morphology of Finnish

Creator:: Pirinen, Tommi A, Listenmaa, Inari, Johnson, Ryan, Tyers, Francis M., and Kuokkala, Juha
Publisher:: University of Helsinki
Type:: tool and toolService
Subject:: morphological analysis and morphological dictionary
Language:: Finnish
Description:: Omorfi is free and open source project containing various tools and data for handling Finnish texts in a linguistically motivated manner. The main components of this repository are: 1) a lexical database containing hundreds of thousands of words (c.f. lexical statistics), 2) a collection of scripts to convert lexical database into formats used by upstream NLP tools (c.f. lexical processing), 3) an autotools setup to build and install (or package, or deploy): the scripts, the database, and simple APIs / convenience processing tools, and 4) a collection of relatively simple APIs for a selection of languages and scripts to apply the NLP tools and access the database
Rights:: GNU General Public Licence, version 3, http://opensource.org/licenses/GPL-3.0, and PUB

1127. Open SDP

Creator:: Flickinger, Dan, Hajič, Jan, Ivanova, Angelina, Kuhlmann, Marco, Miyao, Yusuke, Oepen, Stephan, and Zeman, Daniel
Publisher:: Oslo University and Charles University
Type:: text and corpus
Subject:: semantic dependency and parsing
Language:: English and Czech
Description:: The original SDP 2014 and 2015 data collections were made available under task-specific ‘evaluation’ licenses to registered SemEval participants. In mid-2016, all original data has been bundled with system submissions, supporting software, an additional SDP-style collection of semantic dependency graphs, and additional background material (from which some of the SDP target representations were derived) for release through the Linguistic Data Consortium (with LDC catalogue number LDC2016 T10). One of the four English target representations (viz. DM) and the entire Czech data (in the PSD target representation) are not derivative of LDC-licensed annotations and, thus, can be made available for direct download (Open SDP; version 1.1; April 2016) under a more permissive licensing scheme, viz. the Creative Common Attribution-NonCommercial-ShareAlike license. This package also includes some ‘richer’ meaning representations from which the English bi-lexical DM graphs derive, viz. scope-underspecified logical forms and more abstract, non-lexicalized ‘semantic networks’. The latter of these are formally (if not linguistically) similar to Abstract Meaning Representation (AMR) and are available in a range of serializations, including in AMR-like syntax. Please use the following bibliographic reference for the SDP 2016 data: @string{C:LREC = {{I}nternational {C}onference on {L}anguage {R}esources and {E}valuation}} @string{LREC:16 = {Proceedings of the 10th } # C:LREC} @string{L:LREC:16 = {Portoro\v{z}, Slovenia}} @inproceedings{Oep:Kuh:Miy:16, author = {Oepen, Stephan and Kuhlmann, Marco and Miyao, Yusuke and Zeman, Daniel and Cinkov{\'a}, Silvie and Flickinger, Dan and Haji\v{c}, Jan and Ivanova, Angelina and Ure\v{s}ov{\'a}, Zde\v{n}ka}, title = {Towards Comparability of Linguistic Graph Banks for Semantic Parsing}, booktitle = LREC:16 year = 2016, address = L:LREC:16, pages = {3991--3995} }
Rights:: Creative Commons - Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0), http://creativecommons.org/licenses/by-nc-sa/4.0/, and PUB

1128. Open SDP 1.2

Creator:: Flickinger, Dan, Hajič, Jan, Ivanova, Angelina, Kuhlmann, Marco, Miyao, Yusuke, Oepen, Stephan, and Zeman, Daniel
Publisher:: Oslo University and Charles University
Type:: text and corpus
Subject:: semantic dependency and parsing
Language:: English and Czech
Description:: The original SDP 2014 and 2015 data collections were made available under task-specific ‘evaluation’ licenses to registered SemEval participants. In mid-2016, all original data has been bundled with system submissions, supporting software, an additional SDP-style collection of semantic dependency graphs, and additional background material (from which some of the SDP target representations were derived) for release through the Linguistic Data Consortium (with LDC catalogue number LDC2016 T10). One of the four English target representations (viz. DM) and the entire Czech data (in the PSD target representation) are not derivative of LDC-licensed annotations and, thus, can be made available for direct download (Open SDP; version 1.2; January 2017) under a more permissive licensing scheme, viz. the Creative Common Attribution-NonCommercial-ShareAlike license. This package also includes some ‘richer’ meaning representations from which the English bi-lexical DM graphs derive, viz. scope-underspecified logical forms and more abstract, non-lexicalized ‘semantic networks’. The latter of these are formally (if not linguistically) similar to Abstract Meaning Representation (AMR) and are available in a range of serializations, including in AMR-like syntax. Version 1.1 was released April 2016. Version 1.2 adds the 2015 Turku system, which was accidentally left out from version 1.1. Please use the following bibliographic reference for the SDP 2016 data: @string{C:LREC = {{I}nternational {C}onference on {L}anguage {R}esources and {E}valuation}} @string{LREC:16 = {Proceedings of the 10th } # C:LREC} @string{L:LREC:16 = {Portoro\v{z}, Slovenia}} @inproceedings{Oep:Kuh:Miy:16, author = {Oepen, Stephan and Kuhlmann, Marco and Miyao, Yusuke and Zeman, Daniel and Cinkov{\'a}, Silvie and Flickinger, Dan and Haji\v{c}, Jan and Ivanova, Angelina and Ure\v{s}ov{\'a}, Zde\v{n}ka}, title = {Towards Comparability of Linguistic Graph Banks for Semantic Parsing}, booktitle = LREC:16 year = 2016, address = L:LREC:16, pages = {3991--3995} }
Rights:: Creative Commons - Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0), http://creativecommons.org/licenses/by-nc-sa/4.0/, and PUB

1129. Opening of the Week of Czech Youth

Creator:: Aktualita
Publisher:: Národní filmový archiv
Type:: video and clip
Subject:: akce Týden české mládeže, akce Kuratorium pro výchovu mládeže, Kuratorium pro výchovu mládeže, projev Teuner František zv., Kuratorium, Places::Karlštejn::nádvoří, Places::Karlštejn::celkový pohled, People::Moravec Emanuel (1893-1945), People::Teuner František (1911-1978), People::Fischer Ferdinand (1907-), and Československý zvukový týdeník Aktualita::1944/28
Language:: Czech
Description:: Segment from Český zvukový týdeník Aktualita (Czech Aktualita Sound Newsreel) issue no. 28A, B from 1944 was shot during the official opening of the Week of Czech Youth organised by the Board of Trustees for the Education of Youth and held in the courtyard of Karlštejn Castle on 1 July. The ceremony was attended by Minister of Education and People´s Enlightenment and Chairman of the Board Emanuel Moravec and SS officer Ferdinand Fischer. General Secretary of the Board František Teuner spoke to the participants.
Rights:: http://creativecommons.org/licenses/by-nc-nd/4.0/, PUB, and Creative Commons - Attribution-NonCommercial-NoDerivatives 4.0 International (CC BY-NC-ND 4.0)

1130. OpenLegalData (2022 - Corpus)

Creator:: Rüdiger, Jan Oliver
Publisher:: Rüdiger, Jan Oliver
Type:: text and corpus
Subject:: corpus, legal texts, legal domain, annotated corpus, NLP, and CorpusExplorer
Language:: German
Description:: OpenLegalData is a free and open platform that makes legal documents and information available to the public. The aim of this platform is to improve the transparency of jurisprudence with the help of open data and to help people without legal training to understand the justice system. The project is committed to the Open Data principles and the Free Access to Justice Movement. OpenLegalData's DUMP as of 2022-10-18 was used to create this corpus. The data was cleaned, automatically annotated (TreeTagger: POS & Lemma) and grouped based on the metadata (jurisdiction - BundeslandID - sub-size if applicable - ex: Verwaltungsgerichtsbarkeit_11_05.cec6.gz - jurisdiction: administrative jurisdiction, BundeslandID = 11 - sub-corpus = 05). Sub-corpora are randomly split into 50 MB each. Corpus data is available in CEC6 format. This can be converted into many different corpus formats - use the software www.CorpusExplorer.de if necessary.
Rights:: Creative Commons - Attribution-NonCommercial-ShareAlike 4.0 International (CC BY-NC-SA 4.0), http://creativecommons.org/licenses/by-nc-sa/4.0/, and PUB

1121. On-line Dictionary of medieval latin in the Czech lands

1122. On-Request Concert in Lucerna Palace

1123. Ondřej Sekora (painter)

1124. One Year Since the Death of President Masaryk

1125. onion

1126. Open morphology of Finnish

1127. Open SDP

1128. Open SDP 1.2

1129. Opening of the Week of Czech Youth

1130. OpenLegalData (2022 - Corpus)

Limit your search

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Search

Search Constraints

Search Results

Limit your search

Contributor

Show values starting with

Coverage

Creator

Show values starting with

Language

Show values starting with

Publisher

Show values starting with

Rights

Show values starting with

Subject

Show values starting with

Type

Show values starting with

Date

Original context has metadata only

Harvested from