Language: English - LINDAT/CLARIAH-CZ Catalog Search Results

45204. Woodrow Wilson and the American Diplomatic Tradition. The Treaty Fight in Perspective /

Creator:: Ambrosius, Loyd E.
Publisher:: Cambridge University Prass,
Subject:: Wilson, Woodrow,, politici USA, prezidenti USA, biografie, diplomacie, válka první světová (1914-1918), politické dějiny, politici, světové dějiny 1789-1918, světové dějiny 1918-1945, and USA
Language:: English
Rights:: unknown

45206. Word Importance Dataset

Creator:: Osuský, Adam and Javorský, Dávid
Publisher:: Charles University, Faculty of Mathematics and Physics, Institute of Formal and Applied Linguistics (UFAL)
Type:: text and corpus
Subject:: word importance, ranking, and importance ranking
Language:: English
Description:: This dataset comprises a corpus of 50 text contexts, each about 60 words in length, sourced from five distinct domains. Each context has been evaluated by multiple annotators who identified and ranked the most important words—up to 10% of each text—according to their perceived significance. The annotators followed specific guidelines to ensure consistency in word selection and ranking. For further details, please refer to the cited source. --- rankings_task.csv - This csv contains information about the contexts which are to be annotated: - id: A unique identifier for each task. - content: The context to be ranked. --- rankings_ranking.csv - This csv includes ranking information for various assignments. It contains four columns: - id: A unique identifier for each ranking entry. - score: The score assigned to the entry. - word_order: A JSON detailing the order of words positions. It is essentially the selected word positions and their ordering from an annotator. - assignment_id: A reference ID linking to the assignments. --- rankings_assignment.csv - This csv tracks the completion status of tasks by users. It includes four columns: - id: A unique identifier for each assignment entry. - is_completed: A binary indicator (1 for completed, 0 for not completed). - task_id: A reference ID linking to the tasks. - user_id: The identifier for the user who should complete the task (rank the words). --- Known Issues: Please note that each annotator was intended to rank each context only once. However, due to a bug in the deployment of the annotation tool, some entries may be duplicated. Users of this dataset should be cautious of this issue and verify the uniqueness of the annotations where necessary. --- This dataset is a part of work from a bachelor thesis: OSUSKÝ, Adam. Predicting Word Importance Using Pre-Trained Language Models. Bachelor thesis, supervisor Javorský, Dávid. Prague: Charles University, Faculty of Mathematics and Physics, Institute of Formal and Applied Linguistics, 2024.
Rights:: Creative Commons - Attribution 4.0 International (CC BY 4.0), http://creativecommons.org/licenses/by/4.0/, and PUB

45207. Word Order Variation in John Malalas' Chronicle /

Creator:: Bočková Loudová, Kateřina,
Type:: text
Subject:: Ióannés Malalas,, jazyk řecký, kroniky byzantské, světové dějiny středověku (do r. 1492), and jazyk, písmo
Language:: English
Rights:: unknown

45208. Word representations for multiple languages

Creator:: Müller, Thomas and Schütze, Hinrich
Publisher:: Center for Information and Language Processing, University of Munich
Type:: text and corpus
Subject:: morphological dictionary, morphological analysis, and PoS tagging
Language:: English, German, Latin, Hungarian, Spanish, and Czech
Description:: Dictionaries with different representations for various languages. Representations include brown clusters of different sizes and morphological dictionaries extracted using different morphological analyzers. All representations cover the most frequent 250,000 word types on the Wikipedia version of the respective language. Analzers used: MAGYARLANC (Hungarian, Zsibrita et al. (2013)), FREELING (English and Spanish, Padro and Stanilovsky (2012)), SMOR (German, Schmid et al. (2004)), an MA from Charles University (Czech, Hajic (2001)) and LATMOR (Latin, Springmann et al. (2014)).
Rights:: Creative Commons - Attribution 3.0 Unported (CC BY 3.0), http://creativecommons.org/licenses/by/3.0/, and PUB

45209. Word-final /s/ durations in spoken German

Creator:: Luef, Eva Maria
Publisher:: Charles University and Universität Hamburg
Type:: text and corpus
Subject:: sibilant, acoustics, word-final s, duration, and German
Language:: English
Description:: German has various homophonous sibilant fricatives of phonemic or morphemic nature that can appear in word-final position. In English, the functional status of a word-final \s\ influences its durational properties, with phonemic \s\ being longer than morphemic types. The data set presented here is a small selection of laboratory-elicited German sentences containing various words with final sibilant phonemes (e.g., "das Haus") and morphemes (plural, genitive, clitic, inflection). Durations of the \s\ types were measured and compared across the conditions. An ANOVA between the \s\ types and post-hoc Tukey pair-wise comparisons are presented that show various significant differences. The submission consists of a csv data file, containing a number of variables, and a PDF document detailing the experiment and variables.
Rights:: Creative Commons - Attribution-NoDerivatives 4.0 International (CC BY-ND 4.0), http://creativecommons.org/licenses/by-nd/4.0/, and PUB

45210. Words in Music - Music in Words /

Creator:: Doubravová, Jarmila,
Subject:: Erben, Karel Jaromír,, Dvořák, Antonín,, Novák, Vítězslav,, hudba česká, literatura česká, vztahy literatura-hudba, české země 1848-1914, and hudba, tanec, hudební nástroje
Language:: English
Rights:: unknown

45201. Woodrow Wilson :

45202. Woodrow Wilson and a revolutionary world, 1913-1921 /

45203. Woodrow Wilson and the American diplomatic tradition :

45204. Woodrow Wilson and the American Diplomatic Tradition. The Treaty Fight in Perspective /

45205. Word from the editors

45206. Word Importance Dataset

45207. Word Order Variation in John Malalas' Chronicle /

45208. Word representations for multiple languages

45209. Word-final /s/ durations in spoken German

45210. Words in Music - Music in Words /

Limit your search

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Search

Search Constraints

Search Results

Limit your search

Contributor

Show values starting with

Coverage

Show values starting with

Creator

Show values starting with

Format

Show values starting with

Language

Show values starting with

Publisher

Show values starting with

Rights

Show values starting with

Subject

Show values starting with

Type

Show values starting with

Date

Original context has metadata only

Harvested from