Harvested from: LINDAT/CLARIAH-CZ repository / Original context has metadata only: true

Start Over Original context has metadata only true Harvested from LINDAT/CLARIAH-CZ repository

131. CorpusExplorer

Creator:: Rüdiger, Jan Oliver
Publisher:: Jan Oliver Rüdiger
Type:: tool and toolService
Subject:: Corpus Linguisitics, NLP, conll, tei, XML, nlp, Natural Language Processing, linguistics, Linguistics, Computational Linguistics, corpus processing, tagger, POS tagger, lemmatization, text cleaning, CommonCrawl, epub, JSON, Twitter, Pandoc, Wikipedia, digital data, DTA, DSpin, MySQL, ElasticSearch, TextGrid, text corpora, TigerXML, and WeblichtXML
Language:: German, English, French, Italian, Dutch, Spanish, Polish, Arabic, Chinese, and Portuguese
Description:: Software for corpus linguists and text/data mining enthusiasts. The CorpusExplorer combines over 45 interactive visualizations under a user-friendly interface. Routine tasks such as text acquisition, cleaning or tagging are completely automated. The simple interface supports the use in university teaching and leads users/students to fast and substantial results. The CorpusExplorer is open for many standards (XML, CSV, JSON, R, etc.) and also offers its own software development kit (SDK). Source code available at https://github.com/notesjor/corpusexplorer2.0
Rights:: Not specified

132. Croatian Dependency Treebank

Publisher:: University of Zagreb, Faculty of Humanities and Social Sciences
Format:: application/octet-stream
Type:: corpus
Language:: Croatian
Description:: Manually tagged dependency treebank, analytical layer according to the PDT formalism adapted for Croatian
Rights:: Not specified

133. Croatian Frequency Dictionary

Publisher:: University of Zagreb, Faculty of Humanities and Social Sciences
Format:: application/octet-stream
Type:: lexicalConceptualResource
Language:: Croatian
Description:: 38,573 lemmas, plain text; database file
Rights:: Not specified

134. Croatian Lemmatization Server

Publisher:: University of Zagreb, Faculty of Humanities and Social Sciences
Type:: toolService
Language:: Croatian
Description:: On line service for lemmatization, full POS or MSD tagging of Croatian texts.
Rights:: Not specified

135. Croatian Morphological Lexicon

Publisher:: University of Zagreb, Faculty of Humanities and Social Sciences
Type:: lexicalConceptualResource
Language:: Croatian
Description:: 110,000+ lemmas; 3,900,000+ word-forms, MulText East lexica format
Rights:: Not specified

136. Croatian National Corpus

Publisher:: University of Zagreb, Faculty of Humanities and Social Sciences
Type:: corpus
Language:: Croatian
Description:: This is the reference corpus of standard Croatian. In its 3.0 version, which is accessible via noSketch Engine, it has 216.8 million tokens. In terms of annotation, the corpus is tokenised, lemmatised and tagged for MSDs (morphosyntactic descriptions).
Rights:: Not specified

137. Croatian-English Parallel Corpus

Publisher:: University of Zagreb, Faculty of Humanities and Social Sciences
Type:: corpus
Language:: Croatian and English
Description:: written; domain-specific (newspaper); synchronic; bilingual; parallel; unidirectional; XML; S-alignment
Rights:: Not specified

138. CST's lemmatiser

Publisher:: Center for Sprogteknologi, University of Copenhagen
Type:: toolService
Language:: Danish, Dutch, English, German, Modern Greek (1453-), Icelandic, Norwegian, Russian, Slovenian, and Swedish
Description:: 1) Fully automatic rule based lemmatization of inflected languages 2) Fully automatic training of lemmatization rules based on full form-lemma list
Rights:: Not specified

139. CST's lemmatizer

Creator:: Jongejan, Bart
Publisher:: Københavns Universitet, Center for Sprogteknologi (CST)
Type:: toolService
Description:: 1) Fully automatic rule based lemmatization of inflected languages 2) Fully automatic training of lemmatization rules based on full form-lemma list
Rights:: Not specified

140. Cyril Belica : Kookkurrenzdatenbank CCDB

Publisher:: Institut für Deutsche Sprache
Type:: toolService
Language:: German
Description:: A co-occurrence database, developed by the Institut fuer Deutsche Sprache, for research in the field of collocation analysis in modern German. The database holds over 200,000 analysed words that can be browsed or searched and shown in context.
Rights:: Not specified

« Previous
Next »
1
2
…
10
11
12
13
14
15
16
17
18
…
69
70

131. CorpusExplorer

132. Croatian Dependency Treebank

133. Croatian Frequency Dictionary

134. Croatian Lemmatization Server

135. Croatian Morphological Lexicon

136. Croatian National Corpus

137. Croatian-English Parallel Corpus

138. CST's lemmatiser

139. CST's lemmatizer

140. Cyril Belica : Kookkurrenzdatenbank CCDB

Limit your search

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Search

Search Constraints

Search Results

Limit your search

Contributor

Show values starting with

Coverage

Show values starting with

Creator

Show values starting with

Format

Language

Show values starting with

Publisher

Show values starting with

Rights

Show values starting with

Subject

Show values starting with

Type

Show values starting with

Date

Original context has metadata only

Harvested from