Language: Italian and Spanish - LINDAT/CLARIAH-CZ Catalog Search Results

Start Over Language Italian Language Spanish Date Unknown

25. HamleDT 3.0

Creator:: Zeman, Daniel, Mareček, David, Mašek, Jan, Popel, Martin, Ramasamy, Loganathan, Rosa, Rudolf, Štěpánek, Jan, and Žabokrtský, Zdeněk
Publisher:: Charles University
Type:: text and corpus
Subject:: annotated corpus, morphology, syntax, dependency, treebank, harmonized annotation, and common annotation style
Language:: Arabic, Basque, Bengali, Bulgarian, Catalan, Croatian, Czech, Danish, Dutch, English, Estonian, Finnish, French, German, Modern Greek (1453-), Ancient Greek (to 1453), Hebrew, Hindi, Hungarian, Indonesian, Irish, Italian, Japanese, Latin, Persian, Polish, Portuguese, Romanian, Russian, Slovak, Slovenian, Spanish, Swedish, Tamil, Telugu, and Turkish
Description:: HamleDT (HArmonized Multi-LanguagE Dependency Treebank) is a compilation of existing dependency treebanks (or dependency conversions of other treebanks), transformed so that they all conform to the same annotation style. This version uses Universal Dependencies as the common annotation style. Update (November 1017): for a current collection of harmonized dependency treebanks, we recommend using the Universal Dependencies (UD). All of the corpora that are distributed in HamleDT in full are also part of the UD project; only some corpora from the Patch group (where HamleDT provides only the harmonizing scripts but not the full corpus data) are available in HamleDT but not in UD.
Rights:: HamleDT 3.0 License Terms, https://lindat.mff.cuni.cz/repository/xmlui/page/licence-hamledt-3.0, and PUB

26. Le compagnon de tous ou dictionnaire polyglotte pour les écoles, et pour ceux qui s'occupent de lnfues étrangères et aux Arabes qui étudient les Langues Occidentales: enrichi des termes nouveaux de sciences et arts , choisis ou approuvés dans une réunion de sceïkhs. par Louis Calligaris

Creator:: Calligaris, Luigi
Type:: model:monograph and TEXT
Language:: French, Arabic, Latin, Italian, Spanish, Portuguese, German, English, and Modern Greek (1453-)
Rights:: http://creativecommons.org/publicdomain/mark/1.0/ and policy:public

27. Lesen und Schreiben in Europa 1500-1900 :

Type:: text and sborníky konferenční
Subject:: Literatura. Literární život, vzdělanost, školství, pedagogika, učitelé, péče o mládež, světové dějiny novověku (1492-1918), and zahraniční periodika a sborníky
Language:: German, English, French, Italian, and Spanish
Description:: "Tagung in Ascona, Monte Verità, von 11. bis 14. November 1996"--Rub tit. l.
Rights:: unknown

28. Neologismos económicos en las lenguas románicas a través de la prensa

Publisher:: Institut Universitari de Lingüística Aplicada, Universitat Pompeu Fabra
Type:: lexicalConceptualResource
Subject:: terminology database
Language:: Catalan, French, Galician, Italian, Portuguese, Romanian, and Spanish
Description:: Multilingual terminological resource containing 3.875 entries from the Economics, Finance and Banking domains.
Rights:: Not specified

29. OmegaWiki

Publisher:: Universität Bamberg, World Language Documentation Centre
Format:: application/octet-stream
Type:: lexicalConceptualResource
Language:: Afrikaans, Arabic, Basque, Bulgarian, Catalan, Chinese, Czech, Danish, Dutch, English, Esperanto, Estonian, Finnish, French, Galician, Georgian, Modern Greek (1453-), Hebrew, Hungarian, Icelandic, Indonesian, Interlingua (International Auxiliary Language Association), Irish, Italian, Japanese, Khmer, Norwegian, Polish, Portuguese, Romanian, Russian, Serbian, Slovak, Spanish, Swedish, Turkish, Ukrainian, and Welsh
Rights:: GFDL or CC and http://www.omegawiki.org/Licensing

30. ParaCrawl Corpus version 1.0

Creator:: Koehn, Philipp, Heafield, Kenneth, Forcada, Mikel L., Esplà-Gomis, Miquel, Ortiz-Rojas, Sergio, Sánchez, Gema Ramírez, Cartagena, Víctor M. Sánchez, Haddow, Barry, Bañón, Marta, Střelec, Marek, Samiotou, Anna, and Kamran, Amir
Publisher:: ParaCrawl
Type:: text and corpus
Subject:: ParaCrawl, parallel corpus, CommonCrawl, machine translation, and text corpora
Language:: English, German, French, Spanish, Italian, Portuguese, Dutch, Polish, Czech, Romanian, Finnish, Latvian, Russian, and Estonian
Description:: The January 2018 release of the ParaCrawl is the first version of the corpus. It contains parallel corpora for 11 languages paired with English, crawled from a large number of web sites. The selection of websites is based on CommonCrawl, but ParaCrawl is extracted from a brand new crawl which has much higher coverage of these selected websites than CommonCrawl. Since the data is fairly raw, it is released with two quality metrics that can be used for corpus filtering. An official "clean" version of each corpus uses one of the metrics. For more details and raw data download please visit: http://paracrawl.eu/releases.html
Rights:: Public Domain Dedication (CC Zero), http://creativecommons.org/publicdomain/zero/1.0/, and PUB

« Previous
Next »
1
2
3
4
5
6
7

21. Deep Universal Dependencies 2.8

22. Deltacorpus

23. Deltacorpus 1.1

24. HamleDT 2.0

25. HamleDT 3.0

26. Le compagnon de tous ou dictionnaire polyglotte pour les écoles, et pour ceux qui s'occupent de lnfues étrangères et aux Arabes qui étudient les Langues Occidentales: enrichi des termes nouveaux de sciences et arts , choisis ou approuvés dans une réunion de sceïkhs. par Louis Calligaris

27. Lesen und Schreiben in Europa 1500-1900 :

28. Neologismos económicos en las lenguas románicas a través de la prensa

29. OmegaWiki

30. ParaCrawl Corpus version 1.0

Limit your search

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Show values starting with

Search

Search Constraints

Search Results

Limit your search

Contributor

Show values starting with

Coverage

Creator

Show values starting with

Format

Language

Show values starting with

Publisher

Show values starting with

Rights

Show values starting with

Subject

Show values starting with

Type

Original context has metadata only

Harvested from