Search and Browse – PORTULAN CLARIN

Europarl QTLeap WSD/NED corpus

Europarl QTLeap WSD/NED corpus This corpora is part of Deliverable 5.5 of the European Commission project QTLeap FP7-ICT-2013.4.1-610516 (http://qtleap.eu). The texts are sentences from the Europarl parallel corpus (Koehn, 2005). We selected the monolingual sentences from parallel corpora ...

Resource Type:	Corpus
Media Type:	Text
Languages:	Basque
	Bulgarian
	Czech
	English
	Portuguese
	Spanish; Castilian

QTLeap WSD/NED corpus

QTLeap WSD/NED corpus This corpora is part of Deliverable 5.5 of the European Commission project QTLeap FP7-ICT-2013.4.1-610516 (http://qtleap.eu). The texts are Q&A interactions from the real-user scenario (batches 1 and 2). The interactions in this corpus are available in Basque, Bulgar...

Resource Type:	Corpus
Media Type:	Text
Languages:	Basque
	Bulgarian
	Czech
	English
	Portuguese
	Spanish; Castilian

Manually annotated corpora for teaching and learning purposes of Brazilian Portuguese, Dutch, Estonian, and Slovene

These are manually annotated corpora for teaching and learning purposes of Brazilian Portuguese, Dutch, Estonian, and Slovene, as a contribution to the Manually Annotated Corpora Family available in CLARIN. Sentences are annotated with “problematic” or “non-problematic” labels, from the point of ...

Resource Type:	Corpus
Media Type:	Text
Languages:	Brazilian Portuguese
	Dutch
	Estonian
	Slovene

Multilingual Parallel Discourse Markers Corpus

This new language resource is an ISO-based annotated multilingual parallel corpus for discourse markers. The corpus comprises nine languages, Bulgarian, Lithuanian, European Portuguese, Hebrew, Romanian, Polish, Macedonian, and Italian with English as a pivot language. To represent the meaning of...

Resource Type:	Corpus
Media Type:	Text
Language:	Multiple languages

U-Compare Type system

The resource constitues of a hierarchically-structured system of data types, which is intended to be suitable for describing the inputs and output annotation types of a wide range of natural language processing applications which operate within the UIMA Framework. It is being developed in conjunc...

Resource Type:	Language Description
Media Type:	Text
Language:	English

Termcat Neoloteca

Terms that have (more or less) recently been accepted and normalised by Termcat, mixed fields

Resource Type:	Lexical / Conceptual
Media Type:	Text
Languages:	Basque
	Catalan; Valencian
	English
	French
	Galician
	German
	Italian
	Latin
	Portuguese
	Spanish; Castilian

Termcat Industry

Industry terms

Resource Type:	Lexical / Conceptual
Media Type:	Text
Languages:	Basque
	Catalan; Valencian
	English
	French
	German
	Italian
	Portuguese
	Spanish; Castilian

Termcat Social Webs

Terms of Social Webs

Resource Type:	Lexical / Conceptual
Media Type:	Text
Languages:	Catalan; Valencian
	English
	French
	Galician
	Italian
	Portuguese
	Spanish; Castilian

Termcat Research Thesaurus

Terms of Research Thesaurus

Resource Type:	Lexical / Conceptual
Media Type:	Text
Languages:	Catalan; Valencian
	English
	French
	German
	Italian
	Latin
	Portuguese
	Spanish; Castilian

Termcat Digital Marketing

Terms for Digital Marketing

Resource Type:	Lexical / Conceptual
Media Type:	Text
Languages:	Catalan; Valencian
	English
	French
	Galician
	German
	Italian
	Portuguese
	Spanish; Castilian

Order by:

Filter by: