Termcat Neoloteca

Terms that have (more or less) recently been accepted and normalised by Termcat, mixed fields

Resource Type:Lexical / Conceptual
Media Type:Text
Languages:Basque
Catalan; Valencian
English
French
Galician
German
Italian
Latin
Portuguese
Spanish; Castilian
DA-EN Danish Ministry of Higher Education and Science 2 (Processed)

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. Parallel texts Danish-English from the Danish Ministry o...

Resource Type:Corpus
Media Type:Text
Languages:Danish
English
Polish Ministry of Foreign Affairs Youth 2011 Report (Processed)

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. A parallel Polish-English version of the Youth 2011 repo...

Resource Type:Corpus
Media Type:Text
Languages:English
Polish
DA-EN Danish Ministry of Higher Education and Science 4 (Processed)

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. Parallel texts Danish-English from the Danish Ministry o...

Resource Type:Corpus
Media Type:Text
Languages:Danish
English
Bilingual hr-en parallel corpus from Croatian Mine Action website (Processed)

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. Contents of http://www.hcr.hr website downloaded, aligne...

Resource Type:Corpus
Media Type:Text
Languages:Croatian
English
Hallituskausi 2011-2015 -- Finnish-English Translation Memory (Processed)

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. Information on the "Hallituskausi 2011–" translation mem...

Resource Type:Corpus
Media Type:Text
Languages:English
Finnish
PKN Orlen Dataset (Processed)

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. Dataset of the Polish public sector company PKN Orlen, a...

Resource Type:Corpus
Media Type:Text
Languages:English
Polish
Bilingual hr-en parallel corpus from the National and University Library in Zagreb website (Processed)

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. Contents of http://www.nsk.hr were crawled, aligned on d...

Resource Type:Corpus
Media Type:Text
Languages:Croatian
English
Parallel corpus from Estonian Ministry of Foreign Affairs (Processed)

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. Parallel corpus from content of Estonian Ministry of For...

Resource Type:Corpus
Media Type:Text
Languages:English
Estonian
LexMan-ChunkerTokenizer

LexMan-ChunkerTokenizer is a tokenizer and sentence splitter tool. Marks sentence boundaries, multi-word boundaries. Size: Lemmas verbs: 12 995; Lemmas nouns and adj: 38 180; Lemmas adverbs: 7 250; Compound words: 35 201. Language: Portuguese.

Resource Type:Tool / Service
Language:Portuguese

Order by:

Filter by:

Text (446)
Audio (18)
Image (1)