Romanian Language Resources Repository
Showing 1 - 10 out of 215
Description:
Multilingual thesaurus of Agriculture and food. Searching interface and XML/RDF full download.
Description:
terminological database about anatomy in various languages
Description:
The PARSEME shared task aims at identifying verbal MWEs in running texts. Verbal MWEs include idioms (let the cat out of the bag), light verb constructions (make a decision), verb-particle constructions (give up), and inherently reflexive verbs (se suicider 'to suicide' in French). VMWEs were annotated according to the universal guidelines in 18 languages. The corpora are provided in the parsemetsv format, inspired by the CONLL-U format.
For most languages, paired files in the CONLL-U format - not necessarily using UD tagsets - containing parts of speech, lemmas, morphological features and/or syntactic dependencies are also provided. Depending on the language, the information comes from treebanks (e.g., Universal Dependencies) or from automatic parsers trained on treebanks (e.g., UDPipe).
This item contains training and test data, tools and the universal guidelines file.
Description:
This multilingual resource contains corpora in which verbal MWEs have been manually annotated. VMWEs include idioms (let the cat out of the bag), light-verb constructions (make a decision), verb-particle constructions (give up), inherently reflexive verbs (help oneself), and multi-verb constructions (make do). VMWEs were annotated according to the universal guidelines in 19 languages. The corpora are provided in the cupt format, inspired by the CONLL-U format. The corpora were used in the 1.1 edition of the PARSEME Shared Task (2018).
For most languages, morphological and syntactic information not necessarily using UD tagsets including parts of speech, lemmas, morphological features and/or syntactic dependencies are also provided. Depending on the language, the information comes from treebanks (e.g., Universal Dependencies) or from automatic parsers trained on treebanks (e.g., UDPipe).
This item contains training, development and test data, as well as the evaluation to...
Description:
This multilingual resource contains corpora in which verbal MWEs have been manually annotated, gathered at the occasion of the 1.2 edition of the PARSEME Shared Task on semi-supervised Identification of Verbal MWEs (2020).
VMWEs include idioms (let the cat out of the bag), light-verb constructions (make a decision), verb-particle constructions (give up), inherently reflexive verbs (help oneself), and multi-verb constructions (make do). For the 1.2 shared task edition, the data covers 14 languages, for which VMWEs were annotated according to the universal guidelines. The corpora are provided in the cupt format, inspired by the CONLL-U format. Morphological and syntactic information not necessarily using UD tagsets including parts of speech, lemmas, morphological features and/or syntactic dependencies are also provided. Depending on the language, the information comes from treebanks (e.g., Universal Dependencies) or from automatic parsers trained on treebanks (e.g., UDPi...
Author(s):
Păiș, Vasile; Ion, Radu; Avram, Andrei-Marius; Mitrofan, Maria; Tufiș, Dan
Description:
Models for Stanza, RNNTagger, NLP-Cube, UDPipe, TreeTagger trained on the RRT UD 2.7 corpus. The models were evaluated in the associated paper. Scripts used in training and evaluating the models are available in our GitHub. A working version of the TTL tool is available in the TEPROLIN service repository. For downloading the corpus visit the Universal Dependencies website or directly download UD 2.7 treebanks.
Description:
A parallel corpus created from translations of the Bible containing 102 languages.
Description:
RO-EN Bilingual corpus acquired from https://vaccination-info.eu
Description:
EN-RO Bilingual corpus from the Publications Office of the EU on the medical domain
Description:
EN-RO Bilingual corpus extracted from the Publications Office of the EU on the medical domain. These are sourced from laws, studies, EC announcements, etc. labelled with concepts like epidemiology, epidemic, disease surveillance, health control, public hygiene, freedom of movement, distance learning, etc.
Showing 1 - 10 out of 215