11 Language Resources

Order by:

 A Bilingual English-Ukrainian Lexicon of Named Entities Extracted from Wikipedia    
  • English
  • Ukrainian

ID: ELRA-M0104

ISLRN: 110-617-195-245-4

The bilingual English-Ukrainian lexicon of named entities uses Wikipedia metadata as a source. The extracted named entity pairs are classified into five classes: PERSON, ORGANIZATION, LOCATION, PRODUCT, and MISC (miscellaneous). The lexicon consists of 624,168 pairs and comes in two formats: csv ...

MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use - CC-BY-NC-4.0
0.00 € submit
0.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use - CC-BY-NC-4.0
0.00 € submit
0.00 € submit
 Bulgarian Valency Frame Lexicon    
  • Bulgarian

ID: ELRA-L0132

ISLRN: 188-702-981-369-5

The Bulgarian Valency Frame Lexicon is composed of 9547 lexical entries organized by frames with 960 mappings to Princeton WordNet available in XML format. It is a treebank-driven resource of extracted valency frames from BulTreeBank. The frames were manually curated. The frames followed the surf...

MEMBERacademiccommercial
Licence: Attribution, Share Alike - CC-BY-SA-3.0
0.00 € submit
0.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Share Alike - CC-BY-SA-3.0
0.00 € submit
0.00 € submit
 CALEM (Comprehensive Arabic LEMmas)    
  • Arabic

ID: ELRA-L0133

ISLRN: 462-532-124-988-8

Comprehensive Arabic LEMmas is a lexicon covering a large list of Arabic lemmas and their corresponding inflected word forms (stems) with details (POS + Root). Each lexical entry represents a lemma followed by all its possible stems and each stem is enriched by its morphological features, especia...

MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use, No Derivatives - CC-BY-NC-ND
0.00 € submit
0.00 € submit
Licence: Commercial Use - ELRA VAR
5000.00 € submit
5000.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use, No Derivatives - CC-BY-NC-ND
0.00 € submit
0.00 € submit
Licence: Commercial Use - ELRA VAR
7500.00 € submit
7500.00 € submit
 CroaTPAS    
  • Croatian
  • English

ID: ELRA-M0108

ISLRN: 649-554-159-147-9

CroaTPAS (Croatian Typed Predicate Argument Structures) is a bi-lingual lexicon in Croatian and English. It was created by manual annotation from the Croatian Web as Corpus and pattern creation using the Skema editor on the Sketch Engine platform. CroaTPAS is tailor-made to represent verb polysem...

MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use - CC-BY-NC-4.0
0.00 € submit
0.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use - CC-BY-NC-4.0
0.00 € submit
0.00 € submit
 English-Danish EASTIN-CL Multilingual Ontology of Assistive Technology (Processed)    
  • Danish
  • English

ID: ELRA-M0075

ISLRN: 034-297-263-067-2

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. EASTIN-CL Multilingual Ontology of Assistive Technology ...

MEMBERacademiccommercial
Licence: Attribution, Share Alike - CC-BY-SA-3.0
0.00 € submit
0.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Share Alike - CC-BY-SA-3.0
0.00 € submit
0.00 € submit
 English-Estonian EASTIN-CL Multilingual Ontology of Assistive Technology (Processed)    
  • English
  • Estonian

ID: ELRA-M0073

ISLRN: 367-945-013-309-2

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. EASTIN-CL Multilingual Ontology of Assistive Technology ...

MEMBERacademiccommercial
Licence: Attribution, Share Alike - CC-BY-SA-3.0
0.00 € submit
0.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Share Alike - CC-BY-SA-3.0
0.00 € submit
0.00 € submit
 English-Latvian EASTIN-CL Multilingual Ontology of Assistive Technology (Processed)    
  • English
  • Latvian

ID: ELRA-M0076

ISLRN: 704-517-283-753-9

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. EASTIN-CL Multilingual Ontology of Assistive Technology ...

MEMBERacademiccommercial
Licence: Attribution, Share Alike - CC-BY-SA-3.0
0.00 € submit
0.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Share Alike - CC-BY-SA-3.0
0.00 € submit
0.00 € submit
 English-Lithuanian EASTIN-CL Multilingual Ontology of Assistive Technology (Processed)    
  • English
  • Lithuanian

ID: ELRA-M0074

ISLRN: 133-724-111-130-7

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. EASTIN-CL Multilingual Ontology of Assistive Technology ...

MEMBERacademiccommercial
Licence: Attribution, Share Alike - CC-BY-SA-3.0
0.00 € submit
0.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Share Alike - CC-BY-SA-3.0
0.00 € submit
0.00 € submit
 MADED (Moroccan Arabic Dialect Electronic Dictionary)    
  • Arabic

ID: ELRA-L0134

ISLRN: 977-057-254-691-5

Moroccan Arabic Dialect Electronic Dictionary (MADED) is an electronic lexicon containing almost 11,500 entries. They are written in Arabic script wherein each Modern Standard Arabic (MSA) lemma is provided with its corresponding Moroccan Arabic equivalent. In addition, MADED entries are annotate...

MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use, No Derivatives - CC-BY-NC-ND
0.00 € submit
0.00 € submit
Licence: Commercial Use - ELRA VAR
1000.00 € submit
1000.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use, No Derivatives - CC-BY-NC-ND
0.00 € submit
0.00 € submit
Licence: Commercial Use - ELRA VAR
2000.00 € submit
2000.00 € submit
 MORV (Moroccan Morphological vocabulary)    
  • Arabic

ID: ELRA-L0135

ISLRN: 064-194-729-767-0

The Moroccan Morphological vocabulary is a lexicon containing more than 4.6 M entries describing a given Moroccan Arabic word with fourteen (14) morphological and semantic features: the word orthographic form, the segmentation (prefix and suffix), part-of-speech (POS), gender, number, tense and t...

MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use, No Derivatives - CC-BY-NC-ND
0.00 € submit
0.00 € submit
Licence: Commercial Use - ELRA VAR
6000.00 € submit
6000.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use, No Derivatives - CC-BY-NC-ND
0.00 € submit
0.00 € submit
Licence: Commercial Use - ELRA VAR
12000.00 € submit
12000.00 € submit
 T-PAS    
  • Croatian
  • English

ID: ELRA-M0109

ISLRN: 432-666-503-743-8

T-PAS (Typed Predicate Argument Structures) is a digital lexicon consisting of a corpus-derived collection of Italian verb argument structures, whose arguments have been manually annotated with a set of hierarchically organised semantic labels called Semantic Types. T-PAS is primarily tailored f...

MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use - CC-BY-NC-4.0
0.00 € submit
0.00 € submit
NON MEMBERacademiccommercial
Licence: Attribution, Non Commercial Use - CC-BY-NC-4.0
0.00 € submit
0.00 € submit