Back to Search
Start Over
Zur Darstellung eines mehrstufigen Prototypbegriffs in der multilingualen automatischen Sprachgenerierung: vom Korpus über word embeddings bis hin zum automatischen Wörterbuch
- Source :
- Lexikos, Vol 31, Pp 20-50 (2021)
- Publication Year :
- 2021
- Publisher :
- Woordeboek van die Afrikaanse Taal-WAT, 2021.
-
Abstract
- Towards the Description of a Multi-sided Prototype Concept in Multilingual Automatic Language Generation: From Corpus via Word Embeddings to the Automatic Dictionary. The multilingual dictionary of noun valency Portlex is considered to be the trigger for the creation of the automatic language generators Xera and Combinatoria, whose development and use is presented in this paper. Both prototypes are used for the automatic generation of nominal phrases with their mono- and bi-argumental valence slots, which could be used, among others, as dictionary examples or as integrated components of future autonomous E-Learning-Tools. As samples for new types of automatic valency dictionaries including user interaction, we consider the language generators as we know them today. In the specific methodological procedure for the development of the language generators, the syntactic-semantic description of the noun slots turns out to be the main focus from a syntagmatic and paradigmatic point of view. Along with factors such as representativeness, grammatical correctness, semantic coherence, frequency and the variety of lexical candidates, as well as semantic classes and argument structures, which are fixed components of both resources, a concept of a multi-sided prototype stands out. The combined application of this prototype concept as well as of word embeddings together with techniques from the field of automatic natural language processing and generation (NLP and NLG) opens up a new way for the future development of automatically generated plurilingual valency dictionaries. All things considered, the paper depicts the language generators both from the point of view of their development as well as from that of the users. The focus lies on the role of the prototype concept within the development of the resources.
- Subjects :
- nlg: natural language generation
automatic dictionary
interactive dictionary
language generators
corpus lexicography
ontology
prototype
lexical prototype
semantic prototypical classes
Philology. Linguistics
P1-1091
Languages and literature of Eastern Asia, Africa, Oceania
PL1-8844
Germanic languages. Scandinavian languages
PD1-7159
Subjects
Details
- Language :
- Afrikaans, German, English, French, Dutch; Flemish
- ISSN :
- 16844904 and 22240039
- Volume :
- 31
- Database :
- Directory of Open Access Journals
- Journal :
- Lexikos
- Publication Type :
- Academic Journal
- Accession number :
- edsdoj.812564d82b51487b807f74e40734e495
- Document Type :
- article
- Full Text :
- https://doi.org/10.5788/31-1-1623