Zur Darstellung eines mehrstufigen Prototypbegriffs in der multilingualen automatischen Sprachgenerierung: vom Korpus über word embeddings bis hin zum automatischen Wörterbuch
Autor: | María José Domínguez Vázquez |
---|---|
Jazyk: | Afrikaans<br />German<br />English<br />French<br />Dutch; Flemish |
Rok vydání: | 2021 |
Předmět: |
nlg: natural language generation
automatic dictionary interactive dictionary language generators corpus lexicography ontology prototype lexical prototype semantic prototypical classes Philology. Linguistics P1-1091 Languages and literature of Eastern Asia Africa Oceania PL1-8844 Germanic languages. Scandinavian languages PD1-7159 |
Zdroj: | Lexikos, Vol 31, Pp 20-50 (2021) |
Druh dokumentu: | article |
ISSN: | 1684-4904 2224-0039 |
DOI: | 10.5788/31-1-1623 |
Popis: | Towards the Description of a Multi-sided Prototype Concept in Multilingual Automatic Language Generation: From Corpus via Word Embeddings to the Automatic Dictionary. The multilingual dictionary of noun valency Portlex is considered to be the trigger for the creation of the automatic language generators Xera and Combinatoria, whose development and use is presented in this paper. Both prototypes are used for the automatic generation of nominal phrases with their mono- and bi-argumental valence slots, which could be used, among others, as dictionary examples or as integrated components of future autonomous E-Learning-Tools. As samples for new types of automatic valency dictionaries including user interaction, we consider the language generators as we know them today. In the specific methodological procedure for the development of the language generators, the syntactic-semantic description of the noun slots turns out to be the main focus from a syntagmatic and paradigmatic point of view. Along with factors such as representativeness, grammatical correctness, semantic coherence, frequency and the variety of lexical candidates, as well as semantic classes and argument structures, which are fixed components of both resources, a concept of a multi-sided prototype stands out. The combined application of this prototype concept as well as of word embeddings together with techniques from the field of automatic natural language processing and generation (NLP and NLG) opens up a new way for the future development of automatically generated plurilingual valency dictionaries. All things considered, the paper depicts the language generators both from the point of view of their development as well as from that of the users. The focus lies on the role of the prototype concept within the development of the resources. |
Databáze: | Directory of Open Access Journals |
Externí odkaz: |