Speakers
Description
This paper describes the lexicographic resources developed for the ETAP text analysis and generation system, with particular emphasis on its semantic module, SemETAP. In our approach, semantic analysis is viewed not merely as the construction of a semantic representation, but as the derivation of inferences licensed by linguistic and background knowledge. The semantic component operates in two stages. First, it constructs a Basic Semantic Structure (BSemS), which represents the meaning explicitly conveyed by the text. This representation is then enriched through ontological knowledge and inference rules, yielding an Extended Semantic Structure (EnSemS) that incorporates both explicit and implicit information. To support this process, ETAP relies on a set of large-scale, knowledge-rich lexicographic resources characterized by a high degree of reusability. Linguistic knowledge is distributed across a morphological dictionary, a combinatorial dictionary, and rule-based grammars, whereas background knowledge is represented in an ontology, including a repository of individuals, and a collection of inference rules. The paper describes the information encoded in these resources, shows formal languages used to represent it, and the interaction between linguistic and extralinguistic knowledge during semantic analysis.