Conveners
Digital Lexicography: Design, Data Modelling & Theory
- Ivana Filipović Petrović (Croatian Academy of Sciences and Arts)
Digital Lexicography: Design, Data Modelling & Theory
- Alexander Geyken (Berlin-Brandenburg Academy of Sciences and Humanities)
Digital Lexicography: Design, Data Modelling & Theory
- Hanna Fischer (Research Center Deutscher Sprachatlas)
Digital Lexicography: Design, Data Modelling & Theory
- Margit Langemets (Estonian Language Institute)
-
Lian Chen (LLL, University of Orléans), Damien Nouvel (ERTILM), Huy-Linh Dao (CRLAO-CNRS), Alexander Delaporte (CNRS-CRLAO)29/09/2026, 16:30
This paper presents NeoLex, a multilingual electronic lexical resource dedicated to contemporary Chinese–French and Vietnamese–French neology. Based on diachronic media corpora (2015–2025) and pre-2015 reference lexicons, the project combines corpus linguistics, Natural Language Processing, and computational lexicography to detect emerging lexical units. Candidate neologisms were extracted...
Go to contribution page -
Simon Krek (Jožef Stefan Institute), Iztok Kosem (University of Ljubljana & Jožef Stefan Institute), Polona Gantar (Faculty of Arts, University of Ljubljana)29/09/2026, 17:00
In the era of data-driven linguistics, the transition from traditional, human-oriented lexicography to machine-readable and interoperable language resources is paramount. The Digital Dictionary Database of Slovene (DDDS), developed by the Centre for Language Resources and Technologies (CJVT) at the University of Ljubljana, represents a paradigm shift in how national lexicographical data is...
Go to contribution page -
Lars Trap-Jensen (Society for Danish Language and Literature), Henrik Lorentzen (Society for Danish Language and Literature)29/09/2026, 17:30
This article discusses the ongoing revision of the usage-labelling system in Den Danske Ordbog (DDO, ‘The Danish Dictionary’) in response to increasing public attention on discriminatory and otherwise socially marked language. In recent years, the dictionary has received growing criticism concerning the treatment of identity-related terms, revealing limitations in the existing inventory of...
Go to contribution page -
Ondřej Matuška (Lexical Computing), Miloš Jakubíček (Lexical Computing), Vojtěch Kovář (Lexical Computing)01/10/2026, 10:001
Modern lexicographic workflows increasingly combine automatic dictionary drafting from large corpora with extensive human post-editing and validation. Advances in corpus linguistics and natural language processing make it possible to generate comprehensive dictionary drafts automatically, but the quality of the final product still depends on systematic human editorial work. As a result,...
Go to contribution page -
Michela Bandini (Istituto di Linguistica Computazionale “A. Zampolli”), Silvia Piccini (Istituto di Linguistica Computazionale “A. Zampolli”), Andrea Bellandi (Istituto di Linguistica Computazionale “A. Zampolli”), Emiliano Giovannetti (Istituto di Linguistica Computazionale “A. Zampolli”)01/10/2026, 10:20
This paper presents a computational narrative dictionary of youth language in contexts of high socio-economic vulnerability, where “narrative” is understood, following Bruner’s distinction between paradigmatic and narrative modes of meaning-making (Bruner, 1986), as meaning grounded in experience, examples, values, and self-positioning rather than only in categorisation. The resource is...
Go to contribution page -
Rufus H. Gouws (Stellenbosch University), Theo J.D. Bothma (University of Pretoria)01/10/2026, 10:40
This paper focuses on certain aspects of interface design in online dictionaries. A discussion of a number of examples of interface design in existing online dictionaries is presented. It is emphasised that successful interface design in transformative lexicographic products should be characterised by a human-centred approach that enables intuitive access to the required data. Dictionary...
Go to contribution page -
Pius ten Hacken (University of Innsbruck)02/10/2026, 10:00
The place of theory in lexicography has been the matter of intensive debate. There are different views of what a theory is. In line with most approaches in philosophy of science, it is assumed here that a theory has to provide an explanation. It is argued that lexicography should be considered as an applied science, i.e. the same category as medicine. In the analysis of applied science, the...
Go to contribution page -
Vasyl Starko (Ukrainian Catholic University), Andriy Rysin (Independent Researcher)02/10/2026, 10:20
We have created and implemented a workflow for large-scale semiautomatic detection of new and previously unregistered words in Modern Ukrainian. Crucially, it involves a diverse corpus of Ukrainian (2 billion tokens), a large electronic morphological dictionary of Ukrainian, and a specially crafted NLP toolkit. The key methodological innovation lies in the dual use of the dictionary. On the...
Go to contribution page -
Rukiye Smilla Burkart (Trier University), Susanne Kabatnik (Trier University)02/10/2026, 10:40
The formulation of definitions is a core task of lexicography, traditionally carried out manually. Large language models now offer new opportunities to automate this workflow. While AI-assisted definition generation has been well studied for well-known vocabulary, this is not yet the case for lesser-known historical vocabulary, which is scarcely represented in training data. This paper...
Go to contribution page -
David Lindemann (EHU University of the Basque Country)02/10/2026, 14:00
In this article, we describe several steps taken to convert Gidor Bilbao’s Oinarrizko Hiztegia Latina-Euskara (Basic Latin-Basque Dictionary) from a paper-based resource and move toward presenting it as Linked Data. Starting from a document in MS Word format, we attempted to isolate and tag elements of the dictionary’s mac-ro- and microstructures. The goal of the experiments outlined here...
Go to contribution page -
Peter Meyer (Leibniz Institute for the German Language)02/10/2026, 14:30
Complex digital lexicographical resources often feature rich data structures that are difficult for non-specialist users to query. This paper presents a pilot study on an LLM-driven natural-language interface for the Lehnwortportal Deutsch, a graph-based database of German loanwords in other languages. Rather than translating user questions directly into the database query language, the...
Go to contribution page -
Krzysztof Nowak (Institute of Polish Language PAS), Dorota Mika (Institute of Polish Language PAS)02/10/2026, 15:00
The Institute of the Polish Language (Polish Academy of Sciences) maintains a body of dictionaries and card-file archives assembled over more than a century, whose differences in structure, editorial convention, and descriptive metalan-guage have long kept them difficult to search together. This paper reports on PoliLex, an effort to bring such material under a common, machine-queryable...
Go to contribution page