Toward a corpus-based multilingual terminology database for Intercultural Communication
Publication date
2025
Editors
Kosem, I.
Jakubíček, M.
Medveď, M.
Zgaga, K.
Arhar Holdt, Š.
Munda, T.
Salgado, A.
Advisors
Supervisors
DOI
Document Type
Part of book
Metadata
Show full item recordCollections
License
cc_by_sa
Abstract
This contribution focuses on the methodological aspects of the ICoMuTe project aiming to design a corpus-based multilingual terminology database for Intercultural Communication (ICC). The project seeks to explore how ICC terms relate to each other within six European languages (Dutch, English, German, French, Italian, Spanish), how these terms are connected to their scientific and cultural contexts, and how they can be translated across different languages and cultures while preserving meaning. The selected approach is corpus-based, using comparable corpora of ICC handbooks and a parallel corpus of texts produced by the European Parliament dealing with key questions related to ICC. Using text recognition and data mining tools (e.g., Sketch Engine), the most frequent ICC terms per language are extracted and analysed in context. To account for the culturally specific aspects of terms while achieving a high degree of cultural neutrality, a semantic model based on tags has been developed for comparing and linking terms across languages in a neutral manner, but natural language corpus-based definitions are also provided that reflect the cultural load of each term. The main findings suggest that semantic tags are relevant to balance the cultural specificity and neutrality of ICC terms, and that English acts as a reference linguistic and cultural framework for the emergence and development of terms in other languages.
Keywords
intercultural communication, multilingual terminology, corpus-basedlexicography, lexical functions, semantic primes
Citation
Vázquez, M I, Venema, C & Steffens, M 2025, Toward a corpus-based multilingual terminology database for Intercultural Communication. in I Kosem, M Jakubíček, M Medveď, K Zgaga, Š Arhar Holdt, T Munda & A Salgado (eds), Electronic lexicography in the 21st century (eLex 2025): Intelligent lexicography. Proceedings of the eLex 2025 conference. . Lexical Computing CZ s.r.o., pp. 435-452. < https://elex.link/elex2025/proceedings/ >