LTDaugumos užsienio kalbų leksikografijoje tekstynų lingvistikos metodas jau įprastas: remiantis tekstynais, per tris pastaruosius dešimtmečius parengta daug anglų, vokiečių, ispanų, lenkų ir kitų kalbų leksikografinių išteklių (įvairių spausdintų ir elektroninių žodynų, duomenynų). Lietuvių kalbotyros darbuose tekstynų lingvistikos metodas taip pat plačiai taikomas, tačiau lietuvių leksikografijoje tekstynais vis dar remiamasi nedaug, paprastai rengiant aiškinamuosius žodynus tekstynai pasitelkiami tik vartojimo pavyzdžiams atrinkti. Vis dėlto, kaip įrodo užsienio leksikografijos tyrimai ir praktiniai darbai, tekstynų duomenys apie žodžių ir žodžių junginių vartoseną svarbūs visiems leksikografinio darbo etapams, be to, tekstynų duomenys gali padėti leksikografams sudaryti empiriškai pagrįstą kalbos vartosenos aprašą, vadinasi, išspręsti gana įsisenėjusią aiškinamųjų lietuvių kalbos žodynų problemą, kad žodynai nepakankamai gerai atspindi realią dabartinės kalbos žodžių vartoseną. Šioje monografijoje, remiantis atvejo – „Mokomojo lietuvių kalbos vartosenos leksikono“ – analize, yra išsamiai aprašomos tekstynų lingvistikos taikymo galimybės lietuvių leksikografijoje. Leksikonas yra pirmas leksikografinis išteklius lietuvių kalba, kurio antraštynas ir žodžių vartosenos aprašas pagrįstas konkrečiu tekstynu – „Mokomojo lietuvių kalbos tekstyno“ rašytine dalimi. Monografiją, be būtinų ir mokslo darbams įprastų dalių, sudaro du pagrindiniai skyriai. 1-ajame monografijos skyriuje aptariami su leksikono makrostruktūra susiję praktiniai ir teoriniai klausimai: tekstyno rengimas, antraštyno formavimo principas, naudotos leksikografinio darbo automatizavimo ir tekstyno analizės priemonės.2-ajame monografijos skyriuje dėmesys skirtas mikrostruktūros aspektams: teorinėje apžvalgoje parodyta įvairovė būdų (metodų), kuriais bandoma praktiškai tirti bei aprašyti leksikos ir gramatikos vienovę užsienio tyrėjų darbuose; išsamiai aprašyta vartosenos modelių analizė – metodas, kuris buvo adaptuotas leksikone tiriant tipišką žodžių vartoseną. Šiame skyriuje taip pat aptarti vartosenos modelių analizės adaptavimo principai leksikone ir su išsamiais pavyzdžiais aprašyti svarbiausi vartosenos modelių analizės praktiniai aspektai, pateikti duomenys apie vardažodžių ir veiksmažodžių vartosenos modelius. Monografijoje parengta ir rekomendacijų dalis: joje ne tik pateikiamos leksikono panaudojimo ir tobulinimo gairės, bet ir diskutuojami probleminiai taikytos metodikos aspektai, išdėstoma, į ką turėtų atsižvelgti leksikografai praktikai siekdami taikyti tekstynų lingvistikos metodą ir automatizuodami žodyno rengimo procesus. Svarbi monografijos dalis yra ir 8 priedai, kuriuose pateikiami ankstesniuose tyrimuose neskelbti duomenys iš leksikono duomenų bazės.
ENAlthough the methods of corpus linguistics are common and well-tested in many foreign lexicographic works, they have been underused in Lithuanian lexicography. Corpora can be useful at all stages of lexicographic work, and automatization of lexicographers’ work can enhance the development of existing and compilation of new lexicographic resources for the Lithuanian language. Thus, the methods of corpus linguistics could and should be more widely applied for Lithuanian lexicography. To demonstrate the potential of using corpora in Lithuanian lexicography, we present one of the most recent resources, the Lexical Database of Lithuanian Language Usage (further on, database), as a case study. This electronic resource is the first lexical database of its kind in the Lithuanian language, with a headword list and a description of word usage based on the written part of the Pedagogic Corpus of Lithuanian (further on, corpus), which consists of about 620,000 words. The database will exemplify how corpora can be used in product development for learner lexicography. The database was prepared in 2018-2020, under the project “Lithuanian Academic Scheme for International Cooperation in Baltic Studies“ (No. 09.3.1-ESFA-V-709-01-0002) and, together with other resources, is available at https://kalbu.vdu.lt/. The compilers of the database are Jolanta Kovalevskaitė (work group leader), Agnė Bielinskienė, Loic Boizou, Laima Jancaitė, and Erika Rimkutė. The data about the accentuation and pronunciation and audio files for the database were prepared by Asta Kazlauskienė and Sigita Dereškevičiūtė; user interface was developed by Petras Pauliūnas. The database differs from other Lithuanian language dictionaries by the applied methodology of corpus-driven and corpus-based lexicography: 1) the headword list was generated from the corpus, rather than taken from previous dictionaries.2) the meanings of words were revealed not by providing definitions, but usage patterns of a given word or word meaning, which capture lexical and grammatical patterning. Specifically, these patterns were used to distinguish the meanings of words. Following the corpus linguistic approach, the meaning of a word was associated with a specific lexical and grammatical environment. As the database has a user interface, it can be considered as a prototype of an active learner dictionary. Other Lithuanian dictionaries often lack usage data necessary for active learner dictionaries or teaching resource development focusing on language production skills and collocational competence. The aim of the study is to demonstrate the possibilities of corpus linguistic methods in Lithuanian lexicography for usage description of lexical items taking the Lexical Database of Lithuanian Language Usage as a case study. To achieve this goal, the following tasks were formulated: 1. To describe the compilation of headword list: 1) to introduce the source of the database – the corpus, and to discuss the compilation principles, structure, and linguistic features of the corpus; 2) to explain the principles used to select the lexical items included in the headword list, and to describe the presentation of lexicalized forms, multi-word lexical items, derivatives, and homonyms; 3) to make an initial assessment of the validity of the headword list for the purpose of teaching Lithuanian as a foreign language. 2. To discuss the corpus-based and corpus-driven methods of studying the usage of lexical items based on research in other languages. 3. To present the applied research on the identification of usage patterns of Lithuanian lexical items which were included in the database: 1) to present the theoretical principles of the Corpus Pattern Analysis.2) to describe the two stages of the analysis of corpus patterns: automated procedure to detect word collocability and manual description of corpus patterns by linguists; 3) to provide analysis of the identified corpus patterns with nouns, verbs, adjectives, and adverbs. 4. To provide recommendations for further corpus-driven and corpus-based research in Lithuanian lexicography.