Training, Language and Culture

ISSN 2520-2073 | ISSN 2521-442X


Lexical competition in Kazakh–Russian bilingual media: A corpus-based quantitative study

In recent years, corpus linguistics has provided new opportunities for studying language contact using quantitative and data-driven techniques. For the Kazakh language, situated in a multilingual setting involving Kazakh, Russian, and English, lexical interference represents an important sociolinguistic issue. This study investigates lexical competition from quantitative and functional standpoints, with particular attention to frequency, distribution, and variation in Kazakh-language mass media under Kazakh-Russian language contact. Using corpus methods, the study shows that large collections of texts make it possible to identify borrowed lexemes and competing equivalent forms. Borrowed items may coexist with native equivalents, producing measurable differences in frequency and stylistic preference. In some instances, foreign lexemes are more frequent than standardised national alternatives in the analysed corpus, although established Kazakh norms may continue to coexist with such borrowings. To address this issue, the study introduces the Lexical Competition Coefficient (LCC), a descriptive numerical indicator that measures the relative frequency of competing lexical equivalents in corpus data. When the Kazakh item predominates, the LCC exceeds one (k > 1), indicating that the Kazakh variant is more frequent in the corpus. By contrast, a coefficient below one (k < 1) indicates that the borrowed variant is more frequent. The findings indicate that lexical competition in Kazakh-language media is uneven across lexical pairs, with some borrowed forms substantially outnumbering Kazakh equivalents while other Kazakh variants remain dominant in corpus usage. The study combines scholarship on language contact and quantitative corpus analysis and proposes a systematic procedure for assessing lexical competition in corpus data.

Лексическая конкуренция в казахско-русских билингвальных медиа: корпусное количественное исследование

В последние годы корпусная лингвистика открыла новые возможности для изучения языковых контактов с применением количественных методов и подходов, основанных на данных. Для казахского языка, функционирующего в многоязычной среде с участием казахского, русского и английского языков, лексическая интерференция представляет собой важную социолингвистическую проблему. В настоящем исследовании лексическая конкуренция рассматривается с количественной и функциональной точек зрения, с особым вниманием к частотности, распределению и вариативности в казахскоязычных СМИ в условиях казахско-русского языкового контакта. С использованием корпусных методов в исследовании показано, что крупные массивы текстов позволяют выявлять заимствованные лексемы и конкурирующие эквивалентные формы. Заимствованные единицы могут сосуществовать с национальными эквивалентами, образуя измеримые различия в частотности и стилистических предпочтениях. В отдельных случаях иностранные лексемы в анализируемом корпусе встречаются чаще, чем стандартизированные национальные варианты, хотя устоявшиеся казахские нормы продолжают сосуществовать с такими заимствованиями. Для решения этой задачи в исследовании вводится Коэффициент лексической конкуренции (Lexical Competition Coefficient, LCC), описательный числовой показатель, измеряющий относительную частотность конкурирующих лексических эквивалентов в корпусных данных. Если преобладает казахская единица, значение LCC превышает единицу (k > 1), что указывает на более высокую частотность казахского варианта в корпусе. Напротив, коэффициент ниже единицы (k < 1) указывает на более высокую частотность заимствованного варианта. Исследование объединяет теоретические подходы к языковым контактам и количественный корпусный анализ и предлагает системную процедуру оценки лексической конкуренции на материале корпусных данных.




Volume 10 Issue 2

Address practices in Nigerian English academic discourse: Politeness, relationality and lingua-cultural adaptation

Felicia Bulus, Tatiana V. Larina
READ MORE

Lexical competition in Kazakh–Russian bilingual media: A corpus-based quantitative study

Assel B. Ormanova, Madina L. Anafinova, Kaziza Kozhagulova
READ MORE

Terminological asymmetry in consumer health Q&A discourse: A corpus-based study of MedQuAD

Asya S. Akopova
READ MORE

Visual–verbal blending in social advertising: Conceptual integration and meaning compression in environmental and humanitarian campaigns

Aida Nurbayeva, Gulzhan Tulekova, Gulaim Ospankulova
READ MORE

Intercultural communicative competence in ESP research: A Scientometric Mapping of global trends, theoretical clusters, and pedagogical gaps (2003–2023)

Vrindavanam Sahadeva Kurup Sreelakshmi, Selvaraj Vijayakumar
READ MORE

Disciplinary literacy and English for Academic Purposes in non-Anglophone EMI higher education: A scoping review of methodological trends

Sheikh Nahiyan, Manjet Kaur Mehar Singh
READ MORE

Metaphor comprehension in Arabic-speaking EFL learners: Embodiment types and gender effects

Haneen Eyad Yousef, Aseel Zibin, Abdel Rahman Mitib Altakhaineh
READ MORE

Goal orientation and intercultural sensitivity in EFL learners: Evidence from psychometric assessment and speaking performance

Mohamed Ridha Ben Maad
READ MORE

Creative Construction Grammar (book review)

Lazar Stošić
READ MORE