Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
16504 papers
Статья посвящена сопоставительному анализу образно-символической интерпретации тактильного жеста, вербализуемого глаголами «толкать» (рус.) и «推 (tuī)» (кит.) в русской и китайской лингвокультурах. Актуальность работы обусловлена включенностью проблематики репрезентации тактильных жестов в русской и китайской фразеологии в общенаучный тренд кросс-культурных исследований, а также малой степенью изученности специфики когнитивных моделей метафорического означивания на базе образов силового тактильного воздействия. Цель работы – выявить особенности концептуализации и образно-символического переосмысления фрейма «толкать/推 (tuī)», детерминированные различиями культурных контекстов. Материалом исследования послужили русские и китайские фразеологические единицы, содержащие лексический компонент, описывающий данный жест, отобранные из авторитетных фразеологических и толковых словарей. Методология включает семантический анализ переносных значений и метафор, сопоставительный анализ для выявления параллелей и расхождений в формировании концептуального поля, функциональный анализ употребления единиц в речи, а также выявление культурных особенностей, закрепленных во фразеологии. Результаты исследования демонстрируют, что, несмотря на универсальность использования физического действия как основы для метафорического переноса, русская и китайская фразеологические системы существенно различаются. В русском языке жест «толкать» формирует разветвленную сеть метафор, выражающих представления о социальных конфликтах, негативно окрашенной речевой деятельности, нахождении в тесном пространстве. Ценностный аспект фразеологической семантики характеризуется двойственностью: могут обозначаться как позитивное побуждение, так и различные препятствия. В китайской лингвокультуре жест «推 (tuī)» в значительной степени подвергается символической трансформации в русле конфуцианской этики, становясь маркером моральных норм: абсолютного доверия, уступчивости, иерархии и жертвенности. Грамматическая структура также различается: для китайских идиом характерен параллелизм и использование парных иероглифов, тогда как в русских преобладают глагольно-предложные конструкции. Теоретическая значимость работы заключается в углублении понимания механизмов метафоризации тактильного опыта и этнокультурных различий в языковой картине мира. Исследование вносит вклад в развитие сопоставительной лингвистики, лингвокультурологии и теории фразеологии. Практическая значимость полученных результатов связана с их применением в сфере межкультурной коммуникации, лингводидактики (обучение русскому и китайскому языкам как иностранным), а также в практике перевода, где необходимо учитывать культурные коннотации, стоящие за образами внешне сходных жестов. Полученные результаты открывают перспективы для дальнейших исследований других концептов из сферы тактильного взаимодействия в различных лингвокультурах. The article is dedicated to a comparative analysis of the figurative-symbolic interpretation of the tactile gesture verbalized by the verbs “толкать” (push) (Russian) and “推 (tuī)” (push) (Chinese) in Russian and Chinese linguocultures. The relevance of the work is determined by the inclusion of issues related to the representation of tactile gestures in Russian and Chinese phraseology within the general scientific trend of cross-cultural studies, as well as the limited extent of research on the specifics of cognitive models of metaphorical signification based on images of forceful tactile influence. The aim of the work is to identify the specifics of conceptualizing the frame “толкать/ 推 (tuī)” and the features of its figurative-symbolic reinterpretation, determined by differences in cultural contexts. The research material consists of Russian and Chinese phraseological units containing lexical components describing this gesture, selected from authoritative phraseological and explanatory dictionaries. The methodology includes semantic analysis of figurative meanings and metaphors, comparative analysis to identify parallels and discrepancies in the formation of the conceptual field, functional analysis of the use of units in speech, as well as analysis of cultural features entrenched in phraseology. The results of the study demonstrate that, despite the universality of using physical action as a basis for metaphorical transfer, the Russian and Chinese phraseological systems differ significantly. In the Russian language, the gesture “толкать” develops an extensive network of metaphors encoding social conflicts, competitive practices, negatively charged speech activity, and the experience of existing in confined social spaces. The semantics are characterized by duality, encompassing both positive encouragement and obstruction. In Chinese linguoculture, the gesture “推 (tuī)” undergoes significant symbolic transformation within the framework of Confucian ethics, becoming a marker of moral norms: absolute trust, compliance, hierarchy, and self-sacrifice. Grammatical structures also differ: Chinese idioms are characterized by parallelism and the use of paired characters, whereas Russian idioms predominantly feature verb-prepositional constructions. The theoretical significance of the work lies in deepening the understanding of the mechanisms of metaphorizing tactile experience and ethnocultural differences in the linguistic worldview. The research contributes to the development of comparative linguistics, linguoculturology, and phraseology theory. The practical significance of the obtained results is associated with their application in the field of intercultural communication, language teaching (teaching Russian and Chinese as foreign languages), as well as in translation practice, where it is necessary to consider the cultural connotations behind the images of seemingly similar gestures. The obtained results open prospects for further research on other concepts from the sphere of tactile interaction in various linguocultures.
данная статья посвящена анализу этрусского текста на каменном памятнике из Сьерра-де-Медина (провинция Тукуман, Аргентина), с привлечением кабардино-черкесского языка (абхазо-адыгская языковая группа) к чтению и осмыслению исторических событий этрусков. Выявлено типологическое сходство структуры фраз с прямым порядком слов СПО (подлежащее-сказуемое-объект), что характерно для этрусского и кабардино-черкесского языков: эргативность (маркирование субъекта переходного глагола); агглютинативность (присоединение аффиксов с отдельными грамматическими значениями). Лексико-синтаксическая структура текстов соответствует эпиграфическим нормам древневосточных царских надписей. this article is devoted to the analysis of the Etruscan text on the stone monument from Sierra de Medina (Tucuman Province, Argentina), with the involvement of the Kabardian-Cherkess language (Abkhaz-Adyghe language group) in reading and understanding the historical events of the Etruscans. Typological similarities in the structure of phrases with the direct word order SVO (subject-verb-object) have been identified, which is typical for Etruscan and Kabardian-Cherkessian languages: ergativity (marking the subject of a transitive verb); agglutination (the addition of affixes with specific grammatical meanings). The lexical and syntactic structure of the texts corresponds to the epigraphic norms of ancient Eastern royal inscriptions.
This article analyzes the problems of compliance with literary language norms in students' speech from linguistic and pedagogical perspectives. The literary language norm is interpreted as an important and stable element of the language system, and its role in shaping speech culture is highlighted. During the research process, lexical, phonetic, orthoepic, grammatical, orthographic, and punctuation errors occurring in students' speech are analyzed, and the causes of their emergence are identified. In particular, the influence of dialects and vernaculars, the growing role of mass media and internet speech, as well as methodological shortcomings in the educational process are indicated as main factors. The article scientifically highlights the effectiveness of the communicative approach, text-based work, corrective and analytical exercises, and creative tasks in improving compliance with literary language norms. Additionally, the exemplarity of teacher speech and the importance of shaping language culture in the school environment are substantiated. The research results demonstrate the necessity of a conscious and systematic approach to literary language norms in developing students' speech literacy and communicative competence.
What constrains working memory capacity? Classic theories place visual working memory close to perceptual systems, with fixed limits. Yet, emerging evidence shows that visual working memory capacity is increased for real-world objects compared to simple or abstract stimuli. The present study demonstrates that this memory advantage arises from semantic understanding of real-world objects – contrary to classic perceptual accounts of this cognitive system. Using counterfeit objects generated by generative adversarial networks that match real objects in terms of object form and visual similarity, we show that improvements in behavioral performance and increases in neural delay activity emerge solely for semantically meaningful, real objects. Correlation analyses indicate that subjective familiarity ratings predict memory for real objects, whereas stimulus colourfulness predicts memory for artificial objects, suggesting distinct mechanisms support memory for different stimulus types. Thus, conceptual knowledge exerts strong effects on visual working memory, significantly extending current theories that emphasize low-level perceptual features.
цесами граматикалізації та культурно зумовленими комунікативними The article presents a corpus-based analysis of the grammaticalization of the semi-modal verbs gonna, wanna, gotta in contemporary spoken English, with special emphasis on linguocultural variation between American and British English. The relevance of the study lies in the growing influence of spoken interaction, media discourse, and digital communication on the grammatical system of English, as well as in the need for empirical evidence of cross-varietal differences in the use of grammaticalized forms. The aim of the article is to investigate the grammaticalization of the semi-modal verbs gonna, wanna, gotta, to identify their grammatical and functional-semantic properties in spoken discourse, and to conduct a contrastive analysis of their usage in American and British linguocultures. The empirical data are drawn from the British National Corpus, the Corpus of Contemporary American English, and the NOW Corpus, which ensures the representativeness of the material and enables quantitative comparison across registers and discourse types. The corpus analysis demonstrates that the semi-modal verbs under study emerged through the reduction of the constructions going to, want to, and have got to and display high frequency in spoken language and informal genres, while remaining stylistically marked in formal written registers. The findings also reveal different degrees of grammaticalization: gonna and gotta show a higher level of grammatical abstraction, whereas wanna retains traces of lexical meaning. From a linguocultural perspective, the results indicate that these forms are more frequent and more widely accepted in American English, while in British English they preserve stronger stylistic markedness. The study confirms the close relationship between grammaticalization processes and culturally conditioned communicative norms in contemporary English
This study examines how authorial stance is expressed in academic writing by native English speakers (L1) and non-native English speakers (L2), with a focus on the use of discourse markers such as hedges (markers of mitigation), boosters (markers of epistemic strengthening), attitude markers, and self-mentions. The aim of the study is to identify cross-linguistic and cross-disciplinary differences and evaluate how rhetorical and institutional conventions influence L2 authors’ stance strategies. A comparative corpus-based methodology was employed. The analysis drew on two corpora: the British Academic Written English and the Michigan Corpus of Upper-Level Student Papers, supplemented by original academic texts written by students at the Azerbaijan Medical University. Using Hyland’s metadiscourse model, stance markers were extracted through lexicon-based queries and manually verified in context. Data were compared across disciplines (engineering vs business) and author status (L1 vs L2). The findings reveal that L2 authors, especially in technical disciplines, tend to overuse hedging and avoid self-mentions, often due to rhetorical traditions that discourage personal voice. In contrast, L1 authors exhibit greater lexical diversity and a balanced use of stance markers. In business-related texts, L2 authors show more assertive and expressive stance, though still limited in range compared to native speakers. Stance in academic writing is not only a linguistic but also a culturally and institutionally mediated phenomenon. The study underscores the need for targeted instruction in metadiscourse to enhance L2 authors’ rhetorical awareness and help them align with academic norms of different disciplines.
BACKGROUND AND AIMS: Cannabis cue reactivity paradigms are instrumental in studying the behavioral and neurocognitive mechanisms of cannabis use and cannabis use disorders; however, image sets used for cannabis cue reactivity paradigms vary between studies, and the lack of reliability and validity assessment hinders the quality of evidence they generate. The main aim of this study was to create a novel, open access, standardized and representative database of cannabis use-related images including control images matched by resolution, luminosity and complexity: The Cannabis Research Image Database (CRESIDA). The secondary aim was to examine whether subjective cannabis cue-induced craving was associated with cannabis use severity and whether this relationship was moderated by image type. As an illustrative example of how our open data can be used and how sample characteristics can shape cue reactivity, we also explored the role of cannabis-tobacco mixing by comparing cannabis cue induced cannabis and tobacco craving between individuals who did and did not mix the substances. DESIGN: An online survey was administered to participants recruited via online platforms, community advertisement and snowballing. SETTING: USA, the Netherlands and Australia. PARTICIPANTS/CASES: 689 participants who consumed cannabis monthly to daily (385 men, 298 women, 6 other) were recruited between January 2022 and May 2024. MEASUREMENTS: Out of 93 cannabis images and 93 matched neutral images, participants each rated 31 image pairs for cannabis craving (the primary outcome), arousal, valence and tobacco craving. Participants were characterized for socio-demographic data, level of cannabis use and related problems and mixing cannabis and tobacco. A subset of 78 images was selected for further analysis based on cannabis craving results. Image ratings were evaluated for internal consistency (α). Furthermore, we examined the association between cannabis cravings and cannabis use characteristics, and explored if cannabis craving ratings were affected by image type (i.e. product, paraphernalia and actions) and by using cannabis alone vs. mixing cannabis and tobacco. FINDINGS: The database showed excellent reliability (α = 0.995-0.965). Cannabis craving, valence and arousal discriminated cannabis and control images. More cannabis use days [unstandardized beta (β) = 0.162, P < 0.001] and cannabis use-related problems (β = 0.268, P < 0.001) were statistically significantly associated with higher image-related cannabis craving. Mixing cannabis with tobacco, compared with using cannabis alone, was associated with the presence of tobacco craving in relation to cannabis images, and with greater cannabis craving in relation to cannabis images (β = -0.457, P < 0.001). CONCLUSIONS: Images in the open access Cannabis Research Image Database (CRESIDA, https://osf.io/dc9nz/) appear to be reliable and valid for the scientific study of cue reactivity internationally, providing a broad range of free to use cannabis and control images.
The Canvas Model's equality processor operates in two complementary feed-modes. Feed-backwards (Steering) corrects errors via gradient descent: d\mathcal{E}/d\tau = -\kappa \nabla_{\mathcal{E}} \mathbb{E}. Feed-forward (Driving) anticipates goals via gradient ascent: d\mathcal{E}/d\tau = +\kappa \nabla_{\mathcal{E}} \mathbb{E}. The first four papers in this series explored Feed-backwards—training, convergence, regularization, and modularity. This paper explores Feed-forward. What this paper provides: · A formalization of anticipatory dynamics. The positive-sign dynamics enable the system to project forward, predict optimal trajectories, and act preemptively to maximize anticipated alignment with future goals. The update uses gradient ascent on a projected future state, not gradient descent on the current state.· Three experimental validations across domains: 1. Anticipatory continuous control (point mass positioning). Feed-forward Driving achieves integrated error of 0.72 \pm 0.11 vs 1.00 \pm 0.15 for reactive PD control and 0.78 \pm 0.12 for Model Predictive Control (MPC). Overshoot is reduced from 23% (PD) to 8% (Driving), beating MPC (12%). Driving matches MPC performance without a learned dynamics model or horizon optimization. 2. Lookahead for discrete sequence generation (character-level language modeling on Penn Treebank). Feed-forward Driving reduces test perplexity from 78.2 (autoregressive baseline) to 72.8 \pm 1.3. The lookahead projection anticipates grammatical constraints, avoiding locally probable but globally incoherent choices. Qualitative inspection confirms fewer repeated characters and more plausible word formations. 3. Adaptive guidance for image generation (classifier-guided diffusion on CIFAR-10). Standard classifier guidance uses a fixed scale w. Driving introduces adaptive per-step guidance: w_t = w_0 / (1 + p(y \mid x_t)). When the classifier is uncertain, guidance is stronger; when confident, guidance relaxes. Adaptive Driving achieves Fréchet Inception Distance (FID) of 4.05 \pm 0.14 vs 4.31 \pm 0.18 for standard guidance, and class accuracy of 95.1% vs 94.2% — better image quality and higher fidelity.· Connection to existing methods. Driving is not a new algorithm. It is a unifying principle that explains why Model Predictive Control, classifier-guided diffusion, lookahead optimizers, and RLHF work. All are manifestations of the Feed-forward mode, distinguished only by the choice of projection mechanism and the spectral energy being ascended.· The combined architecture. The full processor operates in both modes simultaneously: d\mathcal{E}/d\tau = -\kappa_b \nabla \mathbb{E}[\mathcal{E}_\tau] + \kappa_f \nabla \mathbb{E}[\mathcal{E}_\tau^{\text{proj}}]. The negative term corrects errors. The positive term anticipates goals. This is the mathematical framework for cognition itself: a system that both learns from its errors and acts on its anticipations.· Limitations acknowledged. Driving's effectiveness depends on the quality of the projection. If the projected future state is inaccurate, Driving can move toward the wrong target. The combined architecture (\kappa_b, \kappa_f > 0) mitigates this by allowing the system to both anticipate and correct. Why this matters: Steering corrects. Driving aims. The same processor, the same meta-time, the same spectral energy. Only the sign changes. The Canvas Model reveals that the fundamental dichotomy in AI—reactive vs. anticipatory, error-correcting vs. goal-seeking, introverted vs. extroverted—is not a philosophical distinction. It is a mathematical one: the sign of the update in meta-time. Keywords: Driving, Feed-forward dynamics, anticipatory control, lookahead generation, classifier guidance, diffusion models, gradient ascent, Model Predictive Control, RLHF, Canvas Model, equality processor, meta-time, Steering, Feed-backwards
The increasing use of artificial intelligence (AI) in academic translation has raised important questions about translation quality beyond grammatical accuracy and lexical fluency. In particular, the pragmatic dimension of translation, which involves the preservation of context-dependent meaning, authorial stance, and discourse conventions, remains underexplored in comparative human and AI translation research. This study investigates pragmatic errors and their impact on translational adequacy in human and AI-generated translations of academic texts. Adopting a qualitative linguopragmatic approach, the study analyses a corpus of academic research article abstracts translated by human translators and an AI-based translation system. The analysis focuses on key pragmatic features, including implicature, hedging and stance, deixis, and register and discourse organisation. The findings reveal systematic differences between human and AI-generated translations. While AI-generated translations demonstrate high levels of formal fluency, they exhibit recurrent pragmatic weaknesses, such as over-explicitation, inappropriate stance calibration, and discourse-level misalignment, which cumulatively reduce translational adequacy in academic contexts. Human translations, although not free from pragmatic deviation, show greater sensitivity to communicative intent and academic discourse norms through context-aware and strategic decision-making. The study contributes to translation quality assessment by highlighting pragmatic errors as a crucial indicator of adequacy and underscores the continued importance of pragmatic competence in AI-assisted academic translation and translator education.
This article examines the sociopolitical and sociocultural catalysts of changes in lingual consciousness and lingual behavior and, more broadly, the formation and strengthening of the lingual identity of Ukrainians during the period of the full-scale invasion. The study demonstrates the dynamics of the Ukrainization of national language practices and establishes a clear interdependence between shifts in language code and the declaration of a Ukrainocentric civic stance. The analysis of the lexical and semantic development of the concepts lingual consciousness and lingual identity, as well as the reconsideration of the identity–language nexus, reveals a significant expansion in their combinability with lexemes carrying a strong publicistic connotation, such as code, marker, sign, symbol, basis, foundation, principle, cornerstone, bearer, and instrument. The study also identifies inhibiting factors that impede the transformation of an emotionally motivated impulse toward the use of Ukrainian into a conscious and stable norm of Ukrainian-language behavior. These factors include lingual compromise, lingual inertia, lingual fatigue, lingual-related fear, lingual insecurity, lingual mimicry, the commercialization of language, and the stigmatization of Russian speakers. The findings suggest that effective instruments for overcoming these challenges are the so-called «grassroots language initiatives». The article concludes that the changes identified and analyzed indicate both a radical renewal of lingual consciousness and lingual identity and the persistence of earlier lingual habits. Nevertheless, the overall vector of these transformations is clearly Ukrainocentric: societal lingual mobilization is gradually becoming a norm of language behavior, while a situational and emotionally motivated impulse is evolving into a conscious need. If this process is supported by a consistent state lingual policy and education, Ukrainian lingual identity will be able to effectively withstand the challenges posed by the realities of war. Keywords: lingual consciousness, lingual identity, lingual choice, lingual compromise, lingual inertia, lingual fatigue, lingual-related fear, lingual insecurity, lingual mimicry, commercialization of language, stigmatization of Russian speakers.
What constrains working memory capacity? Classic theories place visual working memory close to perceptual systems, with fixed limits. Yet, emerging evidence shows that visual working memory capacity is increased for real-world objects compared to simple or abstract stimuli. The present study demonstrates that this memory advantage arises from semantic understanding of real-world objects – contrary to classic perceptual accounts of this cognitive system. Using counterfeit objects generated by generative adversarial networks that match real objects in terms of object form and visual similarity, we show that improvements in behavioral performance and increases in neural delay activity emerge solely for semantically meaningful, real objects. Correlation analyses indicate that subjective familiarity ratings predict memory for real objects, whereas stimulus colourfulness predicts memory for artificial objects, suggesting distinct mechanisms support memory for different stimulus types. Thus, conceptual knowledge exerts strong effects on visual working memory, significantly extending current theories that emphasize low-level perceptual features.
The early posterior negativity (EPN) is a mid-latency event-related potential (ERP) component reliably enhanced by emotionally arousing visual cues. Recent work suggests that modulation of the EPN might depend to some extent on evocative cues featuring animate content. We tested this possibility by recording EEG while 80 participants viewed pleasant, neutral, and unpleasant scenes depicting people, objects, or landscapes. People and object scenes were selected to be comparable in composition and arousal ratings, to enable a direct assessment of the impact of scene animacy, separate from emotional intensity. Results showed robust EPN modulation by emotional content across both people and object scenes, with no significant interaction across arousal-matched scenes. This finding further dissociates the EPN from proximal event-related potential components associated with face and body perception and supports its value as an early marker of emotional perception, reliably driven by emotional intensity across multiple domains of visual cues.
Research on how nonnatives process and learn binomials (black and white) is limited. The present study addresses this gap using online (eye-tracking) and offline (familiarity rating) tasks. Sixty nonnative speakers of English (L1 = Arabic) read six stories seeded with 21 novel binomials in three conditions: one exposure, six exposures, and no exposure (i.e., only in post-test) in a counter-balanced design. Each item was also presented in the reversed order (white and black). The nonnatives read the stories as their eye movements were monitored and answered comprehension questions. In addition to the novel binomials, 12 existing binomials (congruent with Arabic) were included in the passages as a baseline for comparison. After completing the reading task, the participants completed an offline rating task as a measure of declarative knowledge of the binomial configuration (i.e., word order). All items were rated twice, once in the forward direction and once in the reversed direction. Online results showed that nonnatives were not sensitive to the configuration of existing binomials and there was limited evidence of any sensitivity to novel binomials. Offline, nonnatives showed sensitivity to the configuration restrictions of existing binomials but not novel ones.
This article provides a comprehensive analysis of the sociolinguistic and pragmatic foundations of the concept of social distance. Social distance is interpreted as a system of relations between communicants determined by social status, age, gender, professional position, and cultural norms. The study identifies the mechanisms of expressing social distance at the lexical, grammatical, and pragmatic levels of the language system, as well as reveals their functional characteristics in the communicative process. Based on Uzbek language material, the linguocultural nature of social distance and its close connection with national mentality and norms of speech etiquette are substantiated.
This study focused on structural and semantic changes in Kazakh borrowed vocabulary adaption. The study examined how imported concepts were absorbed into Kazakh, altered by the national linguistic system, and contributed to current terminology. Various linguistic methodologies were used, including historical and contemporary text analysis, structural and comparative term analysis, and hybrid word classification. The study covered both present and historical English, allowing vocabulary changes to be tracked. The findings showed that complicated historical and cultural processes in active intercultural exchanges caused Kazakh terminology hybridisation. The Greco-Latin, Persian, and Arabic languages enriched Kazakh lexicon with science, religion, culture, and daily life concepts. Phonetic, morphological, and visual changes were made to borrowed terminology to make them useful and conform to Kazakh linguistic norms. The study showed that hybrid terms are crucial to borrowing integration.
Conventional research on phoney speech in legal circumstances typically views deception as a moral or cognitive defect at the individual level that can be recognised by consistent linguistic indicators. By suggesting that deception is an institutionally created discourse practice that results from the interplay of cognitive load, procedural limitations, and power imbalances in courtroom communication, this research proposes a theoretical reorientation. The study uses a mixed-methods strategy that combines qualitative forensic analysis with natural language processing techniques applied to specific Indian criminal court rulings, drawing on forensic linguistics, discourse analysis, and computational language modelling. Patterns of strategic ambiguity, evasive coherence, emotional modulation, and pragmatic indeterminacy are found in the testimonies of witnesses and accused individuals. To track how institutional forces influence communication behaviour, computational methods such as sentiment trajectory mapping, stance identification, and lexical dispersion metrics are combined with careful language reading. The analysis shows that deceitful discourse in legal contexts functions more as an adaptive, situationally sensible tactic conditioned by juridical norms and interpretive authority than as a sign of personal dishonesty. This study offers a reusable analytical framework for forensic linguistics and legal discourse studies by modelling deception as a situated, procedural, and culturally mediated phenomenon. This framework has implications for judicial interpretation, evidentiary evaluation, and the moral use of computational tools in legal contexts.
The article highlights the necessity of adhering to the literary norms of the Ukrainian language in contemporary medical terminology, particularly the importance of understanding its lexical-grammatical and stylistic levels. The study analyzes term-lexemes that form paronymic relations, as well as the causes and consequences of the unmotivated use of paronyms in scientific discourse (semantic similarity, insufficient understanding of lexical meanings, speakers’ lack of competence). It is found that the erroneous use of paronyms leads to distortion of expressed meaning, linguistic paradoxes, and speech errors. Correct (normative) variants of paronymic medical terms appropriate for the professional language of healthcare practitioners are proposed. The research employs the methods of analysis and synthesis, comparative (contrastive) analysis, and the general linguistic method of scientific description. Conclusions. Quantitative and qualitative changes within the national terminological system result from the interaction of linguistic and extralinguistic factors and regularities. A thorough linguistic analysis of core paronymic pairs (series), their inclusion in lexical minima, and active instructional work with them constitute an important aspect of successfully mastering the language of medicine and improving the quality of specialized medical literature.
This repository contains Anomaly Soul Kit, an open simulation framework for observing the emergence, persistence, and evolutionary inheritance of anomalous behavior in populations of LLM-driven agents. Each agent encodes a numeric internal state — vitality (H) and anomaly intensity (Z) — and expresses that state through LLM-generated text each generation. A detection layer scores each expression against the population across three axes: lexical divergence, structural divergence, and novel vocabulary. Agents whose expressions deviate from the population accumulate anomaly intensity, which feeds back into their fitness and is heritable across generations. The project does not claim these anomalies constitute mind or soul. It provides a reproducible kit for observing whether something — a persistent, evolving deviation — reliably emerges from this process, and what it looks like when it does. --- Update — February 2026 v2 of the anomaly detection layer has been released. Two structural issues identified in early testing have been addressed. First, anomaly score inflation: as the population evolved, an increasing proportion of agents were flagged as anomalous, eventually making the designation meaningless. This has been resolved by replacing absolute scoring with a dynamic baseline — scores are now normalized relative to the population median each generation, making it structurally impossible for the entire population to simultaneously score as anomalous. Second, convergence speed: the original selection pressure caused Z-awakening to saturate too quickly (~90% by generation 50). Scaling has been adjusted to allow slower, more observable divergence dynamics. Two new observational metrics have been added: new_normal_threshold tracks whether what was previously anomalous is becoming the new collective norm, and population_drift measures how much the group as a whole is shifting toward anomalous expression across generations.
The article examines religious axiological units in Uzbek and English as significant linguistic and cultural phenomena. Axiological units—lexical, phraseological, and discursive elements that encode values—are analyzed as carriers of religious worldviews, moral norms, and evaluative meanings. Drawing on axiological linguistics, pragmatics, and discourse analysis, the study compares how Islamic and Christian traditions shape value-laden language in Uzbek and English respectively. The analysis identifies dominant axiological categories such as faith, morality, humility, sin, righteousness, and reward, explores their linguistic realization, and discusses implications for translation and intercultural communication.
This article examines memes related to the Ukrainian language as universal information units in the online environment, as well as the special role of the Ukrainian language in the creation and functioning of virtual memes. The concept of the meme is defined and specified as an integral and coherent unit of Internet communication that has a standardized form and such characteristics as virality, replicability, emotionality, seriality, mimicry, minimalism of form, multimodality, relevance, humor, mediatization, and creativity. It is noted that most people understand a meme in a narrow sense—as an image, video, fragment of text, etc., which spreads rapidly from one Internet user to another, often with small modifications that make it humorous and highly popular. The article analyzes the patterns of emergence of Internet memes, the reasons for their popularity, and the peculiarities of their functioning online. It examines memes about the Ukrainian language related to the adoption of the 2019 orthographic reform; memes that vividly reflect lexical usage; as well as those connected to the observance of linguistic norms at the levels of morphology, syntax, punctuation, accentuation, and orthography. Contemporary modes of online communication have contributed to the significant spread of such memes on social networks such as Telegram and Facebook, and across the broader Google search environment. Ukrainian-language Internet memes are characterized as emotionally charged units of linguistic information whose source lies in multilevel elements of knowledge of the Ukrainian language. These memes are replicated and disseminated by Internet users in the form of text (inscriptions, captions) (textual), images or videos (visual), or a combination of text and iconic elements (drawings, photos, tables) (creolized or verbal-visual) in the process of communication. The Internet environment today is not only a sphere for the creation and consumption of informational products but also a space for satisfying the informational needs of the modern linguistic individual and for forming certain life orientations and values. Memes, as elements of online communication, harmonize perfectly with the individual needs of speakers and are flexible and adaptive.
This article analyzes the transformation of the Uzbek language in the digital environment acrossthree main areas. Digital technologies are driving rapid changes in the language: lexical units borrowed fromEnglish, graphic abbreviations, symbols expressing emotions, and hybrid forms have become actively used onthe internet and social media. At the same time, natural language processing (NLP) systems working with theUzbek language, such as text analysis, machine translation, and speech recognition models, are insufficientlyequipped with corpus and linguistic databases
The translation of Uzbek humor into English presents complex challenges due to the interplay of national-cultural realities, linguistic structures, and pragmatic intentions embedded in humorous discourse. Uzbek humor often relies on wordplay, culturally specific idioms, intertextual references, social norms, and context-dependent pragmatic cues, which do not always have direct equivalents in English. This study examines the main strategies of national-cultural and pragmatic adaptation used in English translations of Uzbek humorous expressions, anecdotes, and conversational jokes. Drawing on linguistic pragmatics, cultural semiotics, and translation theory, the research identifies how translators employ techniques such as cultural substitution, explicitation, pragmatic strengthening, functional equivalence, and compensatory humor creation to preserve both the humorous effect and the communicative intent of the source text. The findings show that successful translation of Uzbek humor requires not only lexical and structural transformation but also deep sensitivity to cultural worldview, sociolinguistic norms, and the interactive functions of humor. The study contributes to current scholarship by highlighting adaptive mechanisms that ensure intercultural comprehensibility while maintaining the original humorous nuance.
This article explores the field of lexical semantics and its significant role in shaping cultural perception within language. Lexical semantics, as a branch of linguistics, examines the meaning of words and their interrelations, revealing how language encodes cultural values, social norms, and collective cognition. The study analyzes various lexical items across languages, demonstrating how differences in word meanings can influence cultural understanding and worldview. Furthermore, the article highlights the dynamic interplay between language and culture, emphasizing that shifts in lexical semantics reflect changes in societal attitudes and cultural identity. This research contributes to a deeper understanding of the interconnection between linguistic meaning and cultural perception, providing insights for cross-cultural communication, translation studies, and cognitive linguistics.
This study is devoted to the analysis of the complex and multifaceted relationship between dialects and the literary language in the history of the Uzbek language based on a historical-linguistic approach. The study considers dialects as an important source in the formation and development of the Uzbek literary language, and the influence of their phonetic, lexical and grammatical features on the norms of the literary language is consistently highlighted. The work provides a comparative analysis of written sources of the ancient Turkic period, samples of the old Uzbek literary language and modern dialect materials, paying special attention to the issues of historical continuity and coherence. Also, the role of regional dialects in the formation of literary language norms, the processes of their selection and assimilation into the national language are studied in connection with sociolinguistic factors. The results of the study show that the interaction between dialects and the literary language is not a one-sided, but a dynamic and complex process. This scientific work serves as a theoretical and practical basis for a deeper understanding of the laws of the historical development of the Uzbek language, as well as for improving literary language norms and the effective use of dialect materials.
Music and visual expressions play a central role in young people’s identity formation and function as important arenas for negotiating norms releted to gender, power, and group identity. In recent years, the Swedish music genre Epa-dunk has emerged as a distinct subcultural phenomenon, particularly rooted in rural areas and closely connected to Epa-culture. The aim of this study is to critically analyze how women are represented in the visual communication of the Epa-dunk genre through a case study of the artist Fröken Snusk’s Instagram posts. By examining the artist’s visual expressions, the study seeks to gain a deeper understanding of how these representations may reproduce, challenge, or renegotiate traditional notions of women, as well as how they can be understood in relation to women’s empowerment. The study is guided by two research questions: (1) Which semiotic elements are used to represent female identity in Fröken Snusk’s Instagram posts, and (2) how do Fröken Snusk’s Instagram posts relate to traditional conceptions of women? The theoretical framework underlying the analysis consists of social semiotics and representation theory, as well as gender and feminist concepts such as the male gaze, sexualisation, objectification, stereotypes, the gender contract and empowerment. Methodologically, the study employs a qualitative semiotic text analysis focusing on five analytical categories: setting, posing, attributes, gaze direction, camera angle and lexical choices, in order to understand how different meaning-making elements interact to create meanings related to gender, power and identity. The results of the study show that Fröken Snusk’s visual expressions are mainly sexualized, even though there is some variation. In most cases, traditional gender norms are repeated through sexualized poses and images that focus strongly on her body, as well as other visual elements that can be connected to the male gaze. At the same time, some images also show signs of women’s empowerment. This can be seen when she appears in active roles in male-coded environments and through the use of objects and captions that signal confidence, independence, and control. Overall, the results show that Fröken snusk are mostly represented in ways that follow traditional gender norms, even if there are a few expressions that can be understood as attempts to challenge them.
Abstract This chapter examines Ntozake Shange’s for colored girls as a revolutionary choreopoem that emerges from Black feminist thought, the Black Arts Movement, and United States Black Language to articulate Black women’s interior lives through embodied language and performance. It situates the work within its historical, political, and theatrical contexts and argues that Shange’s written theatricality renders African American Women’s Language visible on the page through phonetic spelling, syntax, punctuation, rhythm, and structure as practices of cultural memory, resistance, and self-definition. The chapter analyzes form, symbolism, and aesthetic strategies—including the rainbow, the slash, eye dialect, musicality, and lowercase typography—to demonstrate how language functions as choreography that directs reading, hearing, and feeling while rejecting standardized English norms and the white gaze. Through close readings of key phases such as “no more love poems #4,” “somebody almost walked off wid alla my stuff,” and “layin on of hands,” it demonstrates how Black lexical items, discourse practices, and sonic rituals enact rhetorical healing, rememory, and collective restoration. Finally, the chapter argues that Shange’s choreopoem functions as a performative Black feminist theory of language that transforms personal and communal trauma into embodied affirmation, spiritual renewal, and an enduring declaration that Black women’s voices, bodies, and lives are already and fully enough.
Abstract: The integration of artificial intelligence (AI) into English as a Foreign Language (EFL) education has brought about transformative changes in how learners develop intercultural communicative competence (ICC). This systematic literature review examines how AI-mediated language production and adaptive feedback mechanisms reshape ICC among EFL learners. Following PRISMA guidelines, this study analysed 35 peer-reviewed articles published between 2020 and 2025. The review focuses on ELT-relevant dimensions, including automated writing evaluation, generative AI in language learning, and AI-mediated cross-cultural exchange. Findings indicate that AI facilitates ICC by providing real-time adaptive feedback that helps learners negotiate cultural nuances and linguistic norms. The study concludes that AI serves as a "cultural mediator," offering a triadic interaction model that enhances learners' knowledge, skills, and attitudes in intercultural settings.
Reviewer assignment is increasingly critical yet challenging in the LLM era, where rapid topic shifts render many pre-2023 benchmarks outdated and where proxy signals poorly reflect true reviewer familiarity. We address this evaluation bottleneck by introducing LR-bench, a high-fidelity, up-to-date benchmark curated from 2024-2025 AI/NLP manuscripts with five-level self-assessed familiarity ratings collected via a large-scale email survey, yielding 1055 expert-annotated paper-reviewer-score annotations. We further propose RATE, a reviewer-centric ranking framework that distills each reviewer's recent publications into compact keyword-based profiles and fine-tunes an embedding model with weak preference supervision constructed from heuristic retrieval signals, enabling matching each manuscript against a reviewer profile directly. Across LR-bench and the CMU gold-standard dataset, our approach consistently achieves state-of-the-art performance, outperforming strong embedding baselines by a clear margin. We release LR-bench at https://huggingface.co/datasets/Gnociew/LR-bench, and a GitHub repository at https://github.com/Gnociew/RATE-Reviewer-Assign.
In March 2021, the EU Parliament adopted Resolution 2021/2557, a legally binding measure that mandates all 27 member states to recognize the right to gender self-identification and to implement juridical norms aligned with this principle. Among its most transformative provisions, the Resolution calls for eliminating the male-female binary in favor of a more expansive framework that currently recognizes at least twenty-one gender identities - a number expected to grow. It also urges the revision of national languages to dismantle patriarchal structures and ensure that legal and institutional language reflects principles of gender plurality and inclusivity. Widely seen as a landmark victory for trans-feminist individuals and advocacy groups, this measure has sparked both support and controversy. The research examines whether such linguistic reforms foster inclusion or provoke democratic tensions in Italy, where gendered language is deeply rooted in historical, grammatical, and cultural traditions. It further investigates how trans-feminist advocacy - supported ideologically and financially by EU bodies (Commission, Parliament, and Council) - has gained significant influence, particularly as left-wing progressive political forces currently hold the majority within these institutions. These actors play a central role in shaping the narrative and enforcement of gender policies across EU member states. Employing a qualitative case study methodology, the analysis draws on a diverse range of materials, including press articles, televised debates, public messaging, lexical usage, multimedia content, and ideologically charged propaganda to assess the impact of EU gender policy on Italy’s linguistic landscape. Findings suggest that while these interventions promote visibility and recognition for gender-diverse individuals, they also raise concerns about linguistic autonomy, democratic principles, and the broader cultural consequences of ideologically driven legal mandates.
The rise of Artificial Intelligence (AI)-based tools is transforming language education, offering adaptive innovations for both language learners and instructors. However, concerns remain about their ability to represent linguistic diversity. Shaped by the data they process and the priorities of their creators, AI systems risk reinforcing dominant linguistic norms. This study explores these issues using Austrian Standard German (ASG), a distinct variety of German, as a case study. By analysing the behaviour of four AI tools—two models of a chatbot, a text-feedback system, and a grammar-correction tool—we assessed whether they recognised ASG as a legitimate standard or altered it to align with German Standard German (GSG). The evaluation, informed by structured testing, revealed significant shortcomings in how these systems handle linguistic variation. Our findings underscore the risks of erasing linguistic particularities and emphasise the need for AI tools to serve as a means of fostering linguistic identity—particularly in educational contexts.
This study examines the participation experiences of multilingual students speaking English as an additional language (EAL) and their decisions to invest or disinvest in language practices during classroom discussions. Using an embedded multiple case study design, data were collected through classroom observations, weekly reflections, and individual interviews, and analyzed using thematic analysis. The findings show that course characteristics, such as structure and topic, influenced power dynamics and identity negotiation, reinforcing sociocultural and linguistic norms aligned with Western participation practices. The complex interplay of course dynamics, norms, and broader ideologies contributed to multilingual EAL students’ novice identity, creating barriers that made it challenging for them to disrupt participatory norms and invest in the language practices of the course community. This article underscores the importance of reframing participation as a collaborative and critical process and highlights the need to create inclusive classroom environments to support the equitable participation of multilingual EAL students.
We revisit punctuation-aware tree binarization for constituency parsing and ask whether dependency-induced headedness improves binary parser supervision. Although learned heads substantially outperform rule-based heads in intrinsic head prediction, they do not yield consistent parsing gains after debinarization. In particular, punctuation-conditioned evaluation shows that learned headedness underperforms rule-based binarization in macro-average punctuation-sensitive $F_1$, despite a small overall gain on CTB. Similar instability appears under cross-treebank transfer. These results suggest that \ycc{linguistically grounded} headedness is not necessarily parser-optimal when used as a binarization control signal. The paper presents a negative result: better head prediction does not imply better punctuation-sensitive constituency parsing.
The mediatisation of politics is a sustainable trend in the development of modern information and communication space and is implemented, inter alia, through the system of communication of government and society, political actors, institutions and media. In the Spanish-language media, which combines elements of several semiotic systems, journalists implicitly ridicule politicians or socio-political phenomena, deliberately violating the linguistic norm and creating the effect of deceived expectations at the expense of lexical-semantic and graphical transformations of verbal and graphic precedent phenomena by replacing or adding lexical and/or graphic components, homonymic word play, as well as creating neologisms by combining precedent names for the naming of phenomena not previously existing in Hispanic linguistics.
The aim of the article is to uncover the cultural specificity of the meaning of lexical units “kola” and “yam” in Nigerian fiction. Methodology. To solve the tasks at hand, a number of both general scientific and specific methods of investigation were used. Systematic and cluster sampling methods were employed in selecting the linguistic material. The need to describe and analyse the semantic structure of the lexical units, as well as the material of the research, made it possible to resort to the methods of contextual and component analysis. The research deals with modern texts by Igbo authors (1990s – 2010s), as well as the classical works of Chinua Achebe. Results. This article identifies the culturally specific semantic properties of the lexemes “kola” and “yam”. The use of these lexical units is thoroughly analysed, especially in regards to units that can be considered as lexical neologisms as compared to the referent norm. These neologisms are a part of the lexical field “traditions and customs”. It is concluded that the culturally significant lexeme “yam” has gender markings and symbolizes the masculine principle, which is reflected both in the early and modern stages of the development of English-language Igbo literature. Research implications. The article is of theoretical and practical value to philologists and specialists that work with various variants of West African English. It provides recommendations as to the translation of phrases containing the kola unit.
Multilingualism is defined as a mode of communication in contemporary world. The multilingualism teaches us the important values to understand the context. This study analyzes dual point of view about the multilingualism: the foreign languages that appear in it, i.e. explicit multilingualism and the universal aspect or hidden languages that are indirectly described, i.e. implicit multilingualism. Thismay comprise linguistic norms, reader and text interaction, among others. The aim of this study is to highlight the impacts of elements of multilingualism used in Amélie Nothomb’s novels. It focuses essentially on the works of the contemporary Francophone writer, notably, Amélie Nothomb. She articulates the enriching elements of multilingualism in French and Japanese languages through herwritings. Her breakthrough works mainly articulate the diversity of multilingualism and also the essential meaning of understanding the different elements or expressions related to the French and Japanese language through the richness of culture from a geographical point of view and also the other elements. These elements are articulated about expressions which show the impact ofmultilingualism in her writings that refer either to French, Japanese, or other languages.
This study examines how gender norms are expressed and negotiated in WeChat Public Accounts using Feminist Critical Discourse Analysis (FCDA). We built a sampling frame of the ten most active gender-related accounts (June-December 2023) and screened 359 posts for relevance and analyzability. Twenty-two articles were selected for close reading. Two researchers independently coded clause- and sentence-level segments and reached agreement through discussion. Coding followed six dimensions used throughout the paper: lexical choice, modality, intertextuality, voice positioning, affective tone, and strategic silence. Four recurring themes were identified. (1) Maternal discourse and gendered discipline: texts often link women's value to motherhood and domestic duties, combining moral language with advice on correct behavior. (2) The body and mechanisms of shame: discussions of menstruation and sexuality frequently use medical or corrective language that assigns responsibility to individual women. (3) Gender identity in cultural, legal, and policy narratives: educational, media, and policy-adjacent texts describe ideal feminine roles through procedure, quantification, and role models, which stabilize familiar expectations. (4) Resistant discourses and incremental change: some pieces re-label practices, shift speaking positions, or use humor to push back against these norms. Overall, the analysis shows how everyday textual choices can normalize gendered expectations, while limited forms of resistance also appear. The findings are bounded by the small, purposive sample and the focus on text rather than audiences or algorithms. The study provides a transparent description of patterns observed in the 22 articles and clarifies where and how resistance is articulated within them.
Child-directed fingerspelling is an approach used by Deaf parents for communication, language, and literacy development. This study reports on findings from a qualitative intrinsic case study aimed at understanding how Deaf parents use fingerspelling with their young children. The research questions were: (1) What are the cultural beliefs of Deaf parents regarding fingerspelling with young children? (2) What are their patterns of use of child-directed fingerspelling in natural settings? Twenty-one Deaf families with 27 deaf children ages 5 years and under were interviewed via recorded Zoom meetings conducted in American Sign Language. Data were analyzed using grounded theory to develop a new theoretical contribution with the core category: Deaf families socialize their children into Deaf visual-linguistic norms through fingerspelling. This new theoretical insight aligns with Holcomb's Deaf epistemological framework (2010) and Ochs and Schieffelin's (2008, 2011) language socialization theory. Limitations and recommendations for future research are also included.
This paper examines how artificial intelligence (AI), machine learning algorithms, and automated digital systems shape linguistic practices, reinforce or challenge linguistic hierarchies, and influence communication in contemporary society. As digital platforms increasingly mediate human interaction, algorithms determine what content becomes visible, which linguistic varieties are privileged, and how users adapt their language to gain visibility and engagement. The study explores algorithmic bias in search engines, social media feeds, voice assistants, and automated moderation systems, highlighting how these technologies reproduce existing social inequalities related to class, caste, gender, and ethnicity. Drawing on sociolinguistic theories of language ideology, linguistic capital, and digital discourse, the paper argues that AI-driven communication environments are not neutral but deeply ideological. They shape linguistic norms, influence identity performance, and regulate public discourse. The findings underscore the need for critical sociolinguistic engagement with AI systems to ensure equitable, inclusive, and culturally sensitive digital communication.
AI-mediated communication refers to communicative processes in which artificial intelligence systems actively generate, interpret, modify, or facilitate language. With the rapid advancement of language technologies such as large language models, conversational agents, speech recognition systems and machine translation tools, AI has evolved from a supportive tool into an active mediator of human communication. This paper conceptualizes AI-mediated communication as a form of language technology that reshapes traditional models of interaction, meaning-making and authorship. The study examines the technological foundations of AI-driven language processing and their broader socio-cultural, pedagogical and ethical implications. It explores how AI mediation influences linguistic norms, accessibility and power relations, while also raising critical concerns related to bias, surveillance and linguistic homogenization. By situating AI-mediated communication within contemporary digital culture, the paper argues that understanding AI as a language technology is essential for critically engaging with evolving communicative practices and for developing responsible, inclusive and ethically grounded AI-driven communication systems.
The English language, traditionally viewed as monolithic and monocentric, now evolves into diverse varieties embedded in unique communities of practice. The expansion of online communication has further informed this evolution, giving rise to virtual spaces where members negotiate shared meanings, identities, and linguistic norms. This paper explores language use within an online networking business community of practice in the Philippines. Drawing on theories of Communities of Practice (COP) and Cultural Models, it examines how English is positioned in relation to the members’ cultural models and how community-specific jargons function as markers of identities. Through in-depth interviews and analysis of actual online conversations and social media posts, the study reveals that group members use English and local languages in dynamic linguistic strategies, such as code-switching and translingual practices, to construct and perform their identities within the group. The findings further reveal the role of English as a tool for empowerment and a marker of identity in digital spaces, hence emphasizing the need for a more inclusive understanding of language use in contemporary online communities.
Introduction. In the present-day scientific discourse, there is a great number of theoretical research focusing on the problem of objectifying the semiotic nature of law and analysing the functional construct of legal semantics. Whereas, many practical legal issues, such as: interpretative ambiguity in the meaning-formation and meaning-application of normative acts, lexical vagueness and contextual dependence of legal notions and the incoherence of legal terminology across different legal systems, remain neglected, which leads to contradictions and inaccuracies in legal practice. The aim of the study is to define the methodological principles fostering establishment of the acceptable scope of semantic interpretation of legal notions in the context of building a legal thinking culture. Materials and Methods. The research methodology was based on the principle of jurisprudential definition of legal norm meaning-formation in socio-legal discourse. Analytical, systematizing and pragmatic methods were used to reveal a complex nature of the semantics of law in the context of legal thinking development. The semiotic analysis of the objectivity and normativity of legal notions taking into account the contextual differences of legal definitions, was used as a specialised research method. Results. It was established that normative notions are the complex semantic constructs encompassing a conceptual sphere (normativity) and social reality. For building sustainable models of legal behaviour and legal culture, it is necessary to overcome external and internal conflicts in interpretation of law. In this regard, a number of advisory measures were proposed aimed at establishing acceptable scope of semantic interpretation: differentiation between the informational nature of prescriptive and descriptive notions, semantic monitoring of legal phenomena, and implementation of the principle of discourse contextualism, which makes it possible to formulate the normativity of law requirements based on the specific contextual interpretations. Discussion and Conclusion. A justified conclusion about possibility of a properly selected semantic toolkit to determine the objectivity of perception of the legal norms and, consequently, to improve the process of building a legal culture was drawn. The main advantage of the principle of discourse contextualism such as conjunction of the semantics and pragmatics of legal notions was identified, which provides a fruitful foundation for further theorizing on the nature and metaphysics of law.
SRC, an acronym for Stimulus-response correlation, refers to determining the relationship between stimulus and corresponding brain responses. The neural aesthetic resonance hypothesis proposes that the level of enjoyment or familiarity can be distinguishable based on the relationship between stimulus and brain responses. To test this hypothesis, we use EEG data of 20 participants listening to 12 songs with their enjoyment and familiarity ratings. We aim to classify the low and high ratings of familiarity and enjoyment based on SRC. Eighteen musical features are extracted and transformed into the first principal component (PC1). In addition, root mean square (RMS) and spectral flux are used for analysis. Canonical Correlation Analysis (CCA), an unsupervised AI optimization method, is employed to compute the SRC between musical features and ten regions of brain responses, followed by considering four principal CCA features for classification using the Random Forest classifier with cross-subject evaluation. Our results demonstrate that the right frontal and right parietal regions provide significant predictive ability. Our empirical finding suggests that RMS features preserve the predictive ability for familiarity, whereas PC1 is for enjoyment prediction. Maximum familiarity and enjoyment accuracy reach nearly 76% and 73% accuracy. This work leverages AI techniques to decode sensor-derived neural signals, advancing real-time applications in affective computing and wearable EEG devices.
This article examines youth language (Jugendsprache) as a significant and dynamic component of contemporary German. The study aims to analyze its structural, lexico-semantic, and functional features, as well as its role in shaping modern linguistic trends. The methodological framework combines descriptive, comparative, and lexico-semantic analysis, along with the examination of digital discourse, enabling a comprehensive investigation of youth language in its natural communicative environment. The findings demonstrate that Jugendsprache is characterized by a high degree of lexical innovation, driven by anglicisms, neologisms, and abbreviations emerging in digital communication. Word-formation processes, including compounding, affixation, and conversion, exhibit increased creativity and hybridization, often deviating from standard linguistic norms. Youth language also performs important sociolinguistic functions, such as identity construction, emotional expression, and the differentiation of social groups. The study highlights the crucial role of the digital environment in accelerating linguistic change and shaping new communicative practices. While Jugendsprache contributes to the enrichment and adaptability of the German language, it may also lead to challenges related to normativity and intergenerational communication. Overall, youth language is interpreted as a “laboratory of linguistic innovation” and a mediator between linguistic change and standardization, reflecting the broader processes of globalization and digitalization in modern society.
This study investigates the linguistic and communicative functions of abbreviations in English and Karakalpak advertising discourse, focusing on how these compressed forms contribute to message efficiency, stylistic expression, and cultural positioning. Although abbreviations are widely used across global advertising, their structural patterns and pragmatic roles vary according to linguistic norms and audience expectations. Therefore, the research employs a mixed qualitative methodology integrating structural analysis, discourse interpretation, and comparative linguistics. The results demonstrate that English advertising makes extensive and creative use of acronyms, initialisms, blends, and hybrid forms to construct modern, technologically oriented, and globally recognizable brand identities. In contrast, Karakalpak advertising relies more on functional initialisms and borrowed English abbreviations, reflecting both local communicative preferences and growing global influence. The discussion interprets these findings within broader socio-cultural and economic contexts, revealing that abbreviation usage serves as a marker of globalization, cultural continuity, and linguistic innovation. Ultimately, the study contributes to a deeper understanding of how abbreviated forms shape contemporary advertising communication in multilingual environments.
Pure speech language models aim to learn language directly from raw audio without textual resources. A key challenge is that discrete tokens from self-supervised speech encoders result in excessively long sequences, motivating recent work on syllable-like units. However, methods like Sylber and SyllableLM rely on intricate multi-stage training pipelines. We propose ZeroSyl, a simple training-free method to extract syllable boundaries and embeddings directly from a frozen WavLM model. Using L2 norms of features in WavLM's intermediate layers, ZeroSyl achieves competitive syllable segmentation performance. The resulting segments are mean-pooled, discretized using K-means, and used to train a language model. ZeroSyl outperforms prior syllabic tokenizers across lexical, syntactic, and narrative benchmarks. Scaling experiments show that while finer-grained units are beneficial for lexical tasks, our discovered syllabic units exhibit better scaling behavior for syntactic modeling.
While word embeddings derive meaning from co-occurrence patterns, human language understanding is grounded in sensory and motor experience. We present $\text{SENSE}$ $(\textbf{S}\text{ensorimotor }$ $\textbf{E}\text{mbedding }$ $\textbf{N}\text{orm }$ $\textbf{S}\text{coring }$ $\textbf{E}\text{ngine})$, a learned projection model that predicts Lancaster sensorimotor norms from word lexical embeddings. We also conducted a behavioral study where 281 participants selected which among candidate nonce words evoked specific sensorimotor associations, finding statistically significant correlations between human selection rates and $\text{SENSE}$ ratings across 6 of the 11 modalities. Sublexical analysis of these nonce words selection rates revealed systematic phonosthemic patterns for the interoceptive norm, suggesting a path towards computationally proposing candidate phonosthemes from text data.
Relevance of the research. This work is primarily based on the personal scientific project of P. J. Piaseckyj, a Ukrainian diaspora scholar from the USA. It was the project titled "Anglo Surzhyk," which was initiated on February 9, 2018 (and still ongoing), that provided the author with the further impetus to write this paper. The project itself is a study of the prevalence of Anglicisms in everyday, academic, cultural, and professional Ukrainian communication. It is based on articles from Ukrainian media published on the Internet and currently contains over 2,500 borrowings. The project is continuously updated and thus not available in print; an electronic copy may be requested directly from the author via email or by visiting its namesake Facebook page. [2] The author began contemplating the Anglicization of the Ukrainian language as early as 1949, at the age of six, upon arriving in New York and hearing the Anglicized Ukrainian spoken by American Ukrainians (from the first and second waves of emigration). Mr. Piaseckyj himself is fully proficient in English and possesses a "keen sensibility toward our language." On the other hand, Oleh Rudnyk (who is also fully proficient in both mentioned languages) focuses this work on the linguistic purism movement, specifically its historical continuity and its relevance to societal needs within the context of contemporary Ukrainian national realities. This focus is grounded in the ideas and works of American researchers such as Edward Sapir, Benjamin Lee Whorf, and Welsh scholar Rhianwen Daniel [23, 24, 25]. This work also serves as an appeal to the Ukrainian academic community to more urgently address the issue of protecting the Ukrainian language from the phenomenon of creolization in the current conditions of global advancement. The purpose of this work is to analyze the historical preconditions and current manifestations of linguistic distortion—caused by Rossification and the influx of mediated Anglicisms—to fully comprehend this influx. We investigate the consequences of these phenomena for the development of the lexical richness and word-formation capacity of the Ukrainian language. Concurrently, we advocate for the necessity of implementing effective measures for its protection. These conclusions are grounded in the empirical analysis of a corpus of words, gathered within the framework of the “Anglo Surzhyk” project [2], which attests to the Rossian-mediated provenance of a considerable portion of these borrowings. Emphasis is placed on the essential role of governmental involvement in the defense and standardization of the Ukrainian language amid globalization and persistent external influence. Conclusions. The cumulative effect of centuries of Rossification, coupled with the percolation of Anglicisms mediated through the Russian language, poses a considerable threat to the evolution of modern Ukrainian. This pervasive process risks the language's creolization and the potential erosion of its distinct identity. Notwithstanding the substantial lexical richness of the Ukrainian vocabulary, there exists an urgent necessity for proactive measures aimed at linguistic protection and norming. The establishment of a specialized state ministry, such as a Ministry for Language Purity—modeled after the French system—is a pivotal step to ensuring the oversight of linguistic standards, the development of specialized terminology, and the preservation of Ukrainian's uniqueness amid global challenges. Furthermore, it is essential to recover and republish dictionaries dating from the 1920’s from archives, universities, libraries, private collections, and even the Security Service of Ukraine (SBU). Ultimately, the defense of the language is not merely a linguistic pursuit but a national priority that underpins cultural and state identity.
This article presents a scholarly analysis of the translation skills of translator Ozod Sharafiddinov. The vivid presentation of lexical richness, stylistic patterns, and expressions in Ozod Sharafiddinov's translations contributed to a clearer and more understandable content for Uzbek readers. He deeply analyzed every phrase and word combination and translated them in accordance with the norms and requirements of the Uzbek language.
Official style has a significant function in administrative, legal, diplomatic, and institutional communication. This research explores the semantic, grammatical, and functional aspects of official style in English and Uzbek languages. Through comparative linguistic analysis, the study determines both similarities and differences in structural organization, lexical choice, and pragmatic features. The results reveal that both languages demonstrate common characteristics such as formality, accuracy, and standardization, while variations are observed in syntactic patterns, terminology, and cultural norms. This research enhances the understanding of official discourse from a cross-linguistic perspective.
This repository contains GSD-NP and GSD-DiNoS, both derived from Universal Dependencies' (UD) GSD Treebank. GSD-NP (.conllu) is a subset of UD-GSD and comprises its simplex noun phrases (NP): Common nouns (NN/NOUN) and their direct dependents (determiners, adnominal adjectives, nmods, adpositions, adverbs). It consists of 49,425 NPs (119.0k tokens) and has an improved feature annotation coverage (gender, case, number). Breaking with UD annotation, a total of 3,649 APPRART tokens were reconstructed in GSD-NP to restore the original orthographic forms. GSD-DiNoS (.json) is a custom data-driven lexion-like data structure built on GSD-NP, which aggregates NPs with the same head lemma. For each lemma, absolute frequencies of the lemma and its word forms are captured. Moreover, each occurrence feeds into three areas of interest within the word form entry: morphosyntactic features in isolation (gender, case, number), in combination with groups of dependents (collocations), and in combination with the syntactic function (dependency relations). GSD-DiNoS spans 17,433 unique lemmas and 20,190 unique word forms, stemming from 49,416 NPs. Lemmas were relemmatised to assign unique lemmas to nominal compounds, a highly productive and often lexicalised construction in German.