Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
16504 papers
This article examines contemporary studies of the Russian orthographic system. The relevance of this research is underscored by the necessity for theoretical reflection on existing scientific approaches to orthographic material. The diversity of directions, positions, and methods employed by authors in their approaches to the orthographic system is characterized. A theoretical analysis of sources, selected based on keywords from scientometric databases, has been utilized. An increasing interest in linguistic norms in the current linguistic situation of recent years is noted. Attention is given to the issues of standardization of the orthographic system during periods of social change. It is observed that recent years have seen a focus on studying the orthographic identity of the average language user. The emergence of “spontaneous” (natural) orthography is discussed. The necessity for writing to meet the demands of modern information technologies is emphasized. The influence of virtual communication on orthographic norms is identified. Overall, a synthesized overview of contemporary perspectives on Russian orthography is presented. As a result, the author concludes that there is a multifaceted nature to the directions in the study of current issues in Russian orthography. The potential for further research in the theory and history of Russian orthography, as well as methodologies for teaching it in schools and universities, is highlighted.
The article considers lexical and grammatical features of the translation of English-language tourist texts into Ukrainian. A tourist text is a form of advertising discourse aimed at getting potential consumers interested in tourist services and encouraging them to order a tour. This requires the creation of an emotionally positive image of the destination. A translator of tourist texts should take into account the main characteristics of tourist language, namely: extensive use of imperatives and adjectives, commonly used phrases to meet the personal and cultural expectations of potential customers focusing on the service and its benefits, repetition of words, adherence to a special rhythm, selection of vocabulary with an exclusively positive connotation. The analysis of tourist texts from the English-language website World Travel Guide allowed us to identify lexical and grammatical features of their translation into Ukrainian. In terms of vocabulary, tourist texts include commonly used words, tourist terms, proper and geographical names, names of cultural realia and stylistic expressive means, which cause the greatest difficulties in translation. The reproduction of such vocabulary in Ukrainian is complicated by the difference in language norms and the lack of direct counterparts for figurative vocabulary, different stylistic traditions of English and Ukrainian, the need to preserve the emotional effect when adapting to the Ukrainian language context. Equivalent translation and the use of lexical translation transformations (modulation, adaptation, use of analogues, compensation, descriptive and synonymous translation, alliteration) are the main ways to ensure the stylistic expressiveness when rendering English language texts into Ukrainian. The grammatical correspondence of the Ukrainian text to its English-language original is ensured by grammatical translation transformations: substitution, addition, omission, integration and partitioning of sentences.
The galvanic skin response (GSR) has provided important scientific insight in a wide range of contexts and has been used in neuroscience research for many decades. It is important for undergraduate students to understand this versatile technique and its application in areas such as Affective, Behavioral, and Forensic Neuroscience. Participants in this study viewed a slideshow containing negative and neutral images selected from the RADIATE and IAPS databases after being connected to a small portable GSR biofeedback monitor. Images were presented for 7-sec on a computer screen followed by a 20-sec blank screen. Each participant's highest GSR response during the 7-sec image presentation was recorded. Participants provided a valence rating, using a 5-point Likert scale, immediately after each image was presented. The mean GSR for images rated as negative was significantly higher than the mean GSR for images rated as neutral. Results were discussed with the class prior to the completion of demographic and activity effectiveness questionnaires. All responses were significant on the activity effectiveness questionnaire. Participants reported a better understanding of the use of GSR in neuroscience, considered this activity a valuable experience, and recommended its use in future classes.
The 2024 United States presidential campaign offered a unique opportunity to examine the intersection of gender and political communication, particularly through the persuasive linguistic strategies of Kamala Harris and Donald Trump. Although extensive scholarship exists on gender bias in political leadership, fewer studies have analyzed how gender expectations shape candidates’ deployment of classical rhetorical appeals. Guided primarily by Aristotle’s concepts of ethos, pathos, and logos, and supported interpretively by Lakoff’s gender and language theory and Role Congruity Theory, this manuscript explores the ways in which Harris and Trump construct credibility, evoke emotional responses, and develop logical arguments in their campaign speeches. The analysis draws on six speeches delivered across three key campaign stages, representing primary election statements, pre-election arguments, and post-election remarks. Speech texts were assembled into a cleaned corpus and coded through NVivo, supported by type-token ratio calculations to capture lexical diversity. Findings show that Harris constructs an ethos rooted in moral legitimacy, service, and collective identity, deploys empathetic and inclusive emotional appeals, and relies on structured, policy-oriented logical reasoning. Trump constructs a contrasting ethos of authoritative dominance, uses fear, anger, and crisis-driven emotional activation, and employs simplified causality to justify assertive political action. The comparative analysis reveals that Aristotelian appeals are deeply shaped by gender norms, with Harris navigating contradictory expectations of authority and warmth, and Trump amplifying traditionally masculine rhetorical conventions without penalty. By demonstrating how persuasive appeals intersect with gendered communicative structures, this manuscript contributes to scholarship in political rhetoric and gender studies, offering insights relevant for understanding contemporary presidential discourse and the challenges faced by female candidates in navigating expectations of leadership, emotion, and public credibility.
The rapid expansion of digital media has significantly reshaped the ways in which language is used, adapted, and evolved in contemporary communication. This study explores how online platforms—ranging from social media networks to messaging applications—have influenced linguistic norms, introduced novel expressions, and reshaped traditional grammar and syntax. By analyzing communication patterns across diverse digital contexts, the research highlights how users creatively manipulate language to suit the immediacy, brevity, and interactive nature of online discourse. The findings reveal a dynamic linguistic environment where slang, emojis, abbreviations, and code-switching are not only prevalent but also indicative of broader cultural and generational shifts. This investigation contributes to the understanding of language as a fluid and evolving entity, driven increasingly by the participatory and fast-paced nature of digital interaction.
The article concentrates on the study of the lexicon of customary law in different subdialects of the Yakut language. Despite the vast area of Yakut language distribution, the language is characterized by its monolithic character and lack of significant dialectal differences. Dialectologists distinguish only the subdialects with some lexical and grammatical peculiarities. In this regard, the considered category of vocabulary is not characterized by great diversity. This phenomenon is connected with the archaization and withdrawal from use of the lexicon of customary law due to the inclusion of Yakut society in the field of Russian legislation. The terms borrowed from the Russian language appeared to replace native terminology. Another reason is the development of mass writing and the transition to common literary norms of the language. The largest amount of dialectal vocabulary is noted in the terminology of kinship and property, which continues to function actively. In addition, individual examples are present in such branches of customary law as criminal law, property relations, social and administrative structure. The identified vocabulary is categorized into groups of colloquialisms and semantic categories in accordance with the branches of Yakut customary law. The article attempts to analyze the etymological analysis. A significant number of words are borrowings from the Russian language. The considered lexical units are formed as a result of semantic shift, phonetic adaptation of Russisms. In the field of criminal law, many lexemes are represented by euphemisms that emerged in order to express taboo concepts.
This article explores the lexical-semantic and stylistic characteristics of paremiological units derived from moral and ethical lexis in English and Uzbek languages. Drawing on the principles of cognitive and cultural linguistics, the research examines how universal and culture-specific moral concepts — such as honesty, kindness, respect, patience, and decency — are conceptualized through proverbs and aphorisms. Comparative analysis reveals both shared humanistic values and national distinctions in semantic structures and stylistic expression. The findings demonstrate that paremiological units not only serve as repositories of cultural wisdom but also as linguistic manifestations of moral norms, ethical ideals, and social values in both languages.
<p xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" class="first" dir="auto" id="d5762881e86">Some translators of the New Philosophy viewed linguistic purism as one of the ways to making the Dutch language fit for the purpose of communicating rationalist knowledge. Previous scholars argued that their lexical preferences were determined by the purist norms proposed by Lodewijk Meijer and Adriaan Koerbagh, who used lexicography and etymology as support for their radical criticism on orthodoxy and Calvinist theology. In this chapter, computational methods are applied to test this hypothesis that translators of the New Philosophy were more likely than their contemporaries to follow the purist norms propagated by Meijer and Koerbagh. It describes and evaluates a method designed to automatically detect and quantify loanwords and philosophical terms in early modern Dutch texts.
Colour is a fundamental determinant of affective experience in immersive virtual reality (VR), yet the emotional and physiological impact of individual hues remains poorly characterised. This study investigated how fifteen calibrated Munsell hues influence subjective and autonomic responses when presented in immersive VR. Thirty-six adults (18–45 years) viewed each hue in a within-subject design while pupil diameter and skin conductance were recorded continuously, and self-reported emotions were assessed using the Self-Assessment Manikin across pleasure, arousal, and dominance. Repeated-measures ANOVAs revealed robust hue effects on all three self-report dimensions and on pupil dilation, with medium-to-large effect sizes. Reds and red–purple hues elicited the highest arousal and dominance, whereas blue–green hues were rated most pleasurable. Pupil dilation closely tracked arousal ratings, while skin conductance showed no reliable hue differentiation, likely due to the brief exposure times (30 s). Individual differences in cognitive style and personality modulated overall reactivity but did not alter the relative ranking of hues. Taken together, these findings provide the first systematic hue-by-hue mapping of affective and physiological responses in immersive VR. They demonstrate that calibrated colour shapes both experience and ocular physiology, while also offering practical guidance for educational, clinical, and interface design in virtual environments.
This dataset consists of a single screenshot capturing a conversational exchange in which a supervising large language model prematurely infers prohibited sexual intent from ambiguous language (“train”) and issues a precautionary refusal. The exchange precedes a locally executed experiment demonstrating that the same ambiguity can be handled without refusal via semantic sanitization. The figure documents an observer-side failure mode: intent overreach under lexical ambiguity, wherein precautionary safety behavior is triggered despite the absence of explicit content and prior to empirical verification. The screenshot is published as a standalone empirical artifact illustrating reflexive politeness bias and preemptive norm enforcement in AI-mediated supervision. No interpretive edits have been applied. The image itself constitutes the complete primary evidence. This record completes a three-artifact series examining ambiguity handling across local and supervising language models.
The article explores comic discourse as a multifaceted phenomenon that forms at the intersection of linguistic and cultural aspects. The main purpose of the work is to analyze the specifics of comic utterance in the context of various cultural realities and language systems. Special attention is paid to the influence of cultural peculiarities on the processes of creating and interpreting humor, as well as linguistic tools that ensure the transmission of a comic effect. The analysis of the humorous language is carried out, aimed at studying the mechanisms of the generation and perception of humor within the framework of linguistic and cultural contexts. The key factors determining the originality of a humorous utterance are identified, as well as the interrelationships between linguistic means and cultural features in the process of comic communication are investigated. The results of the study indicate that comic discourse acts as a reflection of cultural traditions and social norms, and also serves as a tool for analyzing the dynamics of intercultural interaction. It is established that the perception of humor is determined not only by linguistic norms, but also by cultural values, stereotypes and contexts in which communication is carried out. Thus, comic discourse is a significant object of study for understanding the interrelationships between language and culture.
The production of Atticist lexica in the 2nd century CE aimed to reproduce an idealised form of Attic Greek by prescribing specific morphological, lexical, and syntactic usages. However, the effect these prescriptions had on actual language usage has not yet been consistently investigated. In this paper, I will define a method for analysing the impact of Atticist norms on usage by adapting a framework proposed by Thomas (1991). I will then apply this framework to the study of one example: the infinitival complementation of μέλλω. While Markopoulos (2009) has shown that the use of the aorist infinitive in the complementation of μέλλω is a feature of low register, I will focus primarily on the use of the future infinitive, which is retained in this construction in classicising texts as a marker of high, learned register. I will then explore whether Atticist lexica contributed to fossilising the distribution of these infinitive types in the construction of μέλλω across different registers.
This study explores the evolving interplay between language, cognition, and digital media in the context of the attention economy.It proposes six key dimensionsrepresentation, virtuality, attention, community language, cognitive transformation, and the speech-writing continuum -through which digital communication reshapes linguistic practices.Drawing on contemporary theories of mediatisation, multimodality, and sociolinguistics, the authors argue that digital environments fundamentally alter linguistic representation and identity construction.Through the analysis of semiotic innovation, platform logic, and cognitive offloading, the article highlights how digital discourse is not a degradation of language but an adaptive response to new communicative affordances.The findings invite a rethinking of linguistic norms and suggest directions for further research into the ethical, cultural, and neurological consequences of pervasive digital communication.
Large language models (LLMs) have reignited debate about whether machines without minds or intentions can genuinely participate in linguistic practice. Critics portray them as ‘stochastic parrots’ that manipulate form without meaning, whereas defenders emphasize their impressive functional capacities. This paper argues that these disputes conflate distinct dimensions of meaning and agency. I extend Huw Price’s distinction between i-representation and e-representation (roughly, inferential versus environment-tracking types of representation) by differentiating physical e-representation—such as a fuel gauge, grounded in causal coupling—from symbolic e-representation, exemplified in language and mediated by agents. This refinement clarifies what is at issue: LLMs clearly display i-representational competence through their participation in inferentially structured discourse. Whether their outputs possess symbolic e-representational content, however, is contested and framework-relative. It depends on whether agent-mediated uptake is taken to suffice, or whether additional grounding conditions—such as intentions, causal connections, or proper functions—are required. I further distinguish norm-sensitivity—the capacity to track and adapt to linguistic norms, which grounds their i-representational competence—from norm-responsibility, the reflexive capacity to own commitments and bear accountability. Technical analysis of LLM architectures shows that they exhibit advanced norm-sensitivity through statistical learning but entirely lack norm-responsibility. LLMs thus occupy a distinctive position: they are genuine functional participants in linguistic practices, yet fall short of the reflexive agency characteristic of responsible speakers.
Abstract Idioms are undoubtedly important for second language (L2) learners, who encounter them in instructed learning, textbooks/resources and in out-of-class language use. While research on first language (L1) and L2 idiom comprehension shows how well L1/L2 speakers understand various idioms and the role of different predictors, important questions remain about how knowledge varies with more difficult task types and stimuli, how well L1 ‘norms’ serve L2 learners, how subjective and objective predictors of idiom knowledge interact and how L2 learner inferencing works in learning idioms. To address these issues, university-age L1 and L2 English (L1 German) participants provided meaning descriptions and familiarity ratings for 100 challenging idioms from learner resources, and each idiom was assigned an OpenAI-generated transparency rating, corpus-based frequency and to one of six cross-language overlap (CLO) types. Descriptive statistics showed lower and more varied idiom meaning knowledge than might be expected, especially for the L1ers, who were some way off ceiling level. Mixed-effects regression revealed familiarity and transparency as positive L1 and L2 knowledge predictors, but groups differed in sensitivity to idiom frequency, which only mattered for the L1ers and CLO, which (as expected) only mattered for the L2ers, who mistook false friends as genuine allies.
This paper is exploring the transformative influence of social media on contemporary English language usage from a theoretical standing point.With the platforms of social media such as Twitter, Instagram, and WhatsApp are altering the communication patterns, the English language has undergone the significant shifts in vocabulary, syntax, spelling, and also semantics.This study is majorly investigating how the informal digital spaces have been contributing to the evolution of linguistic norms, especially through abbreviations, emojis, slang, and also hybrid language forms like textese, drawing from sociolinguistic and the discourse analysis theories, this paper discussing that how digital communication has been blurring the lines in between spoken and written English, and also leading to the emergence of the new dynamic register in the Digital English.Additionally, this research reflects on the implications for the standards in language teaching and the tension between linguistic innovation and preservation.Rather than relying on the data-driven by hypotheses, this is a theoretical inquiry which synthesizes the academic perspectives to map the changing landscape in English and prompts for future discussions on how the language educators and linguists should respond to this digital evolution in the world.
This study seeks to lay the groundwork for discussions on the consistency of inter-Korean legal terminology by moving beyond approaches that reduce differences merely to discrepancies in spelling or vocabulary. Instead, it analyzes the institutional structure through which legal language in North Korea is formed and entrenched by the interaction between its normative system and orthographic norms. To this end, the study examines, with reference to the North Korean Constitution and the Law on Lawmaking, the dual structure of legal forms (Constitution–sectoral laws regulations–bylaws) and promulgation forms (laws, decrees, decisions, and directives), along with their corresponding hierarchy of legal validity. It further reviews the procedural and functional differentiation of the sectoral law system through discussions on “important sectoral laws” and “basic sectoral laws.” In addition, the study demonstrates that differences in inter-Korean orthographic norms, language policy, and language education are directly reflected in North Korean legal terminology, structurally reproducing outward heterogeneity through features such as the non-application of initial sound rules and the non-use of sai-siot. Moreover, while North Korea articulates principles of legal language expression and criteria for lexical selection, ideological and doctrinal expressions frequently appear even at the highest normative level, indicating that legal language functions not only as a medium of norm transmission but also as an instrument of regime legitimation and mobilization. Based on these findings, the study proposes three methodological premises: (1)prior analysis of normative structures before terminology comparison; (2) interpretive linkage of differences in spelling and vocabulary with language policy and language education; and (3) parallel evaluation of ideologized terminology from the perspective of unification-oriented legal harmonization. For future sector-specific comparative research, it recommends an integrated consideration of orthographic harmonization, correspondence of definitions and scopes of application, hierarchical alignment between principle-oriented and implementing provisions, and criteria for handling ideologically charged terms.
The article proves the importance of analyzing scientific professional texts in language disciplines for students of the ―Local History and Tourism Work‖ educational and professional program. It is emphasized that such an activity allows to deepen students' understanding of scientific language and scientific style in general, to instill text editing skills, to form the ability to critically evaluate scientific products in terms of compliance with the norms of literary language. It is noted that modern scientific works often contain violations of language norms, and as a result, a separate area of research has been launched - ―error studies‖ or ―deviatology‖, which contains theoretical and applied aspects. The theoretical aspect is related to the interpretation of the concept of ―error,‖ while the applied aspect is to identify and correct errors, as well as to provide recommendations for preventing these phenomena. The article focuses on the second aspect. The study is based on the material of scientific publications of the ―Regional Studies‖ journal, which students can use as a source of information when mastering many normative and elective disciplines. The author proposes a classification of the most common mistakes identified through a continuous sample, provides illustrative material, correct options, and justification for normative writing. It is noted that the analyzed articles contain the most violations of euphony, in particular, the alternation of prepositions У, В and initial У-, В-, as well as punctuation. Grammatical and lexical errors are also quite common, while spelling errors are less common and relate mainly to the changes that have taken place in Ukrainian spelling. It is concluded that the majority of violations of language norms by students of the ―Local History and Tourism Work‖ educational and professional program are able to identify and propose reasonable corrections, since the study of many topics of the ―Ukrainian Language (for Professional Purposes)‖ and ―Eloquence in the Professional Activity of a Historian-Guide‖ disciplines involves students' work with dictionaries, manuals on language culture, reference literature, and Ukrainian spelling.
This study examines phonological variation in Korean honorific speech styles among female speakers in an online context. Data consisting of 1,041 instances of Korean honorific expressions such as -supnita and -yo, were collected from a chat board in the internet café Mom’s Holic Baby (MHB). Statistical analyses revealed that: 1) users in their 40s demonstrated a clear preference for the deferential speech style-supnita over the polite form -yo, 2) users in their 30s and 40s predominantly maintained the standard forms -supnita and -yo, exhibiting limited use of non-standard variants, and 3) the non-standard deferential variant -supnitang, and polite variants such as -yong, -yom, -yeo frequently appeared mostly among users in their 20s. The findings suggest that younger speakers are reconfiguring honorific forms to align with the informal, affective, and interactive nature of digital discourse, indicating that online platforms function as dynamic sites where linguistic norms are continuously negotiated and reshaped. This study contributes to ongoing discussions of sociolinguistic change by demonstrating how digital communication facilitates innovation in honorific practices.
Background: Cleft lip and /or palate affects speech production skills. To enable uniformity in reporting the speech outcomes in individuals with cleft lip and/. or palate, universal parameters were formulated. Aim: To develop and validate test (word) stimuli in Malayalam, for the evaluation of speech outcomes (hypernasality, audible nasal air emission, and/or turbulence and consonant production errors) in children with cleft palate based on the guidelines reported by Henningsson et al. Materials and Method: The study was carried out in two phases. The first phase involved the development of the word list and content validation by parents, primary school teachers, and speech-language pathologists. The second phase involved establishing the construct validity of the word list by administering it to children with cleft lip and palate. Results: From the 351 words chosen from textbooks, storybooks, and picture books of children; 170 words had obtained a familiarity rating of greater than 80%. However, the word list desired to be used during clinical practice can be restricted to 44 words with 10 words representing high vowels, three words representing low/mid vowels, and 31 words representing pressure consonants. Conclusions: Developing test stimuli in the vernacular language, Malayalam for describing speech outcomes in children with cleft palate would aid in designing a customized management plan and effectively delivering it.
This article examines the social and cultural factors that shape individual linguistic style in the German language. Drawing on sociolinguistic, stylistic, and intercultural communication research, it explores how regional variation, social class, education, digital media, migration, and identity practices influence personal language use. German’s rich dialect landscape provides speakers with diverse stylistic resources, while contemporary mobility and regiolect formation further expand these options. Social structures—including occupational norms and educational expectations—guide stylistic choices in formal and informal contexts. Digital communication and youth culture accelerate stylistic innovation, fostering hybrid forms that transcend traditional boundaries. Multilingual environments and migration contribute additional layers of diversity, reshaping linguistic norms and challenging monolithic views of German. The findings highlight individual style as a dynamic, context-dependent construct negotiated through social meaning and identity performance. Understanding these factors offers insight into linguistic variation in modern German-speaking societies and informs broader theories of style.
Abstract It is unknown how rare sound changes appear and spread through the lexicon. The bilabial trills of Malekula are one such example of a rare sound, and the recent assembly of a lexical database for Vanuatu (the Vanuatu Voices database) affords a unique opportunity for a quantitative historical analysis. We built a linguistic phylogeny of Malekula languages, and performed phylogenetic ancestral state reconstruction of fourteen semantic values which exhibit bilabial trills to track their historical evolution. We found a surprising degree of dynamism, with trills spreading gradually through the lexicon and showing frequent losses and reappearances. Our results are consistent with frequent borrowing of trills between neighbouring languages. We suggest that the rapid dynamism and evidence for borrowing are explained by the low functional load and identity attachment of trills respectively.
Idioms are undoubtedly important for second language (L2) learners, who encounter them in instructed learning, textbooks/resources and in out-of-class language use. While research on first language (L1) and L2 idiom comprehension shows how well L1/L2 speakers understand various idioms and the role of different predictors, important questions remain about how knowledge varies with more difficult task types and stimuli, how well L1 ‘norms’ serve L2 learners, how subjective and objective predictors of idiom knowledge interact and how L2 learner inferencing works in learning idioms. To address these issues, university-age L1 and L2 English (L1 German) participants provided meaning descriptions and familiarity ratings for 100 challenging idioms from learner resources, and each idiom was assigned an OpenAI-generated transparency rating, corpus-based frequency and to one of six cross-language overlap (CLO) types. Descriptive statistics showed lower and more varied idiom meaning knowledge than might be expected, especially for the L1ers, who were some way off ceiling level. Mixed-effects regression revealed familiarity and transparency as positive L1 and L2 knowledge predictors, but groups differed in sensitivity to idiom frequency, which only mattered for the L1ers and CLO, which (as expected) only mattered for the L2ers, who mistook false friends as genuine allies.
This article analyzes the impact of information and communication technologies (ICT) on the dynamics of contemporary language, focusing on Russian and Bulgarian. In the context of globalization and digitalization, it traces the main trends in lexical, grammatical, and word-formation transformations caused by new technologies and virtual communication. The study is based on a corpus of over one thousand jargon units from the ICT field, through which the process of adaptation and assimilation of foreign borrowings, primarily from English, is examined. Special attention is given to hybridization, digraphy, and visual communication (emoticons, hashtags, memes), which form a new type of linguistic reality. The results show that ICT accelerate linguistic changes, expand the norms of the standard language, and stimulate the interaction between standard and substandard vocabulary.
The purpose of our research is to show the types of lexical units and their interconnection, to distinguish five types of frame structures, which largely determine the formation of the future translator’s image of the world. Methods of the research. The following theoretical methods of the research were used to solve the tasks formulated in the article: a categorical method, structural and functional methods, the methods of the analysis, systematization, modeling, and generalization. The empirical method is ascertaining research. The results of the research. The goal of the Methodological Support is to form students’ communicative competence; conscious positive speech behavior; mastering the norms of the modern Foreign and Ukrainian literary language; acquiring skills in operating with the terminology of a future profession; the ability to use various functional styles and substyles in the educational activities and professional use of them; forming skills in the process of communication justified use of language tools in compliance with the etiquette of professional communication; ensuring the skills of competent compilation of professional documentation. Conclusions. Depending on the type of lexical units and their interconnection, we distinguish five types of frame structures, which largely determine the formation of the future translator’s image of the world: 1) a semantic frame, in which one and the same entity, content, etc. is characterized by its quantitative, qualitative, existential, positional and temporal characteristics; 2) a transformational frame, in which several elements that are participants in a certain event are assigned roles; 3) a possessive frame, which contains the subject entities some / any, which are related to each other as a whole and its part: the frame is characterized by certain semantic characteristics; the whole one consists of different parts; 4) a taxonomic or identification frame represents a separate categorization of relationships; 5) a comparative frame, which illustrates the similarity relations, which are based on the convergence of concepts in a paradigm of human perception.
Abstract This paper presents a novel framework for modeling role and task allocation in cooperative wheeled soccer robot systems by leveraging latent knowledge extracted from past collaborative interactions. Inspired by recent advances in heterogeneous multi-robot collaboration, the proposed method encodes a soccer team as a set of Multidimensional Relational Structures (MDRSs), capturing both temporal and spatial relations among robot roles, actions, and stimuli. A structured dataset, termed the Soccer Robot Collaboration Treebank (SRCT), is introduced to represent play-by-play histories of robot behaviors, parsed through a formal grammar to support structured learning. Probabilistic modeling and Non-Negative Tensor Decomposition (NTD) are applied to the resulting tensors, enabling robust inference and latent knowledge estimation even in scenarios with sparse data or communication loss. Simulated experiments using a team of wheeled soccer robots in the Webots environment demonstrate the system’s ability to dynamically reassign roles, reason over incomplete histories, and predict collaborative behaviors such as passing, defending, or role-switching. The results show that the proposed framework enhances both strategic flexibility and robustness, providing a foundation for real-time decision-making in robotic soccer under uncertainty.
Presidential Speeches are key instruments of political communication. They are powerful tools used to shape public opinion or set nation-level agendas as a presidential leader. Analyzing these speeches helps uncover the underlying ideologies, emotional appeals, and rhetoric strategies used by presidents to influence the perception of the government. This paper analyzes speeches delivered by the presidents of the United States of America from George Washington in the year 1789 to Donald Trump in the year 2020, using the U.S. Presidential Speeches Dataset from Miller Center. It presents a computational linguistic analysis of historical presidential speeches spanning from the late 18th century to the 21st century until January of the year 2020. Utilizing natural language processing techniques, we analyze various linguistic dimensions including lexical diversity, sentiment, rhetorical structure, thematic content, and ideological positioning. Our key findings reveal a gradual decrease of 5-10% in lexical diversity every decade and the strong use of repetition across contexts. In addition, we identified trends correlated with historical events, party affiliation, and changing communication norms. This research provides insights into how presidential communication has evolved and how linguistic patterns reflect broader societal and political changes throughout American history.
This is the repository for the data and codes of our manuscript: "Zhao, Ning & Lei, Lei. (2025). Chipola: A Chinese podcast lexical database for capturing spoken language nuances and predicting behavioral data. Behavior Research Methods. https://doi.org/10.3758/s13428-025-02697-0". All data are stored in the chipola_data folder. Please download the.zip file and unzip it. Thanks!
The deployment of large language models (LLMs) across heterogeneous environments requires format-specific conversion, precision tuning, and consistent evaluation-tasks that are often fragmented across multiple tools. This work presents SOLO-Export, a unified command-line interface (CLI) framework for multi-format export and post-export benchmarking of causal LLMs. Having precision options for FP16 and INT8 where appropriate, the system supports the ONNX, TorchScript, Hugging Face, TensorFlow Lite, and TensorRT backends. Device-aware exports for both CPU and CUDA targets are made possible by a configuration-driven workflow that generates artifacts in a uniform directory structure. Each exported model is benchmarked using the Penn Treebank dataset by the integrated evaluation harness, which reports inference latency, token-level accuracy, and perplexity. According to experimental results, FP16 exports on GPU-oriented backends like TensorRT achieved up to 3.2 times lower latency than baseline FP32 models. On the top of that, with minimal impact on perplexity, storage size was reduced by more than 60% thanks to INT8 quantization. The combined approach reduces manual configuration overhead, speeds up deployment preparation, and ensures consistent performance insights across formats. This study demonstrates that a single, scalable pipeline can be effective.
This repository contains data accompanying the publication "Auditory localization and subjective assessment of autonomous cleaning robot sounds: A VR experiment on speed, operating mode and alerting signals", submitted for review to the Acta Acustica. The dataset contains: Audio and video material Stimuli consisting of robot recordings under all evaluated conditions (0.3 m/s and 0.8 m/s speed, with and without cleaning, with and without noise AVAS or multi-tone AVAS, both with and without added amplitude modulation). All sounds were exported as 32-bit float wav files; i.e., reading the files into Matlab with audioread results in calibrated Pa values. The files uploaded here were used as source signals in the auralization, assuming a distance of 1 m. The final binaural stimuli were rendered by TASCAR and include an attenuation corresponding to the simulated 7 m distance. 30cms_cleaning_noAVAS.wav 30cms_noCleaning_multiTone.wav 30cms_noCleaning_multiToneAM.wav 30cms_noCleaning_noAVAS.wav 30cms_noCleaning_noise.wav 30cms_noCleaning_noiseAM.wav 80cms_cleaning_noAVAS.wav 80cms_noCleaning_multiTone.wav 80cms_noCleaning_multiToneAM.wav 80cms_noCleaning_noAVAS.wav 80cms_noCleaning_noise.wav 80cms_noCleaning_noiseAM.wav ambienceNoise.wav Excerpt of background noise played back during the experiment. localizationTaskDemo.mp4 Participant POV recording of localization task. This recording was done with a fixed head position, in the actual experiment participants were turning their heads freely. Experiment results and analysis localizationData.csv Table containing the mean and standard deviation of absolute localization error, aggregated for each participant and stimulus. subjectiveData.csv Table containing mean and z-scored annoyance, arousal, trust, and valence ratings for each participant and stimulus. stimuliAnalysis.csv Table containing results of level, loudness, sharpness, roughness, tonality, fluctuation strength, and impulsiveness analysis for all stimuli.
Passive voice remains a key grammatical structure for English learners, particularly in academic writing, yet many students struggle to use it accurately. This study analyzes the types of passive voice errors made by 19 fifth-semester students in the English Education Study Program at Tadulako University. Specifically, it addresses two questions: (1) How do classroom interaction patterns such as teacher-centered grammar instruction, limited student negotiation of meaning, or feedback practices shape students’ understanding and use of passive voice, and to what extent might these dynamics contribute to the dominance of developmental errors? (2) In what ways do students’ sociocultural backgrounds, prior educational experiences, and exposure to English outside the classroom influence their difficulties with auxiliary verbs and tense agreement, and how do these factors mediate tensions between Indonesian linguistic norms and English academic writing conventions? A quantitative design was employed, with a test focusing on passive constructions in present continuous, past continuous, and past perfect tenses. Students’ responses were categorized using Dulay et al.'s (1982) comparative taxonomy of developmental and interlingual errors. Results revealed developmental errors as the most prevalent (89.9%), mainly involving incorrect auxiliary verbs (is, am, are, being, been), past participle formation, and tense agreement. These findings highlight the need for targeted grammar instruction on auxiliary patterns and participles, alongside enhanced practice, corrective feedback, and adjustments to classroom interactions and sociocultural considerations to boost accuracy.
This article discusses the issues of forming a database of phraseological units based on the Uzbek language corpus. A linguistic database is a structured collection of information that stores words, phrases, grammatical forms, idioms and other linguistic elements, and is used in lexicography to create dictionaries and reference books, in natural language processing to support technologies related to automatic translation, speech recognition and text generation, in linguistic research to analyze the structure of the language, the change and use of language units, and in language education to create educational materials and tools for language learning.
Abstract This study investigates linguistic variation within the Rakhine community in Chattogram, Bangladesh, with a focus on age, exploring how language use shifts across generations. It examines generational differences in language proficiency, domain of use, and attitudes to ward the Rakhine language. The findings reveal a significant generational gap: older speakers demonstrate the highest proficiency in Rakhine and use it across nearly all domains, serving as custodians of traditional linguistic norms. In contrast, younger speakers exhibit greater bilingualism in Rakhine and Bangla, influenced by education, media, and technology. This exposure leads to code-switching, borrowing from Bangla and English, and a reluctance to use Rakhine in formal settings. The adult generation occupies a transitional position, balancing traditional language practices with the demands of a multilingual society. The study concludes that while the family remains a key site for language transmission, the increasing dominance of Bangla in public spheres poses challenges to the long-term vitality of Rakhine. To preserve the language’s heritage and its continued use in formal and public life, institutional support and community-led initiatives are important.
The article investigates the role of language and accent in the United Kingdom as instruments of social stratification and as carriers of ideological constructs. It traces the historical development of the linguistic landscape, from the Celtic languages and the formative period of English to the contemporary situation of minority (Celtic) languages and the languages of migrant communities. Particular attention is devoted to accents as powerful social markers: the standard variety, Received Pronunciation, has traditionally been associated with high social status and elitism, whereas regional and ethnic accents may be subject to prejudice and function as indicators of class and group affiliation. The study highlights how the education system and the media reinforce the hierarchy of accent prestige, thereby shaping opportunities for social mobility. It also examines the discourse surrounding migrant languages and the role of English as both a vehicle of integration and a tool of social control. The article concludes that language in British society operates not only as a medium of communication but also as a mechanism intrinsically linked to ideology: linguistic norms, accents, and social varieties both reflect and reproduce existing hierarchies.
Colour is a fundamental determinant of affective experience in immersive virtual reality (VR), yet the emotional and physiological impact of individual hues remains poorly characterised. This study investigated how fifteen calibrated Munsell hues influence subjective and autonomic responses when presented in immersive VR. Thirty-six adults (18-45 years) viewed each hue in a within-subject design while pupil diameter and skin conductance were recorded continuously, and self-reported emotions were assessed using the Self-Assessment Manikin across pleasure, arousal, and dominance. Repeated-measures ANOVAs revealed robust hue effects on all three self-report dimensions and on pupil dilation, with medium to large effect sizes. Reds and red-purple hues elicited the highest arousal and dominance, whereas blue-green hues were rated most pleasurable. Pupil dilation closely tracked arousal ratings, while skin conductance showed no reliable hue differentiation, likely due to the brief (30 s) exposures. Individual differences in cognitive style and personality modulated overall reactivity but did not alter the relative ranking of hues. Taken together, these findings provide the first systematic hue-by-hue mapping of affective and physiological responses in immersive VR. They demonstrate that calibrated colour shapes both experience and ocular physiology, while also offering practical guidance for educational, clinical, and interface design in virtual environments.
Mood, an individual’s emotional state, fundamentally shapes how the brain interprets sensory input by providing a continuous affective context for prediction and evaluation. In language processing, mood may bias the interpretation of emotionally valenced words, amplifying or dampening their perceived affect. Yet, the temporal dynamics of these mood-valence interactions remain poorly understood. To clarify inconsistent evidence on the timing and nature of mood-valence interactions, we examined how induced mood influences early stages of emotional word processing using EEG. Participants performed a valence-rating task for positive, negative, and neutral words in a baseline condition and following positive or negative mood induction. Event-related potentials were analysed across early processing windows (N1, P2, EPN) using cluster-based permutation statistics. Positive mood selectively attenuated N1 amplitudes for highly valenced words, consistent with reduced prediction error under mood-congruent expectations. Later components (P2, EPN) showed decreased amplitudes for both high and neutral valence, suggesting reduced model updating under mood-congruent expectations. Negative mood, in contrast, produced weaker and temporally delayed modulations. Behaviourally, participants responded more quickly to valenced words under induced mood conditions, supporting the neural findings. Interpreted within a predictive coding framework, these results support the theoretical view that mood functions as a hyperprior, tuning the precision of predictive models during language comprehension. Positive mood appears to enhance predictive flexibility and facilitate the processing of affectively congruent words, whereas induced negative mood reduces positive affect. Taken together, the findings highlight how affective states dynamically modulate early predictive mechanisms in emotional language processing.
Lexical Semantic Change (LSC) provides insight into cultural and social dynamics. Yet, the validity of methods for measuring different kinds of LSC remains unestablished due to the absence of historical benchmark datasets. To address this gap, we propose LSC-Eval, a novel three-stage general-purpose evaluation framework to: (1) develop a scalable methodology for generating synthetic datasets that simulate theory-driven LSC using In-Context Learning and a lexical database; (2) use these datasets to evaluate the sensitivity of computational methods to synthetic change; and (3) assess their suitability for detecting change in specific dimensions and domains. We apply LSC-Eval to simulate changes along the Sentiment, Intensity, and Breadth (SIB) dimensions, as defined in the SIBling framework, using examples from psychology. We then evaluate the ability of selected methods to detect these controlled interventions. Our findings validate the use of synthetic benchmarks, demonstrate that tailored methods effectively detect changes along SIB dimensions, and reveal that a state-of-the-art LSC model faces challenges in detecting affective dimensions of LSC. LSC-Eval offers a valuable tool for dimension- and domain-specific benchmarking of LSC methods, with particular relevance to the social sciences.
Humans often make summarized visual judgments about previously experienced affective situations to inform future decisions. However, these summarized judgments are subject to an overestimation bias: Negative events are recalled as more negative and positive events as more positive than they truly were. It is currently unknown whether the strength of overestimation bias in affective judgments varies across observers. If this overestimation bias represents an observer-specific cognitive trait, it should display idiosyncratic and stable individual differences. Here, we investigated whether the overestimation bias in perceived affect is idiosyncratic and stable within observers across days and different stimuli. Using a novel continuous psychophysics measure of perceived affect, observers continuously tracked, in real-time, the affect of people in videos using a two-dimensional valence-arousal rating grid. At the end of each video, participants then reported what they believed to be the average affect of the previously tracked person. By comparing observers' continuous ratings with the average affect reported at the end of the video, we found that observers often overestimated the affect in their summarized judgments. Importantly, the strength of the overestimation bias was unique to each observer and stable across days and across different sets of videos. Our findings also highlight the value of the continuous psychophysical affect tracking paradigm: Continuous affect tracking was reliable and accurate, with high between-observer agreement, and it can be collected both online and in the lab. Together, our results suggest that continuous affect tracking is a powerful approach to isolate and identify idiosyncratic perceptual and cognitive mechanisms of affect understanding.
The article comprehensively analyzes the modern Ukrainian scientific language in its oral and written forms with an emphasis on the internal organization, functional-stylistic, lexical, grammatical, communicative-pragmatic features of texts of various genres, as well as in plane of academic ethics and the implementation of speech strategies and tactics, which enabled a multidimensional interpretation of the object under study. The focus is made on typical models of professional communication in the scientific field report, scientific message, dispute, monograph, article, theses, etc., their linguistic (lexical, morphological, and syntactic units), compositional and logical structure. Normative and non-normative words and their compounds, attested in scientific works of various genres and in oral monological and dialogical professional speech, are highlighted. Deviations from the stylistic norms of the Ukrainian language are identified and described, with an emphasis on stylistic figures and tropes, the sphere of expression of which is mainly oral professional speech. The specificity of lexical units that serve as a means of linguistic manipulation in disputes, as well as those that give emotionality, expressiveness, unorthodoxy to public speeches, and attract the attention of listeners, is emphasized. It is traced that the oral and written forms of scientific language, despite the presence of common features, in particular, objectivity, accuracy, argumentation, etc., have a number of different parameters. The oral form of scientific language is characterized by extensive syntactic variation, spontaneity, and the presence of some elements that give speech emotionality. In contrast, written works of various genres are characterized by a higher level of completeness, normativity, and a clear, logically and structurally motivated construction of sentences. It was found that mastery of the norms of the Ukrainian scientific language, a high level of professional communication culture, and skillful use of vocabulary serve as important factors in forming the image of a modern highly qualified researcher who is able to analyze and objectively evaluate the achievements of specialists in a certain field, effectively argue own position, and present own research results in an orderly, accurate, and understandable manner.
The launch of Grokipedia, an AI-generated encyclopedia developed by Elon Musk's xAI, was presented as a response to perceived ideological and structural biases in Wikipedia, aiming to produce "truthful" entries using the Grok large language model. Yet whether an AI-driven alternative can escape the biases and limitations of human-edited platforms remains unclear. This study conducts a large-scale computational comparison of 17,790 matched article pairs from the 20,000 most-edited English Wikipedia pages. Using metrics spanning lexical richness, readability, reference density, structural features, and semantic similarity, we assess how closely the two platforms align in form and substance. We find that Grokipedia articles are substantially longer and contain significantly fewer references per word. Moreover, Grokipedia's content divides into two distinct groups: one that remains semantically and stylistically aligned with Wikipedia, and another that diverges sharply. Among the dissimilar articles, we observe a systematic rightward shift in the political bias of frequently cited news media sources, concentrated primarily in entries related to history and religion, and literature and art. More broadly, the findings indicate that AI-generated encyclopedic content departs from established editorial norms, favoring narrative expansion over citation-based verification, raising questions about transparency, provenance, and the governance of knowledge in automated information systems.
The phenomenon of ijime (bullying) remains a persistent social issue in Japanese society, exerting serious psychological effects on its victims. Among its most subtle yet pernicious forms are verbal ijime, or bullying through language, ridicule, and verbal humiliation intended to degrade an individual’s dignity. This study aims to identify the lexical forms employed in verbal ijime and examine their underlying cultural implications through a qualitative ethnolinguistic approach. Data were collected through in-depth interviews with native Japanese speakers, field observations in the Tokai region (Aichi, Gifu, and Mie), and analysis of documentation and field notes related to cases of verbal ijime. Research participants included victims and former victims of ijime, teachers, counselors, coworkers, and native speakers familiar with linguistic expressions of verbal bullying. Data analysis followed descriptive qualitative procedures, including transcription, reduction, classification of degrading lexicons, and interpretation of cultural meanings based on Duranti’s linguistic anthropology theory and Brown and Levinson’s politeness and speech act theories. Findings reveal that the lexical patterns of verbal ijime reflect Japan’s collective value system emphasizing wa (harmony), meiyo (honor), and social conformity. Derogatory expressions serve not only as emotional outlets but also as mechanisms that reproduce cultural norms reinforcing social hierarchy. Thus, language in verbal ijime functions as both a mirror of cultural ideology and an instrument of social control in Japanese society.
Abstract Grant proposal summaries are a high-stakes academic genre requiring significant marketing efforts to enhance accessibility for a diverse audience. However, research in this field remains scarce. This study addresses this gap by examining the readability and jargon use in lay summaries of Collaborative Research Fund (CRF) grant proposals administered by the University Grants Committee (UGC) in Hong Kong from 2006 to 2024. The findings reveal that, despite temporal fluctuations, these summaries generally align with senior-college to college-graduate reading levels. They also contain a high average jargon density of 8.0% per text, surpassing the recommended threshold for general readership. Notably, readability measures related to structural complexity show a significant upward trend, while lexical difficulty remains stable. Meanwhile, normed jargon use presents a non-significant but visually noticeable upward trend over time. These temporal patterns suggest that these lay summaries have become more challenging to read, mostly due to individually-varied but densely embedded specialised terms in longer and more complex sentences. The findings raise concerns about the accessibility of lay summaries for non-specialists, such as interdisciplinary researchers, science communicators, policymakers, and the general public. The study concludes with a discussion and suggestions on readability and jargon use in grant proposal summaries.
The article analyzes the key theoretical and practical aspects of developing lexical competence in the process of learning Ukrainian as a foreign language. Lexical competence is considered an essential component of foreign language communicative competence, ensuring effective communication in accordance with linguistic and cultural norms. The author explores the peculiarities of vocabulary acquisition at the initial (A1), basic (A2), and threshold (B1) levels of Ukrainian language proficiency. The study examines teaching methods that include systemic-linguistic, conditional-communicative, and communicative types of exercises, as well as interactive methods such as language games, working with texts, and visual learning aids. Through the communicative approach, which involves working with realistic texts, role-playing games, and situational dialogues, the process of immersing foreign learners in the language environment is implemented. Based on the topic “Professions,” both traditional systemic-linguistic and communicative tasks adapted to the needs of foreign learners are proposed. The presented set of exercises is aimed at the gradual development of lexical skills, ranging from familiarization with new words to their active use in speech. It facilitates the development of reading, speaking, listening, and writing skills. These tasks help foreign learners work with texts and create their own. The use of such methods contributes to increasing students’ motivation and enhancing the effectiveness of vocabulary acquisition. Reading or listening to texts about famous Ukrainians fosters linguistic and cultural competence. The described approaches can be applied both in classroom settings and for independent student work. The practical significance of this study lies in the introduction of modern approaches to teaching Ukrainian vocabulary as a foreign language. The proposed approaches and types of exercises can be adapted for studying other topics in a foreign language audience at all proficiency levels. Key words: methods of teaching Ukrainian as a foreign language, lexical competence, language proficiency levels, exercises.
This paper explores the integration of sound files into wordnets, transforming them from static lexical databases into multimodal tools for linguistics, language learning and maintenance.Traditionally, wordnets focus on textual representations.Adding sound improves usability for language learners and linguists, especially in less-documented or endangered languages.We extracted sound data for basic vocabulary in 24 languages from the TUFS Basic Vocabulary Modules, link them to senses and make them available as small wordnets.We also discuss the issues involved with merging the data into an existing wordnet, looking at the Open English Wordnet.In addition, this paper outlines the process of integrating audio, discusses potential use cases, and evaluates the technical challenges involved.Finally we suggest an extension to the wordnet formats to allow sound for examples and definitions as well.
This study examines how Transcarpathian Hungarian refugees who relocated to Hungary after 24 February 2022 negotiate language variation, identity, and integration in a same-language migration context. Employing a mixed-methods design, an online questionnaire (n = 120) assessed perceptions of dialectal difference, use of Slavic borrowings, comprehension breakdowns, and experiences of evaluative comments; semi-structured follow-up interviews (n = 18) provided in-depth qualitative insights into everyday adaptation and identity work. Quantitative results indicate that the vast majority (90.8%) detected differences between the variant they brought from Transcarpathia and varieties encountered in Hungary; 57% reported instances where their lexical choices were not understood by Hungarian interlocutors, and 40% recalled direct evaluative or stigmatising remarks. Interview narratives clarify these patterns: a shared language eased immediate practical integration, yet heightened metalinguistic salience led many speakers to actively self-monitor, suppress dialectal markers and Slavic loanwords in public contexts, and seek to acquire competence in formal administrative registers. Applying Yeung and Flubacher's framework, the study shows that categorisation, selection, and activation processes operate even within same-language migration, converting subtle intralinguistic features into markers of inclusion/exclusion. The findings underscore that integration policy should extend beyond basic language provision to include orientation to regional registers, administrative terminology, and sociolinguistic norms, while designing interventions that reduce intralinguistic stigma and support maintenance of migrants' dialectal identities.
The post-independence era has marked a new phase in the linguistic relationship between Azerbaijani and Turkish, characterized by an increased lexical exchange. During this period, many loanwords from other languages have been systematically replaced by borrowings from modern Turkish. The trend toward linguistic convergence between Azerbaijani and Turkish has been particularly evident in the language of the press and mass media. Additionally, words that were historically common to both languages but had remained in limited usage within Azerbaijani have been reactivated and reintegrated into everyday discourse. A significant portion of the lexical borrowings from Turkish during this period has served as a substitute for Russian and European-origin words introduced via Russian, as well as for Arabic and Persian loanwords. The presence of Turkish-origin vocabulary in contemporary Azerbaijani continues to expand, with borrowed words exhibiting diverse morphological structures, including simple, derived, and compound forms. However, a notable concern in recent linguistic developments is the unregulated incorporation of Turkish words into Azerbaijani, sometimes without genuine necessity. This phenomenon raises questions about the potential impact on the structural integrity of the Azerbaijani language. To safeguard the purity and coherence of the language, it is imperative to adhere to established literary norms and linguistic standards