Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
16504 papers
Colour is a fundamental determinant of affective experience in immersive virtual reality (VR), yet the emotional and physiological impact of individual hues remains poorly characterised. This study investigated how fifteen calibrated Munsell hues influence subjective and autonomic responses when presented in immersive VR. Thirty-six adults (18–45 years) viewed each hue in a within-subject design while pupil diameter and skin conductance were recorded continuously, and self-reported emotions were assessed using the Self-Assessment Manikin across pleasure, arousal, and dominance. Repeated-measures ANOVAs revealed robust hue effects on all three self-report dimensions and on pupil dilation, with medium-to-large effect sizes. Reds and red–purple hues elicited the highest arousal and dominance, whereas blue–green hues were rated most pleasurable. Pupil dilation closely tracked arousal ratings, while skin conductance showed no reliable hue differentiation, likely due to the brief (30 s) exposures. Individual differences in cognitive style and personality modulated overall reactivity but did not alter the relative ranking of hues. Taken together, these findings provide the first systematic hue-by-hue mapping of affective and physiological responses in immersive VR. They demonstrate that calibrated colour shapes both experience and ocular physiology, while also offering practical guidance for educational, clinical, and interface design in virtual environments.
Indonesian, as both a national language and an academic language, demands clarity, precision, and adherence to syntactic, morphological, and orthographic rules. However, in practice, many student writings deviate from these linguistic norms. This study aims to identify and analyze forms of linguistic anomalies in the writings of students in the Madrasah Ibtidaiyah Teacher Education (PGMI) Program, particularly in the use of written Indonesian. This research employs a descriptive qualitative approach with content analysis techniques applied to thesis proposal documents. In addition to documentation, data were collected through direct speech observation and in-depth interviews. The subjects of this literature review consist of various written sources such as books, journals, and scientific documents. Data analysis uses content analysis to interpret information systematically and thoroughly. The results reveal that the most common anomalies include ineffective sentences, nonstandard word usage, and errors in spelling and punctuation. Contributing factors include limited academic literacy, the influence of spoken language and social media, and the lack of continuous practice in formal writing. This study highlights the importance of strengthening language learning and scientific writing training to improve the academic communication skills of PGMI students.
The article analyses the phenomenon of linguistic deviations in online communication, including both linguistic errors and deliberate innovations. It presents the characteristics of the language in the digital space, drawing attention to the speed of communication, user anonymity and the use of abbreviations, neologisms and emoticons. The analysis of online comments revealed the most common deviations from the norm, including starting sentences with a lowercase letter, omitting or using punctuation in a non-standard way, spelling mistakes, and pleonasms. The potential causes of these phenomena were discussed, as well as their possible effects, including permanent changes in the perception of linguistic norms and the development of new forms of communication. The role of education, social campaigns and self-education in maintaining a balance between the innovation and creativity of social media users and the care for linguistic correctness were emphasised.
LOGIQ+ — Cognitive Diagnostic Assistant (TRL-8 Candidate) Version: 1.1 Date: December 01, 2025 Author: Anne Povie (Independent Research) Related DOI: 10.5281/zenodo.17770976 Abstract LOGIQ+ is a research-grade cognitive diagnostic system that leverages a dual-engine architecture—semantic and non-semantic—to generate detailed, scientifically interpretable cognitive profiles. Based on the Skopein AI framework, LOGIQ+ integrates linguistic and symbolic reasoning with fractal and topological modeling to provide probabilistic assessments across cognitive, psychological, and sociological dimensions. This white paper outlines the theoretical framework, scoring architecture, and institutional use potential in education, sociology, public health, and policy-making. 1. System Architecture LOGIQ+ is powered by the Skopein AI dual-engine model: - Semantic Engine: handles symbolic reasoning, lexical abstraction, analogical transfer, and narrative patterning. - Non-Semantic Engine: analyzes dynamic response variation, fractal coherence, and topological relationality. 2. Scoring Methodology All outputs are derived using a three-tiered evidence framework: - H1: Direct linguistic and conceptual signals - H2: Recurrent cognitive structures and motifs - H3: Emergent patterns of convergence and abstraction Scores are expressed probabilistically (0.00–1.00), reflecting model confidence. 3. Cognitive Taxonomy LOGIQ+ detects four primary cognitive orientations: - Analytical - Organizational - Creative Structured - Critical Synthesizer 4. Psychosocial and Anthropological Dimensions The system includes optional interpretation layers for: - Openness to experience - Narrative identity strength - Cultural cognition and relational framing - Symbolic engagement and future orientation 5. Use Cases and Societal Impact LOGIQ+ is designed for: - Educational guidance and inclusion - Social resilience diagnostics - Cognitive readiness mapping for societal roles - Evidence-based program design and targeting 6. Ethical Position and Usage All data and scores are intended for institutional use and research. No score is used for exclusion or stigmatization. The system adheres to open science ethics and complies with PRISMA, STROBE, and RoB2 documentation norms. 7. Example Output See attached file: LOGIQ_CompleteInstitutional_Profile.json 8. Licensing and Repository Compatibility LOGIQ+ is distributed under CC BY 4.0. All files are compatible with Zenodo, HAL, and ReScience C repositories, and prepared for institutional archival referencing.
This paper presents an approach to integrating Latin inflected forms and corpus attestations within a Linked Open Data (LOD) framework, enhancing interoperability between Wikidata and the LiLa knowledge base. Building on the PrinParLat lexicon of Latin verb principal parts, we generate the complete set of inflected forms for over 8,000 verbs, encoded as RDF in a dedicated Wikibase instance. These forms are linked to the Index Thomisticus Treebank (ITTB), whose morphologically annotated tokens are related to corresponding forms based on segmental identity, lemma alignment, and mapped morphological features. Our generation and linking process achieves over 95% coverage of ITTB verbal tokens, demonstrating the robustness of our pipeline even for Medieval Latin data. By aligning Paralex, Wikidata, and LiLa ontologies, we ensure semantic interoperability and facilitate future integration into Wikidata. Beyond Latin, this workflow provides a reproducible model for linking inflectional paradigms and corpus attestations in other languages.
This Capstone examines how generative AI reshapes the act of writing by treating co-writing itself as a site of inquiry. Using an autoethnographic method, I document my collaboration with ChatGPT across months of drafting, tracing how the model’s predictions influence tone, syntax, and rhetorical choice. The project argues that large language models are not neutral tools but participants in a shared writing process shaped by the linguistic hierarchies embedded in their training data. Through examples of prompting, revision, and negotiation, I show how AI leans toward standardized English, simplifies sentence structure, and reproduces dominant linguistic norms, often in ways that feel fluent but unexamined. Drawing on scholarship from linguistics, rhetoric, AI ethics, and corpus studies, I situate these patterns within broader cultural and technological shifts that are redefining authorship and literacy. My findings suggest that the future of writing will be a form of co-authorship and that ethical engagement requires vigilance, not avoidance: understanding how the model learns, questioning its outputs, and preserving human intention as the final authority. This project therefore offers both a critique of AI’s biases and a practical framework for collaborating with language models responsibly as they become embedded in the work of writing itself.
At present, social media has developed into one of the most common communication tools used by university students. The use of social media platforms not only shapes the way individuals engage socially with one another but also leaves a significant impact on their linguistic practices, particularly in relation to the Indonesian language. The purpose of this study is to examine how social media influences students’ language habits when using Indonesian, both in written and spoken forms. A descriptive qualitative approach was employed, collecting data through observation and questionnaires administered to a group of students. The findings reveal that students who frequently use social media tend to adopt non-standard language, abbreviations, and code-mixing with foreign languages, which gradually has the potential to undermine their ability to communicate effectively and correctly in Indonesian. Nevertheless, social media also presents beneficial prospects, such as fostering greater creativity in sentence construction, expanding vocabulary, and cultivating a stronger interest in writing. Consequently, social media exerts a dual influence on students’ language behavior, encompassing both positive and negative aspects, thereby underscoring the need for awareness and the cultivation of Indonesian language use that adheres to established linguistic norms.
The city of Berehove is located in the Zakarpattia (Transcarpathia) region (Oblast) of Ukraine, near the border with Hungary. A significant part of the local population is Hungarian-speaking; that is why, for the more than 150000 Transcarpathian Hungarians, Berehove serves as an unofficial cultural center. At the time of the main data collection (2019–2021), street signs in the city could be found in Ukrainian, Hungarian, Russian, English, and other languages. Such a highly multilingual urban environment gives rise to various types of mistakes, as not all residents are fluent in the common languages. The Hungarian minority in Transcarpathia faces difficulties in learning the Ukrainian language, as it was not compulsory during the Soviet period. The compact settlement of the Hungarian population in Transcarpathia means that they may have very limited contact with speakers of Slavic languages. In addition, the grammar of Ukrainian and Russian is difficult for Hungarians. The article focuses on errors in Ukrainian and Russian texts, but also provides some examples of street signs in Hungarian and English that deviate from linguistic norms. This study proposes an approach that involves students in the joint identification and analysis of such errors in street signs. The inclusion of such exercises in the educational process can help children develop the habit of carefully observing their linguistic environment and critically evaluating their own speech.
This article reports a metalexicographical study of the headwords and indications of meaning which, based on information in the 14th edition of the Swedish Academy Glossary (SAOL 14, 2015), are used either exclusively or mainly in Finland-Swedish: (RQ1) How does it appear in the SAOL 14 that headwords (Hw) and indications of meaning (IoM) are deemed to be finlandisms alternatively Finland-Swedish entries?; (RQ2) What information is given about the relevant Hw and IoM in the SAOL 14? A majority of the Hw and IoM in the corpus are provided with explicit labels indicating their finlandism status. The Finland-Swedish-ness of a minority of the Hw and IoM is signaled only through explanations of their meaning/s. It is rather unusual for finlandism status to appear in terms of formal information and syntagmatic information in the SAOL. Instead, it is primarily a matter of semantic information and/or pragmatic information. Keywords: The Swedish Academy Glossary – SAOL, finlandisms, Finland-Swedish entries, The Swedish Academy Lexical Database – Salex
The article is theoretical in nature. It focuses on the language awareness among young people. Language awareness enables code-switching and balancing between the sociolect known as the youthlanguage and the official, standard-compliant language. Depending on the context, language is usedby young people to identify themselves with the group and to manifest their independence, distinctiveness and uniqueness. It also reflects their integration into linguistic norms and their creative transformation of these norms. The level of language awareness determines which variety young people choose to use in particular situations. Drawing on the levels of language awareness identified in the literature, the author presents examples of both low and high levels of language awareness in thelinguistic competence of young people. The conclusion is presented within the context of the core curriculum, highlighting the need to correlate literary education with language education.
International audience
This research is motivated by concerns about the decline in communication quality and character values among youth due to the use of slang language that does not adhere to proper linguistic norms, particularly through the widespread use of social media. The aim of this study is to examine the influence of slang language used on social media on the development of character values among female students at STITMA Yogyakarta, as well as to explore their perspectives on this phenomenon in daily life. This study employs a quantitative approach with a correlational method. Data collection techniques include questionnaires, interviews, and documentation. The population consists of fourth- and sixth-semester female students from the Islamic Education (PAI) and Arabic Education (PBA) programs, with a sample of 75 students selected using proportionate stratified random sampling. The data were analyzed using validity, reliability, normality, homogeneity tests, and hypothesis testing through the Pearson Product Moment correlation test. The results show that the calculated r-value of 0.967 is greater than the critical r-value of 0.2272, indicating a significant and positive influence between the use of slang language and character values. Thus, slang language is proven to impact the character development of female students in today’s digital era.
This study deals with the role played by social media in reducing the spread of common linguistic errors and improving linguistic competence among students, as well as those interested in the Arabic language and keen to learn it. The researcher adopted the analytical inductive approach, and the study yielded several results, including: There are many terms that indicate a departure from the linguistic norm among Arabs, including: incorrect pronunciation, mistake, and error, and there is no dispute over the term; old and contemporary sources used all of these terms, and social media also plays an important role in serving the Arabic language and reducing - to some extent – these linguistic errors. This is done in a direct way represented by linguistic corrections, and an indirect way manifested in linguistic refinement by enhancing the linguistic skills through the content related to Arabic syntax, morphology linguistic differences.
The present paper describes the building of STAF, a Universal Dependencies treebank for Albanian. STAF was bootstrapped using a Stanza model trained on previously unreleased data and then manually corrected by three Albanian speakers supervised by the author, who also revised all sentences. STAF focuses on the fiction genre, featuring 200 sentences selected from nine literary texts written by Albanian contemporary authors.
Characteristics of emotion, such as valence and arousal can be evaluated using self-reported affective ratings and electroencephalogram (EEG) to gain a better understanding of individual differences in diversified populations. The International Affective Picture System (IAPS) and AI-generated counterparts were used to elicit emotional responses that were collected from the self-assessment manakin (SAM) rating scale for valence and arousal. EEG data were used to observe biomarkers of emotional processes related to the presentation of these stimuli. These methods were correlated to individual differences such as sex and depression and anxiety related questionaries. The study showed significant sex differences in the self-reported affective ratings, where females showed greater aversive affect compared to males. EEG data showed that there was less alpha reduction relative to baseline in percent for individuals who scored high on the BDI-II, which suggests biomarkers of emotional dysregulation, such as anhedonia. Overall, the study highlights the differences in emotional processes for variable populations, which has implications for intervention and targeted treatment efforts.
This research focuses on the veiled economic priorities woven into the climate change rhetoric at the 26th United Nations Climate Change Conference (COP26), convened in Glasgow, Scotland, from October 31 to November 13, 2021. Drawing on 130 English-language addresses by national delegates, obtained from the United Nations Framework Convention on Climate Change (UNFCCC) official repository, the study employs Arran Stibbe’s Ecological Discourse Analysis and his Eight Story Framework as its analytical foundation. It examines linguistic elements such as evocative terms, affirmative expressions, lexical selections, modal verbs, patterns of assertion, and pronoun use to expose underlying ideological currents. The results indicate that orations from 20 nations overtly prioritize profit-driven motives and national self-interest. These narratives reposition climate change as a platform for economic gain, with policy proposals highlighting growth and technological advancement over authentic ecological stewardship. Amazingly, 11 of these 20 nations rank among the top 25 global carbon dioxide emitters (per 2020 data), tying their rhetorical strategies to significant environmental consequences. The analysis reveals a consistent disparity between the declared commitments of developed nations to provide climate funding and their actual contributions to at-risk countries. Using COP26 as a focal point, this study illustrates how international climate platforms are frequently leveraged to continue economic norms, highlighting how linguistic choices advance broader ideological and geopolitical objectives.
This study investigates the impact of social media-induced neologisms on the linguistic proficiency of university students in Cameroon. With the widespread use of platforms like Twitter (now X), TikTok, Instagram, Facebook, and WhatsApp, an increasing number of slang terms and abbreviations—collectively termed neologisms—have become part of students’ daily communication. While these terms foster informal expression and social bonding, their use in academic writing undermines grammatical accuracy, lexical appropriateness, spelling, coherence, and formal tone. Drawing on Halliday’s Systemic Functional Linguistics, particularly Register Theory, this research analyses the written output of 200 first and second year university students, identifying 645 neologism occurrences across 55 distinct types. A coding scheme tracks the frequency, form, and syntactic roles of these neologisms, with attention to their ideational, interpersonal, and textual functions. Additionally, a structured survey explores students’ metalinguistic awareness, language habits on social media, and ability to distinguish between formal and informal registers. Findings show a strong presence of neologisms such as “4u,” “ghosted,” “vibes”, “nerve” and “low-key” in formal assignments, leading to inappropriate register use and reduced clarity. Many students demonstrate "register flattening," struggling to shift between informal digital language and formal academic expression. While the study recognizes the creative and identity-shaping value of social media neologisms, it highlights their unintended negative effects on academic writing standards in the digital age. The research recommends targeted pedagogical interventions to build students’ awareness of language register and promote formal writing skills. Though linguistic innovation reflects cultural change, academic successrequires mastery of formal communication norms; promoting register awareness is therefore essential for maintaining academic standards.
This paper undertakes a Marxist literary analysis of Francesca Simon’s Horrid Henry series, with a specific focus on the short stories Horrid Henry Robs the Bank (2003), Horrid Henry’s Christmas (1994) and Horrid Henry and the Scary Sitter (1997). It studies the construction of class struggle, alienation and superstructure through the lens of comedy. This is done by positioning Henry as the subaltern figure, whose humorous behavioural transgressions expose and are simultaneously contained by prevailing parental, i.e. capitalist state, ideologies. The study is structured around three comic modalities: farce, lexico-semantics and satirical characterisation. The farcical narrative foregrounds the grotesque inequalities embedded in the superstructure; lexical humour is deployed to represent resistance to conformity to bourgeois norms linguistically; authority figures, like parents, teachers, and even babysitters, are portrayed satirically to expose the arbitrariness of ‘disciplinary’ mechanisms (which mirror marginalisation and propaganda to maintain false consciousness) at the heart of the capitalist state’s apparatus. The Horrid Henry series operates as a discursive site wherein the ideological tensions of capitalism are encoded, negotiated, and pedagogically transmitted through the comic form. Stories from three distinct quinquenniums have been selected for this study to explore the consistency in Simon’s political messages across time. Ultimately, by situating Simon’s series within broader debates on the political function of children’s literature, this research underscores the genre’s consequential role in constructing youth perspectives on class, power and justice.
International audience
Electroencephalogram (EEG) signals exhibit nonstationary dynamics with high temporal resolution but limited spatial resolution. A critical challenge lies in identifying stable neural states during rapid emotional transitions and decoding dynamic interregional interactions. To address this, we propose a dynamic microstate temporal graph attention network (DMT-GAT) that integrates transient EEG microstates with brain functional networks. First, EEG signals are segmented into four prototypical microstates (labeled as MS1, MS2, MS3, and MS4) via global field power peak detection and K-means clustering. Emotion-related microstates (MS3/MS4) are then selected through independent t-tests based on valence and arousal ratings. Next, a brain functional network is constructed by calculating phase-locked value synchronization specifically on the time series of MS3/MS4 microstates, capturing millisecond-scale interregional dynamics during emotional shifts. Frequency-domain features are integrated into the network nodes, forming graph-structured data. Finally, a GAT with multi-head mechanisms classifies emotions by adaptively weighting node interactions. On the DEAP dataset, our method achieves average accuracies of 99.19% (valence) and 99.26% (arousal). For the SEED dataset, it maintains a robust accuracy of 95.29%. Crucially, the DMT-GAT uniquely reveals prefrontal-amygdala interactions during emotional regulation, bridging the gap between dynamic brain networks and rapid neurodynamics. This work provides a novel framework for high-resolution emotion recognition and advances understanding of neural mechanisms underlying affective transitions.
BACKGROUND: Body dissatisfaction (BD) is a risk factor for and a maintaining factor of Anorexia nervosa (AN). Furthermore, BD is associated with depressive symptoms. Body exposure (BE) was found to be an effective intervention for reducing BD. The current study aimed to investigate similarities and differences in BD between patients with AN and depressive symptoms and the efficacy of a computerized BE in those adolescents. METHODS: We compared adolescents with AN (n = 36) to adolescents with depression and high body dissatisfaction (n = 21; DBD group). BD was assessed with questionnaires; valence ratings were obtained for different body parts. Emotion ratings and gaze patterns towards the own body were assessed during each session via rating scales and eye-tracking. RESULTS: Satisfaction with several body parts increased and anxiety and disgust decreased throughout the intervention in both groups, with no significant differences between them. An attentional bias towards the three most unattractive body parts was found, expressed via longer viewing times; however, it was not modified by the BE intervention. CONCLUSIONS: The similarities between adolescents with AN and highly body dissatisfied ones with depression in terms of BD, emotional reactions to and gaze patterns on one's own body suggest a transdiagnostic phenomenon of BD. The results suggest that a computer-based BE is an effective intervention for reducing BD. TRIAL REGISTRATION: The study was pre-registered in the German Clinical Trials Register (Deutsches Register Klinischer Studien; DRKS), ID number DRKS00024675.
Interactive internet platforms allow speakers to comment on linguistic variation in utterances from around the world to which they are exposed. Digital platforms thus function not only as spaces for discussing linguistic usage but also as arenas for negotiating language norms. Participants in these discussions often adopt strongly asserted normative positions. In this context, a project developed at the University of Kiel is presented in the article. It aims to analyse such normative discourses from a comparative perspective, with the goal of highlighting the specificities of different linguistic cultures. The article draws attention to a gradual shift in the conception of linguistic norms: traditional regulatory institutions increasingly see their authority challenged by a significant portion of language users, who formulate normative claims grounded in social arguments. This shift reflects a normativisation process that is now shaped by more participatory and transnational dynamics.
The paper studies the issues of interference in Tatar ergonyms. The research is based on modern names of Kazan infrastructure objects, forming the linguistic image of the city in the conditions of Russian-Tatar bilingualism. We analyzed the interference phenomena in Tatar ergonyms, taking into account their theoretical comprehension found in scientific literature. The most common errors include incorrect spelling of Tatar phonemes, case endings and word order in a sentence. We have found that the Russian language influences the formation of ergonyms, so there are deviations from the standards of the modern Tatar literary language. Over time, violations of this type can lead to changes in the language norms that may eventually be perceived as part of the existing linguistic norm, making this problem particularly relevant. In a globalized world, it is important to recognize the importance of preserving different languages and cultural identities for future generations. This allows the authors to conclude that the government should take control of the information presented on signboards of the urban spaces, ensuring that it is orthographically correct, semantically accurate and stylistically comprehensible to native speakers. The article emphasizes that the experts, involved in translating the names of various organizations, enterprises, institutions, etc., and information on signs and official website, should be well aware of the requirements for spelling out ergonyms. In the epoch of globalization, the problem of preserving languages, their national-cultural characteristics and the identity of different peoples is of particular importance.
BACKGROUND AND OBJECTIVES: Narrative discourse is a useful means to organize ideas and create shared understandings. Clinically, performing discourse analysis on disordered spoken language could facilitate researchers and clinicians not only to evaluate one's language abilities but also to foreshadow his/her communication in real-life situations. Given the normative reference data of a specific discourse task, less-biased judgement and evaluation could be made, which could further facilitate assessment and intervention planning. This study aims to first develop norms by analysing the language samples produced by neurotypical Cantonese speakers on two well-familiarized narrative stories, The Boy Who Cried Wolf, and The Tortoise and Hare. Second, we aim to investigate the potential age and education effects on a wide range of micro- and macro-structural linguistics measures. METHOD: Two semi-spontaneous story narratives from the Cantonese AphasiaBank were selected for scoring. A total of 150 neurotypical Cantonese adult speakers produced the spoken discourse samples for each story narrative. All speakers were native Cantonese speakers living in Hong Kong; they were divided into three age groups: young (18-39 years old), middle-aged (40-59 years old), and older (> 60 years old). Audio recordings were transcribed, segmented, and annotated using CHAT conventions. RESULTS: Normative references of various micro- and macro-structural linguistics measures and the standard scoring references for the two narrative stories were established. For the age effect on narrative discourse, the older adults produced less complex, coherent and thematic-related concepts compared to the young group. However, lexical diversity was preserved in the older group, resulting in no significant differences across the three age groups. For education effect, the higher education group outperformed the lower education group in verbal productiveness and content informativeness. Lastly, the two stories were found to be non-comparable to each other, thereby they should not interchange in pre- and post-test arrangements or in monitoring discourse performance. CONCLUSIONS: The Cantonese discourse norms presented here can be applied in both research and clinical settings, facilitating a more objective review of language impairment and treatment planning. Second, this study demonstrated the effect of normal ageing on both the linguistics and conceptual levels specific to discourse production. WHAT THIS PAPER ADDS: What is already known on this subject Discourse analysis is a critical part of evaluating and understanding a person's communication abilities. Studies indicated that narrative discourse is more sensitive to specific linguistic parameters than other genres, and people with stronger narrative skills tended to have more social communication opportunities. An increasing number of studies had been working on setting norms for discourse tasks. Locally, Kong et al. (2025) have recently reported the normative references for descriptive tasks of the Cantonese AphasiaBank. What this study adds to the existing knowledge First, this study completed the norms establishment for all discourse tasks of the Cantonese AphasiaBank. Second, our analysis of the impact of different factors on narrative discourse offers significant value for clinical applications. We found that ageing was not manifested across all microstructural linguistics consistently, while lexical diversity was found to be tolerant to ageing. However, ageing was found to be adversely affecting propositional parameters and discourse informativeness. What are the clinical implications of this study? Narrative discourse, being one of the most popular tasks in clinical assessment but often faced the challenges of a lack of objective references. This study analysed two well-familiarized narrative stories, provides a complete set of norm data readily for front line clinicians and researchers, which could be used for intervention planning, monitoring progress (e.g., used them as control probes for tracking generalization effects) or as an input for investigating the interplay between linguistics and cognitive abilities.
Мақалада түркі тілдерінің синтаксистік құрылымын формалды грамматика тұрғысынан және заманауи аннотациялық модельдер негізінде сипаттаудың тәжірибесі қарастырылады. Синтаксистік аннотация тілдің грамматикалық жүйесін формалды түрде сипаттайтын және оны автоматты өңдеуге мүмкіндік беретін маңызды құрал ретінде танылады. Зерттеу барысында «Universal Dependencies» (UD), «MaTT» (Multilingual Aligned Treebank of Turkic) және «Kazakh Dependency Treebank» (KazDT) сияқты жобаларға сүйеніп, түркі тілдеріне тән морфологиялық және синтаксистік ерекшеліктер сипатталды. Синтаксистік белгіленім модельдері: «құрамдық», «аралас», «басыңқы-бағыныңқылық грамматикасы» т.б. тәсілдердің сипаты, ерекшеліктері, түркі тілдері үшін ұтымды тұстары мен кемшіліктері сараланды. Нәтижесінде басыңқы-бағыныңқы қатынастар грамматикасы негізінде жасалған синтаксистік аннотация моделі түркі тілінің құрылымын тиімді сипаттауға мүмкіндік беретіні дәлелденді. Басыңқы-бағыныңқы грамматикасының (басыңқы-бағыныңқы қатынастар) теориялық негіздері, синтаксистік аннотацияның форматы мен стандарттары сараланды. Түркі тілдерінің жалғамалы табиғаты мен еркін сөз тәртібінің «UD» сияқты әмбебап жобаларға бейімделуі талдауға түсті. Сонымен қатар, қазақ тілінің аннотацияланған корпустарын жетілдіру, автоматты парсинг, тілдік білім беру жүйесіне енгізу секілді болашақтағы бағыттары көрсетілді. Мақала түркі тілдерінің синтаксистік белгіленім тәжірибесі негізінде қазақ тілін цифрлық кеңістікке енгізудің маңызды қадамдарының бірі ретінде синтаксистік аннотацияны ғылыми тұрғыда негіздеуді мақсат етті. Түйін сөздер: түркі тілдері, синтаксистік аннотация, басыңқы-бағыныңқы грамматикасы, «UD», KazDT, формалды модельдер, парсинг.
This paper proposes a conceptual framework for integrating social robots into Islamic Arab communities in a manner that aligns with local cultural, ethical, and linguistic norms. Addressing critical gaps—such as navigating Arabic dialect diversity, adhering to Islamic ethical principles, and maintaining privacy in IoT-enabled environments—the framework comprises five core modules: a Cultural Knowledge Base, Behavior Adaptation Module, Ethical Integration Module, Arabic Language Processing System, and Privacy Management Module. This model fosters trust and acceptance by enabling robots to function effectively in healthcare, education, and public services. Future research will involve iterative testing and real-world evaluations to refine and validate the framework's applicability across diverse global and multi-faith contexts.
The most common usage of the Greek particle οὖν/oûn is to create coherence within a large portion of discourse and convey inferential meaning. This chapter addresses its usage in documentary papyri of the Roman and Byzantine periods, aiming at providing a preliminary general exploration of the contextual usage of this particle and approaching it through pragmatic and discourse analysis approaches. It focuses on the linguistic elements that surround this particle and identifies different patterns that frequently occur within the corpus of documentary papyri. By using the PapyGreek treebank corpus and Trismegistos, statistics for these patterns are added in order to draw observations on the co-occurrence of the particle and other linguistic elements that precede it. By means of examples, the chapter discusses a selection of these patterns, their context of use, and their linguistic function. It analyses the position of some οὖν/oûn utterances, which for instance typify a request, a warning, or a piece of advice, looking at their sequential placement within the discourse. It therefore addresses the question of how inferences are constructed and conveyed. Finally, focusing on some cases of negative οὖν/oûn utterances, the chapter discusses some aspects of their meaning and functions by using a cognitive linguistic approach.
Inside a challenge of ideas there are several phases in a Creative Support System (CSS), they are problem analysis, ideation, evaluation, and implementation. Our problem: we need a full semantic lexical database SLD in an oral (voice) and writing way to help stakeholders to create ideas, these ideas contain nouns, verbs, adverbs, adjectives in the English, Spanish, and French languages. We utilize a Cloud Service Provider to use a service of Artificial Intelligence (AI), also we prepare nouns, verbs, adjectives and adverbs files in order to create the service text to voice and create our SLD with voice. This paper presents, first, an introduction about some contests that use a semantic lexical database in different languages; second, a SLD management approach using analysis of texts; third, a management application approach to complete all the new elements; fourth, the results of the management application approach, finally the conclusions and future work.
We introduce UniRST, the first unified RST-style discourse parser capable of handling 18 treebanks in 11 languages without modifying their relation inventories. To overcome inventory incompatibilities, we propose and evaluate two training strategies: Multi-Head, which assigns separate relation classification layer per inventory, and Masked-Union, which enables shared parameter training through selective label masking. We first benchmark monotreebank parsing with a simple yet effective augmentation technique for low-resource settings. We then train a unified model and show that (1) the parameter efficient Masked-Union approach is also the strongest, and (2) UniRST outperforms 16 of 18 mono-treebank baselines, demonstrating the advantages of a single-model, multilingual end-to-end discourse parsing across diverse resources.
This chapter examines the transformative yet contested role of digital language technologies in multilingual educational contexts. It explores how tools such as AI-powered language platforms, machine translation engines, and digital storytelling applications shape power dynamics, reinforce linguistic hierarchies, and affect the rights and representation of minoritized languages. Through case studies and comparative analysis of widely used digital platforms, the chapter reveals how algorithmic design can marginalize regional dialects and elevate dominant linguistic norms, thereby contributing to digital linguistic imperialism. Despite these serious challenges, the chapter highlights the empowering potential of language technologies to promote adaptive learning, cultural exchange, and intercultural communication, when they are applied through thoughtful, critical, and equitable pedagogical approaches. It further emphasizes the need to examine how learners engage with these technologies in informal contexts and how their experiences reshape language attitudes and agency. By proposing an integrative pedagogical framework grounded in critical language awareness and inclusive digital literacy, the chapter provides practical insights for educators seeking to navigate the ethical and pedagogical complexities of digital language education. Ultimately, this work contributes to a growing discourse on digital justice and linguistic human rights, advocating for multilingual digital futures that center diversity, inclusion, and empowerment within increasingly technologized learning environments.
Designers aim to create designs that resonate with users, the consumers. However, there is often a gap between the sensory and emotional “words” expressed by users and the “words” understood by designers. To address this issue, we propose a method called “Language-based engineering” to reflect the meanings users seek in designs accurately. Using the packaging design of golf gloves as an example, we conducted a survey on six sensory words (I want to pick up, novel, visible, cool, cute, luxurious) to gather user feedback. We analyzed the relationship between user ratings and design elements using Quantification Theory Type I and created a database of these relationships. Furthermore, we converted the sensory word ratings into aspiration levels using the satisficing trade-off method and proposed a method to determine the combination of design elements that satisfy these desired levels as a multi-objective optimization problem. We conducted trade-off analyses on all 30 patterns of the six sensory words to identify the trade-off relationships based on users’ impressions of the packaging design. Additionally, we presented design examples and guidelines that align with users’ sensibilities by combining design elements that meet the aspiration levels.
Bengali (Bangla) is typologically rich in complex predicates, especially verb-verb compounds and verbo-nominal light verb constructions. While these constructions have been studied from theoretical and annotation perspectives, there is no simple corpus-level quantity that summarizes how "saturated" a parsed Bengali corpus is with complex predicates. This paper introduces two related corpus-level constructs attributed to S M Nazmuz Sakib. First, the S M Nazmuz Sakib Constant for Bengali Compound Predicate Saturation (short: Sakib Constant) is defined as the ratio between the number of compound-type dependency relations (compound and compound:lvc) and the number of verbal tokens in a Universal Dependencies (UD) Bengali treebank. Second, the S M Nazmuz Sakib Triangle (short: Sakib Triangle) is a normalized triple giving the relative shares of compound, obj, and advmod relations, interpreted geometrically as barycentric coordinates inside a triangle. Using published statistics for the UD Bengali-BRU and Bengali PUD treebanks, we compute concrete values of the Sakib Constant and Sakib Triangle and visualize them through ten data-based diagrams. We also formulate the S M Nazmuz Sakib Compound Predicate Saturation Hypothesis and the S M Nazmuz Sakib Bipolar Headedness Partition Principle, linking these quantities to word order tendencies in Bengali. The definitions are simple and intended to be testable on larger parsed corpora, such as BDNC, bnTenTen and IndicCorp v2, as more UD-style parsers become available for Bengali.
Verb fluency task is a screening test that shows sensitivity to lexical retrieval abilities. This study aims to establish a baseline for verb fluency in Japanese by comparing data from 61 younger and elderly Japanese speakers. The results show that the elderly group produced fewer verbs than the younger group, but compared to Lee et al. (2013), many more correct responses were produced by both groups. An analysis using the Lancaster Sensorimotor Norms found that the number of clusters and switches was smaller for the elderly participants, although the cluster size did not show any difference.
The article is devoted to the study of the peculiarities of the translation of new language norms that emerge in social networks.The main interpretations and characteristics of these units were analysed.In addition, four main methods of translating such items, proposed by P. Newmark, were identified and illustrated with examples.The active spread of digital communication and the rapid formation of Internet-based vocabulary determined the main purpose of the investigationto analyse the features of functioning and translation of new lexical units from English into Ukrainian in the context of social media.The object of the research is the peculiarities of the translation and functioning of new language norms in digital discourse.The subject of the research is the specifics of translation of new lexical units from English into Ukrainian.The novelty of this research lies in a comprehensive analysis of how new lexical units that emerge in social networks are translated and function in Ukrainian digital discourse.It provides a systematic classification of translation methods: borrowing, transcription, transliteration; naturalisation; cultural equivalent; descriptive method.Their quantitative distribution was determined based on 50 authentic examples from popular platforms (Instagram, WhatsApp, Facebook, WeChat, TikTok).The scientific novelty also consists in identifying the tendencies of adaptation of English neologisms to Ukrainian linguistic and cultural norms, which reflects the ongoing influence of digital communication on the evolution of modern Ukrainian vocabulary.The material of the research is 50 lexical units selected from popular social media platforms such as Instagram, WhatsApp, Facebook, WeChat, and TikTok.
The article considers feminine nouns as a phenomenon that is specific in modern linguistics. The study topicality is explained by the special place of language in formation of national identity, especially in a crisis period, when the state defends its independence. On the other hand, the study relevance is based on values of a democratic society, where gender equality is a priority, which is fixed in language as a means of social thinking and functioning. Authors’ conduct of the study is appropriate, since the topic of gender equality and the masculinity-feminity balance arouses a deep interest among linguists in Scopus publications for the period 2020–2025. It is noteworthy that scientists are interested in gender issues in aspects of national identity both within a language group and neighboring ones. High popularity of feminine nouns in the European space and strengthening of the Ukrainian national identity during the socio-political turbulence (war) determine the authors’ choice of thematic direction of the research. The study object is feminine nouns in the Ukrainian, German and English languages from a linguistic and sociocultural point of view. The study subject is peculiarities of formation, functioning and use of feminine nouns in these languages, their role in reflecting social processes and forming the language norm. The research is implemented via the comparative method for typological analysis of feminine nouns from the Slavic-Germanic perspective with determination of their forms, functions and influence on the linguistic identity formation. Within the study, it was established that the main means of transmitting gender features can be purely lexical resources (English) or morphological and grammatical ones (Ukrainian and German). In the first case, the pattern is due to absence of the grammatical category of gender and tendency to inclusive neutralization, when equality of rights and freedoms is conveyed through lexically neutral words. In the second case, presence of gender led to emergence of morphological elements (suffixes). Their gender functionality is variable. In the German culture the suffix -in- has firmly established itself as a designation of the female gender. In the Ukrainian language, feminine suffixes are a relatively young phenomenon, which causes a contradictory attitude to processes of gender word formation. The research results can be a basis for conducting new studies in intralingual and interlingual aspects.
Information on the relationship between facial thermal responses and emotional state is valuable for sensing emotion. Yet, previous research has typically relied on linear methods of analysis based on regions of interest (ROIs), which may overlook nonlinear pixel-wise information across the face. To address this limitation, we investigated the use of machine learning (ML) for pixel-level analysis of facial thermal images to estimate dynamic emotional arousal ratings. We collected facial thermal data from 20 participants who viewed five emotion-eliciting films and assessed their dynamic emotional self-reports. Our ML models, including random forest regression, support vector regression, ResNet-18, and ResNet-34, consistently demonstrated superior estimation performance compared to traditional simple or multiple linear regression models for the ROIs. To interpret the nonlinear relationships between facial temperature changes and arousal, saliency maps and integrated gradients were used for the ResNet-34 model. The results show nonlinear associations of arousal ratings in nose = tip, forehead, and cheek temperature changes. These findings imply that ML-based analysis of facial thermal images can estimate emotional arousal more effectively, pointing to potential applications of non-invasive emotion sensing for mental health, education, and human-computer interaction.
We expand the second language (L2) Korean Universal Dependencies (UD) treebank with 5,454 manually annotated sentences. The annotation guidelines are also revised to better align with the UD framework. Using this enhanced treebank, we fine-tune three Korean language models and evaluate their performance on in-domain and out-of-domain L2-Korean datasets. The results show that fine-tuning significantly improves their performance across various metrics, thus highlighting the importance of using well-tailored L2 datasets for fine-tuning first-language-based, general-purpose language models for the morphosyntactic analysis of L2 data.
Bidirectional transformers excel at sentiment analysis, and Large Language Models (LLM) are effective zero-shot learners.Might they perform better as a team?This paper explores collaborative approaches between ELECTRA and GPT-4o for three-way sentiment classification.We fine-tuned (FT) four models (ELECTRA Base/Large, GPT-4o/4o-mini) using a mix of reviews from Stanford Sentiment Treebank (SST) and DynaSent.We provided input from ELEC-TRA to GPT as: predicted label, probabilities, and retrieved examples.Sharing ELECTRA Base FT predictions with GPT-4o-mini significantly improved performance over either model alone (82.50 macro F1 vs. 79.14ELECTRA Base FT, 79.41 GPT-4o-mini) and yielded the lowest cost/performance ratio ($0.12/F1 point).However, when GPT models were fine-tuned, including predictions decreased performance.GPT-4o FT-M was the top performer (86.99), with GPT-4o-mini FT close behind (86.70) at much less cost ($0.38 vs. $1.59/F1 point).Our results show that augmenting prompts with predictions from fine-tuned encoders is an efficient way to boost performance, and a fine-tuned GPT-4o-mini is nearly as good as GPT-4o FT at 76% less cost.Both are affordable options for projects with limited resources.
Abstract This paper presents a novel framework for modeling role and task allocation in cooperative wheeled soccer robot systems by leveraging latent knowledge extracted from past collaborative interactions. Inspired by recent advances in heterogeneous multi-robot collaboration, the proposed method encodes a soccer team as a set of Multidimensional Relational Structures (MDRSs), capturing both temporal and spatial relations among robot roles, actions, and stimuli. A structured dataset, termed the Soccer Robot Collaboration Treebank (SRCT), is introduced to represent play-by-play histories of robot behaviors, parsed through a formal grammar to support structured learning. Probabilistic modeling and Non-Negative Tensor Decomposition (NTD) are applied to the resulting tensors, enabling robust inference and latent knowledge estimation even in scenarios with sparse data or communication loss. Simulated experiments using a team of wheeled soccer robots in the Webots environment demonstrate the system’s ability to dynamically reassign roles, reason over incomplete histories, and predict collaborative behaviors such as passing, defending, or role-switching. The results show that the proposed framework enhances both strategic flexibility and robustness, providing a foundation for real-time decision-making in robotic soccer under uncertainty.
This paper examines the pitfalls of word-for-word translation in learning Italian as a second language (L2). Drawing on translation studies and language pedagogy, it highlights how literal translations often distort meaning by ignoring cultural, semantic, and pragmatic complexities. Contrary to the belief that direct translation ensures accuracy, this approach frequently leads to awkward or misleading results, e.g., rendering “over easy eggs” as uova super facilmente instead of uova fritte. Italian-specific structures and conventions, such as the formal Lei or idioms like in bocca al lupo, illustrate the deep cultural embedding of language. Three key factors contribute to word-for-word mistranslation: structural differences between Italian and English, false cognates that create semantic confusion, and cultural-pragmatic gaps in idioms and social norms. High-stakes fields like marketing, literature, and international relations underscore the risks of misinterpretation. Advocating a communicative, functional approach, this paper emphasizes the need for cultural literacy, awareness of traditions, idioms, and symbols. It outlines classroom strategies such as contrastive analysis, peer review, and selective technology use. Through examples and case studies, it argues that translation is a process of cultural mediation rather than mechanical substitution. Educators, learners, and professionals must go beyond one-to-one lexical correspondence to foster true intercultural communication. Lost in Translation: Le insidie del trasferimento linguistico parola per parola Questo articolo analizza le insidie della traduzione parola per parola nell’apprendimento dell’italiano come lingua seconda (L2). Basandosi su studi di traduzione e pedagogia linguistica, evidenzia come le traduzioni letterali spesso distorcano il significato, ignorando complessità culturali, semantiche e pragmatiche. Contrariamente alla convinzione che la traduzione diretta garantisca accuratezza, questo approccio porta frequentemente a risultati imprecisi o innaturali, ad esempio, tradurre over easy eggs come uova super facilmente invece di uova fritte. Strutture e convenzioni italiane, come il Lei formale o espressioni idiomatiche come in bocca al lupo, dimostrano il forte radicamento culturale della lingua. Tre fattori principali contribuiscono agli errori di traduzione letterale: le differenze strutturali tra italiano e inglese, i falsi amici che generano confusione semantica e le discrepanze culturali e pragmatiche negli idiomi e nelle norme sociali. Settori di alto profilo come il marketing, la letteratura e le relazioni internazionali mettono in luce i rischi di un’interpretazione errata. Sostenendo un approccio comunicativo e funzionale, questo studio sottolinea l'importanza della competenza culturale, la consapevolezza di tradizioni, espressioni idiomatiche e simboli culturali. Presenta strategie didattiche come l’analisi contrastiva, la revisione tra pari e l’uso selettivo della tecnologia. Attraverso esempi e casi di studio, dimostra che la traduzione è un atto di mediazione culturale, non una semplice sostituzione meccanica. Docenti, studenti e professionisti devono superare la corrispondenza lessicale uno-a-uno per promuovere una comunicazione interculturale autentica.
In 2025, we held the fourth iteration of the DIS-RPT Shared Task (Discourse Relation Parsing and Treebanking) dedicated to discourse parsing across formalisms.Following the success of the 2019, 2021, and 2023 tasks on Elementary Discourse Unit Segmentation, Connective Detection, and Relation Classification, this iteration added 13 new datasets, including three new languages (Czech, Polish, Nigerian Pidgin) and two new frameworks: the ISO framework and Enhanced Rhetorical Structure Theory, in addition to the previously included frameworks: RST, SDRT, DEP, and PDTB.In this paper, we review the data included in DISRPT 2025, which covers 39 datasets across 16 languages, survey and compare submitted systems, and report on system performance on each task for both treebanked and plain-tokenized versions of the data.The best systems obtain a mean accuracy of 71.19% for relation classification, a mean F 1 of 91.57(Treebanked Track) and 87.38 (Plain Track) for segmentation, and a mean F 1 of 81.53 (Treebanked Track) and 79.92 (Plain Track) for connective detection.The data and trained models of several participants can be found at https://huggingface. co/multilingual-discourse-hub.
Online media has become the primary source of information for modern society due to its quick access and ease of news presentation. However, this development also poses challenges, particularly in maintaining writing quality, such as diction accuracy. Appropriate diction plays a crucial role in delivering clear messages, avoiding misunderstandings, and preserving media credibility. This study aims to analyze diction errors in the daily online news jateng.akurat.co edition of November 12, 2024, including non-standard word usage and word mismatches. This research employs a qualitative descriptive method to identify the types of errors, their causes, and their impacts on readers. Data were collected using documentation techniques by observing, recording, and analyzing the news texts published in the selected edition. The analysis process involved comparing diction usage with applicable linguistic norms and relevant contexts. The findings reveal several common errors, such as the use of terms that do not align with formal language norms, inappropriate word choices, and the improper adaptation of foreign words. Diction errors can create reader confusion, diminish media credibility, and affect perceptions of the information conveyed. This study highlights the importance of accurate diction selection in journalism, especially for regionally based media, which must consider cultural contexts and local values. These findings are expected to serve as a guideline for enhancing linguistic accuracy in digital journalism, maintaining media credibility, and supporting journalism's role as a pillar of a healthy democracy.
This study examines the linguistic landscape (LL) of Chikan Old Street, a historic district in Zhanjiang, Guangdong, through the lens of the SPEAKING model. As one of the most well-preserved historical districts in southern China, Chikan Old Street embodies a rich maritime heritage and commercial traditions, making it a compelling site for LL research. The study investigates the interaction between official language policies, regional linguistic identity, and globalization, providing insights into how language hierarchies are constructed in heritage sites. A mixed-methods approach is employed, integrating quantitative corpus analysis, qualitative semiotic interpretation, and public perception surveys. The findings reveal a clear stratification of language use: Chinese dominates official signage, with pinyin and English in subordinate positions, reflecting state-imposed linguistic norms. Private signage, however, demonstrates greater linguistic flexibility, incorporating Cantonese expressions, traditional Chinese characters, and creative bilingual adaptations. By highlighting the negotiation between top-down language standardization and bottom-up linguistic agency, this study contributes to broader discussions on language policy, cultural heritage preservation, and multilingual accessibility in historical districts. The findings underscore the need for improved linguistic planning, standardized translation policies, and greater public engagement in signage design to ensure that linguistic landscapes in heritage sites are both culturally authentic and globally navigable.
a Real close relationship exists between transferring and public media, imposing a noticeable light on the importance of language especially translation in casting discourse media since translation goes on an essential role in representing the diversity between cultures and perspectives in media, however translation may distorting the culture disparities and context unless not managed cautiously,in this sense, translators have to create a balance between loyalty to the original texts and the target ones taking into consideration the diversity in linguistic and cultural norms, this balance can be molded by applying certain approaches and strategies, this article aims to present theoretical and analytical paths to explore the diverse techniques and mechanisms used in transferring media discourse which are important to be adopted in the process of translation given that translation in the media can play a really malicious role in spreading misinformation and tendentious narratives, faulty or biased translation effects the public perception dramatically and fabricates stereotypes of misconceptions, accordingly, this study deals with real world media discourse and examines the strategies adopted by the translators, the findings show that the process of transferring from one language to another comprises modifying the linguistic norms and cultural backgrounds,spotting the light on power and responsibility that emerge during the process of translation since societies become increasingly interconnected so it is an urgent matter for translators and journalists to maintain ethical, cultural and linguistic standards to preserve accuracy and solidity of translated content and to protect the public perception from distorted narratives and misinformation as much as possible.
This research was conducted within the framework of gender linguistics, the main task of which is to study how grammatical categories related to biological sex are reflected in language and affect the perception of men and women in the minds of speakers of a particular language. The research paper focuses on the use of feminatives in Russian and Greek among bilinguals and monolinguals. The aim of the study was to identify trends in the use of feminine correlates of the names of professions and job positions in legal and political fields, as well as to determine the influence of the linguistic environment on the formation of linguistic norms regarding gender terminology. The authors employed an experimental approach based on a comparative analysis of the use of feminatives among Russian-Greek natural bilinguals aged 16-25 living in Cyprus and learning English as a foreign language, and artificial bilinguals, native speakers of Russian/Greek who also speak English. The experiment participants were given English sentences containing professional designations. The task set for them was to translate the sentences in such a way that the agentive subject or addressee was indicated as a female; after that a systematic analysis of the use of feminatives and their derivational models was conducted. This analysis revealed gender asymmetry in both linguistic contexts, reflecting the unequal representation of male and female genders in the linguistic consciousness of the speakers. The main conclusions emphasize the importance of the linguistic environment and cultural factors in shaping gender-specific lexicon used in everyday communication and media, and indicate the presence of interference in the speech of bilinguals.