Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
Child-directed speech, as a specialized form of speech directed toward young children, has been found across numerous languages around the world and has been suggested as a universal feature of human experience. However, variation in its implementation and the extent to which it is culturally supported has called its universality into question. Child-directed speech has also been posited to be associated with expression of positive affect or "happy talk." Here, we examined Canadian English-speaking adults' ability to discriminate child-directed from adult-directed speech samples from two dissimilar language/cultural communities; an urban Farsi-speaking population, and a rural, horticulturalist Tseltal Mayan speaking community. We also examined the relationship between participants' addressee classification and ratings of positive affect. Naive raters could successfully classify CDS in Farsi, but only trained raters were successful with the Tseltal Mayan sample. Associations with some affective ratings were found for the Farsi samples, but not reliably for happy speech. These findings point to a complex relationship between perception of affect and CDS, and context-specific effects on the ability to classify CDS across languages.
In this paper, we leverage pre-trained language models (PLMs) to precisely evaluate the semantics preservation of edition process on sentences. Our metric, Neighbor Distribution Divergence (NDD), evaluates the disturbance on predicted distribution of neighboring words from mask language model (MLM). NDD is capable of detecting precise changes in semantics which are easily ignored by text similarity. By exploiting the property of NDD, we implement a unsupervised and even training-free algorithm for extractive sentence compression. We show that our NDD-based algorithm outperforms previous perplexity-based unsupervised algorithm by a large margin. For further exploration on interpretability, we evaluate NDD by pruning on syntactic dependency treebanks and apply NDD for predicate detection as well.
The article describes issues on the communicative and pragmatic features of colloquial words in newspaper style. At present, the main source of updating the literary norm is the mass media, including periodicals. It actively responds to the events of each day and thereby reflects changes in the vocabulary of the language. As we know, the newspaper style is currently characterized by a irony, emotional intensity and expressiveness, which makes the language of journalism mobile, subtly responsive to situations in society. At the same time, the opposite phenomena are also observed - the loss of literary purity and a general decline in the style of the mass press. Consequently, while selecting lexical means, a journalist must remember that he/she is responsible for the future of the language, for developing the language taste of readers.
The article is devoted to the consideration of one of the aspects of the problem of the interlingual translation process: transformations. The relevance of studying the use of transformations is due to the goal of achieving equivalence and the condition of observing the norms of the target language. Based on the translation of J. Bowen's novel «A Street Cat Named Bob» by comparing the source text with the translation text, the article examines the lexical transformations used in the translation of descriptions of the urban environment static elements from English into Russian. Particular attention is paid to the issues of defining a translation transformation, the choice of a convenient classification, as well as the reasons for using certain transformations. The study confirms the opinion of the majority of linguists that the use of translation transformations is inevitable: it is due to the differences in the lexical and grammatical systems of the two languages. E. I. Kolyabina, translator of the novel “A Street Cat Named Bob”, resorts to a large number of transformations and sometimes uses them in combination. Analysis of the translation text shows that the choice of transformation can be forced, or dictated by the translator’s personal preferences.
Abstract: The article offers a brief overview of the projects developed by the Hispanic Seminary of Medieval Studies since the publication of Kasten's article in 1978, where he briefly explains the decision to abandon the compilation of the paper-based Tentative-2, in favor of a computer based lexical database, and the origins of the Dictionary of the Old Spanish project. The article pays particular attention to two projects related to the paleographical transcriptions produced by HSMS's collaborators: the Digital Library of Old Spanish Texts, which provides interactive access to a the paleographic transcriptions, as well as to a series of indexes (alphabetical, frequency, reverse alphabetical), and concordances in KWIC format; and the Old Spanish Textual Archive, a lemmatized and morphologically tagged linguistic corpus of about 35,000,000 words of medieval texts written in Spanish, Asturian, Leonese, Navarro-Aragonese, and Aragonese.
Introduction. Along with such important issues of speech culture as mastering the rules of spelling and punctuation, lexicology, it is important to study and use the means of expression depending on the purpose and content of expression, governed by grammatical norms of language, deviations from which are morphological and syntactic levels. Violations of morphological norms indicate insufficient attention of media workers to the culture of professional broadcasting, as well as unprofessional editorial processing of materials presented to readers. The purpose of the article is to describe morphological abnormalities in online media, identify the most typical of them and indicate the main erroneous places within the lexical and grammatical class of nouns, numerals, verbs, adverbs formed from numerals. The source base of the study was journalistic texts of online publications of national and regional significance ("Ukrainian Truth", "Doba"). Results. Analysis of morphological deviations within the lexical and grammatical class of nouns showed violations in its formation. Significant difficulties in the inflection of nouns is the genitive singular of the second declension of the masculine gender, namely the correct choice of the ending a (-я) or -у (-ю). In journalistic texts, these endings are interchangeable. he grammatical design of the dative singular of masculine nouns is normative in the analyzed texts. Although most normative sources indicate that the parallel endings овіi (-івi, -євi) / -у (-ю) can be used to denote them, it is stated that the inflections - ові (-eвi, -євi) should predominate in the names of creatures), and in the names of non-beings mostly use -у (-ю). Improper formation of compound numerals also leads to morphological errors. Within the lexical-grammatical class of verbs, the frequency associated with the formation of the imperative mood with the help of the verb tracing paper from the Russian davay / davayte. A powerful source of language pollution of the analyzed online publications is the use of active verbs not peculiar to the Ukrainian language. Forms of active verbs of the present tense in -uch (-yuch), -ach (-yach) are not typical for the Ukrainian language. Conclusion. Thus, the highest frequency in the media texts of the analyzed online publications at the level of morphology are errors associated with the use of uncharacteristic of the Ukrainian language active verbs of the present tense on -уч (-юч), -aч (-яч), with the formation of nouns.
The article examines the features of grammatical adaptation of the borrowing “art” in the modern Ukrainian language, identifies typical violations of current spelling rules in words with this component, suggests the correct options and clarifies the feasibility of its frequent use in modern Ukrainian speech. It has been found that the English borrowed component “art” is well adapted to the recipient language. This is evidenced by its active use in both oral and written speech, the ability to be used independently with a fixed meaning, high word-forming productivity. It is revealed in the texts of the mass media that English borrowing “art” functions in the modern Ukrainian as a inflective masculine noun of the second declension of the hard group with the semantics “art, fine arts”. It is observed that having entered the modern Ukrainian lexical system, it takes an active part in word formation, showing high productivity in the process of forming compound nouns. Two models of forming new nouns with the component “art” are differentiated, in which “art” with the meaning “artistic” is the first part of a compound word and “art” with the semantics “art” is the second part of a new lexeme. It is established that the first model is the most productive because of the largest number of these nouns in Ukrainian. It is proved that according to their grammatical nature lexemes with “art” as a reference word are not Ukrainian juxtaposites, as they are not formed on the basis of the Ukrainian, but are borrowed in it like ready-to-use integral lexical units, which speakers adapted to the grammatical structure of modern Ukrainian. In accordance with the current spelling rules the nouns with borrowed component “art” have to be written together. Found in the texts of mass media and oral speech the new lexemes with two prepositive foreign attributives are classified as juxtaposites with a declension of the second part; it is recommended to write them in accordance with the norms of the current “Ukrainian spelling” with a hyphen. It was found that the names of establishments, events, places of cultural events, etc., which consist of proper names or abbreviations combined with borrowing “art”, completely copied from the English word formation, are not typical for Ukrainian and artificially adapted to it in violation of current spelling rules. The article offers their alternative names that meet the grammatical and spelling rules. The analysis of the peculiarities of the adaptation of English borrowed component “art” in the modern Ukrainian will promote a faster study of the grammatical nature of borrowings, their word-forming potential, the correct linguistic qualification of new derivatives and their further normalization.
Changes in language are largely a process of loosening the old norm and gradually creating a new one. It is believed that the media contribute to the preservation of the norm and the maximum slowdown of changes. Newspapers and magazines fix the graphic appearance of the word, and radio and television — sound. Together they set grammatical, syntactic and other patterns, which guides the people who read and listen to them. The author of the article studied the speech features of the interview genre in modern Tatar journalism. Namely 3012 fragments were analysed and the following conclusions were made: due to the interview genre, phonetic, morphological, lexical, syntactic, orthoepic enrichment of the language occurs; many modern elements borrowed from the Internet are used in print magazines in the Tatar language; there is a tendency to reduce words, use abbreviations, typed language constructions, phraseological units, borrowings, terms, professionalisms, dialectisms, slang expressions, and it is also observed that in different printed publications in the Tatar language, different variants of using common terms are offered in interviews.
The article’s main aim is to emphasize that mastering the skills of literary pronunciation in reading lessons in primary school should be carried out consistently and systematically in the lessons of the Azerbaijani language. The Azerbaijani language subject provides pupils with an initial understanding of phonetic, spelling, orthoepic, lexical norms, and grammatical structure of words. At the same time, it helps pupils master the art of communication. To meet the aim of the idea, a descriptive and data gathering methods are utilized. Based on the results acquired, it can be concluded that learning of reading teaches a simple course of literature in elementary school, lays the foundation for developing pupils' reading and speaking skills. Moreover, traditional and active (interactive) teaching methods should be used in reading lessons to develop literary pronunciation skills.
The article examines lexical-semantic, syntactic and stylistic features of a female author’s self-presentation in a Russian-language dating advertisement. At first, the essence of the strategy of self-presentation is observed briefly, as well as the compositional and semantic structure of the dating advertisement. Further, the main linguistic and stylistic tools, contributing to the implementation of the strategy of self-presentation, are analyzed. The research was carried out on the material of advertisements posted on Russian-language dating sites, as well as on the material of completed profiles on social dating networks (more than 200 ads and profiles were studied in total). As a result of the research, it was determined that the dating advertisement includes the subtext of self-presentation, or the subtext of the addresser, the subtext of the addressee and the subtext of the future. The presence of the addresser's subtext in the structure of a dating advertisement is invariant. The self-presentation of the female author is realized via the stylization of the text with the help of lexemes, describing character traits and appearance, figurative language, euphemisms, as well as low colloquial speech. The use of expressive units that are not the part of the standard language is more peculiar to authors in the age group of people 18 to 25 years. The researched material indicates the presence of a relation between the author`s age and the frequency of using low colloquial and invective speech, as well as non-verbal means (e.g. emoji). The punctuation spelling norms of the Russian language are often ignored in dating advertisements. However, no obvious link has been established between the frequency of punctuation and spelling errors in the advertisement and the author’s age group. In the analyzed advertisements, there is significant variability of syntactic structures, through which the strategy of self-presentation is implemented. Elliptical sentences with a bright author's expression are the most common.
В статье рассматриваются специфические черты терминологии хендмейд - творческой деятельности, связанной с созданием изделий ручной работы. В современном русском языке это динамически развивающаяся область терминообразования, которая требует лингвистического описания и словарной фиксации. Объектом исследования являются лексические единицы тематической группы «Хендмейд» в количестве около 200 слов. Целью исследования является определение состава и особенностей данной терминологии. С одной стороны, она демонстрирует основные черты отраслевой терминологии, с другой, - для терминологии хендмейд характерны образность, экспрессивность, полисемия, что показано на примерах. Данная терминология формируется преимущественно стихийно в рамках разговорной профессиональной коммуникации в интернет-дискурсе. Выделяются шесть основных тематических групп рассматриваемых терминов. Терминологию хендмейд можно рассматривать как систему систем, так как вокруг центрального понятия «техника или вид рукоделия» формируется разветвленная система номинаций, связанных с этой понятийной областью. Терминология хендмейд не имеет лексикографической фиксации в виде специального словаря, и нормы правописания многих слов еще только формируются. Очевидно, что в дефиницию подобных лексических номинаций желательно включать элементы толкования, а сам словарь считать толково-дефиниционным. The article studies the specific features of hand-made (creative activity associated with handwork) terminology. This is a dynamically developing area of the term formation in modern Russian, which requires linguistic description and dictionary fixation. The object of the research is the lexical units of the thematic group "Handmade" in the amount of about 200 words. The purpose of the research is to determine the composition and features of this terminology. On the one hand, it demonstrates the main features of industry terminology. On the other hand, handmade terminology is characterized by imagery, expressiveness, polysemy, as shown in examples. This terminology is formed mainly spontaneously in the framework of spoken professional communication in the Internet discourse. There are six main thematic vocabulary groups. Handmade terminology can be viewed as a system of systems. Around the central concept of “technique or type of handiwork” is formed a ramified system of nominations. Handmade terminology does not have a lexicographic fixation in the special dictionary, and the spelling norms of many words are still being formed. Obviously, that the definition of such lexical nominations should include elements of interpretation. And the dictionary itself should be considered explanatory and definitional.
AbstractThe article examines the influence of the Arabic, Persian-Tajik, Russian languages on the structure of the dictionary of the Uzbek language, interference errors in language learning, their causes, views on interference, types of interference, elimination of speech errors associated with interference, methods are described. The development of vocabulary in the dictionary of the Uzbek language took a strong place in the dictionary arsenal of our language by two factors: through the colloquial language and the biblical language. Such non-zero factors, due to the development of Science and technology, are now entering our lexicon very rapidly, as a result of using their own language without finding an alternative to this word, lexical interperience takes place from colloquial speech. Over time, the words adopted from the Arabic, Persian-Tajik, Russian languages have lost their interperensiveness and have passed into the literary norm.
Постановка задачи. Современный этап обучение иностранному языку в техническом университете основывается на компетентностном подходе. Немаловажную роль в языковой подготовке будущих специалистов играет формирование социолингвистической компетенции обучающихся, которую в отечественной методике принято включать в состав социокультурной компетенции. Формирование социолингвистической компетенции студентов в процессе занятий по английскому языку имеет свои особенности и связано с ознакомлением обучающихся с культурой и историей стран изучаемого языка, правилами и нормами речевого поведения языковых носителей, явлениями территориально и социально обусловленной языковой вариативности. Результаты. Проведенное исследование дает основание утверждать, что типичные ошибки и трудности студентов, возникающие в процессе иноязычной коммуникации и связанные с выбором социально и территориально обусловленных фонетических и лексических вариантов, свидетельствуют о недостаточной сформированности у обучающихся социолингвистической компетенции. Выводы. Направленность внимания преподавателя на формирование данной компетенции в учебном процессе, объяснение студентам природы явления языковой вариативности, анализ причин неоднозначности номинации в каждом конкретном случае позволяют обучающимся успешно преодолевать проблемы выбора языкового варианта в процессе коммуникации, становиться более уверенными пользователями английского языка. Statement of the problem. At present, the teaching of a foreign language at a technical university rests on a competency-based approach. An important role in the linguistic education of the future specialists plays the formation of the students’ sociolinguistic competency, which is traditionally included in the structure of the sociocultural competency according to the national methodology.The formation of the sociolinguistic competency of the technical university students at the lessons of English has its distinctive features and is associated with introducing them to the culture and history of the country of the studied language, working with authentic professionally oriented materials, studying the norms of the native speakers’ communication behavior, the phenomena of territorially and socially determined linguistic variability. Results. The research that has been carried out gives reason to argue that the students’ typical mistakes and difficulties associated with the choice of the geographically determined phonetic variants and the geographically and socially determined lexical variants in the process of communication indicate that the sociolinguistic competency of the students is formed insufficiently. Conclusion. The focus of the teacher’s attention on the formation of the sociolinguistic competency in the educational process, explaining to the students the nature of the language variation phenomena, analyzing the reasons for the ambiguity of the nomination in each special case help the students to successfully overcome the problems of choosing a language variant in the communication process, and to become more confident users of English.
The article is devoted to the study of lexical and grammatical features of epistolary addresses (on the material of “Letters to Oles Honchar” compiled by M. Stepanenko). The address is interpreted as one of the manifestations of human communication needs which serves to establish and maintain speech contact, as well as to express the emotional and evaluative characteristics of the interlocutor. An epistolary address is a word or phrase by which the author of a letter nominates his addressee in the text of a written message to establish contact with him.
 We processed 895 letters to Oles Honchar, in which 1185 addresses in Ukrainian and about 200 units in other languages had been recorded. Lexical features of addresses represent their belonging to the following semantic groups: addresses-anthroponyms (name, patronymic and surname); traditional etiquette forms (пан, товариш); general addresses (names of persons by generic or gender feature; names of persons by kinship in the indirect sense; names of persons by friendly relations); special addresses (names by profession, type of activity, position, academic titles); occasional addresses. Most often, senders address Oles Honchar by patronymic or by name, using it in full or in short form, and sometimes by surname.
 The lexical and semantic content of addresses depends on the intention of the speaker, his politeness, knowledge of language etiquette and the peculiarities of the relationship with the writer. In order to strengthen the address, attributive distributors expressed by honorific and emotional-evaluative adjectives аre used. Honorific adjectives (шановний, високошановний, найшанованіший, глибокошановний, вельмишановний, високоповажний, etc.) convey a polite attitude and perform etiquette function. Emotional-evaluative adjectives (дорогий, славний, щирий, незабутній, рідний, любий, коханий, etc.) denote sincerity, friendliness, friendly affection and perform an evaluative function. We reveal a significant proportion of constructions in which adjectives of both groups are used. This causes a change in the tonality of the communicative situation and reduces interpersonal distance. Possessive pronouns мій, наш, which have partially lost the meaning of possessiveness, strengthen the intimacy, cordiality and sincerity of the relationship. Addresses in Russian, Belarusian, Polish and English are described.
 It is found that the grammatical differentiation of addresses directly depends on lexical and grammatical features (proper or common names and substantivized parts of speech) and morphological means of their expression. It is confirmed that the typical morphological form of addresses is the vocative case of the noun, as well as the homonymous nominative case in letters written during the Soviet period. Violations of morphological norms (different case forms of lexical phrase components, a non-normative form Олесе) and orthographic mistakes in spelling of the writer’s patronymic are revealed. The non-normative form of the nominative case as a means of expressing the address in letters dated 1990–1995 is substantiated.
 The results of the research show that the most frequent lexeme is Олесю Терентійовичу. Forms Олесь Терентійович and Олесю are less used. Quantitative indicators of addressing forms are summarized in the table.
 We see the prospect of further scientific research in deepening other vectors of analysis of addresses, in particular in the study of their functional and stylistic potential.
Abstrak: Tulisan ini merupakan penelitian mengenai nyanyian daerah, di Bima Nusa Tenggara Barat yang disebut dengan Ntoko Mbojo. Tulisan ini mengkaji Ntoko Mbojo dengan menggunakan analisis wacana dalam paradigma etnolinguistik. Metode yang digunakan adalah teknik dasar berupa teknik pancing dan teknik lanjutan yang berupa teknik cakap bertemu muka. Selain itu, pengumpulan data juga menggunakan metode observasi partisipatoris. Struktur wacana Ntoko Mbojo yang dipentaskan semalaman tersebut terdiri dari 17 ntoko (irama) dengan 2 irama yang berulang. Unsur pembentuk Ntoko Mbojo terdiri dari dua unsur yang saling bertautan yaitu unsur irama dan unsur verbal. Sedangkan unsur verbalnya menggunakan bahasa mbojo umum, tetapi terkadang ada sedikit campur kode dengan bahasa Indonesia. Konteks wacana Ntoko Mbojo juga berpengaruh antara lain dari aspek latar, partisipan, hasil, pesan, cara, sarana, norma dan genrenya. Tulisan ini juga menunjukkan bahwa terdapat kohesi gramatikal (4) dan leksikal (6) dalam wacana Ntoko Mbojo. Sedangkan koherensi yang memadukan sejumlah irama dengan syairnya ialah wacana tentang cinta. Kemudian aspek kebahasaan yang menonjol adalah gaya bunyi aliterasi dan asonansi, gaya bahasa ironi, metafora, repetitio dan pilihan kata. Kata kunci: Ntoko Mbojo; Bima; analisis wacana; kohesi; koherensi Abstract: This research is a study of the oral traditions, folk song, in the Province of West Nusa Tenggara Bima called Ntoko Mbojo. This study examines the Ntoko Mbojo using discourse analysis in the paradigm of ethno-linguistic. The method used is the basic techniques such as fishing techniques and continue techniques in the form of a conversation face to face. In addition, collecting data is also using participatory observation method. Discourse structure of Ntoko Mbojo performed was consists 17 ntoko (rhythm) with two repetitive rhythm. Ntoko Mbojo elements consist of two intertwine elements, namely the elements of rhythm and verbal elements. Rhythm analysis looks at the short length of the tone, and sometimes steady or not steady. Whereas the verbal element mbojo common language, but sometimes there is a bit of code-mixing with Indonesia language. Contexts of Ntoko Mbojo also influences among aspects like as background, participants, ends, art sequences, keys, instrumentalities, norms and genre. The research also showed that there were grammatical cohesion (4) and lexical (6) in a Ntoko Mbojo discourse. While coherence that combines a number of rhythms with his lyric is a discourse on love. Then the linguistic aspect that stands out is the sound style alliteration and assonance, style irony, metaphor, repetitive and diction. Keyword: Ntoko Mbojo; Bima; discourse; cohesion; coherence
Ovaj se rad bavi člancima hrvatskoga jezikoslovnog časopisa Jezik, točnije onim člancima koji govore o leksičkoj razini hrvatskoga jezika u razdoblju od 1952. do 1990. godine. Proučavani korpus sadrži teorijske članke, ali u najvećoj mjeri one savjetodavne prirode. Široka tema leksičke norme podijeljena je u nekoliko tematskih podskupina koje su se pokazale najzastupljenijima. To su neologija, sinonimija, leksičko posuđivanje, nazivlje i onomastika. Zajedničko je svim navedenim temama što su u velikoj mjeri obilježene purističkim tendencijama, čemu su uzroci velik prodor posuđenica iz drugih jezika (prvenstveno iz ruskoga i engleskoga) i specifična politička situacija u kojoj se hrvatski jezik nalazio u Jugoslaviji. Na kraju rada nalazi se rječnik leksema koje su iz različitih razloga autori u Jeziku smatrali problematičnima te prijedlozi za njihovo zamjenjivanje, kao i rječnik sinonimnih parova koje su autori popratili dodatnim objašnjenjima.
<strong>Introduction</strong> Quotebank is a dataset of 235 million unique, speaker-attributed quotations that were extracted from 196 million English news articles (127 million containing quotations) crawled from over 377 thousand web domains (15 thousand root domains) between September 2008 and April 2020. The quotations were extracted and attributed using Quobert, a distantly and minimally supervised end-to-end, language-agnostic framework for quotation attribution. For further details, please refer to the description below and to the original paper: Timoté Vaucher, Andreas Spitz, Michele Catasta, and Robert West<br> "Quotebank: A Corpus of Quotations from a Decade of News"<br> Proceedings of the 14th International ACM Conference on Web Search and Data Mining (WSDM), 2021.<br> https://doi.org/10.1145/3437963.3441760 When using the dataset, please cite the above paper (Note that the above numbers differ from those listed in the paper, as the updated data in this repository has been computed from an expanded set of input news articles). <strong>Dataset summary</strong> The dataset consists of two versions: <strong>Quotation-centric version</strong> (<em>quotes-YYYY.json.bz2</em>)<br> An aggregated set of unique quotations with the most likely speaker. Each unique quotation occurs only once in this version of the data and the probabilities of the candidate speakers to which the quotation can be attributed are aggregated over all occurrences of the quotation. This version of the data is a minimal - but complete - list of attributed quotations that is aimed at users who only require quotation-speaker attributions, but no individual contexts for these quotations from the original articles. <strong>Article-centric version</strong> (<em>quotebank-YYYY.json.bz2</em>)<br> A complete set of all individual quotation mentions with associated speaker as well as the article context in which they are mentioned. This larger version contains one entry per article in the news data. Each entry contains all speakers that appear in the news article as well as the (attributed) quotations, alongside a context window surrounding the quotations. Both versions are split into 13 files (one per year) for ease of downloading and handling. <strong>Dataset details</strong> The following formatting applies to both versions of the dataset: All data is made available in JSON format that has been compressed using bzip2. The data is split per year (i.e., there is one file for each calendar year). The offsets of quotations, contexts, and speaker annotations are given in units of Penn TreeBank Tokenizer tokens. Offsets are zero-based and are computed from the start of the article. When pairs of offsets are provided, the end offset is non-inclusive (e.g. in Python you can call tokens[start:end] without having to do end+1). The Spinn3r data from which Quotebank was extracted had been collected over the course of over a decade. During this time, the client-side code used for collecting the data changed several times, and various character-encoding-related issues led to different representations of the original text at different times. We thus divide the 12 years spanned by the Spinn3r corpus into five phases (Phases A through E). A detailed description is available on GitHub; the key takeaways are that (1) text was lowercased in Phases A, B, and C, whereas the original capitalization was maintained in Phases D and E, and that (2) non-ASCII characters are properly represented only in Phase E. <br> <strong>Version 1: Quotation-centric data</strong> In this version of the dataset, the quotations are aggregated across all their occurrences in the news article data, and assigned a probability for each speaker candidate. We consider two quotations to be equivalent and suitable for aggregation if they are identical after lower-casing and removing punctuation. <pre><code>Quotation-centric data |-- quoteID: Primary key of the quotation (format: "YYYY-MM-DD-{increasing int:06d}") |-- quotation: Text of the longest encountered original form of the quotation |-- date: Earliest occurrence date of any version of the quotation |-- phase: Corresponding phase of the data in which the quotation first occurred (A-E) |-- probas: Array representing the probabilities of each speaker having uttered the quotation. The probabilities across different occurrences of the same quotation are summed for each distinct candidate speaker and then normalized |-- proba: Probability for a given speaker |-- speaker: Most frequent surface form for a given speaker in the articles where the quotation occurred |-- speaker: Selected most likely speaker. This matches the the first speaker entry in `probas` |-- qids: Wikidata IDs of all aliases that match the selected speaker |-- numOccurrences: Number of time this quotation occurs in the articles |-- urls: List of links to the original articles containing the quotation </code></pre> Note that for some speakers there can be more than one Wikidata ID in the `qids` field. To access Wikidata information about those speakers it is necessary to disambiguate them, i.e., select one of the listed Wikidata IDs that most likely corresponds to the respective speaker. Speaker disambiguation can be done using scripts available in the quotebank-toolkit repository. Additionally, the repository contains useful scripts for cleaning and enriching Quotebank. <strong>Version 2: Article-centric data</strong> In this data set, individual quotations are not aggregated. For each article, one JSON entry contains all speakers that appear in the news article, the (attributed) quotations, and the text within a context window surrounding each of the quotations. <pre><code>Article-centric data |-- articleID: Primary key |-- articleLength: Length of the article in PTB tokens |-- date: Publication date of the article |-- phase: Corresponding phase in which the article appeared (A-E) |-- title: Title of the article |-- url: Link to the original article |-- names: List of all extracted speakers that occur in the article |-- name: Surface form of the first occurrence of each speaker in the article |-- ids: List of Wikidata IDs that have `name` as a possible alias |-- offsets: List of pairs of start/end offset, signifying positions at which the speaker occurs in the article (full and partial mention of the speaker) |-- quotations: List of all the quotations that appear in the article |-- quoteID: Foreign key of the quotation (from the quotation-centric dataset) |-- quotation: Text of the quotation as it occurs in this article |-- quotationOffset: Index where the quotation starts in the article |-- leftContext: Text in the left context window of the quotation (used for the attribution) |-- rightContext: Text in the right context window (used for the attribution) |-- globalProbas: Array representing the probabilities of each speaker having uttered the quote *at the aggregated level*. Same as `probas` for a given `quoteID` |-- globalTopSpeaker: Most probable speaker *at the aggregated level*. Same as `speaker` for a given `quoteID` |-- localProbas: Array representing the probabilities of each speaker having said the quote *given this article context*. |-- proba: Probability for a given speaker |-- speaker: Name of the speaker as it first occurs in this article |-- localTopSpeaker: Selected speaker. Same name as the first entry in `localProbas` |-- numOccurrences: Number of times this quotation occurs in any article </code></pre> <strong>Code repository</strong> The code of Quobert that was used for the extraction and attribution of this data set is available and managed in a Github repository, which you can find here.
Background. In linguistics, euphemisms are the subject of much research. The scientific literature has given different descriptions of this phenomenon, which suggests that euphemisms are multifaceted, of a changing nature.A modern linguistic approach to euphemisms used in place of words that are morally and culturally inappropriate among members of society, especially medical euphemisms in doctor-patient communication, to come to new scientific and theoretical conclusions, their lexical-semantic, methodological-functional, linguopragmatic, gender and structural It is important to explain the features. This article identifies the structural groups of medical euphemisms. We expressed them in the following order. Euphemism is characterized by a high degree of mobility. It also serves to replace the tabooed(tabulashtirilgan) word. For example: to go to the place of death; like a tumor at the site of the tumor. The use of euphemisms in language has been shaped as a historical ethnographic phenomenon associated with the taboo phenomenon. Euphemisms are associated with the development of customs, cultural levels, aesthetic tastes, and ethnic norms in nations. As language develops, so does the euphemistic layer within it. On the basis of new morals, new worldviews, new forms of it emerge. A euphemism is an occasional, individual, contextual unit that replaces a certain word or phrase with a purpose, softens the word that expresses the original essence, and “wraps it in paper”Medical euphemisms have the forms of word-euphemism, compound-euphemism (free compound and fixed compound (phrase)), speech-euphemism, depending on the form of expression. Word-euphemism, phrase-euphemism, compound-euphemism serve to name diseases, body parts, some physiological processes (and others) in speech.\n\nMethods. The article uses methods of description, classification, contextual analysis.\n\nResults. The occurrence of medical-related physician speech euphemisms at the language level was revealed, and structural groups of euphemisms in physician speech were classified.\n\nConclusion: Medical euphemisms have the forms of word-euphemism, compound-euphemism (free compound and fixed compound (phrase)), speech-euphemism, depending on the form of expression. Word-euphemism, phrase-euphemism, compound-euphemism serve to name diseases, body parts, some physiological processes (and others) in speech. Speech-euphemism occurs in medical speech as a component of a compound sentence (subordinate clause), in the form of a simple sentence ([WPm] device), in the participle structure (circle) and performs its function.
In this article, we analyse the objects of wine-drinking taken from The Zone story by Sergei Dovlatov and their translations into Spanish (the translated book was published by the Ikusager Ediciones publishing house in 2009), and propose our own translation solutions. For this purpose, we considered three types of national realities. First of all, there are concepts that appeared in the original ethnoculture and bear the semantic load which is well-known to native Russian speakers and often unknown to Spanish native speakers (samogon, braga, bormotukha, zveroboy). Secondly, there are names of alcoholic beverages borrowed from other cultural traditions. They were incorporated into the structure of the national Russian language and are either unknown to foreigners (chacha, shnaps, shartrez), or have other (additional) meanings in the original culture, which do not exist in the target language and cannot be perceived by its readers (portveyn, vermut, los’on, odekolon, spirt). To translate fiction, one needs to have not only deep background knowledge, but also knowledge of the culture of a foreign language. This will help in the process of identifying the differences between cultural traditions of the source and target languages. In particular, lexical units related to national realities can be determined as accurately as possible, and the ways to translate them into a foreign language can be found. We propose, following Lawrence Venuti (1994), two main ways of their transmission, namely: domestication (deliberate disregard of the linguistic and cultural norms of the target language and culture) and foreignization (orientation of the target text to the language system and values of the target culture), and different translation techniques depending on the characteristics of the national reality, knowledge of which is necessary for the translator and the extent to which the national reality is understandable in the target language to a reader not specialized in Russian culture. In general terms, in this article, we show that, when translating a text, it is more reasonable to choose the method of domestication or to use native words denoting close or semantically similar meaning even though it is not absolutely identical (it can range from a stylistically neutral designation of a word, its partial neutralization to absolute neutralization and descriptive translation). In some cases, one can also resort to foreignization, using such techniques as transliteration in the form of calque or borrowings, which in some cases require textual explanations or notes.
The purpose of this study is to investigate German native speakers' use of exclamation marks in L2 Danish texts by comparing their use to that of Danish and German native speakers respectively. The comparison thus aims to identify points of attention in L2 Danish writing. Writing skills in a foreign language include not only lexical, grammatical and textual knowledge, but also knowledge of the sociopragmatics in the foreign language. The norm-adequate use of the exclamation mark appears to embody such knowledge in several respects. As part of the exclamation as a rhetorical device, the exclamation mark serves to increase intensity in the text and accentuate the writer's emotional involvement in his text. At the same time, the exclamation mark also influences and affects the recipient. It thus has the function of an interactive sign. Based on a comparison of the standards in the two language systems and on speech act theory, the use of the sign is examined in e-mails and commentaries. The study shows that the use of the sign is genre dependent and that the L2 writers overuse the sign in the expressive and evaluative acts of expressing thanks and approval in the e-mails. In terms of didactic implications, the study draws attention to the formulaic language used in the co-text of the exclamation mark.
This article explores how the death of the elderly is represented by examining the death notices of 3,160 people aged 65 and over that were published in two daily newspapers in French-speaking Switzerland. Notices no longer merely announce death, but have become more extensive since the 1950s, providing information about the death and the dead person. The words chosen by the family to announce that an elderly relative has passed away reveal several social norms. Using textual statistics, our analyses show that representations of death differ according to age, sex, religion, and place of death. Thus, a plural vision of death in old age emerges, one that goes beyond a simple opposition between good and bad death. The age criterion is important, reiterating the distinction found in old age between the young-old and the oldest-old. We identify what is perceived as an “unfair” age and a “normal” age to die. The death notices document the use of a vocabulary that associates death in the third age with the fight against disease, justifying an early departure. On the other hand, when the deceased leaves us at the age of 85 or older, the vocabulary borrows figures and metaphors from the lexical field of sleep, such as “fell asleep,” or “fell asleep peacefully.”
[EN] This article explores how the death of the elderly is represented by examining the death notices of 3,160 people aged 65 and over that were published in two daily newspapers in French-speaking Switzerland. Notices no longer merely announce death, but have become more extensive since the 1950s, providing information about the death and the dead person. The words chosen by the family to announce that an elderly relative has passed away reveal several social norms. Using textual statistics, our analyses show that representations of death differ according to age, sex, religion, and place of death. Thus, a plural vision of death in old age emerges, one that goes beyond a simple opposition between good and bad death. The age criterion is important, reiterating the distinction found in old age between the young-old and the oldest-old. We identify what is perceived as an “unfair” age and a “normal” age to die. The death notices document the use of a vocabulary that associates death in the third age with the fight against disease, justifying an early departure. On the other hand, when the deceased leaves us at the age of 85 or older, the vocabulary borrows figures and metaphors from the lexical field of sleep, such as “fell asleep,” or “fell asleep peacefully.”
In this article I explore the construction of singing child characters in Isaac Watts’ Divine and Moral Songs for Children (1715) and Christopher Smart’s Hymns for the Amusement of Children (1771). The first part focusses on the nature of the lyrical persona within the lexical fields »voice and vocal sound« and »religion« and also looks at the possible addressees. The second part examines stylistic, phonetic, and formal elements, and explores their role in constructing the ›singing I.‹ To show the potential of Watts’ »Against Quarrelling and Fighting« to function as an invitation to playfully adopt behaviour opposed to Christian norms, the article examines a performance of Let Dogs Delight to Bark and Bite, a chorale by Matthew J. Zimnoch, whose text is taken from Watts’ hymn. Combining approaches from research on children’s poetry with ones from the interface of children’s literature and hymnody, the article also integrates a digitally supported close reading. The hymn texts were inputted into f4analyse, a software used in text linguistics and the social sciences, which allows for the assignment of categories, such as positive self-connotation of the ›singing I‹ or rhyme patterns. In conclusion, the article evaluates the potential of such a digitally supported research methodology for future research at the intersection of children’s literature and digital humanities.
Introduction As a literary technique, hybridization seeks to break the tradition of marriage within the family of words and the creation of new networks of companionship. It has a significant role in the beauty and structural coherence of literary texts. Since words lose their effectiveness due to constant association with each other, by deviating from the norms governing the standard language, words can be selected from one family and placed next to other families to highlight literature, open a new horizon, and affect the audience. This literary device became popular in literature via Saint-John Perse, the eminent French poet (1887-1975), in the works of other contemporary Iranian and Arab poets. Before him, however, the Russian theorist Mikhail Bakhtin had laid the foundations of hybridization by proposing polyphony and monophony theory in fiction. Among the poets who are influenced by this literary technique; Shāmloo and Adūnīs can be mentioned. The two poets have a relatively similar worldview due to their almost common social and political conditions. Methodology Conducted via a descriptive-analytical research method, the present study aims to examine the aesthetic dimension of this literary technique and its implications and effectiveness by examining the use of hybridization in Adūnīs’s and Shāmloo’s poetry, which has led to the departure of the language of their poems from the standards of natural language. This article is an attempt to comparatively study Adūnīs’ and Shāmloo’s poetry for introducing them to the reader, and to answer the following questions: 1- What is the use of hybridization in the two poets’ poetry? 2- What are the implications of this literary device in the two poets’ poetry? Discussion One of the techniques of "pretending to be familiar with familiar things" (Shafi'i Kadkani, 2013, p. 308) is hybridization via which words from a family are combined with a new language family, creating a kind of new poetic language. Among the poets who have been influenced by the Perse and who tried this literary technique are Shāmloo and Adūnīs. Part of hybridization in Shāmloo’s and Adūnīs’ poetry arises from the fusion of words from the family of music with the family of nature. Shāmloo takes the word "symphony" from the world of music and the word "night" from nature, and by combining them in “the sunset of the Siahrud River”, he creates the following combination: Symphony of the night drips / quiet / over the evening’ sadness (Shāmloo, 2004, p. 326). In this poem, Shāmloo highlights the composition (symphony of the night) by quoting the verb dripping and has also become unfamiliar with the verb dripping. The combination of the word “symphony” with “night” evokes the pleasant feeling of being in unison with the calm of the night. Since the word night alone is devoid of such a meaning, the poet gives it a new meaning by using the technique of hybridization. This has also led to lexical and semantic aberrations in the form of similes. In an atmosphere of emotional resemblance to Shāmloo, Adūnīs also combines the word “bell” from the family of music with the word “night” from the network of nature in the poem “the Divine Wolf”: The morning is burning and displaced / And I am the death of the moon / Under my face the bell (song) of the night breaks / And I am the new wolf of inspiration (Adūnīs, 1996, Vol. 1, p. 227). In the hybridization of night and bell, it is a kind of lexical aberration in the companionship axis, which creates a new combination of clichéd words in the form of simile and has led to semantic aberration, indicating signs of darkness and sorrow. The linguistic-grammatical network and the nature network are other networks that promote the poems of these two poets. Concerning this network, Shāmloo chooses the word syllable from the system of language and the wave from the network of nature in his poem “the anthem of the one who left and the one who remained” and uses hybridization as follows: In the darkness of the salty beach / We listened to the repeated syllables of the wave (Shāmloo, 2004: 550). The metaphorical hybridization of repeated syllable of the wave prepares the ground for lexical and semantic aberrations. Using repeated adjectives, the poet emphasizes the fluidity and movement of the waves, and by adding a wave to the syllable, he de-familiarizes the verb to listen. The wave alone has no meaning, but it is meaningful in terms of the poet's intention when it is put in a hybridization. In the surrealist atmosphere of the poem “limbo”, Adūnīs combines the word abjadiyyat (alphabets) with the word stars and creates a metaphorical combination of the merging of these two different linguistic families as follows: I was returned/My body is a gentle book/ The alphabet of stars and clouds wrote it/ My body moves towards the light at night and the bodies are galaxies (Adūnīs, 1996, Vol. 2, p. 370). The attribution of the word abjadiyyat (alphabet) to stars is because the body of the poet in limbo is a hybridization of the book of deeds and a galaxy of the results of those deeds. So the alphabet of the stars must be its writer. Because the poet considers his body to be a combination of books and galaxies, the hybridization of the alphabet of the stars, which are in complete harmony with these two, is necessary. This hybridization leads to the formation of such repetitive words in the form of similes, lexical, and semantic aberrations, as well as de-familiarization with the act of writing and highlighting this practice. The network of the supernatural and the network of nature are other networks that promote their poems. To express his feelings towards the beloved, Shāmloo in the poem (the sixth hymn) takes the tree from the family of nature and the miracle from the other side, and by mixing the two, the garments of the following are covered on it: I am not a miracle tree / only one of my trees (Shāmloo, 2004: 1036) Shāmloo is a tree like other ones and that tree is not the expected miracle. The hybridization of a miracle tree is a composition in the form of the past tense. Although the word tree and the concept of miracle go back to different language families, since the tree has long had a special and sacred place in mythology; therefore, the choice of these two families with obvious differences in networks is defensible in terms of communication. It has the poet’s intention of rejecting the expectation of something extraordinary from him. Adūnīs, like Shāmloo in his poem The Land of Magic, chooses the words earth and magic from different families, preparing such a capacity to express his feelings: And my land is constantly a land of magic / I confuse the air / I injure the face of the water / I go out of my house to the sea. (Adūnīs, 1996, vol. 1, p. 188). In this poem, Adūnīs wears the mask of an adventurous Sinbad and goes to war only with the unhealthy social conditions, to achieve his utopia. The poet is not disappointed; because his land, like the land of magic, announces unexpected events. The use of this eloquent combination in the noun phrase, which implies continuity, and the use of the verb (لم تزل: continuous) is a double emphasis on the expectation of being extraordinary from the poet's land. The network of religion and the network of nature is also one of the networks considered by the two poets. As Shāmloo chooses the word Sajjadeh (prayer rug) from the family of religion and soil from the words of nature in the poem "Rainy Design" and by combining them, he creates such works of art: Then / the holy silence of the sun setting on the soil prayer rug, / and the heavy hesitation of the bloody knife (Shāmloo, 2004, p. 1003). The poet believes in the sanctity of the sun and the earth; therefore, the composition of the soil prayer rug, which is a favorably compound in the form of simile, has created lexical and semantic aberrations. Despite the apparent distance of these two words in terms of lexical family and semantic infrastructure, since "in many languages, it is called the human being of the earth" (Eliade, 1996, p. 168). Accordingly, in terms of the concept of sacredness between the prayer rug and the soil is remarkable. In his poem “dream”, Adūnīs mentions the glorious past of the Arab countries with longing, remembering the past. Hybridization in this section, by combining the words of Surah and verse from religion and the words of cloud and stone from the family of nature, expresses the concepts intended by the poet. I entered the religious ceremonies of the Caliph / The womb of the waters and the sedition of the tree / I saw the trees that want me / And I saw rooms between its branches / And the boards and the windows are hostile to me / And I saw children for whom I recited / I heard, I recited for them / verses of the cloud and the verse of the stone (Adūnīs, 1996, Vol. 1, pp. 262-263). At this time, referring to the customs of the caliphate, Adonis called the seditions of that period the sedition of the tree. The trees of sedition, amid their foliage, are terrifying rooms. But inside these horrible rooms, there are clean and good children who are ready to receive the truth; Therefore, the poet recites for them verses of the cloud and verses of stone. Hybridization of verses of cloud and verse of stone indicate the sacredness of the elements of nature that the poet in the form of similes of two different language families to better convey their intended concepts, and thus he uses repetitive words for lexical and semantic aberrations. Conclusion Part of the literary value of the poetry of these two poets, from the point of view of novel verbal knowledge, depends on the use of hybridization networks. The same tendency towards hybridization-like combinations has increased the frequency of use of metaphorical arrays in their poetry. Accordingly, in addition to lexical anomalies, it has also caused semantic anomalies. Sometimes the second word
The article investigates functional techniques of extralinguistic expression in multimedia texts; the effectiveness of figurative expressions as a reaction to modern events in Ukraine and their influence on the formation of public opinion is shown. Publications of journalists, broadcasts of media resonators, experts, public figures, politicians, readers are analyzed. The language of the media plays a key role in shaping the worldview of the young political elite in the first place. The essence of each statement is a focused thought that reacts to events in the world or in one’s own country. The most popular platform for mass information and social interaction is, first of all, network journalism, which is characterized by mobility and unlimited time and space. Authors have complete freedom to express their views in direct language, including their own word formation. Phonetic, lexical, phraseological and stylistic means of speech create expression of the text. A figurative word, a good aphorism or proverb, a paraphrased expression, etc. enhance the effectiveness of a multimedia text. This is especially important for headlines that simultaneously inform and influence the views of millions of readers. Given the wide range of issues raised by the Internet as a medium, research in this area is interdisciplinary. The science of information, combining language and social communication, is at the forefront of global interactions. The Internet is an effective source of knowledge and a forum for free thought. Nonlinear texts (hypertexts) – «branching texts or texts that perform actions on request», multimedia texts change the principles of information collection, storage and dissemination, involving billions of readers in the discussion of global issues. Mastering the word is not an easy task if the author of the publication is not well-read, is not deep in the topic, does not know the psychology of the audience for which he writes. Therefore, the study of media broadcasting is an important component of the professional training of future journalists. The functions of the language of the media require the authors to make the right statements and convincing arguments in the text. Journalism education is not only knowledge of imperative and dispositive norms, but also apodictic ones. In practice, this means that there are rules in media creativity that are based on logical necessity. Apodicticity is the first sign of impressive language on the platform of print or electronic media. Social expression is a combination of creative abilities and linguistic competencies that a journalist realizes in his activity. Creative self-expression is realized in a set of many important factors in the media: the choice of topic, convincing arguments, logical presentation of ideas and deep philological education. Linguistic art, in contrast to painting, music, sculpture, accumulates all visual, auditory, tactile and empathic sensations in a universal sign – the word. The choice of the word for the reproduction of sensory and semantic meanings, its competent use in the appropriate context distinguishes the journalist-intellectual from other participants in forums, round tables, analytical or entertainment programs. Expressive speech in the media is a product of the intellect (ability to think) of all those who write on socio-political or economic topics. In the same plane with him – intelligence (awareness, prudence), the first sign of which (according to Ivan Ogienko) is a good knowledge of the language. Intellectual language is an important means of organizing a journalistic text. It, on the one hand, logically conveys the author’s thoughts, and on the other – encourages the reader to reflect and comprehend what is read. The richness of language is accumulated through continuous self-education and interesting communication. Studies of social expression as an important factor influencing the formation of public consciousness should open up new facets of rational and emotional media broadcasting; to trace physical and psychological reactions to communicative mimicry in the media. Speech mimicry as one of the methods of disguise is increasingly becoming a dangerous factor in manipulating the media. Mimicry is an unprincipled adaptation to the surrounding social conditions; one of the most famous examples of an animal characterized by mimicry (change of protective color and shape) is a chameleon. In a figurative sense, chameleons are called adaptive journalists. Observations show that mimicry in politics is to some extent a kind of game that, like every game, is always conditional and artificial.
The article is a study of the formation of a methodology for studying the stylistics of the Russian language. The goal is to trace the origin of the methodology for studying stylistics in the works of authors who chose it as the subject of their research. In the course of the work, in the process of studying various sources of information, it became known that many scholars-philologists turned to this topic, for example: V.V. Vinogradov, A.I. Efimov, M.N. Lotina, A.N. Veselovsky, A.A. Potebnya and many others. Stylistics received the status of an independent science in the XX century, at the same time, its study was of interest to people much earlier. Particular attention is paid to the exercise systems of T.I. Chizhova and S.N. Ikonnikov, designed to study the stylistic features of the text. S.N. Ikonnikov drew attention to the combination of analytical and creative exercises. The system of exercises developed by T.I. Chizhova. There are two types of exercises in her system: exercises to observe the use of phonetic, lexical and other means of speech; exercises to observe speech styles and their key features. Noteworthy is the fact that T.I. Chizhova is assigned a special role as a stylistic exercise in her didactic complex. It includes the following types of exercises: exercises for the analysis and study of individual sections of science related to language (phonetics, vocabulary, etc.); exercises for mastering certain stylistic principles of the language and in its sections; exercises for teaching coherent speech; exercises for understanding the norms, characteristics of individual styles of speech (scientific, journalistic, etc.).
Nowadays, no works in Tuvinian linguistics consider the specificity of anthroponyms borrowed from the Mongolian language, peculiarities of their adaptation and functioning in Tuvan Buddhist texts. Meanwhile, studying this word group can shed light on the formation and functioning of the Tuvan Buddhist vocabulary and reveal additional data on the history of the formation of lexical and phonetic features of the Tuvinian language associated with the Tuvan- Mongolian language and cultural contacts. It is worth studying Mongolian borrowed anthroponyms in the Tuvan translation of the Buddhist work “Üleger-Dalay” – the sutra “Sea of Proverbs,” the only Tuvan Buddhist source not influenced by the Russian-speaking Buddhist literature actively published and translated into Tuvan since the 1990s. The specificity of the Mongolian anthroponyms analyzed is that their nominative function simultaneously characterizes the referent and reflects its essence. They are divided into six thematic groups: names-epithets of Buddha, names associated with Buddhist practices, names indicating the inner qualities of a person, names of celestials, names associated with natural objects, names with somatismatic components indicating the appearance or associated with the circumstances of the referent’s birth. They are divided into three structural types: 1-component, 2-component (most of them), and 3-component. All borrowed Mongolian names have mostly been adapted to the phonetic norms of the Tuvinian language. The main ways of phonetic transformation are assimilation, formation of long vowels, replacement of some sounds and sound combinations with other sounds, simplification of vowels.
Abstract Engagement in a wide array of mental, social, and physical leisure activities confers several health benefits. Indeed, theories of successful aging argue that an active lifestyle serves as an important criterion for maintaining high levels of psychological, functional, and physical well-being in old age. Findings from parallel studies also show that people who hold positive (self-)views of aging exhibit higher and maintained levels of well-being over time. Yet, whether views of aging enhances the link between activity engagement and well-being - and whether they do so on a daily basis – remains unknown. This study therefore sought to extend prior literature by examining the relationship between activity engagement, subjective age, and affective ratings within-person over several days. Old adults (N = 115; Age: Range = 60 – 90, M = 64.65, SD = 4.86) in the Mindfulness and Anticipatory Coping Every Day (MACED) study completed an 8-day daily diary. Participants reported on their positive and negative affect, the age they subjectively felt compared to their actual age, and the number and types of leisure activities in which they engaged. Results from multilevel analyses indicate that people felt more positive on days when they also engaged in more activities (total across mental, social, physical types) than usual. Moreover, the effect of activity engagement was most pronounced on days when people felt younger than usual. No effects were found for negative affect. Preliminary findings suggest that people benefit psychologically from daily leisure activities and a positive self-view of aging.
The article presents educational, informative and methodical materials necessary for studying the theme “Professional communication a physician and a patient with symptoms diseases of the digestive system’s organs” in classes on the discipline “Professional medical communication of a doctor with a patient in Ukrainian language”. The complex of tasks is aimed at the development of students’ communicative skills and abilities: to study the vocabulary to denote the organs of the digestive system, gastrointestinal diseases; be able to build monologue and dialogic expressions that describe the causes and symptoms of the digestive system’s diseases, using learned lexical units and phrases; to develop skills of collecting the anamnesis of gastroenterological diseases; memorize phrases and sentences for first aid in food poisoning, in emergencies during stomach pain; to create modern informative and expert systems for developing lesson’s materials. The proposed system of tasks will help to master the skills and abilities to communicate orally and in writing in accordance with the goals and social norms of speech behavior in typical spheres and situations. Taking into account different methods and forms of work, the tasks with which it was possible to ensure the active participation of each student in the class, to stimulate interest and desire to study medical terminology in accordance with the topic of the class are singled out. It was suggested a variety of test and creative tasks to check the knowledge of the student’s studied topic. The materials described in the article are intended for foreign students of medical specialties who speak the language at a sufficient level. Tasks selected on the basis of four main types of speech activity (listening, speaking, reading, writing) will help foreign students not only to expand their vocabulary, but also to achieve social interaction in a foreign language professional sphere.
<p class="MsoNormal" style="margin-bottom:.0001pt; text-align: justify; line-height: 115%;"><strong><span style="font-size: 16.0pt; line-height: 115%; font-family: "Jameel Noori Nastaleeq"; mso-bidi-language: ER;"><span style="mso-spacerun: yes;"> </span></span></strong><span style="font-size: 12.0pt; line-height: 115%; font-family: "Jameel Noori Nastaleeq"; mso-bidi-language: ER;">Stylistics is defined as the study of style, Phonetics, Morphology, Syntax and Semantics as tools of criticism, relatively modern sources of discovering hidden meaning and layers of expression. According Nills Erik Enkvist the style of a poet is analyzed through Language he used. He suggests this analysis upon the basis of two aspects of style. (1) Choice between alternative expressions. (2) Derivation from the norm. Majeed Amjad having unique style and poetics must be analyzed on modern parameters of stylistics. Majeed Amjad has his own distinguish style among the urdu poets. He carved out a new style other. He experimented with metrical forms and rhythms. Stylistic analysis of Amjad’s poetry Lenlighteen the individual traits of his style. Stylistics choice of lexical items, phonetics, Morphological choices and grammatical sources shows distinctive features of his poetics and literary style.</span>
Abstract Background The Montessori Method underpinned by the principle of person-centered care has been widely adopted to design activities for people with dementia. However, the methodological quality of the existing evidence is fair. The objectives of this study are to examine the feasibility and effects of a culturally adapted group-based Montessori Method for Dementia program in Chinese community on engagement and affect in community-dwelling people with dementia. Methods This was a two-arm randomized controlled trial. People who were aged 60 years or over and with mild to moderate dementia were recruited and randomly assigned to the intervention group to receive Montessori-based activities or the comparison group to receive conventional group activities over eight weeks. The attendance rates were recorded for evaluating the feasibility. The Menorah Park Engagement Scale and the Apparent Affect Rating Scale were used to assess the engagement and affect during the activities based on observations. Generalized Estimating Equation model was used to examine the intervention effect on the outcomes across the sessions. Results A total of 108 people with dementia were recruited. The average attendance rate of the intervention group (81.5%) was higher than that of the comparison group (76.3%). There was a significant time-by-group intervention effect on constructive engagement in the first 10 minutes of the sessions (Wald χ 2 = 15.21–19.93, ps = 0.006–0.033), as well as on pleasure (Wald χ 2 = 25.37–25.73, ps ≤ 0.001) and interest (Wald χ2 = 19.14–21.11, p s = 0.004–0.008) in the first and the middle 10 minutes of the sessions, adjusted for cognitive functioning. Conclusions This study provide evidence that Montessori-based group activities adapted to the local cultural context could effectively engage community-dwelling Chinese older people with mild to moderate dementia in social interactions and meaningful activities and significantly increase their positive affect. Trial registration ClinicalTrials.gov, NCT04352387. Registered 20 April 2020. Retrospectively registered.
Language Variation in Bernese Swiss German In the atlas of German-speaking Switzerland (SDS) (cf. Hotzenkocherle et al. 1962-2003) we find data on the greater area of Bern (Berner Mittelland), collected around 1944. Since then, only very specific factors of this particular linguistic variety have been examined, e.g. Hodler 1969 on Bernese German syntax, Marti 1976 on Bernese German grammar more generally or Siebenhaar 2000 on social varieties in the city of Bern, but the dialect has not been examined in its entirety. Therefore, developments which origin in language contact or speaker mobility and have effectively influenced the dialects of this region, have not been documented to the present day. In my project the focus is on language change in the research area and on reasons for the present changes. Currently I collect new data for Bern and its greater area according to selected variables already surveyed in the SDS, and I then compare the new data to the original data. In addition to the variables originating in the SDS, also some new variables are taken into account. Of special interest are borrowings from foreign languages, e.g. the realization of engl. steak (stɛɪk vs. ʃtɛik(x) vs. ʃti:kx) or recent lexical changes as from Swiss German Nidle (ni:dlə) [cream] to Rahm (rɑ:m), a variant which is mainly used in southern Germany and Austria. My survey includes 20 places in the greater area of Bern where I record 4 speakers per place. The speakers are classified in three age groups (18-35, 35-65, 65+) and I also take an agriculturalist into account. This occupational group is meant to be more traditional in respect of language, as the language atlas of Middle Franconia (cf. Mang 2004) shows repeatedly. In this respect, my project differs clearly from the SDS, where mainly NORMs have been taken into account (one or two per place). I suggest the main reasons for language change in my research area in speaker mobility and migration movement. Already the present, relatively small set of data shows tendencies, which support my suggestions, as e.g. the decline of pronunciation according to Staub's law (/n/--> o_fricative; with vowel lengthening and/or diphthongisation) shows. Whereas the pronunciation of the variable Fenster [window] was represented by (faeiʃtəɾ) in the majority of places examined in the SDS within the greater area of Bern, today the realization (faenʃtəɾ) is common to be found. Currently, more variables and places are examined in order to present the shift of the isogloss as soon as possible. References: Baumgartner Heinrich, Hotzenkocherle Rudolf (1962-2003). Sprachatlas der deutschen Schweiz. Bern, Basel: Francke Verlag Baumgartner, Heinrich (1940). Stadtmundart: Stadt- und Landmundart: Beitrage zur bernischen Mundartgeographie. Bern: Lang Christen Helen, Glaser Elvira und Friedli Matthias (2012). Kleiner Sprachatlas der deutschen Schweiz. Frauenfeld: Verlag Huber Hodler, Werner (1969). Berndeutsche Syntax. Bern: Francke Verlag Mang, Alexander (2004). Sprachatlas von Mittelfranken. Sprachregion Nurnberg (Band 6). Heidelberg,: Universitatsverlag C. Winter. Marti, Werner (1985). Berndeutsch-Grammatik fur die heutige Mundart zwischen Thun und Jura. Bern: A. Francke Siebenhaar Beat, Staheli Fredy, Ris Roland (2000). Stadtberndeutsch: Sprachportrats aus der Stadt Bern. Murten: Licorne-Verlag
This article presents the results of a study of the Mari numeral phrases in terms of the influence of the Russian language. The aim of this work is to trace the role of Russian borrowings in the formation of Mari numeral phrases, primarily in the expression of their components, and to reveal other changes that have arisen under the influence of similar phrases and structures of the Russian language. The study was conducted on the basis of the lexical card index of the MarNIIYALI (Mari Scientific Research Institute of Language, Literature and History), which is based on written sources of the meadow-eastern literary norm, namely, its electronic part in the amount of about one thousand author's sheets. During the collection and analysis of material, elements and techniques of the following research methods were applied: descriptive and analytical (observation with the identification of the studied facts in sources, their generalization, interpretation and classification, description), comparative (regular comparison of Mari models with Russian ones for identity and non-identity), comparative-historical (in other cases, indications of the origin of the words), quantitative (counting models of various groups containing Russianisms). According to the results of the research, Russian borrowings may play a role of a head word (3 units in 4 models) and a dependent component (mainly substantive case forms and postpositional constructions, numerals, as well as some pronouns and adverbs of degree in 14 models). 3 models with cardinal numbers as a head very rarely can be represented by phrases with Russianisms in both components. As a result the syntactic units in some models and the models of numeral phrases themselves were replenished, the last ones by 3 units. Also the shifts in the forms of grammatical number of dependent nouns in some models appeared.
В данной статье рассматривается в семантическом контексте с обрядовым действом лексика в сибирских тюркских и монгольских языках, связанная с культом гор, земли и воды. Основными языками исследования являются алтайский, бурятский и якутский с привлечением монгольских, хакасских, тувинских параллелей. У тюрко-монгольских народов прослеживаются общие принципы организации сакрального пространства, концептуально схожие ритуальные действа в коллективном обряде, посвященном духам-хозяевам местности, присутствуют универсальные атрибуты и символы, характерные для шаманизма и буддизма. Впервые проведен сопоставительный анализ лексем, семантики и символики слов, связанных с обрядовыми действами и сопровождающий их вербальный контекст. Целью работы является выявление сохранности, распространения и трансформации культурных универсалий, имеющих вербальное выражение. Актуальным представляется исследование древнейших культурных кодов, сохранившихся в предметной и акциональной сферах обрядности. Понятие культурного кода в изучении обрядности позволит получить ключ к пониманию культурной картины мира и расшифровать глубинный смысл составных частей обряда (смыслов, знаков, символов, норм и т. д.). В итоге констатируем, что некогда существовала единая тюрко-монгольская традиция шаманизма, имевшая общую культурно знаковую систему, что подтверждается лексическим материалом и их обрядовым контекстом. Рассмотренные два основных ритуала жертвоприношения (кровавые и бескровные) доказывают как древние связи, так устойчивую универсальную последовательность и сохранность акциональных кодов в обрядовом событии монгольских и тюркских народов Сибири. Выявлено, что ключевые лексемы, используемые в предметном коде, имеют универсальную семантическую нагрузку в обрядовом событии. Лексические соответствия и схожие ритуальные предметы и действа, скорее всего, доказывают восхождение обряда к единым корням с последующими региональными и временными трансформациями. Установлено, что одинаковые атрибуты ритуала со схожим или разным лексическим обозначением являются архетипами, отражающими общие культурные коды тюркских народов Сибири и монгольских этносов. The authors consider the vocabulary of Siberian Turkic and Mongolian languages, related to the cult of mountains, land and water in a semantic context with the ritual action. The main languages of the study are Altai, Buryat and Yakut with the involvement of Mongolian, Khakas, and Tuvan parallels. In rites of Turkic-Mongolian people devoted to the host spirits of the area, the general principles of the organization of the sacred space, conceptually similar ritual actions are traced. There are universal attributes and symbols which are illustratory of shamanism and Buddhism. For the first time, a comparative analysis of lexemes, semantics and symbolism of words related to ritual actions and accompanying their verbal context was carried out. The purpose of the work is identifying the preservation, dissemination and transformation of cultural universals with verbal expression. A study of the oldest cultural codes preserved in the subject and actional spheres of rite seems relevant. The concept of a cultural code in the study of rite will provide a key to understanding the cultural picture of the world and allows deciphering the deep meaning of the components of the rite (meanings, signs, symbols, norms, etc.). As a result, we state that there was once a united Turkic-Mongolian tradition of shamanism, which had a common culturally significant system. It is confirmed by lexical material and their ritual context. The considered two main rituals of sacrifice (bloody and bloodless) prove both ancient ties and a stable universal sequence and the preservation of national codes in the ritual event of the Mongolian and Turkic peoples of Siberia. It is revealed that the key lexemes used in the subject code have a universal semantics in the ritual event. Lexical correspondences and similar ritual objects and actions most likely prove the ascension of the rite to single roots with subsequent regional and temporary transformations. It was established that the same attributes of the ritual with a similar or different lexical designation are archetypes reflecting the general cultural codes of the Turkic peoples of Siberia and Mongolian ethnic groups.
The aim of the paper is twofold: (1) to automatically predict the ratings assigned by viewers to 14 categories available for TED talks in a multi-label classification task and (2) to determine what types of features drive classification accuracy for each of the categories. The focus is on features of language usage from five groups pertaining to syntactic complexity, lexical richness, register-based n-gram measures, information-theoretic measures and LIWC-style measures. We show that a Recurrent Neural Network classifier trained exclusively on within-text distributions of such features can reach relatively high levels of overall accuracy (69%) across the 14 categories. We find that features from two groups are strong predictors of the affective ratings across all categories and that there are distinct patterns of language usage for each rating category.
The purpose of this research is to study issues of interdisciplinary approach to teaching the Englishlanguage interaction to multinational non-immigrant masters of Arts within legal domain, to identify key communicative skills essential for lawyers' collaboration.This research uses a qualitative approach.The methodological basis is constituted by conceptual provisions of worldwide researchers of theory and methodology of teaching foreign languages, intercultural communication, competence approach, content and language integrated learning (CLIL), English for specific purposes (ESP), theoretical phonetics in linguistics.The novelty is in the interdisciplinary approach to teaching professionally-focused English-language intercultural interaction within legal domain to multinational non-immigrant masters of Arts at RUDN University which is based on three basics: received pronunciation (RP), legalese, and communicative competence formation.179 texts at an intermediate level selected from numerous authentic resources reflecting law realities in Russia, the USA and UK were chosen; stress location factors in poly-stressed words were studied in the selected lexical corpus.Compiled by multinational non-immigrant students trilingual mini-vocabularies: triplets of polysyllabic English, Russian and their mother-tongue words (English -Russian -Serbian/ Lezghian/ Kyrgyz, etc.) brought originality to this research.By the method of comparison and grouping, the study analysed a series of surveys of 50 masters of Arts aged 21-27 years to identify the resultativeness of the chosen interdisciplinary approach.This research practical value is in deepening multinational non-immigrant law-students' knowledge of the English and Russian languages through mastering their RP proficiency, writing and speaking skills of legal opinions' argumentation and justification, the norms of law competent elucidation.The results of the presented interdisciplinary approach can be applied by researchers of law and linguistics.
The aim of the investigation is to study word formation processes in the semantics of weather lexis in the online narrative of weather news stories in British newspapers.Methods. The source base of the study is represented by a corpus of electronic texts of the new genre “weather news story” in British electronic quality (www.thetimes.co.uk, www.theguardian.com/uk) and mass (www.www.thesun.co.uk, www.thedailymail.co.uk) newspapers (2014–2017). The research is methodologically based on structural-semantic and functional research methods. The application of a functional approach in our study enables the identification of the specifics of functioning and organization of lexical units in the contextual environment of the online narrative taking into account their role in the formation of content and meaning of weather news stories. At the same time, the use of descriptive and comparative methods allows revealing more deeply the individual (author) features of the texts under study.Results. A review of the literature confirms the idea that the creation of new words is based on the models and word formation types that are already established in the language or re-emerging. The modern word formation of weather lexis in weather news stories is characterized by typical methods of word formation, among which we distinguish the affixation, compounding, blending, shortening, and conversion. The common types of word formation in the genre under research are affixation and compounding. The most common prefixes are prefixes with the contrasting semantics un-, anti-, de- and the diminutive prefix mini-, among the noun suffixes -ation, -er, -ism, ity, -ness and adjective suffixes y, -ly, -al, -ic, -ish, -ent, -ous. Both dictionary (normative) and occasional (author) forms are distinguished. The use of occasional forms serves as an expressive means of reflecting the author’s worldview, where the author deliberately violates language norms to provide a certain expressive background to the story.Conclusions. The analysis of word formation of weather lexis in weather news stories shows that its organization occurs through the prism of word formation features of the narrative and diversifies weather news stories.Key words: word formation, lexis, abbreviation, shortening, compounding, blending, conversion, weather news story. Мета. Мета статті – дослідження дериваційних процесів у семантиці номінацій погоди в онлайн-наративі про погодні новини британських газет.Методи. Джерельна база дослідження представлена корпусом електронних текстів неожанру weather news story інтернет-версій британських періодичних якісних (www.thetimes.co.uk, www.theguardian.com/uk) і масових (www.www.thesun.co.uk, www.thedailymail.co.uk) видань (2014–2017 рр.). Дослідження методологічно ґрунтується на положеннях структурно-семантичного та функційного методів. Застосування функційного підходу у проєкції на наше дослідження дає змогу виявити спе-цифіку функціювання й організації лексичних одиниць у контекстуальному середовищі онлайн-наративу з огляду на їх роль у процесі змісто- та смислотворення погодних новин. Водночас застосовано описовий і порівняльний методи, які допомогли глибше розкрити індивідуально-авторські особливості текстів онлайн-наративу про погодні новини.Результати. Досліджено дериваційні процеси у семантиці номінацій погоди. Здійснений огляд літератури підтвердив ідею про те, що творення нових слів відбувається за тими моделями, за тими словотворчими типами, які вже встановилися в мові або знову виникають. Сучасному дериваційному процесу номінацій погоди в ОНПН властиві типові способи словотворення, серед яких виокремлюємо афіксальний спосіб, спосіб словоскладання, телескопію, скорочення і конверсію. Найпоширені-шими способами словотворення у корпусі дослідження є афіксальний і словоскладання. Найпоширенішими префіксами, які слугували утворенню нових номінацій погоди, є префікси з контрарною семантикою un-, anti-, de- та зменшувальний префікс mini-, серед суфіксів ‒ відіменникові -ation, -er, -ism, ity, -ness і відприкметникові y, -ly, -al, -ic, -ish, -ent, -ous. У корпусі досліджуваних текстів виокремлено узуальні/словникові й оказіональні/авторські форми. Визначено, що вживання оказіональних форм слугує експресивним засобом вираження індивідуально-авторської мовної картини, де автор свідомо порушує мовні норми, аби надати певного експресивного фону оповіді.Висновки. Проведений аналіз дериваційних процесів у семантиці номінацій погоди в онлайн-наративі про погодні новини показав, що організація лексичних одиниць на позначення погоди відбувається крізь призму словотвірних особливостей оповіді. Дериваційні процеси в семантиці номінацій погоди зумовили наявність широкої палітри лексики в ОНПН, яка демонструє авторський чинник його організації та урізноманітнює оповідь.Ключові слова: словотвір, лексика, абревіація, скорочення, словоскладання, телескопія, конверсія, weather news story.
Abstract This article presents a typology of phonological, morphosyntactic, and lexical features illustrative of factors conditioning the usage of speakers and writers of Revived Manx, including substratal influence from English; language ideologies prevalent within the revival movement, especially forms of linguistic purism; and language-specific features of Manx and its orthography. Evidence is taken primarily from a corpus of Revived Manx speech and writing. The observed features of Revived Manx are situated within Zuckermann's (2009, 2020) framework of ‘hybridization’ and ‘revival linguistics’, which takes Israeli Hebrew as the prototypical model of revernacularization of a non-L1 language. However, Manx arguably provides a more typical example of what to expect when a revived minority language remains predominantly an L2 for an indefinite period, with each new cohort of speakers able to reshape the target variety in the absence of a firmly established L1 norm. (Manx, Celtic, language revival, language ideology, language shift, language contact)*
To understand the causes of differences in language ability we must measure the specific and separable processes that contribute to natural language comprehension. Specifically, we need measures of the three language subsystems – semantics, syntax, and phonology – as they are used during the comprehension of real speech. Event-Related Potentials (ERPs) are a promising approach to reaching this level of specificity. Previous research has identified distinct ERP effects for each of the subsystems – the N400 to semantic anomalies, the Anterior Negativity and P600 to syntactic anomalies, and the Phonological Mapping Negativity to unexpected speech sounds. However, these studies typically use stimuli and tasks that encourage processing that differs from real-world language comprehension. Further, previous ERP studies indexing language processing in young children not only use unfamiliar tasks, but also typically exclude data from the large proportion of children. We need to measure language-related ERPs in a context as close as possible to real-world processing, and in a manner that includes data from representative rather than highly-selected samples of children. The experiments described in this dissertation achieve that goal. Adults and five-year-old children listened to a child-directed story while answering comprehension questions. Infrequent violations were included to independently probe the three language subsystems. In children and adults, the canonical N400 response was evident in response to semantic violations. Morphosyntactic violations elicited a long-duration Anterior Negativity without a later P600. Phonological violations on suffixes elicited a Phonological Mapping Negativity in adults. This is the first report of this phonological effect outside of highly-predictable lexical contexts. Popular normed behavioral assessments were also administered to the children who participated in this study. Results from these assessments confirmed that performance on tasks claiming to measure categorically different abilities are correlated with one another, and that language measures correlate with so-called nonverbal measures. ERPs indexing different language subsystem did not correlate with each other or with measures of nonverbal cognitive ability. Using multiple ERP measures during natural language comprehension, we are able to isolate specific aspects of language processing, increasing the possibility of making meaningful connections between biology, experience, and resulting language ability.
The present study investigates how variation in acoustic measures and lexical semantic properties affect the comprehension of threatening speech. Two experiments were conducted in British English, where participants had to rate the threat of three types of sentences: sentences neutral in prosody and semantics, sentences containing only prosodic threat, and sentences containing only semantic threat. They also rated the sentences for arousal and valence, and categorised them by type of emotional expression. Threat ratings were analysed via Bayesian ordered-logistic regressions using acoustic measures and affective norms as regressors. Classification was assessed via simple Bayesian categorical (Dirichlet prior) models. Acoustic properties of stimuli were compared through a Bayesian method for comparison of means to test whether neutral stimuli differ, on average, from their threatening counterparts. Acoustic but not semantic properties affected ratings of threatening prosody (increased pitch and decreased voice quality predicting increased threat), while the reverse was true for ratings of threatening semantics (increased arousal and decreased valence predicting increased threat). Using data from the first speaker (first experiment) the models were able to predict ratings given to sentences produced by the second speaker (second experiment). Furthermore, threatening sentences of both types tended to be categorized as angry or enraged. These results are interpreted as supporting the sufficiency of the selected measures (pitch, voice quality, arousal, valence) to characterise threat. Although multidimensional models may be better suited for describing threatening prosody/semantics, the present approach demonstrates that selected bidimensional features are sufficient to identify and predict threat comprehension efficiently and accurately.
The article discusses the features of academic writing in English based on the recommendations from the British Council in Ukraine in the framework of the “Researcher Connect” project, aiming to facilitate the transition to academic standards of English and improve the academic discourse produced by non-native language users. The authors outline major tendencies in the modern English language as pertains to written discourse and provide recommendations for rendering academic writing persuasive. It is a well-established fact that academic writing in English possesses unique features, which must be respected and taken into account. Hence, a transfer of academic norms from a person's mother tongue to English can be a challenge, which may impair the quality of academic writing. Presenting the research results without consideration of academic norms, grammar, and lexical features of English academic writing can lead to mistakes and misunderstanding, and result in a written work of poor quality, even if the research findings are valid. The mechanisms of improving the academic writing skills during the study of English for Academic Communication with due account for relevant grammar and lexical peculiarities have been explored. Therefore, the major challenge for researchers is the difficulty in transition to academic standards of a foreign language. The article discusses the surface and the deeper purposes in any academic writing; the significance of understanding one’s audience; the concepts of persuasion, clarity, and conciseness, as well as grammar and lexical means for achieving them. Developing the communication skills of Ukrainian scientists is crucial for successful international communication and cooperation. The study of potential difficulties, which the Ukrainian medical professionals may face in the process of academic writing in English, is important for developing the guidelines to eliminate possible mistakes and avoid misunderstanding in a medical setting. Further study of the peculiarities of academic writing in English will contribute to the optimization of international professional communication, the expansion of inter-institutional dialogue, and the integration of Ukraine into the world community.
The study aims to determine the features charactering the process of feminisation of the French lexis denoting professions. The article analyses the main ways of feminisation, semantics and stylistic colouring of lexical units. Scientific novelty of the study lies in conducting a comprehensive analysis and developing a classification of ways for forming new feminine units of the French lexis denoting professions. As a result, the researcher has identified the features and problems peculiar to feminisation of profession names which are related to both linguistic factors and the issues pertaining to women’s position in the French linguistic culture. It is proved that the predominant factors for the language phenomena being accepted as a linguistic norm are social conditions and frequency of using units in speech.
S{\\o}gaard (2020) obtained results suggesting the fraction of trees occurring\nin the test data isomorphic to trees in the training set accounts for a\nnon-trivial variation in parser performance. Similar to other statistical\nanalyses in NLP, the results were based on evaluating linear regressions.\nHowever, the study had methodological issues and was undertaken using a small\nsample size leading to unreliable results. We present a replication study in\nwhich we also bin sentences by length and find that only a small subset of\nsentences vary in performance with respect to graph isomorphism. Further, the\ncorrelation observed between parser performance and graph isomorphism in the\nwild disappears when controlling for covariants. However, in a controlled\nexperiment, where covariants are kept fixed, we do observe a strong\ncorrelation. We suggest that conclusions drawn from statistical analyses like\nthis need to be tempered and that controlled experiments can complement them by\nmore readily teasing factors apart.\n
Colexification is a linguistic phenomenon that occurs when multiple concepts are expressed in a language with the same word. Colexification patterns are frequently used to estimate the meaning similarity between words, but the hypothesis that these are related is still missing direct empirical validation at scale. Here, we show for the first time that words linked by colexification patterns capture similar affective meanings. Using pre-existing translation data, we extend colexification databases to cover much longer word lists. We achieve this with an unsupervised method of affective lexicon extension that uses colexification network data to interpolate the affective ratings of words that are not included in the original lexicon. We find positive correlations between network-based estimates and empirical affective ratings, which suggest that colexification networks contain information related to affective meanings. Finally, we compare our network method with state-of-the-art machine learning, trained on a large corpus, and show that our simple linguistics-informed unsupervised algorithm yields comparable performance with high explainability. These results show that it is possible to automatically expand affective norms lexica to cover exhaustive word lists when additional data are available, such as in colexification networks. Supplementary Information: The online version contains supplementary material available at 10.1007/s42761-021-00033-1.
We present the enrichment of a French treebank of various genres with a new annotation layer for multiword expressions (MWEs) and named entities (NEs).1 Our contribution with respect to previous work on NE and MWE annotation is the particular care taken to use formal criteria, organized into decision flowcharts, shedding some light on the interactions between NEs and MWEs. Moreover, in order to cope with the well-known difficulty to draw a clear-cut frontier between compositional expressions and MWEs, we chose to use sufficient criteria only. As a result, annotated MWEs satisfy a varying number of sufficient criteria, accounting for the scalar nature of the MWE status. In addition to the span of the elements, annotation includes the subcategory of NEs (e.g., person, location) and one matching sufficient criterion for non-verbal MWEs (e.g., lexical substitution). The 3,099 sentences of the treebank were double-annotated and adjudicated, and we paid attention to cross-type consistency and compatibility with thesyntactic layer. Overall inter-annotator agreement on non-verbal MWEs and NEs reached 71.1%. The released corpus contains 3,112 annotated NEs and 3,440 MWEs, and is distributed under an open license.
Abstract This paper proposes to study the contrastive syntax of French and Chinese through the lens of syntactic mismatches, and by making use of parallel treebanks. A syntactic mismatch is the non-similarity between the syntactic structures of one linguistic unit and its translation. Syntactic mismatches are formalized using the notion of paraphrase from the Meaning-Text Theory, which allows for capturing mismatches at different levels of the linguistic description (e.g. Semantic, Deep-Syntactic, and Surface-Syntactic). In this paper, we report in details on the types of paraphrases found in the seed corpus used, demonstrating that the Deep-Syntactic paraphrases constitute the best starting point for our study. Then, we show how, starting from the seed corpus, we semi-automatically constructed a multi-layer parallel treebank with the alignment and annotation of paraphrases.
CorefUD is a collection of previously existing datasets annotated with coreference, which we converted into a common annotation scheme. In total, CorefUD in its current version 0.1 consists of 17 datasets for 11 languages. The datasets are enriched with automatic morphological and syntactic annotations that are fully compliant with the standards of the Universal Dependencies project. All the datasets are stored in the CoNLL-U format, with coreference- and bridging-specific information captured by attribute-value pairs located in the MISC column. The collection is divided into a public edition and a non-public (ÚFAL-internal) edition. The publicly available edition is distributed via LINDAT-CLARIAH-CZ and contains 13 datasets for 10 languages (1 dataset for Catalan, 2 for Czech, 2 for English, 1 for French, 2 for German, 1 for Hungarian, 1 for Lithuanian, 1 for Polish, 1 for Russian, and 1 for Spanish), excluding the test data. The non-public edition is available internally to ÚFAL members and contains additional 4 datasets for 2 languages (1 dataset for Dutch, and 3 for English), which we are not allowed to distribute due to their original license limitations. It also contains the test data portions for all datasets. When using any of the harmonized datasets, please get acquainted with its license (placed in the same directory as the data) and cite the original data resource too. References to original resources whose harmonized versions are contained in the public edition of CorefUD 0.1: - Catalan-AnCora: Recasens, M. and Martí, M. A. (2010). AnCora-CO: Coreferentially Annotated Corpora for Spanish and Catalan. Language Resources and Evaluation, 44(4):315–345 - Czech-PCEDT: Nedoluzhko, A., Novák, M., Cinková, S., Mikulová, M., and Mírovský, J. (2016). Coreference in Prague Czech-English Dependency Treebank. In Proceedings of the Tenth International Conference on Language Resources and Evaluation (LREC'16), pages 169–176, Portorož, Slovenia. European Language Resources Association. - Czech-PDT: Hajič, J., Bejček, E., Hlaváčová, J., Mikulová, M., Straka, M., Štěpánek, J., and Štěpánková, B. (2020). Prague Dependency Treebank - Consolidated 1.0. In Proceedings of the 12th International Conference on Language Resources and Evaluation (LREC 2020), pages 5208–5218, Marseille, France. European Language Resources Association. - English-GUM: Zeldes, A. (2017). The GUM Corpus: Creating Multilayer Resources in the Classroom. Language Resources and Evaluation, 51(3):581–612. - English-ParCorFull: Lapshinova-Koltunski, E., Hardmeier, C., and Krielke, P. (2018). ParCorFull: a Parallel Corpus Annotated with Full Coreference. In Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018), Miyazaki, Japan. European Language Resources Association. - French-Democrat: Landragin, F. (2016). Description, modélisation et détection automatique des chaı̂nes de référence (DEMOCRAT). Bulletin de l’Association Française pour l’Intelligence Artificielle, (92):11–15. - German-ParCorFull: Lapshinova-Koltunski, E., Hardmeier, C., and Krielke, P. (2018). ParCorFull: a Parallel Corpus Annotated with Full Coreference. In Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018), Miyazaki, Japan. European Language Resources Association - German-PotsdamCC: Bourgonje, P. and Stede, M. (2020). The Potsdam Commentary Corpus 2.2: Extending annotations for shallow discourse parsing. In Proceedings of the 12th Language Resources and Evaluation Conference, pages 1061–1066, Marseille, France. European Language Resources Association. - Hungarian-SzegedKoref: Vincze, V., Hegedűs, K., Sliz-Nagy, A., and Farkas, R. (2018). SzegedKoref: A Hungarian Coreference Corpus. In Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018), Miyazaki, Japan. European Language Resources Association. - Lithuanian-LCC: Žitkus, V. and Butkienė, R. (2018). Coreference Annotation Scheme and Corpus for Lithuanian Language. In Fifth International Conference on Social Networks Analysis, Management and Security, SNAMS 2018, Valencia, Spain, October 15-18, 2018, pages 243–250. IEEE. - Polish-PCC: Ogrodniczuk, M., Glowińska, K., Kopeć, M., Savary, A., and Zawisławska, M. (2013). Polish coreference corpus. In Human Language Technology. Challenges for Computer Science and Linguistics - 6th Language and Technology Conference, LTC 2013, Poznań, Poland, December 7-9, 2013. Revised Selected Papers, volume 9561 of Lecture Notes in Computer Science, pages 215–226. Springer. - Russian-RuCor: Toldova, S., Roytberg, A., Ladygina, A. A., Vasilyeva, M. D., Azerkovich, I. L., Kurzukov,M., Sim, G., Gorshkov, D. V., Ivanova, A., Nedoluzhko, A., and Grishina, Y. (2014). Evaluating Anaphora and Coreference Resolution for Russian. In Komp’juternaja lingvistika i intellektual’nye tehnologii. Po materialam ezhegodnoj Mezhdunarodnoj konferencii Dialog, pages 681–695. - Spanish-AnCora: Recasens, M. and Martí, M. A. (2010). AnCora-CO: Coreferentially Annotated Corpora for Spanish and Catalan. Language Resources and Evaluation, 44(4):315–345 References to original resources whose harmonized versions are contained in the ÚFAL-internal edition of CorefUD 0.1: - Dutch-COREA: Hendrickx, I., Bouma, G., Coppens, F., Daelemans, W., Hoste, V., Kloosterman, G., Mineur, A.-M., Van Der Vloet, J., and Verschelde, J.-L. (2008). A coreference corpus and resolution system for Dutch. In Proceedings of the Sixth International Conference on Language Resources and Evaluation (LREC’08), Marrakech, Morocco. European Language Resources Association. - English-ARRAU: Uryupina, O., Artstein, R., Bristot, A., Cavicchio, F., Delogu, F., Rodriguez, K. J., and Poesio, M. (2020). Annotating a broad range of anaphoric phenomena, in a variety of genres: the ARRAU Corpus. Natural Language Engineering, 26(1):95–128. - English-OntoNotes: Weischedel, R., Hovy, E., Marcus, M., Palmer, M., Belvin, R., Pradhan, S., Ramshaw, L., and Xue, N. (2011). Ontonotes: A large training corpus for enhanced processing. In Handbook of Natural Language Processing and Machine Translation: DARPA Global Autonomous Language Exploitation, pages 54–63, New York. Springer-Verlag. - English-PCEDT: Nedoluzhko, A., Novák, M., Cinková, S., Mikulová, M., and Mírovský, J. (2016). Coreference in Prague Czech-English Dependency Treebank. In Proceedings of the Tenth International Conference on Language Resources and Evaluation (LREC’16), pages 169–176, Portorož, Slovenia. European Language Resources Association.
Abstract Emotion concepts are representations that enable people to make sense of their own and others’ emotions. The present study, theoretically driven by the conceptual act theory, explores the overall spectrum of emotion concepts in older adults and compares them with the emotion concepts of younger adults. Data from 178 older adults (⩾55 years) and 176 younger adults (20–30 years) were collected using the Semantic Emotion Space Assessment task. The arousal and valence of 16 discrete emotions – anger, fear, sadness, happiness, disgust, hope, love, hate, contempt, guilt, compassion, shame, gratefulness, envy, disappointment, and jealousy – were rated by the participants on a graphic scale bar. The results show that (a) older and younger adults did not differ in the mean valence ratings of emotion concepts, which indicates that older adults do not differ from younger adults in the way they conceptualise how pleasant or unpleasant emotions are. Furthermore, (b) older men rated emotion concepts as more arousing than younger men, (c) older adults rated sadness, disgust, contempt, guilt, and compassion as more arousing and (d) jealousy as less arousing than younger adults. The results of the present study indicate that age-related differentiation of conceptual knowledge seems to proceed more in the way that individuals understand how arousing their subjective representations of emotions are rather than how pleasant they are.
Arabic is a highly systematic language where its words exhibit elegant and rigorous logic. The field of Arabic word semantic similarity becomes more challenging due to its higher complexity and subtlety. This research is concerned with investigating the development of free open-source frameworks containing packages to calculate the semantic similarity between two Arabic words or concepts. These packages are known as AWN-ConceptSimilarity and AWN-WordSimilarity. The developed packages implement seven semantic similarity algorithms. One of these algorithms was proposed for Arabic and the rest were proposed for English where successfully adapted to Arabic using an Arabic lexical database, Arabic wordnet. The functionality of the developed packages is validated using two-word similarity benchmarks datasets previously produced for Arabic. The results of the validation process indicate that the developed frameworks represent an important contribution to the Arabic semantic similarity field. Moreover, the developed packages are reliable to use and embed them with Arabic researchers' projects for improving or comparing their methodologies.
Reviewed by: Männliche Hauptfiguren im ‘Tristan’ Gottfrieds von Strassburg. Charakterisierung, Konstellation und Rede by Anna Karin Charles Taggart anna karin, Männliche Hauptfiguren im ‘Tristan’ Gottfrieds von Strassburg. Charakterisierung, Konstellation und Rede. Berlin: Walter de Gruyter, 2019. Pp. 388. isbn: 978–3–11–057225–4. €99.95. In the published version of her dissertation, Anna Karin provides a series of close readings centered on verbal and non-verbal communication in the thirteenth-century Tristan by Gottfried von Strassburg. The ‘primary male figures’ [männliche Hauptfiguren] referred to in the monograph’s title are Tristan, a knight, courtier, and minstrel, and King Marke, Tristan’s uncle, sovereign, and husband of their shared love interest, Isolde. Karin presents these two figures as polar opposites in how, what, and why they communicate. Tristan is ‘the director, who calls the shots’ [der Regisseur, der die Anweisungen gibt] (p. 260), exploiting courtly discourse and conventions as he assumes myriad different roles. Marke, by contrast, remains ‘trapped in the courtly code of behavior’ [Verhaftetsein im höfischen Verhaltenskodex] (p. 300) and adheres to this code long after his nephew and disreputable courtiers have undermined its ideals. Karin’s sharp focus on communication in Gottfried’s text demonstrates ‘just how sophisticated individual figures’ speech [...] can be even in the Middle High German epic despite being strongly shaped by rhyme and meter’ [wie differenziert [End Page 173] die individuelle Figurensprache bereits in der mittelhochdeutschen Epik trotz ihrer Überformung durch Reim und Metrik sein kann] (p. 363). This study devotes great attention to detail, and the arguments are clear and cogent. However, Karin is swimming with the tide. Much attention has already been paid to mimetic and diegetic language as well as the resulting slipperiness of meaning and interpretation in Tristan, as Karin’s own bibliography and use of secondary sources demonstrate. More fruitful is Karin’s decision to treat Tristan’s linguistic cunning as an integral part of his heroic status, calling him ‘a hero—and a linguistic hero’ [ein Held–und ein Sprachheld] (p. 260). Gottfried endows his primary figure with boundless linguistic prowess, and his Tristan impeccably mimics courtly behavior and discourse. At the same time, Tristan is unrepentantly savage toward his adversaries, plays fast and loose with the truth by taking on new identities, and shows little remorse for his transgressions. Rather than trying to reconcile Tristan’s courtliness with his lies and savagery, Karin allows these characteristics to exist side by side in the figure of the ‘linguistic hero.’ Karin builds on her lengthy analysis of the figure of Tristan and contrasts him with Marke, persuasively arguing that each figure represents competing communicative trends in Gottfried’s text: ‘Linguistically, the two male figures come across differently, as downright complete opposites: Where what Tristan says seems atypical, innovative, manipulative and proactive, Marke’s language comes across as restrictive, conventional and reactive’ [Die beiden Männerfiguren werden sprachlich als different wahrgenommen, regelrecht als Gegenpole: Da, wo Tristans Sprechen als außergewöhnlich, innovativ, manipulativ und aktiv erscheint, wirkt Markes Sprache begrenzt, konventionell und reaktiv] (p. 5). Tristan is a radical who ‘repeatedly creates new identities’ [immer wieder neue Identitäten konstruiert] (p. 50) through his words and actions, while Marke embodies a stable code of linguistic norms and behaviors as an ‘agent of courtly values’ [Vermittler höfischer Werte] (p. 285). There has been a tendency in the abundant research on Tristan to either condemn or exonerate the figure of Tristan based on a broad conception of courtliness. In much the same vein, Marke’s stringent adherence to courtly values, even when upended by his court and nephew, has frequently been interpreted as a sign of the king’s weakness. But Karin refuses to square the circle in either case. In her analysis, the figure of Tristan is cloaked in and employs the language of courtly literature, and he sets himself apart because of ‘his charming and, at the same time, horrifying nature’ [sein einnehmendes und gleichermaßen erschreckendes Wesen] (p. 261). Marke is confronted with the unenviable task of attempting to harmonize his nephew’s deviancy with a grossly inadequate courtly code. Gottfried endows each figure with a particular approach to language, and it...