Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
Abstract This article describes the steps and results of the lemmatization of the derived anomalous verbs of Old English. The data have been retrieved from The Dictionary of Old English Web Corpus, searched through the lexical database from the Nerthus Project called Norna. The methodology comprises several steps combining automatic searches on the lemmatizer and manual revision. Part of the results, including the verbs starting with the letters A to H, are compared with the Dictionary of Old English, while the rest of the lemmas are checked with the standard Old English dictionaries (Clark-Hall, Sweet and Bosworth-Toller). The discussion leads to the conclusion that the lemmatization of the verbs of Old English, a language with a remarkable degree of spelling variation, requires considerable manual revision. However, the progressive improvement of automatic searches, based on the comparison of the initial results with the available lexicographical sources, minimizes the need for manual adjustment.
This research examined whether the semantic relationships between representational gestures and their lexical affiliates are evaluated similarly when lexical affiliates are conveyed via speech and text. In two studies, adult native English speakers rated the similarity of the meanings of representational gesture-word pairs presented via speech and text. Gesture-word pairs in each modality consisted of gestures and words matching in meaning (semantically-congruent pairs) as well as gestures and words mismatching in meaning (semantically-incongruent pairs). The results revealed that ratings differed by semantic congruency but not language modality. These findings provide the first evidence that semantic relationships between representational gestures and their lexical affiliates are evaluated similarly regardless of language modality. Moreover, this research provides an open normed database of semantically-congruent and semantically-incongruent gesture-word pairs in both text and speech that will be useful for future research investigating gesture-language integration.
Випуск 13.Том 2 and pragmatics is only a small sample of the coinages based on the analysed non-combined and combined word-formation models, used in Modern English media discourse.Media is an amazing source of data for any linguistic research area.It reflects the dynamic changes occurring in the language.An increasing number of lexical innovations, customary found in Modern English media to represent the totality of facts in various spheres of life, are used to organize the message in the media discourse.These findings provide the following insights for further research: to establish the cognitive mechanisms in the process of forming lexical innovations involving a particular word-formation model; to determine the role of lexical innovations in organization of other types of discourse. ФОНЕТИЧНА СТРУКТУРА ФРАНЦУЗЬКИХ ЗАПОЗИЧЕНЬ В АСПЕКТІ БРИТАНСЬКОЇ ВИМОВНОЇ НОРМИ
The article discusses the issues of constructing the identity concept that is relevant for cross-cultural communication in the frame of literary discourse. The author considers the literary text as a symbiosis of the content and features of individual creativity, which reflects the author's ethnic and cultural identity. This is due to the poly-code nature of the literary text, which has an ethnic specificity, implemented in the dichotomy ‘Us’ vs. ‘Them’. While studying the relationship between ‘Us’ vs. ‘Them’ in Russian and German linguistic cultures embodied in lexical units, the author comes to the conclusion that these language units of literary texts with ethnic coloration act as specific indicators of communicative behavior of men and women belonging to different ethnic groups. The article emphasizes that in the mentality of the ethno-cultural community the concepts of ‘Us’ vs. ‘Them’ are also reflected by means of stereotypes through which the characteristic of ‘Us’ in comparison with ‘Them’ is revealed. Analyzing the literary works by Eugene Vodolazkin, the author pinpoints the heterostereotypes underlying the discreteness of ‘Us’ vs. ‘Them’. Besides, the author focuses on the fact that ethnic self-identification in this opposition is based on relations within such categories as norms of behavior, traditions, food, drinks, clothing, language and habits. The main aim is to determine how the self-identification of the personality created by the writer is expressed linguistically based on existing stereotypes in the society regarding the ethnic community of different cultures representatives.
Mindful meditation, an exercise which encourages its practitioners to be present in the moment and to be aware of their current emotions, thoughts, and sensations, has been shown to affect the processing of emotional information (Sobolewski et al., 2011) and to increase empathy (Tan, Lo, & Macrae, 2014). The prefrontal cortex has been implicated in these processes (Seitz, Nickel, & Azari, 2006). We sought to investigate the neural mechanisms which underlie how mindful meditation affects emotional processing and to determine whether any changes in brain activity could be linked to changes in empathy. Participants in our experimental group practiced mindful meditation for ten minutes. Those in the control groups either listened to an unexciting newscast or sat quietly for ten minutes. Next, participants in all conditions viewed a series of emotionally valenced images from the International Affective Picture System (IAPS) (Lang, Bradley, & Cuthbert, 2008). As they viewed the images, activity in the prefrontal cortex was monitored with functional near-infrared spectroscopy (fNIRS). Following the presentation of the IAPS images, participants completed questionnaires and inventories that measured empathy, previous experience with meditation, and personality traits. As well as differences in empathy between participants who meditated and participants who did not, we expect to discover any variations in patterns of prefrontal cortical activity during the viewing of positive, negative, and neutral imagery. We hope that our results will elucidate how mindful meditation affects emotional processing and the development of empathy, and to determine whether certain personality traits, social conservatism, or lack of sleep may predict neural and behavioral responses to mindful meditation. Lang, P.J., Bradley, M.M., & Cuthbert, B.N. (2008). International affective picture System (IAPS): Affective ratings of pictures and instruction manual. Technical Report A-8. University of Florida, Gainesville, FL. Seitz, R.J., Nickel, J., & Azari, N.P. (2006). Functional modularity of the medial prefrontal cortex: Involvement in human empathy. Neuropsychology, 20(6), 743-751. Sobolewski, A., Holt, E., Kublik, E., & Wróbel, A. (2011). Impact of meditation on emotional processing—A visual ERP study. Neuroscience Research, 71(1), 44–48. doi: 10.1016/j.neures.2011.06.002 Tan, L.B.G., Lo, B.C.Y., & Macrae, C.N. (2014). Brief Mindfulness Meditation Improves Mental State Attribution and Empathizing. PLoS ONE, 9(10). doi: 10.1371/journal.pone.0110510
It has been frequently observed in the literature that assertions of plain sentences containing predicates like fun and frightening give rise to an acquaintance inference: they imply that the speaker has first-hand knowledge of the item under consideration. The goal of this paper is to develop and defend a broadly expressivist explanation of this phenomenon: acquaintance inferences arise because plain sentences containing subjective predicates are designed to express distinguished kinds of attitudes that differ from beliefs in that they can only be acquired by undergoing certain experiences. Its guiding hypothesis is that natural language predicate expressions lexically specify what it takes for their use to be properly ‘grounded’ in a speaker's state of mind: what state of mind a speaker must be in for a predication to be in accordance with the norms governing assertion. The resulting framework accounts for a range of data surrounding the acquaintance inference as well as for striking parallels between the evidential requirements on subjective predicate uses and the kind of considerations that fuel motivational internalism about the language of morals. A discussion of how the story can be implemented compositionally and of how it compares with other proposals currently on the market is provided.
The article is devoted to topical problems of translation of modern English-language film discourse (based on the TV series "The Big Bang Theory"). Stylistic features of English-language film discourse are characterized. It is noted that the language of English-language film discourse has certain features and directions for certain categories of speakers and recipients. In the course of the analysis of English-language film discourse, a direct relationship between the degree of complexity of the selected language tools and socio-cultural specific features of the target audience is proved. The irony is a move to challenge norms, rules, common sense. It easily turns into a paradox and a joke. Comparisons of two or more textual worlds, styles, paradoxical comparisons, quotations, parodies lead to endless possibilities for variations in the understanding of ironic means. Socio-cultural linguistic elements that complicate film translation include realities (non-equivalent vocabulary), proper names, idiomatic expressions and jargon, dialect and variant features, humor. Non-linguistic features of the socio-cultural character contained in the visual and sound plans of a motion picture can also affect translation. At the same time, empirical studies of the application and methods of film translation testify to the existence of specific operational rules, that is, the patterns of behavior of the translator in some sociocultural situation – the situation of translation for film screening. Thus, the conducted study shows that "film discourse" is a "broader concept than cinema text” and “film dialogue”, which includes various correlations with other fields of science, such as literature, theater, art, etc. In addition, it is in the cinema discourse that the final interpretation of sen sous, embedded in the movie. In this case, cinema text is a fragment of cinema discourse and includes two heterogeneous semiotic systems: linguistic and non-linguistic, the film dialogue appears as the linguistic component of the film. Therefore, the audiovisual images operated by cinema become an indispensable element of the new discourse of modernity, which is the source of social, cultural, psychological as well as linguistic knowledge. This is why book adaptations are so popular because they save time and effort. However, as a result of a comparative lexical stylistic analysis of the book and its TV version, it was found that the number of lexical means and stylistic figures in the literary work is much greater and striking in its diversity, especially when it comes to descriptions. In the film, the language becomes poorer because dialogic speech is a predominantly spoken-and-everyday style characterized by general vocabulary and changes in the syntactic structure of sentences. The sentences in the movie are usually simple, not complex, full of exclamations and pauses and easy to hear. Film discourse should be understood as the process of play and perception of a film, the meaning of which is the mutual influence of several semiotic systems. Cinema discourse involves participants in the discourse, time and space of their interaction. The main difficulty in translatable movie text is the possibility and degree of adaptation of the text to a foreign language culture, built on a different system of values and concepts, and this factor causes the inevitable loss in the perception of translatable cinema with other subjects and / or incompatible with other linguistic culture failures of a large number of films.
Kinship is a fundamental and universal aspect of the structure of human society. The kinship category of 'grandparents' is socially salient, due to grandparents' investment in the care of the grandchildren as well as to older generations' control of wealth and cultural knowledge, but the evolutionary dynamics of grandparent terms has yet to be studied in a phylogenetically explicit context. Here, we present the first phylogenetic comparative study of grandparent terms by investigating 134 languages in Pama-Nyungan, an Australian family of hunter-gatherer languages. We infer that proto-Pama-Nyungan had, with high certainty, four separate terms for grandparents. This state then shifted into either a two-term system that distinguishes the genders of the grandparents or a three-term system that merges the 'parallel' grandparents, which could then transition into a different three-term system that merges the 'cross' grandparents. We find no support for the co-evolution of these systems with either community marriage organisation or post-marital residence. We find some evidence for the correlation of grandparent and grandchild terms, but no support for the correlation of grandparent and cross-cousin terms, suggesting that grandparents and grandchildren potentially form a single lexical category but that the entire kinship system does not necessarily change synchronously.
Latvian Radio offers an exchange of opinions and discussions on various subjects. Channel 1 has a show “Kā labāk dzīvot” (‘How to live better’), which among other issues addresses the use of Latvian. This paper is based on the questions covered in 15 broadcasts of the years 2017–2019. What are the listeners worried about? Usually, it is the question of whether the word or phrase is wrong. Does it correspond to the norms and conventions, can it be found in the dictionaries and how it is defined and explained. There is often a clash of opinions on the use between people of different generations. Many questions relate to grammar norms, their application and explanation. These are issues of declining of proper names, use of singular and plural, and gender. There is much uncertainty about the use of lexis often governed by rigid and conservative views. There seems to be more unanimity on issues of style. The impact of English and separate English loans attract numerous questions. English affects various levels of Latvian today. The paper views phonetic, morphological, lexical, phraseological, and syntactic influence. Lexical impact of English is the one felt most: nonce words, loans, translation loans, idioms can be met in both translated and original texts. A semantic broadening of many Latvian terms under the influence of English is widespread, often without any need and additional stylistic value. Borrowed synonyms for Latvian words are in no way detrimental, but they should not oust Latvian words. Many listeners enjoy the opportunity of gaining knowledge on the use of Latvian and thus improving their language competence. We cannot and should not try to control the development of the language, but every speaker can contribute to its perfection, enrichment, and innovation.
The article deals with the concepts of "insulting potential" and "insult" and their representation in con-flict-ridden texts. The author tries to answer the questions which an expert faces when conducting a linguistic expertise and states the necessity to take into account pragmatic factors, as well as the extralinguistic situation as a whole, when analyzing the fact of insult. The urgency of the topic can be attributed to the need for an in-depth study of the interaction of various definitions of offensiveness, creating difficulties in qualifying the legal norm "insult"..
The issues of Russian lexical borrowings (rusisms) in the Bashkir language dialects and subdialects have not been addressed yet. Dictionaries and monographs on the Bashkir language dialects and subdialects describe specific dialectal loanwords without providing a dialectal analysis of loanwords and the specific features of their adaptation and functioning in the Bashkir language dialects and subdialects. Meanwhile, studying rusisms in dialects and subdialects can elucidate both the dialectal lexicology and the formation history of the lexical, phonetic, and grammatical features of a particular Turkic language. Investigating rusisms in dialects and subdialects of Turkic languages, including Bashkir, is also relevant for the Russian language dialectology: the chronology of individual borrowings. It is worth studying the Bashkir language southern dialect widespread in the southern regions of modern Bashkortostan, Bashkir-speaking regions of Orenburg, Samara, and Saratov regions of Russia. Historically located in the very center of the Orenburg province, this territory bordered the provincial city of Orenburg and by the late 18th and early 19th centuries became one of the administrative, political, economic, and trade centers. It was then that Russian loanwords and lexemes of European languages began to actively penetrate the Bashkir dialects. These borrowings constitute a considerable group, thematically related to household, administrative and managerial, military- marching, and agricultural spheres. All rusisms underwent adaptation to the norms of the Bashkir language Southern dialect, e.g., Russian lexemes with hard-row vowels in the southern dialect have front-row vowels. South Russian dialects are considered the dominant source of the Bashkir language southern dialect lexical borrowings.
We address the problem of unsupervised extractive document summarization, especially for long documents. We model the unsupervised problem as a sparse auto-regression one and approximate the resulting combinatorial problem via a convex, norm-constrained problem. We solve it using a dedicated Frank-Wolfe algorithm. To generate a summary with k sentences, the algorithm only needs to execute k iterations, making it very efficient. We explain how to avoid explicit calculation of the full gradient and how to include sentence embedding information. We evaluate our approach against two other unsupervised methods using both lexical (standard) ROUGE scores, as well as semantic (embedding-based) ones. Our method achieves better results with both datasets and works especially well when combined with embeddings for highly paraphrased summaries.
English has become ‘the world’s default mode’ (McArthur, 2002: 13) for communication. As a de facto lingua franca, English and its associated cultures are increasingly pluralistic. According to Kachru (1996: 135), ‘the term “Englishes” is indicative of distinct identities of the language and literature. “Englishes” symbolizes variation in form and function, use in linguistically and culturally distinct contexts, and a range of variety in literary creativity.’ As far as Chinese English is concerned, Kirkpatrick & Xu (2002: 278) suggest that since ‘the great majority of the estimated 350 million Chinese’ who have been learning English are far more likely to use it with other speakers of world Englishes, the development of Chinese English ‘with Chinese characteristics’ will be ‘an inevitable result’. Kirkpatrick & Xu also predict that such a variety of English will be characterized by linguistic and cultural norms derived from Chinese. This chapter will review the definitions of Chinese English, and then identify a selection of lexical, syntactic, discourse and pragmatic features of Chinese English based on an analysis of a variety of data including interviews, newspaper articles, and literary works. The chapter will conclude by considering the likelihood of Chinese English becoming a powerful variety of English.
The article is devoted to the communicative competence of a doctor as a component of professional ethics. Knowledge of norms of the modern Russian literary language, compliance with these standards in the oral and written speech of a medical worker helps to establish contact between doctor and a patient. To identify the level of knowledge of Russian language norms, readiness for professional speech a scientific research was made, during which the most typical mistakes were revealed: orthoepic, morphological, lexical, stylistic. Following the norms of the language and ethics of communication contributes to the achievement of the main aim of a medical activity - recovering of a patient.
Dans cet article, nous proposons un modele de representations vectorielles de paire de mots, obtenues a partir d’une adaptation du modele Skip-gram de Word2vec. Ce modele est utilise pour generer des vecteurs de paires de verbes, entrainees sur le corpus de textes anglais Ukwac. Les vecteurs sont evalues sur les donnees ConceptNet & EACL, sur une tâche de classification de relations lexicales. Nous comparons les resultats obtenus avec les vecteurs paires a des modeles utilisant des vecteurs mots, et testons l’evaluation avec des verbes dans leur forme originale et dans leur forme lemmatisee. Enfin, nous presentons des experiences ou ces vecteurs paires sont utilises sur une tâche d’identification de relation discursive entre deux segments de texte. Nos resultats sur le corpus anglais Penn Discourse Treebank, demontrent l’importance de l’information verbale pour la tâche, et la complementarite de ces vecteurs paires avec les connecteurs discursifs des relations.
The article presents the results of a linguistic and cultural study devoted to the problem of verbalization of the concepts of birds, which in the Ukrainian worldview are given special opportunities to predict human destiny. The relevance of the study is related to the key role of birds and their stereotypes as important components of the national language picture of the world and the works of M. Stelmakh in particular. The aim of the article is to find out the peculiarities of verbalization of bird concepts, which are associated with negative perceptions in the Ukrainian ethnic group, to outline their traditional symbol meanings, and new associations due to the historical era, as well as to substantiate their estimated value. Proverbs and sayings as apt pearls of folk wisdom convince that birds have been an object of admiration for Ukrainians, a subject for comparison, formulation of their own conclusions about the world of nature and its impact on human life and activity. The crow, the owl, and the bat are not just concepts of birds, but concepts-symbols, concepts-soothsayers. The most relevant national and cultural stereotypes about them in the Ukrainian-language picture of the world do not lose their relevance and acquire a kind of manifestation in literary texts. They are based on primary religious beliefs, national experience, and folk traditions. In M. Stelmakh’s literary works the folk symbolism of these birds acquires a special relevance, which is best manifested in the syntagmatic ties of their key nominations. The general negative evaluation inherent in the outlined concepts is due to the relevant features of denotations, their external, behavioral characteristics, beliefs and observations established in the Ukrainian language picture of the world. In context, the negative features sometimes turn into positive ones, and lexical means of exposing the evil, become a means of humor, fascination with birds as realities of nature, to which Ukrainians have always shown respect and love. The ambivalence of the concepts of soothsayers is a national-cultural stereotype that embodies the evaluative norms and values of the ethnos at the level of relationships and interaction of two important concept spheres – nature and man. In literary texts, they become evaluative examples for comparing the objects of society and their properties.
This article analyses the typical features of the influence of the Russian language on Belarusian stage speech. Contemporary theatre, concerts, and other cultural events have played an important role in the popularization in the Belarusian language of misbalanced bilingualism. According to the transcripts of modern actors, emcees, and musicians recorded in theatres, music clubs, literature museums, and culture centres between 2012 and 2019, different fragments and types of mixing of Belarusian and Russian are revealed. Firstly, we notice switching of language code, which is used for famous phrases from literature, cinema, etc., or shows someone else’s speech. Secondly, we have examples of Russian language interference in Belarusian at the lexical, grammatical, and phonetic levels. So-called trasianka-distorted Russian language according to the special features of the Belarusian orthoepyis observed as a specific kind of interfered speech. In addition, various variants of hard mixing of systems and subsystems, functional styles of two languages (especially including of dialect words and forms), jargon lexis, and special units of alternative norms of the Belarusian language tarashkevitsa are shown. This research notes that mixed speech is a result not only of bilingual interaction, but the decline in the standard of stage speech. It is noted that trasianka helps to create a comic or negative character image on stage.
In the last decade, intralingual translation has started to gain momentum amongst a number of translation academics. Nevertheless, some types of intralingual translation remain largely undiscovered, such as the process of abridgement in the production of simplified versions of classic literary works (i.e. graded readers). This article subjects three chapters of the abridged version of And Then There Were None by Agatha Christie to qualitative analysis using Descriptive Translation Studies theory. The aim is to contribute to bridging a research gap in Translation Studies by examining the norms and laws governing the process of abridgement. Translation norms and laws are detected by situating the source and the target texts in their respective socio-cultural backgrounds and by analysing translation shifts. Relevant shifts are identified by means of a check-list of features elaborated on the basis of theory on graded readers, which classifies them into lexical, structural and information shifts. The results of the analysis showcase the vast research potential of intralingual translation for language learning purposes. Keywords: Intralingual translation, abridgement, graded readers, DTS, translation norms, translation laws, Agatha Christie, And Then There Were None
Summary This paper compares the romanization of Gaul in the 1st century BC and the gallicization of the island of Martinique during 17th-century French colonial expansion, using criteria set out by Muf- wene's Founder Principle. The Founder Principle determines key ecological factors in the formation of creole vernaculars, such as the founding populations and their proportion to the whole, language varieties spoken, and the nature and evolution of the interactions of the founding populations (also referred to as “colonization styles”). Based on the comparison, it will be claimed that new languages arise when a language undergoes vehicularization and subsequently shifts from one speech community to another. In other words, linguistic genesis would be a complicated case of language contact, where not only one, but sev- eral dialects of both superstrate and substrate varieties are involved, in a historical context where the identity function of language, or the norm, is overriden by the need to communicate. Research also indicates that language varieties spoken at the time of the shift did not pertain to normative usage, but to popular varieties, dialects, or both, since the emerging vernaculars - in Gaul, as well as in Martinique - preserved some of their phonological and lexical particularities.
Binary rating methods are known to allow easier and quicker ratings, yet they do not allow users to express how much they like/dislike an item. In this study, we argue that response times in ratings can be used to gauge the degree or strength of user ratings. We asked users to rate a set of movies while collecting individual response times, confidence levels, and reasons for user ratings. We investigate (1) the possibility of utilizing response time as a way to further distinguish likes and dislikes, and (2) the various factors that affect rating time. We find that response time can be used to distinguish between sure and unsure answers, as well as the mental process users take before rating an item. We believe that our study will be informative to better find user preferences and relations between explicit ratings and implicit data.
The purpose of this study was to find out the correlation between the level of academic background and the rules of Hangul orthography by examining the compliance of Korean spelling. To examine this, six universities were divided into three divisions by level into A, B, and C, and the status of marking by university was compared with the free bulletin board of Everytime, a college student community. As a result of the survey, A-grade K universities best observed the Hangul Hangul orthography followed by C-grade A universities. The university that did not keep the Hangul orthography well was a C-grade D university. By educational level, the grade with the lowest mislabeling rate is grade A. It can be seen that the status of compliance with the lexical norms is related to each level of education. However, it cannot be generalized because there are not many vocabulary in common from six universities, but there is some correlation between knowledge and actual notation.
This paper deals with the topic of lexical modality in Norwegian as a second language. Basing on data obtained from the ASK-corpus -The Norwegian Language Learner Corpus containing second language texts written in a language examination, the authors analyse the use of two modal particles, jo and nok, by three groups of Norwegian as a second language writers: English, Polish and German. The focus of the study is on analysing lexical patterns for co-occurrence of the modal particles. The patterns used by the learners are compared with the ones used by the native speakers of Norwegian and between the learner groups. The discrepancies found in the data are discussed within the broader framework of second language development and second language writing. The findings suggest that the second language writers' use of modal particles is influenced by several factors, such as general interlanguage tendencies, transfer from the learners' first languages and the perception of textual norms.
SloIE is a manually labelled dataset of Slovene idiomatic expressions. It contains 29,400 sentences with 75 different expressions that can occur with either a literal or an idiomatic meaning, with appropriate manual annotations for each token. The idiomatic expressions were selected from the Slovene Lexical Database (http://hdl.handle.net/11356/1030). We selected only expressions that can occur with both a literal and an idiomatic meaning. The sentences were extracted from the Gigafida corpus. For each sentence, the file first contains the text of the sentence prefixed by #. This is followed by a line of numbers indicating the positions of tokens that belong to the expression. The numbers also indicate the word order for expressions where the word order is flexible. They are ordered according to the dictionary form of the expression (e.g., the first number indicates the position where the first word of the expression - in its dictionary form - occurs). Each token is labelled with either 'DA', indicating tokens in an expression that have an idiomatic meaning, 'NE', indicating tokens in an expression that have a literal meaning, or '*', indicating tokens outside the expression. Additionally, 'NEJASEN ZGLED' indicates tokens where the annotators could not determine the meaning from the example sentence. Each token is also tagged with the dictionary form of the expression that is present in the sentence. Key reference: Škvorc, Tadej, Polona Gantar, and Marko Robnik-Šikonja. "MICE: Mining Idioms with Contextual Embeddings." arXiv preprint arXiv:2008.05759 (2020).
In a theoretical article on the current issue of how to teach the culture of speech of future primary school teachers, an attempt is made to describe some well-known views, definitions of linguistic scholars by this term, to find out how linguists are the only ones in the interpretation of this category. And also on the basis of the concepts presented, the author characterizes what the speech of the future primary school teacher should be in order to be a model for a younger student in his education in a general educational institution. Analysis of various pedagogical sources makes it possible to talk about the traditionally established communicative qualities of speech culture in general and teachers in particular - this is correctness, richness, relevance, sufficiency, logic, accuracy, clarity, brevity, simplicity and emotional expressiveness, imagery, colorfulness, euphony, purity, emotionality, multifunctionality and the like. Among the aspects of the manifestation of the culture of speech, we highlight aesthetics: the use of expressive-stylistic means of language that make speech rich and expressive; communicative expediency: it is not enough to speak and write correctly, one must still be able to use words and expressions in accordance with the communicative situation. It is noted that the teacher’s speech culture should be ensured by: the correctness of the professional utterance associated with the observance of the language norm at the phonetic, lexical, grammatical, word-formation, stylistic levels; the accuracy of the wording (that is, a strict correspondence between the word and the concept that is designated by this word), since the concept of “accuracy” includes two aspects: accuracy in the reflection of actions and accuracy of the expression of thought in the word; the purity of speech, that is, the absence of elements in it that are not related to the norms of the literary language: stamps, semantic-non-self-dependent words (in fact, in fact, in a certain way, in fact, etc.), dialectisms, colloquial words, as well as words and phrases, in which interfering errors are made due to the influence of other languages in the Ukrainian language of students.
Lexical normalization is the task of translating non-standard social media data to a standard form. Previous work has shown that this is beneficial for many downstream tasks in multiple languages. However, for Italian, there is no benchmark available for lexical normalization, despite the presence of many benchmarks for other tasks involving social media data. In this paper, we discuss the creation of a lexical normalization dataset for Italian. After two rounds of annotation, a Cohen’s kappa score of 78.64 is obtained. During this process, we also analyze the inter-annotator agreement for this task, which is only rarely done on datasets for lexical normalization,and when it is reported, the analysis usually remains shallow. Furthermore, we utilize this dataset to train a lexical normalization model and show that it can be used to improve dependency parsing of social media data. All annotated data and the code to reproduce the results are available at: http://bitbucket.org/robvanderg/normit.
Stylistics is a science born from the works of Charles Bally at the start of the 20th century. It centers on literary texts. To decipher them, stylistics focuses on methods and concepts from the language sciences. It is therefore effective in bringing to light the various subtleties of the African novel which continues to regenerate from the resources of African language and linguistic heritage. This regeneration results in an Africanization of the writing language. It follows that re-lexicalization is at the heart of verbal praxis. It is materialized by the embedding of oral ethno-texts in the novel. This creativity is a form of discursive subversion which is characterized by a polyphonic and transgeneric aesthetics. Le-fils-de-la-femme-mâle by Maurice Bandaman conforms to this standard. The typographical configuration shows that the enunciative foundations of the work rests on the structure of the traditional African tale. As a result, the novel resists the norms of traditional storytelling. Furthermore, the insertion of the African tale into the structure of the novel is not based on any stable rule. The discourse is based on an intertextual practice that upsets the architectonics of the classic novel. It breaks with the canon and imposes another reception modality which provides obvious enjoyment to the reader. This work highlights all of the enunciative, linguistic and discursive phenomena that contribute to the authenticity of this work.
The present study is an approach to Aymara ethnogeometry that aims to identify mathematical terminology on Aymara geometry, under the epistemological support of Ethnomathematics and interculturality, hoping that it can have a positive impact on the learning and identity of students from rural schools in Puno, where there are problems due to linguistic interference. Within the framework of the ethnographic method, the information was obtained through visits and interviews with the Aymara speakers of the communities of the Moho and El Collao province of the Puno region - Peru, contrasted and supplemented with a documentary source Vocabulary of the Aymara language from 1612 and current specialized literature (books and dictionaries). The geometric terms were identified by equivalence and conceptual approximation, and this process showed that the Aymara language has a rich and flexible mathematical lexical background of its own to adapt to current scientific and pedagogical requirements, whether by creating neologisms or borrowing from raw or foreign languages, in case of gaps. The terminology presented in tables was written respecting the linguistic norms of Aymara, and it is expected that it will be standardized and socialized by the authorities of the Ministry of Education, and serve as the basis for future discussions and research.
-Currently, the coverage of the studies of language levels between related languages is expanding in terms of historical continuity in linguistics, thus much attention is paid to the problem of finding complicity in related languages. It is because of a sign system, which saved the nation’s history, culture, cognition in a lexical richness of each language. Linguistic signs that have emerged as the norm and widespread among the people in the structure of any language are very common in the language of other people. This feature is very often found inside the Turkic languages themselves. This is a key factor that shows historical relations between Turkic peoples. Integral selection and study of pronouns, which constitute a large-scale part of the lexical fund, their comparison from the point of view of a separate linguistic phenomenon, are the definition of language and cognition of a related ethnic group and is also considered a spiritual and cultural source of accurate information about their relationship. The article discusses the similarities and differences of Turkic pronouns through the study by the method of historical comparison of pronouns in the Turkic languages.
OBJECTIVE: Superselective pseudocontinuous arterial spin labeling (ss-pCASL) is an MRI technique in which individual vessels are labeled to trace their perfusion territories. In this study, the authors assessed its merit in defining feeding vessels and gauging preoperative embolization feasibility for patients with meningioma, using digital subtraction angiography (DSA) as the reference method. METHODS: Thirty-one consecutive patients with meningiomas were prospectively recruited, each undergoing DSA (and embolization, if feasible) before resection. All ss-pCASL imaging studies were performed 1 day prior to DSA. Two neuroradiologists independently reviewed ss-pCASL images, rating the contribution of each labeled vessel to tumor blood supply as none, minor, or major. Two neuroradiologists also gauged the feasibility of embolization in each patient, based on ss-pCASL images. Interobserver and intermodality agreement were determined using Cohen's kappa statistic. The diagnostic performance of ss-pCASL was assessed in terms of discerning tumor blood supply and the potential for embolization. RESULTS: Interobserver agreement in the rating of blood supply by ss-pCASL was very good (κ = 0.817, 95% CI 0.771-0.863), and intermodality agreement (consensus ss-pCASL readings vs DSA findings) was good (κ = 0.688, 95% CI 0.632-0.744). In delineating tumor blood supply, ss-pCASL showed high sensitivity (87.1%) and specificity (87.2%). The positive and negative predictive values for embolization feasibility were 85.2% and 100%, respectively. CONCLUSIONS: In patients with meningiomas, feeding vessels are reliably predicted by ss-pCASL. This noninvasive approach, involving no iodinated contrast or radiation exposure, is particularly beneficial if there are no prospects of embolization.
This study aims to examine whether there are linguistic differences in College Scholastic Ability Test (CSAT) English using corpus linguistic programs such as Readability Formulas and Lexical Complexity Analyzer (LCA) on the basis of the criterion-referenced assessment introduced in the 2018 academic year. The study compares CSAT English reading from 2015 to 2017 academic year (norm-referenced assessment) and 2018 to 2020 academic year (criterion-referenced assessment). There are no significant differences on any of the measures. As level of difficulty in CSAT English is to be maintained, we see that the exam purpose has been achieved. However, the results of evaluation with the Readability Formulas web-site indicate that the level of CSAT reading texts is too high for twelfth grade students in Korea. Previous studies have found that the policy of linking EBS and CSAT has produced difficult reading passages in CSAT English. In accord with the aim of the national curriculum to enhance students’ communicative skills, decreasing the difficulty level of reading texts and removing some word families or vocabulary at too high a level may cause teachers and students to focus more on speaking and writing; and speaking and writing skills need to be directly assessed.
In the modern information society, the term communicative culture combines such concepts as the culture of dialogue in business and everyday communication, mutual understanding, and tolerance.Сommunicative culture сharacteristics include value, normative, and informational components.We understand the information component as the content of informationtext, since in the last century, text studies consider text as a speech production of a certain genre form, which accumulates norms and qualities of speech and is aimed at communication, that is, text is an integral component of speech and communication culture.Textual analysis allows students not only to master the full range of knowledge about the vocabulary of the language and its norms, about the grammatical structure and lexical compatibility of words, but also to develop speech skills.The article deals with the development of students and schoolchildren's communicative culture in the process of analyzing professionally oriented text, offers didactic material and tasks that ensure the development of communication skills in the Russian language lessons of schoolchildren and first-year university students of philology in the lessons of methodology, such as "Workshop on Spelling and Punctuation", "Difficult Cases of Spelling and Punctuation", "Speech Practices", etc.
The article provides an overview of the lexical and grammatical features as well as the sociopolitical environment of Marollien that originated in the 18th century as a dialect on the territory of Brussels. Marollien is essentially the Dutch language in its Brabantian dialect, strongly influenced by French. There are literary works, performances, and musicals written and staged in Marollien, as well as dictionaries and journals published in it. Historically, the Marollien dialect is a sociolect: it was generally used by Belgians coming to Brussels from Wallonia in search of a job and settling in one of the districts of Brussels — Marolles. A special emphasis is placed on lexical features of the dialect: gastronomic and everyday vocabulary are looked at and the examples of French loanwords and Southern Dutch language norm deviations are provided. Standard Dutch calques in French, when translating idioms in particular, are also identified. The differences between Dutch, French, and Marollien place names are illustrated. In the field of morphology and word formation, there is a regular mixture of Germanic and Romanic stems which is indicated. Examples of Marollien phonetic features are also provided. The article acknowledges frequent code switching in Marollien speech, which by and large resembles the phenomenon of linguistic interference. Due to the fact that Marollien is rapidly disappearing, the Brussels-Capital region is trying to support the dialect: various activities are being organized in order to propagate its use and enhance its prestige. Nevertheless, Marollien is not included in the well-known citizen initiative “Marnix Plan”, aimed at developing the methodology for the sequential study of several languages for all segments of the population in Brussels. This initiative is also discussed in the article.
As in any field of inquiry that depends on experiments, the verifiability of experimental studies is important in computational linguistics. Despite increased attention to verification of empirical results, the practices in the field are unclear. Furthermore, we argue, certain traditions and practices that are seemingly useful for verification may in fact be counterproductive. We demonstrate this through a set of multi-lingual experiments on parsing Universal Dependencies treebanks. In particular, we show that emphasis on exact replication leads to practices (some of which are now well established) that hide the variation in experimental results, effectively hindering verifiability with a false sense of certainty. The purpose of the present paper is to highlight the magnitude of the issues resulting from these common practices with the hope of instigating further discussion. Once we, as a community, are convinced about the importance of the problems, the solutions are rather obvious, although not necessarily easy to implement.
The study was a comparison of general students of promise affect and mathematical students of promise affect after doing a mathematical modeling activity. Participants’ gender (n=160), in grades 7-8, were nearly equal in number (81 girls & 79 boys). After completing a Model-eliciting Activity (MEA) in groups of three, participants completed the 31-item Chamberlin Affective Instrument for Mathematical Problem Solving, hereafter referred to as CAIMPS (Chamberlin, Moore, & Parks, 2017). Using four subconstructs, it was determined that the only statistically significant difference in student affect among the groups was self-esteem and self-efficacy (SS) with the general students of promise group having a mean of 3.43 and the mathematical students of promise group having a mean of 3.76. Implications are that the difference in SS may have surfaced because of the mathematical demands of the problems that ultimately influenced participants’ ratings. Three subconstructs (Attitude Value Interest [AVI], Anxiety [ANX], and Aspiration [ASP]) may not have realized a statistically significant difference because they were not as contingent upon mathematical content knowledge as was SS. The final implication is that similar affective ratings may be an indication that MEAs are similarly suitable for use with groups containing individuals with varying talents.
Focus of the CONcreTEXT task is conceptual concreteness: systems were solicited to compute a value expressing to what extent target concepts are concrete (i.e., more or less perceptually salient) within a given context of occurrence. To these ends, we have developed a new dataset which was annotated with concreteness ratings. Interestingly, these works extend information on conceptual concreteness available in existing (non contextual) norms derived from human judgments with new knowledge from recently developed neural architectures, in much the same multidisciplinary spirit whereby the CONcreTEXT task was organized.
This study investigated the interference of Bahasa Indonesia passive voice norm on English sentence. There are many studies that investigated the interference of native language on the learning of target language. Most of the studies talked about interference in the level of lexical, grammatical, phonetic, syntactical, and many more. However, the study about interference of a norm have never been discussed before. Thus, it is important to conduct this study to give some prove that norm of languages may interfere language learning. This study involved 50 students of Tour and Travel Business Department at Sekolah Tinggi Pariwisata (STP) AMPTA Yogyakarta. The data was collected by giving students 3 sentences in Bahasa Indonesia and they had to write them in English. The sentences that the students had produced were compared to the correct one. The finding shows that most of the students� sentences were interfered by the norm of passive voice in Bahasa Indonesia. It is due to the lack of students� understanding toward the concept of passive voice norms in both of Bahasa Indonesia and English. Thus, the teacher must give clear explanation about the norm of passive voice in both of languages.
New language phenomena are driven by social and political shifts at the global level. Even though the traditional literary norm is being destroyed, these linguistic innovations fulfil a language compensatory function. Internet communication and the new speech processes found in it provoke a keen research interest and are extensively explored by linguists. Major global changes in our life (cloud-based technologies, ecology, post-truth, the problem of generations, Big Data, etc.) were bound to transform communication itself. Therefore, we see changes in genres, functional styles, texts and our traditional ideas of various forms of the Russian national language usage. The Russian Internet (Runet) reveals language potential, fulfils the compensatory function of the language filling in all the elements missing so far and language shortcomings (neologisms denoting feminine gender-specific job titles, deviant verbal forms, new structures in comparative forms of adverbs and adjectives, etc.). The speech system of the Internet communication should be considered not as a double-sided one (oral and written) but as a conceptually new digital form of language use. In the democratic environment of pluralism, tolerance and the freedom of language use, lexical and lexical-grammatical innovations, “the new vernacular”, irregular grammar and lexical collocability, as well as the direct and conscious intention to break the norm of the literary language, should be justified and deemed a manifestation of the compensatory language function. Special attention is given to the acute problem of fundamental transformations in teaching practice.
The article deals with the linguistic-poetic and lexical-semantic features of historical epics and heroes of the Nogai era, which left a fundamental trace in the political and legal development of the nomadic society of Eurasia.Nogai Orda was one of the largest structures that appeared on the territory of Kazakhstan after theweakening of the Ak Orda and the collapse of the great AltynOrda.Ten Nogai tribes brought with them a rich literary heritage at the time when they became part of theKazakh people. The poetic and rich poetry of the Nogai people developed and improved the norms ofthe oral Kazakh literary language. As well as the epic works of the Nogai era of the Kazakh, Nogai andBashkir tribes speaking the same Turkic languages, left a big mark in the lexical basis and in the composition of these languages.The lexical and semantic features of the Nogai epics are rich in didactic contents, have a deep philosophical meaning, various expressions and emotions, words of edification, aphorisms, periphrases andsyntactic parallelisms.Epic tales and historical epics of the Nogai era have a rich lexical composition, as it includes suchthematic groups as toponomic names, anthroponyms, ethnonyms, military vocabulary, zoonymies, socio-political vocabulary, theological vocabulary and abstract concepts, household vocabulary.
Przedmiotem badań jest jednojęzyczna leksykografia elektroniczna. Celem artykułujest ukazanie wpływu technik komputerowych na organizację, rozmiar, przeznaczeniei zawartość słowników. W swych badaniach autorka koncentruje się na elektronicznychbazach danych. Definiuje, czym są, oraz objaśnia, jak ich budowa i sposób organizacjizgromadzonych w nich danych wpływają na postać słowników elektronicznych. W artykulezostały poddane analizie trzy współczesne słowniki języka polskiego: Uniwersalny słownikjęzyka polskiego PWN, Wielki słownik języka polskiego PAN oraz Słownik gramatycznyjęzyka polskiego. Autorka dowodzi, że sposób organizacji i prezentacji wiedzy w omówionychdziełach umożliwia użytkownikom korzystanie z nich w sposób zaawansowany,co oznacza sprawne dotarcie do szczegółowych informacji o jednostkach leksykalnych,grupowanie ich, jak również doraźne kompilowanie „podsłowników”, spełniających określoneoczekiwania odbiorców.
Corpus-driven valency (subcategorization) lexicon automatically extracted from the Ancient Greek Treebank.
The paper reports on the implementation of an NLP pipeline for Bulgarian, developed within the spaCy framework and based on BulTreeBank as main source for training and test data. This new end-to-end pipeline aims to ensure easier technical maintenance and synchronization between modules, superior processing speeds -for use in real applications, and greater flexibility of adaptation. We discuss the challenges encountered in the implementation process and the solutions adopted, including the architecture itself, as well as its quantitative and qualitative evaluation.
The article deals with the features of linguo-pragmatical contents and ways of its expression in the text of the professional ethics code. It is described the main linguo-pragmatical categories of the ethical code and come to light modal meanings and means of their expression. The ethical code represents the specific genre establishing norms and rules of the office (professional) behavior based on the conventional system of moral ideals. The main objective of the ethical code creation is prevention of conflict situations and an illegal behavior. Standard code of ethics is considered as the coherent text representing a dichotomizing division of a speech product into the dynamic process of the language activity and result of this activity. The significant elements of pragmatics of Standard code of ethics are some categories of a modality and an appraisal as they form pragmatical contents of texts of the similar genres. Means of expression of the modality are the verbs with the meaning of need, the lexical and phraseological means being based on the modal meaning of approval / disapproval. It is allocated the values which became a basis for formation of a certain corporate picture of the world of public servants of the Russian Federation and municipal employees, for example, impartiality, conscientiousness, correctness, tolerance, and respect. It is come to the conclusion that the understanding of a lexical meaning of a word and its pragmatical opportunities are defined with an individual (corporate) picture of the world and cannot coincide with nationwide.
Abstract The purpose of Language among human beings is to communicate ideas feelings or thoughts. However, human beings are found in groups characterized by various shades of linguistic habits which control their interactions. Hence, it is not farfetched that they become creative when using language in any given context. In view of this, this paper takes a pragmatic analysis of lexical creativity in the use of Nigerian English. Data were gathered from focused discussions among Higher National Diploma Students of Federal Polytechnic, Kebbi randomly selected from two departments of the school. It became evident that Nigerian English contains some lexical items through some morphological processes like borrowings, compounding, acronyms, among others, in a bid to make themselves understood as not break the sociocultural norms that rule the Nigerian linguistic context. Hence, their speeches most often can best be understood from the perspective of pragmatics. Keywords: Pragmatics, Nigerian English, Lexical Creativity, Usage.
This article presents a series of recent studies on the mistakes and difficulties - leading to mistakes - that are manifested at every linguistic level: phonetic, morphologic, syntactic, lexical and semantic. Among the studies described here, there are dictionaries as well as other books written in a form resembling dictionaries to a greater or lesser extent, which discuss different aspects of standard language. The main section dedicated to these studies is preceded by a short overview presenting the preoccupations for the standardization of Romanian. The overview implicitly refers to aspects that facilitate understanding the relation between norm and deviation from the norm, language dynamics and the norm itself.
This paper deals with some criteria of stylistic marking of the words or their meanings as colloquial or vernacular in academic explanatory dictionaries starting with Ushakov Dictionary and ending with the latest lexicographic works, such as Large Dictionary of the Russian Language ed. by S.A. Kuznetsov, Active Dictionary of the Russian Language ed. by Ju.D. Apresjan, Academic Explanatory Dictionary of the Russian Language ed. by L.P. Krysin, et al. The parameters of the colloquial speech and vernacular, which are formulated explicitly in the prefaces of the dictionaries, are based on the speech usage and on the language norm (in particular, the use in live and mainly oral speech, as well as compliance / non-compliance with the norms of literary use). Analyses of stylistic marks “colloquial” and “vernacular” in academic explanatory dictionaries shows that, in addition to these characteristics, lexicographers were also guided by some implicit criteria, such as: 1) figurative (metaphorical, metonymic); 2) emotional-evaluative connotation; 3) presence of the word neutral lexical equivalent. The article discusses some controversial cases of stylistic marking of the words as colloquial and vernacular in academic explanatory dictionaries based on these criteria.