Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
Listening to music prompts strong emotional reactions in the listeners but relatively little research has focused on individual differences. This study addresses the role of musical preference and familiarity on emotions induced through music. A sample of 50 healthy participants (25 women) listened to 42 excerpts from the FMMS during 8 s while their autonomic and facial EMG responses were continuously recorded. Then, affective dimensions (hedonic valence, tension arousal, and energy arousal) and musical preference were rated using a 9-point scale, as well as familiarity using a 3-point scale. It was hypothesized that preferred and familiar music would be evaluated as more pleasant, energetic and less tense, and would prompt an increase of autonomic and zygomatic responses, and a decrease of corrugator activity. Results partially confirmed our hypothesis showing a strong effect of musical preference but not familiarity on emotion correlates. Specifically, musical preference predicted valence ratings, as well as HR acceleration and facial EMG activity. Overall, current findings suggested a great influence of musical preference on music-induced emotions, particularly modulating hedonic valence correlates. Our findings add evidence about the role of individual differences in the emotional processing through music and suggest the importance of considering those variables in future studies.
Humans have systematic and reliable color preferences. The dominant account of color preference is that individuals like some colors more than others due to the valence of objects that they associate with colors (Ecological Valence Theory). In support of this theory, Palmer and Schloss show that the average valence of objects associated with a color, when weighted (the WAVE), explains up to 80% of the variation in color preference for adults from the United States (US). Here we investigate whether Ecological Valence Theory can account for the color preferences of female and male adults from Saudi Arabia to test how well the theory generalizes across cultures and how well it accounts for sex differences in color preference. We also extend the investigation of EVT by investigating whether abstract concept associations as well as object associations can account for preference. Saudi adults' color preferences, color object and concept associations, and association valence ratings were collected, and the WAVE was computed and correlated with preference ratings. The WAVE accounted for no more than half of the variance in Saudi color preferences, although there was some degree of sex specificity in the relationship of the WAVE and color preference. Adding abstract concept associations did not account for more variance than object associations alone, but the number of abstract concept associations did account for a significant amount of the variance in color preference for females, but not males. The findings converge with other cross-cultural studies in suggesting that the success of EVT in accounting for color preference varies across cultures and indicates that additional factors other than color associations are likely also at play.
Discourse parsing is an essential upstream task in Natural Language Processing with strong implications for many real-world applications. Despite its widely recognized role, most recent discourse parsers (and consequently downstream tasks) still rely on small-scale human-annotated discourse treebanks, trying to infer general-purpose discourse structures from very limited data in a few narrow domains. To overcome this dire situation and allow discourse parsers to be trained on larger, more diverse and domain-independent datasets, we propose a framework to generate "silver-standard" discourse trees from distant supervision on the auxiliary task of sentiment analysis.
We present an incremental syntactic representation that consists of assigning a single discrete label to each word in a sentence, where the label is predicted using strictly incremental processing of a prefix of the sentence, and the sequence of labels for a sentence fully determines a parse tree. Our goal is to induce a syntactic representation that commits to syntactic choices only as they are incrementally revealed by the input, in contrast with standard representations that must make output choices such as attachments speculatively and later throw out conflicting analyses. Our learned representations achieve 93.72 F1 on the Penn Treebank with as few as 5 bits per word, and at 8 bits per word they achieve 94.97 F1, which is comparable with other state of the art parsing models when using the same pre-trained embeddings. We also provide an analysis of the representations learned by our system, investigating properties such as the interpretable syntactic features captured by the system and mechanisms for deferred resolution of syntactic ambiguities.
Abstract Ambiguity in communicative signals may lead to misunderstandings and thus reduce the effectiveness of communication, especially in unpredictable interactions such as between closely matched rivals or those with a weak social bond. Therefore, signals used in these circumstances should be less ambiguous, more stereotyped and more intense. To test this prediction, we measured facial movements of crested macaques (Macaca nigra) during spontaneous social interaction, using the Facial Action Coding System for macaques (MaqFACS). We used linear mixed models to assess whether facial movement intensity and variability varied according to the interaction outcome, the individuals' dominance relationship and their social bond. Movements were least intense and most variable in affiliative contexts, and more intense in interactions between individuals who were closely matched in terms of dominance rating. We found no effect of social bond strength. Our findings provide evidence for a reduction in ambiguity of facial behaviour in risky social situations but do not demonstrate any mitigating effect of social relationship quality. The results indicate that the ability to modify communicative signals may play an important role in navigating complex primate social interactions. This article is part of the theme issue ‘Cognition, communication and social bonds in primates’.
Concern over partisan resentment and hostility has increased across Western democracies. Despite growing attention to affective polarization, existing research fails to ask whether who serves in office affects mass-level interparty hostility. Drawing on scholarship on women’s behavior as elected representatives and citizens’ beliefs about women politicians, we posit the women MPs affective bonus hypothesis: all else being equal, partisans display warmer affect toward out-parties with higher proportions of women MPs. We evaluate this claim with an original dataset on women’s presence in 125 political parties in 20 Western democracies from 1996 to 2017 combined with survey data on partisans’ affective ratings of political opponents. We show that women’s representation is associated with lower levels of partisan hostility and that both men and women partisans react positively to out-party women MPs. Increasing women’s parliamentary presence could thus mitigate cross-party hostility.
Dependency parsing has become the norm for its advantages of representing syntactic information for numerous tasks of natural language processing (NLP). In Vietnamese, a challenging problem which arises in this domain is the insufficiency of the training resource. Our work presents a new method to automatically convert a Vietnamese constituency treebank into dependency trees. We designed new dependency labels for Vietnamese treebank. Furthermore, in this research, we proposed new head-percolation rules and dependency relations. The experimental results on two state-of-the-art parsers, MaltParser and MSTParser, indicated that our treebank were roughly 13% UAS and 21% LAS higher than previous works.
This paper introduces an algorithm to convert Universal Dependencies (UD) treebanks to Combinatory Categorial Grammar (CCG) treebanks.As CCG encodes almost all grammatical information into the lexicon, obtaining a high-quality CCG derivation from a dependency tree is a challenging task.Our algorithm relies on hand-crafted rules to assign categories to constituents, and a non-statistical parser to derive full CCG parses given the assigned categories.To evaluate our converted treebanks, we perform lexical, sentential, and syntactic rule coverage analysis, as well as CCG parsing experiments.Finally, we discuss how our method handles complex constructions, and propose possible future extensions.
Producing easily usable, professional-looking descriptive dictionaries on a shoestring budget in a short time span is a priority for documentation, but hard to achieve. The usual procedure for field dictionaries is to compile a target language-to-contact language lexical database (e.g. Chechen-English, in our case) and generate a contact-to-target dictionary or index from the glosses, This is economical but not always fully satisfactory. Here we describe our solutions to some common problems of field lexicography based on several years’ experience at compiling, editing, and publishing dictionaries of Chechen and Ingush, close sister languages of the Nakh-Daghestanian language family spoken in the central Caucasus. They are languages with large, literate speech communities for which dictionaries need to be sizable, attractive, and linguistically sophisticated and for which two different alphabets are needed, and we hope that our experience in trying to meet these goals will be helpful to linguists embarking on lexical documentation.
The amygdala, orbitofrontal cortex (OFC) and medial prefrontal cortex (mPFC) form a crucial part of the emotion circuit, yet their emotion induced responses and interactions have been poorly investigated with direct intracranial recordings. Such high-fidelity signals can uncover precise spectral dynamics and frequency differences in valence processing allowing novel insights on neuromodulation. Here, leveraging the unique spatio-temporal advantages of intracranial electroencephalography (iEEG) from a cohort of 35 patients with intractable epilepsy (with 71 contacts in amygdala, 31 in OFC and 43 in mPFC), we assessed the spectral dynamics and interactions between the amygdala, OFC and mPFC during an emotional picture viewing task. Task induced activity showed greater broadband gamma activity in the negative condition compared to positive condition in all the three regions. Similarly, beta activity was increased in the negative condition in the amygdala and OFC while decreased in mPFC. Furthermore, beta activity of amygdala showed significant negative association with valence ratings. Critically, model-based computational analyses revealed unidirectional connectivity from mPFC to the amygdala and bidirectional communication between OFC-amygdala and OFC-mPFC. Our findings provide direct neurophysiological evidence for a much-posited model of top-down influence of mPFC over amygdala and a bidirectional influence between OFC and the amygdala. Altogether, in a relatively large sample size with human intracranial neuronal recordings, we highlight valence-dependent spectral dynamics and dyadic coupling within the amygdala-mPFC-OFC network with implications for potential targeted neuromodulation in emotion processing.
Siyao Peng, Yang Janet Liu, Amir Zeldes. Proceedings of the 2nd Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics and the 12th International Joint Conference on Natural Language Processing (Volume 2: Short Papers). 2022.
In this study, we aim to offer linguistically motivated solutions to resolve the issues of the lack of representation of null morphemes, highly productive derivational processes, and syncretic morphemes of Turkish in the BOUN Treebank without diverging from the Universal Dependencies framework. In order to tackle these issues, new annotation conventions were introduced by splitting certain lemmas and employing the MISC (miscellaneous) tab in the UD framework to denote derivation. Representational capabilities of the re-annotated treebank were tested on a LSTM-based dependency parser and an updated version of the BoAT Tool is introduced.
This study sought to determine if hues overlayed on a video recording of a piano performance would systematically influence perception of its emotional arousal level. The hues were artificially added to a series of four short video excerpts of different performances using video editing software. Over two experiments 106 participants were sorted into 4 conditions, with each viewing different combinations of musical excerpts (two excerpts with nominally high arousal and two excerpts with nominally low arousal) and hue (red or blue) combinations. Participants rated the emotional arousal depicted by each excerpt. Results indicated that the overall arousal ratings were consistent with the nominal arousal of the selected excerpts. However, hues added to video produced no significant effect on arousal ratings, contrary to predictions. This could be due to the domination of the combined effects of other channels of information (e.g., the music and player movement) over the emotional effects of the hypothesized influence of hue on perceived performance (red expected to enhance and blue to reduce arousal of the performance). To our knowledge this is the first study to investigate the impact of these hues upon perceived arousal of music performance, and has implications for musical performers and stage lighting. Further research that investigates reactions during live performance and manipulation of a wider range of lighting hues, saturation and brightness levels, and editing techniques, is recommended to further scrutinize the veracity of the findings.
An increasing amount of research has recently focused on dimensional sentiment analysis that represents affective states as continuous numerical values on multiple dimensions, such as valence-arousal (VA) space. Compared to the categorical approach that represents affective states as distinct classes (e.g., positive and negative), the dimensional approach can provide more fine-grained (real-valued) sentiment analysis. However, dimensional sentiment resources with valence-arousal ratings are very rare, especially for the Chinese language. Therefore, this study aims to: (1) Build a Chinese valence-arousal resource called Chinese EmoBank, the first Chinese dimensional sentiment resource featuring various levels of text granularity including 5,512 single words, 2,998 multi-word phrases, 2,582 single sentences, and 2,969 multi-sentence texts. The valence-arousal ratings are annotated by crowdsourcing based on the Self-Assessment Manikin (SAM) rating scale. A corpus cleanup procedure is then performed to improve annotation quality by removing outlier ratings and improper texts. (2) Evaluate the proposed resource using different categories of classifiers such as lexicon-based, regression-based, and neural-network-based methods, and comparing their performance to a similar evaluation of an English dimensional sentiment resource.
= 192, 179) were assigned to the certain (CG) or uncertain group (UG) and presented with 100% (CG) or 50% (UG) S1-S2 congruency between visual stimuli. During the test phase, participants were presented with a new 75% S1-S2 paradigm and visual (Experiment 1) or auditory (Experiment 2) S2s. Participants were asked to rate the expected valence of upcoming S2s (expectancy ratings) or valence and arousal to S2s. In both experiments, the CG reported more extreme expectancy ratings than the UG, suggesting that experiencing previous reliable S1-S2 associations led CG participants to subsequently predict similar associations. No group differences emerged on valence and arousal ratings, which were more prominently influenced by the new 75% contingencies of the test phase rather than by previous learned contingencies. Last, comparing the two experiments, no significant group by experiment interaction was found, supporting the hypothesis of cross-modality generalization at the subjective level. Overall, our results advance knowledge about the mechanisms by which previous learned contingencies shape subjective affective experience. (PsycInfo Database Record (c) 2023 APA, all rights reserved).
Abstract This study examines the distinction between knowing the meaning of a word and experiencing the feelings associated with it. We collected affective ratings for a set of emotional and neutral English words from a group of English native speakers and a group of European Portuguese–English bilinguals. Half of the emotional words named emotions (emotion words) and the other half did not name emotions but could provoke them (emotion-laden words). Some participants were asked to focus on the meaning of words while others were asked to focus on the feeling produced by the words. Native speakers of English produced more intense affective ratings that Portuguese–English bilinguals. Such difference was larger when participants focused on their feelings than when they focused on the words’ meaning. Accordingly, such distinction should be considered in the study of bilingual affective language processing. Finally, the type of emotional word (emotion vs. emotion-laden) had only modest effects.
"Words and pictures stimuli are often used in study of perception, language, and memory. More and more studies are being done on how emotional words or pictures influence different cognitive processing. However, the emotional rating process of these stimuli has rarely been studied in young children. Especially, no study has investigated emotional rating process on pre-schoolers. This research examines how young children process emotional words and pictures stimuli. More precisely, we measured age (4, 5, and 6-years-old) and sex differences (girls and boys) in emotional valence rating of pictures and words. A corpus of 90 words and 90 pictures was selected from among the emotional databases compiled by Alario & Ferrand (1999), Bonin et al. (2003), Cannard et al. (2006) Syssau & Monnier (2009). This corpus was rated by 92 French children (28 four-years-old children, 16 girls and 12 boys; 34 five-years-old children, 14 girls and 20 boys; and 30 six-years-old children, 13 girls and 17 boys). These ratings were made using a three points emotional valence rating scale (negative, neutral, and positive) based on AEJE scale (Largy, 2018). To keep the rating task simple for the children, the scale labels were using drawings of faces. The 90 Words and 90 pictures were divided in sets of 15 stimuli. Each child rated all sets of stimuli in separate sessions. These sessions were in a random order between words and pictures stimuli sets. Good response reliability was observed in the three age groups. We assessed age differences in the valence ratings: Four-year-old children shown lower mean scores in valence rating (positive, neutral, and negative) than did five-year-old ones who shown lower mean scores in valence rating than did six-year-old ones. Despite a lack of consensus in the literature, we found sex differences in the valence ratings. Girls in each age groups shown higher mean scores in valence rating than did boys. Moreover, results shown a significant difference between pictures and words ratings. Children better rated words than pictures in each age group and sex. Besides, analyses revealed significant differences in emotional valence rating between negative, neutral, and positive words and pictures stimuli. Positive words and pictures stimuli were better rated by children than negative ones which were better rated than neutral ones. Future research will compile this corpus in a database, and it could become a worthwhile tool to control emotional verbal and visual stimuli in experimental design for children."
LeiKo is a comparable corpus of German easy-to-read news texts. This freely available resource is systematically compiled and linguistically annotated for linguistic and computational linguistic research. LeiKo consists of 216 news and newspaper texts (approx. 56,600 tokens) and their meta data structured in four subcorpora according to the websites they were published on. All texts are tokenized, lemmatized, part-of-speech tagged and dependency parsed and can be queried in ANNIS (Krause/Zeldes 2016). A core corpus of 40 texts is manually corrected.<br> <br> Version 0.9 contains only the core corpus with lemma and pos annotations and can be queried here: https://corpora.uni-hamburg.de/hzsk/de/hzsk_access/annis/leiko<br> <br> Version 1.0 comprises all 216 texts and not only lemma and pos annotations, but also syntactic annotations and metadata. The corpus is provided in the annis format, which can be directly imported into ANNIS Kickstarter. Version 1.1 is identical to version 1.0, but in addition to the annis files contains all corpus texts in the conll format. In version 1.3, some parts in the tazleicht subcorpus were added that were missing due to a mistake in the Python script written to gather the texts from the websites. In addition, the syntactic segmentation for lists was revised and the URL for each text was added to the set of meta data. Version 1.4 contains additional coreference annotations. Version 1.5 contains corrected meta data annotations as well as annotations of discourse relations according to the Penn Discourse Treebank guidelines (Webber et al. 2019) in separate csv files (currently not searchable in ANNIS).<br> <br> Further versions with additional manual annotation levels will follow. If you use the corpus, please cite: Jablotschkin, Sarah / Heike Zinsmeister (2020): „LeiKo: A corpus of easy-to-read German“. Poster presentation at the Computational Linguistics Poster Session in the course of the 42nd annual conference of the Deutsche Gesellschaft für Sprachwissenschaft (DGfS) in Hamburg. (The version of the poster you find in the download section was corrected for typos.)
The emergence of world Englishes (WE) has revolutionized the way scholars and linguists view the English language. Before Kachru introduced the concentric circles model of WE, there were only two varieties of English that were widely accepted in the world, namely, American English (AmE) and British English—both of which fall in the category of Inner Circle Englishes. Despite having its own linguistic norms, PhE is yet to be recognized and accepted in the discipline of language assessment. This chapter argues that language assessment in the Philippines must be WE paradigm-informed. The resistance to include appropriate linguistic norms is one of the issues in testing WE. Despite the fact that the variety of English being used in Philippine classrooms is PhE, the assessment tasks and standards taught are still those of AmE.
Cite the source of the dataset as: Fabrício Ferraz Gerardi, Stanislav Reichert, Carolina Aragon, Johann-Mattis List, & Tim Wientzek. (2021). TuLeD: Tupían lexical database. Max Planck Institute for Evolutionary Anthropology: Leipzig
With the advent of the WordNet (a lexical database), the field of linguistics ushered into the new realm where this structured database, based on ontology can be used for a variety of operations such as Word Sense Disambiguation (WSD), checking Semantic Similarity, performing Text Based Summarization etc. The development of various knowledge bases/lexicons like WordNet, with graphically structured information allowed many graph based approaches to be used for choosing the best sense in the Word Sense Disambiguation process. In these graph based approaches, local graph connectivity measures like Fuzzy Degree Centrality ($\mathrm{V}_{\mathrm{F}\mathrm{D}\mathrm{C}}$), Fuzzy Closeness Centrality ($\mathrm{V}_{\mathrm{F}\mathrm{C}\mathrm{C}}$), Fuzzy Betweenness Centrality ($\mathrm{V}_{\mathrm{F}\mathrm{B}\mathrm{C}}$), Fuzzy PageRank Centrality ($\mathrm{V}_{\mathrm{F}\mathrm{P}\mathrm{C}}$) and global graph connectivity measures like Fuzzy Compactness ($\mathrm{V}_{\mathrm{F}\mathrm{C}}$) and Fuzzy Graph Entropy ($\mathrm{V}_{\mathrm{F}\mathrm{G}\mathrm{E}}$) are widely used to perform WSD.Proposed and implemented here is a novel fuzzy graph connectivity measure named as Fuzzy Expected Force (FExF) to choose the correct word sense, by taking into consideration the influential as well as less influential nodes for the node under consideration in the graph formed using Fuzzy Hindi WordNet. FExF performs less computations than $\mathrm{V}_{\mathrm{F}\mathrm{G}\mathrm{E}}$ to give accurate results; thus becoming a reliable local graph connectivity measure, as overall structure of graph is not necessary for finding correct sense of the word.
For fans of popular cultural products, digitization has meant the configuration of affinity spaces online and opportunities for learning in the “digital wilds,” including incidental language learning and identity development. Through online multiparty written interaction, we explored how 15 Catalan-speaking gamers organized themselves to translate games from English into Catalan. Results indicated distributed roles, routinized translation practices, and, through translating and commenting on new members’ translation tests, the emergence of folk linguistic attitudes and beliefs regarding Spanish-Catalan interference, changes to Catalan linguistic norms, or positionings towards the group’s English-Catalan translation strategy: They favored creative expressions and what they believed were idiosyncratic language choices with symbolic power to signal what Catalan is or ought to be. The group of fan translators configured a site for metapragmatic discussion and identity development which navigates across monoglot/heteroglossic, prescriptive/descriptive/idiosyncratic conceptions of language, manifesting grassroots linguistic activism fueled by the diglossic status of the Catalan language.
The present study investigated the effect of background luminance on the self-reported valence ratings of auditory stimuli, as suggested by some earlier work. A secondary aim was to better characterise the effect of auditory valence on pupillary responses, on which the literature is inconsistent. Participants were randomly presented with sounds of different valence categories (negative, neutral, and positive) obtained from the IADS-E database. At the same time, the background luminance of the computer screen (in blue hue) was manipulated across three levels (i.e., low, medium, and high), with pupillometry confirming the expected strong effect of luminance on pupil size. Participants were asked to rate the valence of the presented sound under these different luminance levels. On a behavioural level, we found evidence for an effect of background luminance on the self-reported valence rating, with generally more positive ratings as background luminance increased. Turning to valence effects on pupil size, irrespective of background luminance, interestingly, we observed that pupils were smallest in the positive valence and the largest in negative valence condition, with neutral valence in between. In sum, the present findings provide evidence concerning a relationship between luminance perception (and hence pupil size) and self-reported valence of auditory stimuli, indicating a possible cross-modal interaction of auditory valence processing with completely task-irrelevant visual background luminance. We furthermore discuss the potential for future applications of the current findings in the clinical field.
Abstract The affective variability of Bipolar Disorder (BD) is thought to qualitatively differ from that of Borderline Personality Disorder (BPD), with changes in affect persisting for longer in BD. However, quantitative studies have not been able to confirm this distinction. It has therefore not been possible to accurately quantify how treatments like lithium influence affective variability in BD. We assessed the affective variability associated with BD and BPD as well as the effect of lithium using a novel computational model that defines two subtypes of variability: affective changes that persist (volatility) and changes that do not (noise). We hypothesized that affective volatility would be raised in the BD group, noise would be raised in the BPD group and that lithium would impact affective volatility. Daily affect ratings were prospectively collected for up to 3 years from patients with BD, BPD and non-clinical controls. In a separate experimental-medicine study, patients with BD were randomized to receive lithium or placebo, with affect ratings collected from week -2 to +4. We found a diagnostically specific pattern of affective variability. Affective volatility was raised in patients with BD whereas affective noise was raised in patients with BPD. Rather than suppressing affective variability, lithium increased the volatility of positive affect in both studies. These results provide a quantitative measure of the affective variability associated with BD and BPD. They suggest a novel mechanism of action for lithium, whereby periods of persistently low or high affect are avoided by increasing the volatility of affective responses.
The prevailing practice in the academia is to evaluate the model performance on in-domain evaluation data typically set aside from the training corpus. However, in many real world applications the data on which the model is applied may very substantially differ from the characteristics of the training data. In this paper, we focus on Finnish out-of-domain parsing by introducing a novel UD Finnish-OOD out-of-domain treebank including five very distinct data sources (web documents, clinical, online discussions, tweets, and poetry), and a total of 19,382 syntactic words in 2,122 sentences released under the Universal Dependencies framework. Together with the new treebank, we present extensive out-of-domain parsing evaluation utilizing the available section-level information from three different Finnish UD treebanks (TDT, PUD, OOD). Compared to the previously existing treebanks, the new Finnish-OOD is shown include sections more challenging for the general parser, creating an interesting evaluation setting and yielding valuable information for those applying the parser outside of its training domain.
In this paper, we investigate to which extent contextual neural language models (LMs) implicitly learn syntactic structure. More concretely, we focus on constituent structure as represented in the Penn Treebank (PTB). Using standard probing techniques based on diagnostic classifiers, we assess the accuracy of representing constituents of different categories within the neuron activations of a LM such as RoBERTa. In order to make sure that our probe focuses on syntactic knowledge and not on implicit semantic generalizations, we also experiment on a PTB version that is obtained by randomly replacing constituents with each other while keeping syntactic structure, i.e., a semantically ill-formed but syntactically well-formed version of the PTB. We find that 4 pretrained transfomer LMs obtain high performance on our probing tasks even on manipulated data, suggesting that semantic and syntactic knowledge in their representations can be separated and that constituency information is in fact learned by the LM. Moreover, we show that a complete constituency tree can be linearly separated from LM representations.
Modern Irish is a minority language lacking sufficient computational resources for the task of accurate automatic syntactic parsing of usergenerated content such as tweets. Although language technology for the Irish language has been developing in recent years, these tools tend to perform poorly on user-generated content. As with other languages, the linguistic style observed in Irish tweets differs, in terms of orthography, lexicon, and syntax, from that of standard texts more commonly used for the development of language models and parsers. We release the first Universal Dependencies treebank of Irish tweets, facilitating natural language processing of user-generated content in Irish. In this paper, we explore the differences between Irish tweets and standard Irish text, and the challenges associated with dependency parsing of Irish tweets. We describe our bootstrapping method of treebank development and report on preliminary parsing experiments.
While scholars increasingly link affective polarization to the rise of populist parties, existing empirical studies are limited to the effects of radical right parties, without considering the possible effects of leftist populist parties or of parties' varying degrees of populism. Analyzing novel survey data across eight European publics, we analyze whether citizens' affective party evaluations broadly map onto these parties' varying degrees of populism, along with their Left-Right ideologies. We scale survey respondents' party feeling thermometer evaluations and social distance ratings of rival partisans using multidimensional scaling (MDS) to estimate a two-dimensional affective partisan space for each mass public, finding that in most (though not all) publics our mappings are strongly related to the parties' varying degrees of populism, as well as to Left-Right ideology. We substantiate these conclusions via analyses regressing respondents' affective ratings against exogenous measures of the parties' Left-Right ideologies and their degrees of populism. Our findings suggest that in many European publics, populism structures citizens' affective ratings of parties (and of their supporters) to roughly the same degree as Left-Right ideology.
We analyze Recurrent Neural Network (RNN) architectures to handle the problem of Part-of-Speech (POS) Tagging. When linguistic rules are inserted ad-hoc into the decision algorithm, there is a difficulty in understanding the role of prior information and learning. The real potential of recurrent networks is demonstrated in this paper on the Italian language in a purely data-driven approach, where we can reach the state-of-the-art on the UD_Italian-ISTD (Italian Stanford Dependency Treebank) dataset in comparison to TINT. We propose a methodology for splitting words that are mapped to embedding spaces and fed to forward-backward networks.
This dataset contains all image, complexity and distinctiveness data that was used for: Han, S. J, Kelly, P., Winters, J., & Kemp, C. (2022). Simplification is not dominant in the evolution of Chinese characters. <em>Open Mind</em>. The code for this project can be found here. The file uploaded here is intended to replace the sample data folder that is available in the code repository. Our dataset includes data scraped from hanziyuan.net, as well as data from the following sources: Sun, C. C., Hendrix, P., Ma, J., & Baayen, R. H. (2018). Chinese lexical database (CLD): A large-scale lexical database for simplified Mandarin Chinese. Behavior Research Methods, 50(6), 2606–2629. Wikimedia Commons. (2021). Chinese characters decomposition. https://commons.wikimedia.org/wiki/Commons:Chinese_characters_decomposition Liu, C.-L., Yin, F., Wang, D.-H., & Wang, Q.-F. (2011). CASIA online and offline Chinese handwriting databases. In 2011 international conference on document analysis and recognition (pp. 37–41). https://doi.org/10.1109/ICDAR.2011.17 Chen, P.-C. (2020). Traditional Chinese handwriting dataset. GitHub. https://github.com/AI-FREE-Team/Traditional-Chinese-Handwriting-Dataset
With the development of deep learning, neural networks are widely used in various fields, and the improved model performance also introduces a considerable number of parameters and computations. Model quantisation is a technique that turns floating-point computing into low-specific-point computing, which can effectively reduce model computation strength, parameter size, and memory consumption but often bring a considerable loss of accuracy. This paper mainly addresses the problem where the distribution of parameters is too concentrated during quantisation aware training (QAT). In the QAT process, we use a piecewise function to statistics the parameter distributions and simulate the effect of quantisation noise in each round of training, based on the statistical results. Experimental results show that by quantising the Transformer network, we lose less precision and significantly reduce the storage cost of the model; compared with the full precision LSTM network, our model has higher accuracy under the condition of a similar storage cost. Meanwhile, compared with other quantisation methods on language modelling task, our approach is more accurate. We validated the effectiveness of our policy on the WikiText-103 and PENN Treebank datasets. The experiments show that our method extremely compresses the storage cost and maintains high model performance.
The aim of this article is to take the first steps toward the compilation of a treebank of Old English compatible with the framework of Universal Dependencies (UD). Such a treebank will comprise morphological and syntactic annotation of Old English texts adequate for cross-linguistic comparison, diachronic analysis and natural language processing. The article, therefore, engages in four tasks: (i) identifying the Old English exponents of UD lexical categories; (ii) selecting the Old English exponents of UD morphological features; (iii) finding the areas of Old English morphology that require token indexing in the UD format; and (iv) checking on the relevance of the universal set of dependency relations. The data have been extracted from ParCorOEv2, an open access annotated parallel corpus Old English-English. The main conclusions are that the annotation format calls for two additional fields (gloss and morphological relatedness) and that enhanced dependencies are required in order to account for some syntactic phenomena.
Exploration of the physiological signals associated with subjective emotional dynamics has practical significance. Previous studies have reported that the dynamics of subjective emotional valence and arousal can be assessed using facial electromyography (EMG) and electrodermal activity (EDA), respectively. However, it remains unknown whether other methods can assess emotion dynamics. To investigate this, EMG of the trapezius muscle and fingertip temperature were tested. These measures, as well as facial EMG of the corrugator supercilii and zygomatic major muscles, EDA (skin conductance level) of the palm, and continuous ratings of subjective emotional valence and arousal, were recorded while participants (n = 30) viewed emotional film clips. Intra-individual subjective–physiological associations were assessed using correlation analysis and linear and polynomial regression models. Valence ratings were linearly associated with corrugator and zygomatic EMG; however, trapezius EMG was not related, linearly or curvilinearly. Arousal ratings were linearly associated with EDA and fingertip temperature but were not linearly or curvilinearly related with trapezius EMG. These data suggest that fingertip temperature can be used to assess the dynamics of subjective emotional arousal.
BACKGROUND: The understanding of the cerebral neurobiology of anorexia nervosa (AN) with respect to state- versus trait-related abnormalities is limited. There is evidence of restitution of structural brain alterations with clinical remission. However, with regard to functional brain abnormalities, this issue has not yet been clarified. METHODS: We compared women with AN (n = 31), well-recovered female participants (REC) (n = 18) and non-patients (NP) (n = 27) cross-sectionally. Functional magnetic resonance imaging was performed to compare neural responses to food versus non-food images. Additionally, affective ratings were assessed. RESULTS: Functional responses and affective ratings did not differ between REC and NP, even when applying lenient thresholds for the comparison of neural responses. Comparing REC and AN, the latter showed lower valence and higher arousal ratings for food stimuli, and neural responses differed with lenient thresholds in an occipital region. CONCLUSIONS: The data are in line with some previous findings and suggest restitution of cerebral function with clinical recovery. Furthermore, affective ratings did not differ from NP. These results need to be verified in intra-individual longitudinal studies.
The aim of the study is to identify the hierarchical organization of the GARBAGE cluster, one of the segments of the metaphorical model of dirt in the political discourse. The article shows that the GARBAGE cluster is singled out on the basis of the feature of cognitive deviation from the norm and includes useless and unnecessary remnants of human life, which, projected onto the sphere of politics, serve for a negative moral and ethical assessment of political actors. The scientific novelty of the work lies in the approach to the study of the metaphorical cognitive segment GARBAGE from the point of view of the conceptual and semantic continuity of the metaphor “politics is dirt”. As a result, a synonymic-gradual group of lexemes has been identified and characterized, which verbalizes the nuclear cognitive meanings of the concept GARBAGE; lexical representatives of specific varieties of garbage have been determined, their axiological potential has been described.
The Magyarization of Romanian toponymy in historical Transylvania was achieved in three different ways: 1) adapting the onomastic material to the Hungarian orthographic and phonetic system; 2) translating the toponymic items; 3) adopting the specific Hungarian morphosyntactic rules. The Magyarization of microtoponymy did not have repercussions on the morphosyntactic and lexical levels; the adoption of Hungarian orthography ensured only the formal assimilation of the toponyms. The Hungarian orthographic principles and norms, used inconsistently, reflect numerous oscillating contexts in which the sounds ă, î, u have as graphic correspondents both labial and non-labial vowels. The Magyarization of Romanian toponymy in historical Transylvania did not obscure specific dialectal features, which highlight important information on the age and strata of populations and the relationships among them
This article performs a comparative analysis of diachronic translations of Agatha Christie’s novel Ten Little Niggers/Ten Little Indians/And Then There Were None (intralingual translations into American English and interlingual translations into French, German and Russian) in terms of adequacy of literary translation and following the norms of inclusive language. The paper found that the intralingual translation into American English is a culturally determined adaptation of the original text excluding culture-specific concepts perceived differently in the American culture. While classical interlingual translations of the novel into German, French and Russian follow the norms of adequacy and equivalence, modern interlingual translations reflect the desire of translators to make various lexical and lexico-grammatical transformations in order to avoid invective vocabulary, which often leads to the loss of text-forming dominants of the novel (the modern translation into French is such an example). Further, the importance of preserving the intertext, which reflects the author’s artistic intention, is demonstrated. The symbolic dominants of the novel are considered here in their historical and sociocultural context. It is concluded that these dominants are closely related to the collective ideas about racial differences in the second half of the 19th and early 20th centuries. Moreover, it is shown that following new language norms in translation should not violate the unity of a literary work; hence the importance of conveying the author’s intention and the historical and sociocultural context of the work. Finally, recommendations are given to translators as to finding a balance between the requirements of political correctness and adequacy of translation.
The importance of criminal law protection of honor and reputation and imprecise legal conceptual determination of insult, as a basic and general criminal offense against honor and reputation, pointed to the need to determine in theory and court practice the parameters that will help the courts when deciding whether in the particular case the criminal offense of insult exists or not. On this occasion, an objective criterion is applied, according to which an insulting statement is assessed from the aspect of existing customary, moral, and other norms in specific time and space. The variability of the concept of honor, which often changes its content and scope, also creates the need for language analysis of the degrading statements, which can sometimes be helpful in assessing whether in a particular case there is a criminal offense of insult or not. Connecting law with linguistics provides an interdisciplinary overview of the relationship of language, style, and composition of legal documents and their conditionality by specifics of individual fields of law. The attitude to the use of language which exists in legal theory and practice is significant and worth studying, and in the focus of interdisciplinary contribution to this issue, there is a criminal law overview of the use of language tools in criminal offenses against honor and reputation. The legal part of the analysis was used, above all, the dogmatic method in order to determine the true meaning of the analyzed norms, and the normative method as a method of studying the social function of the norms. The historical method was used to show the criminal protection of honor and reputation in various historical periods, which was accompanied by the use of the sociological method to explain social factors of occurrence and development of the phenomena. In order to assess existing normative solutions, the axiological method was used. The language analysis of the criminal offense of insult was performed using a descriptive method, and the content analysis was used as a research technique. For the purposes of research, a special sample was formed, which consists of thirty judgments for the criminal offense of insult. In the corpus of the court judgments of Kragujevac courts, the repertoire of lexical assets used for the purpose of injury to honor and reputation was separated. Sublings of the mother, pejoratives, vulgarism, metaphorical nominations with negative connotation, and lexemes marked by the bearer of socially unacceptable traits and ethnicities used with derogatory meaning are among the most common funds. Public insults in writing, in the form of comments on social networks or forums, have greater weight than orally imposed insults in the presence of several faces.
Введение. Синтез искусств в культуре становится на рубеже XIX–XX столетий одной из доминирующих идей. Эта направленность в полной мере способствует раскрытию многочисленных талантов такого яркого представителя эпохи Серебряного века, как М. А. Волошин.Целью данной статьи является анализ лингвистических и художественных особенностей текстового материала, сопровождающего акварели М. Волошина.Материал и методы. В статье приводятся данные анализа лирических зарисовок разных лет, служащих сопровождением акварелей М. Волошина. Внимание к данному жанру обусловлено его несомненной значимостью для определения особенностей творческой манеры поэта и художника, а также для понимания его мировосприятия.В статье использованы методы семантико-стилистического, контекстологического, мотивного анализа, позволяющие раскрыть специфику поэтической картины мира автора, отраженную в надписях к акварелям М. А. Волошина.Результаты и обсуждение. Киммерия занимает особое место в творчестве М. Волошина – поэта, художника, переводчика, искусствоведа, мыслителя. Конгениальность М. Волошина как мастера кисти и слова нашла свое отражение в его надписях к акварелям. Данные лирические миниатюры – отдельный жанр, восходящий к античности, который роднит творчество поэта с искусством Востока.Образный строй поэтических миниатюр М. Волошина, организующий их смысловое пространство, включает в себя реалии земные (камень, вода) и небесные (облака, луна, солнце) и обнаруживает при их детальном рассмотрении синкретизм «земного» и «небесного». Цветовая картина мира, представленная многообразием цветообразов, в сочетании со звуковым оформлением передает синестезию авторского мировосприятия. Языковой и образный строй поэтических зарисовок М. Волошина отличается богатством и разнообразием: автор использует многочисленные сравнения, метафоры, эпитеты, оксюморонные сочетания, отступления от грамматических норм, позволяющие передать особенности творческой манеры мастера-творца.Заключение. Рассмотрение лингвистических и художественных особенностей лирических миниатюр М. Волошина позволило выявить их основные черты: метафоричность, синкретизм и синестетичность при создании образов, эмотивный и прагматический потенциал цветовой символики – и сделать вывод о своеобразии поэтической картины мира автора. Introduction. At the turn of the 19th and 20th centuries, the synthesis of arts in culture becomes one of the dominant ideas. This orientation fully contributes to the disclosure of the many talents of such a brilliant representative of the Silver Age as M. A.Voloshin. The purpose of this article is to analyze the linguistic and artistic features of the text material accompanying M. Voloshin’s watercolors. Material and methods. The article presents the data of the analysis of lyrical sketches of different years, serving as an accompaniment to M. Voloshin’s watercolors. Attention to this genre is due to its undoubted importance for determining the characteristics of the creative manner of the poet and artist, understanding his worldview. The article uses the methods of semantic-stylistic, contextological, motivational analysis, allowing to reveal the specifics of the author’s poetic picture of the world, reflected in the inscriptions on the watercolors of M. A. Voloshin. Results and discussion. Cimmeria occupies a special place in the work of M. Voloshin – a poet, artist, translator, art critic, thinker. The congeniality of M. Voloshin as a master of brush and word is reflected in his inscriptions for watercolors. These lyrical miniatures are a separate genre dating back to antiquity, which makes the poet’s work related to the art of the East. The figurative structure of M. Voloshin’s poetic miniatures, organizing their semantic space, includes earthly (stone, water) and heavenly (clouds, moon, sun) realities and reveals, when examined in detail, the syncretism of “earthly” and “heavenly”. The color picture of the world, represented by a variety of color images, in combination with sound design, conveys the synesthesia of the author’s perception of the world. The linguistic and figurative structure of M. Voloshin’s poetic sketches is rich and diverse: the author uses numerous comparisons, metaphors, epithets, oxymoric combinations, deviations from grammatical norms, which make it possible to convey the peculiarities of the creative manner of the master-creator. Conclusion. Consideration of the linguistic and artistic features of M. Voloshin’s lyrical miniatures made it possible to identify their main features: metaphoricity, syncretism and synestheticism in creating images, emotive and pragmatic potential of color symbolism – and to draw a conclusion about the originality of the author’s poetic picture of the world.
, 2020) has claimed there is no strong evidence that multi-letter graphemes are used in reading tasks with proficient adult readers, with most studies being statistically weak or having confounds in the stimuli used. Here, I used Monte Carlo simulation with data from reading mega-studies to examine the extent to which the number of multi-letter graphemes matters in words when letter length is held constant. This was done by simulating thousands of experiments using different sets of items for each of a small number of comparisons (e.g., words with only single-letter graphemes versus words with one multi-letter grapheme). The results showed that words with two multi-letter graphemes tended to cause slower reaction times than words with one or no multi-letter graphemes, with effects found in both naming and lexical decision tasks. Interestingly, when words with no multi-letter graphemes were compared with words with one multi-letter grapheme, the differences were much weaker. Simulations of naming results using two computer models, the connectionist dual-process (CDP) model and the dual-route cascaded (DRC) model, showed only CDP predicted this pattern. Since CDP learns simple associations between graphemes and phonemes whereas DRC uses a set of grapheme-phoneme rules, this suggests that the results may have been caused by simple associations between spelling and sound being relatively easy to learn with words with one compared with two multi-letter graphemes. More generally, the results suggest that graphemes are used when reading, but they often produce relatively weak effects and thus differences in some studies may not have been found due to a lack of power.
The purpose of the article is to study the structural (syntactic) aspect of pleonastic formations, build a structural concept and thereby prove their semantic significance, as well as informative interpretation. The object of analysis is pleonastic expressions and their pragmatic functions based on the material of publicist and artistic discourses. The subject of analysis is the comparison of the transaction (transmission) of the author’s and the “communicative isolation” of the meaningful lexical unit – pleonasm. The result of the study is the recognition of the normality of the use of this phenomenon, but subject to certain conditions: only models with homogeneous components with the meaning of a person, phenomenon, quality, state, basically, complement, clarify the conceptual content of the expression. The semantics of the object or subject is redundant in models with predicative units. These constructions are usually found in colloquial and publicistic styles, which does not allow them to be attributed to the norm of the modern Russian literary language. Structural models with a subordinate relationship with the meaning of an object, attribute, action (agreement) are also considered superfluous.
The existence of language varieties has a considerable impact on communication. They influence the interaction between language users from various centres due to the number of linguistic differences observed on the level of phonetics, spelling, grammar, lexis, and pragmatics. On the one hand, pluricentric languages connect people from various centres by using the “same” language, and on the other hand, they separate them by developing national norms. This article aims to demonstrate the importance of teaching language varieties in foreign language classes because the knowledge of national norms of pluricentric languages is essential in communication with people from various centres. Both English and German are pluricentric languages. Advanced language users should be aware of the differences between language varieties and be able to use the appropriate variety according to the communicative situation. The research undertaken in this article is meant to verify the undergraduate students’ knowledge of English and German varieties, emphasising terminology used in everyday life and their abilities to communicate in languages other than English or German.
This work categorizes Japanese Sign Language (JSL) toponyms, or place names, and examines factors that potentially affect their structure. Exonyms, influenced by the source Japanese name, and endonyms, independent JSL names, contrast structurally in that exonyms tend to emerge as compounds while endonyms conform more closely to canonical monomorphemic JSL lexemes. Based on the contrast, one would expect the structurally more favorable, less marked, endonymic outputs to emerge as the norm; however, structurally inefficient exonymic forms do so. This study analyzes a collection of over 900 names from the 2009 Japan National Federation of the Deaf (JFD) National Toponym Sign Language Map to determine structural influences on JSL toponym outputs. A spreadsheet tracked morpheme counts and relationships between source name characters and their outputs to find categorical distributions. This study finds that JSL toponyms tend to disproportionally borrow Japanese morphemes associated with JSL character signs that map to the source characters. A character sign is a sign isomorphic in some respect to a Chinese-origin orthography. Favored indexation to such source morphemes demonstrates that source name familiarity can support the persistent spread of exonymic toponyms.