Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
Abstract Stress, anxiety, and depressive symptoms can be reduced by listening to music, but the underlying mechanisms remain unclear. To address this gap, we measured brain connectivity while participants listened to songs of different genres: ambient, pop, and metal. Additionally, affective ratings were obtained while participants ( n = 30) listened to the six different songs, and subjective ratings of state anxiety were solicited at the terminus of each song. Electroencephalography (EEG) connectivity combining weighted Phase Lag Index and graph theory was utilised to document brain activity during listening. Repeated-measures ANOVA indicated that listening to more pleasant and less arousing songs was associated with lower self-reported state anxiety levels than songs rated unpleasant and highly arousing. Of interest, EEG alpha connectivity differed across two ambient songs, particularly in the frontal lobes, despite being from the same genre and rated as highly pleasant and low in arousal. We also observed a sex effect on EEG results, where female participants ( n = 18) displayed stronger connectivity than male participants ( n = 12). Combined, these results suggest that ambient songs reduce state anxiety but have divergent brain responses, possibly reflecting the complex nature of music listening, including sensory processing, emotion and cognition.
Any approach to classifying the different types of dictionaries may have absolute or relative validity, as long as the classification is the result of a rigorous execution of the criterion or criteria underpinning the method chosen. If we adhere to, for example, the traditional distinction (based on a semiotic model) of linguistic lexicography/encyclopaedic lexicography, the first pair of prototypes would be linguistic dictionary/encyclopaedia. If, on the other hand, we embrace a strictly linguistic model, the initial prototypes would be monolingual dictionary/bilingual dictionary (thus, multilingual). Applying a strictly material criterion (format), determined by the application of new technologies to the creation of dictionaries, the distinction would be printed lexicography/digital lexicography, with the corresponding duality being printed dictionary/digital dictionary. The typology of Spanish dictionaries that we propose is based on the principle of immanence applied to the general monolingual dictionary. We recognize, however, that the lexicography of Europe’s modern languages has its roots in bilingual lexicography. Therefore, we start from the basis of universal classifiers, but assign pre-eminence to the identity factors that shape the lexicographic history of any language. The macrostructure of the dictionary (nomenclature: entry) allows us to justify and classify the different types of paradigmatic dictionaries. The microstructure of the lexicographical article gives rise to the grouping of dictionaries depending on each of their elements. In accordance with the nature of the general dictionary: normative dictionary (academic and pedagogical)/descriptive dictionary (manual, basic and fundamental). If we focus upon usage labels and variations, the groupings are: temporal (etymological, historical), geolinguistic (regional), socio-cultural (specialized, jargon), diaphasic (formal/ colloquial) and semantic changes (ambiguous, humor). If, instead, we concentrate on quotes, or examples, we justify the long list of syntagmatic dictionaries. The typology of the bilingual (multilingual) dictionary is largely determined by the preeminence of the monolingual dictionary, and also by the specific characteristics of its microstructure (semi-bilingual). The digital dictionary or lexical and linguistic database addresses the needs of users through the atomistic classification generated by the wide variety of print dictionaries.
The paper traces the dynamics of the interpretation of the grammatical nature of the vocative in Ukrainian grammars from the 16th century until the present. The subject of the analysis is the content and presentation of this category in two sections of Ukrainian grammar books: (1) morphological, which clarifies the status of the vocative in the inflectional paradigm of the noun, and (2) syntactic, in which the means of expressing address are characterized. Based on the findings of the research, various trends in the description of the vocative in different historical periods have been identified, in particular: (1) until the beginning of the 20th, it was unequivocally qualified as an equal member of the inflectional paradigm of the noun, equal to other cases; (2) from the beginning of the 20th century to 1933 was a period of competition between two theories (the vocative is a case the same as others or the vocative is not a true case, but a “special” form in the inflectional paradigm of the noun); (3) the canonization of the “fake case” status theory; (4) from 1991 to the present there has been an unanimity of authors in qualifying the vocative as a case. Comparing the stages of fundamental changes in the scientific definition of the vocative in grammars with defining events in the history of Ukraine provides the basis for discussions about the influence of socio–political factors on the representation of linguistic theories and the codification of the linguistic norms.
The Wall Street Journal section of the Penn Treebank has been the de-facto standard for evaluating POS taggers for a long time, and accuracies over 97\% have been reported. However, less is known about out-of-domain tagger performance, especially with fine-grained label sets. Using data from Elder Scrolls Fandom, a wiki about the \textit{Elder Scrolls} video game universe, we create a modest dataset for qualitatively evaluating the cross-domain performance of two POS taggers: the Stanford tagger (Toutanova et al. 2003) and Bilty (Plank et al. 2016), both trained on WSJ. Our analyses show that performance on tokens seen during training is almost as good as in-domain performance, but accuracy on unknown tokens decreases from 90.37% to 78.37% (Stanford) and 87.84\% to 80.41\% (Bilty) across domains. Both taggers struggle with proper nouns and inconsistent capitalization.
Numerous studies have been conducted on the interpretation and translation of English terms into other languages. The purpose of this study was to identify the adequate Indonesian equivalent terminology for hotel amenities, services, and facilities applied in English and the strategies utilized by both domestic and international hotel guests in understanding the equivalent terms in their native language. Qualitative research methodology was used. The subjects included 10 domestic guests from a 5-star hotel, 10 domestic guests from a 4-star hotel, 5 international guests from a 3-star hotel, and 2 hotel staff from a 5-star hotel, 3 staff from a 4-star hotel, and 1 staff from a 3-star hotel. The findings demonstrated that some of the English terms commonly used in hotels had Indonesian equivalents, and some did not. The international guests strategies were: 1) searching in an online dictionary or a Google search; 2) asking people they met nearby immediately; and 3) guessing the meaning. Domestic guests’ strategies included: (a) asking other guests or hotel staff for clarification; and (b) guessing the meaning. Future research should overcome the limitations of this study, considering translations and linguistic norms training strategies.
The rapid development of information technology has made the amount of information in massive texts far exceed human intuitive cognition, and dependency parsing can effectively deal with information overload. In the background of domain specialization, the migration and application of syntactic treebanks and the speed improvement in syntactic analysis models become the key to the efficiency of syntactic analysis. To realize domain migration of syntactic tree library and improve the speed of text parsing, this paper proposes a novel approach-the Double-Array Trie and Multi-threading (DAT-MT) accelerated graph fusion dependency parsing model. It effectively combines the specialized syntactic features from small-scale professional field corpus with the generalized syntactic features from large-scale news corpus, which improves the accuracy of syntactic relation recognition. Aiming at the problem of high space and time complexity brought by the graph fusion model, the DAT-MT method is proposed. It realizes the rapid mapping of massive Chinese character features to the model's prior parameters and the parallel processing of calculation, thereby improving the parsing speed. The experimental results show that the unlabeled attachment score (UAS) and the labeled attachment score (LAS) of the model are improved by 13.34% and 14.82% compared with the model with only the professional field corpus and improved by 3.14% and 3.40% compared with the model only with news corpus; both indicators are better than DDParser and LTP 4 methods based on deep learning. Additionally, the method in this paper achieves a speedup of about 3.7 times compared to the method with a red-black tree index and a single thread. Efficient and accurate syntactic analysis methods will benefit the real-time processing of massive texts in professional fields, such as multi-dimensional semantic correlation, professional feature extraction, and domain knowledge graph construction.
Introduction A study was conducted to investigate if an individual’s trust in law enforcement affects their perception of the emotional facial expressions displayed by police officers. Methods The study invited 77 participants to rate the valence of 360 face images. Images featured individuals without headgear (condition 1), or with a baseball cap (condition 2) or police hat (condition 3) digitally added to the original photograph. The images were balanced across sex, race/ethnicity (Asian, African American, Latine, and Caucasian), and facial expression (Happy, Neutral, and Angry). After rating the facial expressions, respondents completed a survey about their attitudes toward the police. Results The results showed that, on average, valence ratings for “Angry” faces were similar across all experimental conditions. However, a closer examination revealed that faces with police hats were perceived as angrier compared to the control conditions (those with no hat and those with a baseball cap) by individuals who held negative views of the police. Conversely, participants with positive attitudes toward the police perceived faces with police hats as less angry compared to the control condition. This correlation was highly significant for angry faces ( p < 0.01), and stronger in response to male faces compared to female faces but was not significant for neutral or happy faces. Discussion The study emphasizes the substantial role of attitudes in shaping social perception, particularly within the context of law enforcement.
- output-{ciep,treebanks}-full.csv: frequency and entropy for all the categories, using four types of combinations of layers;<br> - plots.R: R script to draw plots from the output files;<br> - readReport-{CIEP+,treebanks}.R: R script to extract frequency and compute entropy from the report files (not included);<br> - ud-wordorder.py: Python script to extract word order pairs from conllu files and write them in report files. Unfortunately, I cannot include the report files, as CIEP+ is protected by copyright; the analysis can be however replicated with respect to the UD Treebanks.
The goal of this contribution is to present The Digital Rosetta Stone, which is a project developed at Leipzig University by the Chair of Digital Humanities and the Egyptological Institute/Egyptian Museum Georg Steindorff in collaboration with the British Museum and the Digital Epigraphy and Archaeology Project at the University of Florida. The aims of the project are to produce a collaborative digital edition of the Rosetta Stone, address standardization and customization issues for the scholarly community, create data that can be used by students to understand the language and content of the document, and produce a high-resolution 3D model of the stone. First, the three versions of the text were transcribed and encoded in XML according to the EpiDoc guidelines. Next, the versions were aligned with the Ugarit iAligner tool that supports the alignment of ancient texts with modern languages, such as English and German. All three texts were then parsed syntactically and morphologically through Treebank annotation. Finally, the project explored new 3D-digitization techniques of the Rosetta Stone in the British Museum in order to enhance traditional archaeological methods and facilitate the study of the artifact. The results of this work were used in different courses in Digital Humanities, Digital Philology, and Egyptology.
Embodied cognition research identifies mechanisms by which our cognitive activity is connected to body experiences. This approach encompasses not only experimental manipulations but also the quantification of variables related to group and individual differences, i.e., participant-related variables. Moreover, stimuli-related characteristics, such as sensorimotor word ratings, can either be used for the selection of experimental materials or can be the main output of a study themselves. This quantitative information about individuals or stimuli can be collected through non-experimental methods, such as questionnaires and cognitive tests. This chapter gives an overview of questionnaires and cognitive tests often used in embodied cognition research. A questionnaire is a list of questions asking participants to provide information on certain aspects, such as their sociodemographic or medical status. A test is a series of tasks which participants perform for further evaluation by researchers, such as tests of mathematical ability, reading speed, or counting direction. Rating studies collect subjective evaluations of various parameters, typically for large sets of items. The present chapter is divided into two main sections: Participant-related variables and stimuli-related characteristics. We present examples from cognitive linguistics, psycholinguistics, psychophysics, as well as from research on numerical cognition, peripersonal space, and attitudes towards social robots.
We present an approach for assessing how multilingual large language models (LLMs) learn syntax in terms of multi-formalism syntactic structures. We aim to recover constituent and dependency structures by casting parsing as sequence labeling. To do so, we select a few LLMs and study them on 13 diverse UD treebanks for dependency parsing and 10 treebanks for constituent parsing. Our results show that: (i) the framework is consistent across encodings, (ii) pre-trained word vectors do not favor constituency representations of syntax over dependencies, (iii) sub-word tokenization is needed to represent syntax, in contrast to character-based models, and (iv) occurrence of a language in the pretraining data is more important than the amount of task data when recovering syntax from the word vectors.
The purpose of the article is to consider the morphological peculiarities of the system of the nouns in New Bulgarian translation of the “Catechismos” written by Theodore the Studite, which is a part of the manuscripts no. 1/154 kept in Odessa National Scientific Library. The subject of the research is the morphological specifics of nouns in Odessa copy of the “Catechismos” dating from the 18th. The morphological peculiarities of nouns is considered in the context of the formation of a linguistic norm, which allowed the combination with different intensity of linguistic means of several language systems functioning at the time (traditional Middle Bulgarian written language, Church Slavonic Eastern recensions, and vernacular language form). The analysis proposed in this paper presents the extensive system of cases, which does not reflect the real vernacular Bulgarian speech in the 18th; the specifics of the functioning of the gramemes of the case paradigm of masculine, feminine and neuter nouns in the singular and plural forms is analyzed. Usage case endings mistakes, which indicates their artificial nature, are considered. The lack of article of nominal parts of speech is noted; the predominance of compound declension form of the adjectives and participles over short forms is revealed; the relatively high frequency of use of active present participles is registered. The results of the study make it possible to outline some probable factors that determine the writer’s preference for using the linguistic tools of the so-called “bookish”, “traditional”, “archaic” writing systems. An another reason which to some extent explains he usage of case inflections in the text of this relatively late stage of the historical development of the Bulgarian language might be the use of East Slavic copies of the Studite’s sermons by the scriber. Еhe comparison of “Catechismos” copies of South and East Slavic origin is necessary for verification of this assumption, in which we see prospects for further research.
An analysis of robots (simulators) in education is provided. Promising directions for their development are highlighted, such as realism, interactivity, adaptation and personalization. The features of using simulators in dentistry are considered. The main disadvantages of existing simulators in dentistry have been identified, namely the lack of a communicative component and imitation of patient behavior. The anthropomorphic dental simulator is based on the Robo-C robot, which is a unique combination of advanced technologies and human facial expressions, which allows it to communicate with people, reproduce movements of different parts of the body and express emotions. As dental components, the following components were created and implemented into the Robo-C control system: a Smart jaw, including cameras and a temperature sensor, and a Smart tooth, including a pressure sensor. The Robo-C control system has been upgraded taking into account the Smart jaw and Smart tooth, which made it possible to connect the dental treatment process with the robot’s servos through its linguistic base. The process of analyzing data obtained from Smart jaw cameras using a neural network is described. A two-stage classification scheme for dental defects has been proposed and its effectiveness has been proven. The linguistic base contains a set of rules with the help of which devices (microphone, speakers, servos, Smart jaw, Smart tooth) interact with each other. An example of compiling a linguistic database rule is given. The linguistic base, Smart-jaw and Smart-tooth are configured for one of four cases: caries treatment, tooth preparation for a crown, tooth extraction, endodontic treatment. Treatment quality control is carried out using a comprehensive assessment of communication interaction with the robot and analysis of Smart-jaw and Smart-tooth data. An example of work in one of the cases is given. The anthropomorphic dental simulator presented in the article allows the use of new technologies in the training of dentists, as well as the simulation of various dental procedures, which will significantly improve the practical preparation of students for working with patients.
In this paper, we propose a method for removing linguistic information from speech for the purpose of isolating paralinguistic indicators of affect. The immediate utility of this method lies in clinical tests of sensitivity to vocal affect that are not confounded by language, which is impaired in a variety of clinical populations. The method is based on simultaneous recordings of speech audio and electroglotto-graphic (EGG) signals. The speech audio signal is used to estimate the average vocal tract filter response and amplitude envelop. The EGG signal supplies a direct correlate of voice source activity that is mostly independent of phonetic articulation. These signals are used to create a third signal designed to capture as much paralinguistic information from the vocal production system as possible-maximizing the retention of bioacoustic cues to affect-while eliminating phonetic cues to verbal meaning. To evaluate the success of this method, we studied the perception of corresponding speech audio and transformed EGG signals in an affect rating experiment with online listeners. The results show a high degree of similarity in the perceived affect of matched signals, indicating that our method is effective.
BACKGROUND: Cognitive behavioral therapy (CBT) is a moderately efficacious treatment for hoarding disorder (HD), with most individuals remaining symptomatic after treatment. The Joining Forces Trial will evaluate whether 10 weeks of in-home decluttering can significantly augment the outcomes of group CBT. METHODS: A randomized controlled trial of in-home decluttering augmentation of group CBT for HD. Adult participants with HD (N = 90) will receive 12 weeks of protocol-based group CBT for HD. After group CBT, participants will be randomized to either 10 weeks of in-home decluttering led by a social services team or a waitlist. The primary endpoint is 10 weeks post-randomization. The primary outcome measures are the self-reported Saving Inventory-Revised and the blind assessor-rated Clutter Image Rating. Participants on the waitlist will cross over to receive the in-home decluttering intervention after the primary endpoint. Data will be analyzed according to intention-to-treat principles. We will also evaluate the cost-effectiveness of this intervention from both healthcare and societal perspectives. DISCUSSION: HD is challenging to treat with conventional psychological treatments. We hypothesize that in-home decluttering sessions carried out by personnel in social services will be an efficacious and cost-effective augmentation strategy of group CBT for HD. Recruitment started in January 2021, and the final participant is expected to reach the primary endpoint in December 2024. TRAIL REGISTRATION: ClinicalTrials.gov NCT04712474. Registered on 15 January 2021.
The principle of DEPENDENCY LENGTH MINIMIZATION, which seeks to keep syntactically related words close in a sentence, is thought to universally shape the structure of human languages for effective communication. However, the extent to which dependency length minimization is applied in human language systems is not yet fully understood. Preverbally, the placement of long-before-short constituents and postverbally, short-before-long constituents are known to minimize overall dependency length of a sentence. In this study, we test the hypothesis that placing only the shortest preverbal constituent next to the main-verb explains word order preferences in Hindi (a SOV language) as opposed to the global minimization of dependency length. We characterize this approach as a least-effort strategy because it is a cost-effective way to shorten all dependencies between the verb and its preverbal dependencies. As such, this approach is consistent with the bounded-rationality perspective according to which decision making is governed by "fast but frugal" heuristics rather than by a search for optimal solutions. Consistent with this idea, our results indicate that actual corpus sentences in the Hindi-Urdu Treebank corpus are better explained by the least effort strategy than by global minimization of dependency lengths. Additionally, for the task of distinguishing corpus sentences from counterfactual variants, we find that the dependency length and constituent length of the constituent closest to the main verb are much better predictors of whether a sentence appeared in the corpus than total dependency length. Overall, our findings suggest that cognitive resource constraints play a crucial role in shaping natural languages.
Nowadays, tree-structured deep learning classifier models have been widely used in different applications to ensure effective feature representation and learning. Amongst, dimensional sentiment analysis is the most interactive research field, which intends to identify continuous numerical values in the valence-arousal (VA) space. To achieve this, a tree-structured regional convolutional neural network with long short-term memory (T-CNN-LSTM) model was developed, which predicts the VA ratings of the texts for sentiment analysis. In contrast, the effect of a low prediction rate and difficulty of feature learning in a small number of class samples was not analyzed. Hence, this manuscript proposes an adversarial T-CNN-LSTM (A-T-CNN-LSTM) model for predicting the VA to achieve more fine-grained sentiment analysis. This model develops a semantic-enabled frequency-aware generative adversarial network (SFGAN) to produce more adversarial samples using the generator network and decrease the spectral data loss of the discriminator. It embeds the frequency-aware categorizer (FAC) into the discriminator to determine the input veracity in the spatial and spectral domains. Besides, semantic restricted sampling is employed in SFGAN for synthesizing the image subject to a semantic mask. Further, the created samples are classified by the T-CNN-LSTM for predicting the VA scores of sentences. Finally, the experimental results exhibit that the A-T-CNN-LSTM on stanford sentiment Treebank (SST) and CIFAR-10 databases achieves 90.12% and 91% accuracy than the other tree-structured CNNs.
Automated Multiple-Choice Question (MCQ) generation is a rapidly growing field in Natural Language Processing (NLP) that aims to assist educators and trainers in creating high-quality, efficient, and effective assessment materials. This is achieved by analyzing large amounts of textual data, such as educational content, and identifying key concepts and relationships between them. The process of generating MCQs automatically can be broken down into several steps, including text summarization, keyword extraction, distractor generation, and sentence mapping. A new approach is proposed for Text summarization is based on NLP models like XLNet and YAKE is used for keyword extraction. Distractor generation is done using lexical databases such as ConceptNet and WordNet. Sentence mapping is used to identify the main concepts and relationships within a text, which can then be used to formulate questions and options for the multiple-choice questions. The output is a set of MCQs that are semantically related to the input text and can be used for educational and training purposes.
Primary visual cortex (V1) is generally thought of as a low-level sensory area that primarily processes basic visual features. Although there is evidence for multisensory effects on its activity, these are typically found for the processing of simple sounds and their properties, for example spatially or temporally-congruent simple sounds. However, in congenitally blind individuals, V1 is involved in language processing, with no evidence of major changes in anatomical connectivity that could explain this seemingly drastic functional change. This is at odds with current accounts of neural plasticity, which emphasize the role of connectivity and conserved function in determining a neural tissue's role even after atypical early experiences. To reconcile what appears to be unprecedented functional reorganization with known accounts of plasticity limitations, we tested whether V1's multisensory roles include responses to spoken language in sighted individuals. Using fMRI, we found that V1 in normally sighted individuals was indeed activated by comprehensible spoken sentences as compared to an incomprehensible reversed speech control condition, and more strongly so in the left compared to the right hemisphere. Activation in V1 for language was also significant and comparable for abstract and concrete words, suggesting it was not driven by visual imagery. Last, this activation did not stem from increased attention to the auditory onset of words, nor was it correlated with attentional arousal ratings, making general attention accounts an unlikely explanation. Together these findings suggest that V1 responds to spoken language even in sighted individuals, reflecting the binding of multisensory high-level signals, potentially to predict visual input. This capability might be the basis for the strong V1 language activation observed in people born blind, re-affirming the notion that plasticity is guided by pre-existing connectivity and abilities in the typically developed brain.
Abstract Background and Objective Dyspnoea is a debilitating symptom in individuals with chronic obstructive pulmonary disease (COPD) and a range of other chronic cardiopulmonary diseases and is often associated with anxiety and depression. The present study examined the effect of visually‐induced mood shifts on exertional dyspnoea in individuals with COPD. Methods Following familiarization, 20 participants with mild to severe COPD (age 57–79 years) attended three experimental sessions on separate days, performing two 5‐min treadmill exercise tests separated by a 30‐min interval on each day. During each exercise test, participants viewed either a positive, negative or neutral set of images sourced from the International Affective Picture System (IAPS) and rated dyspnoea or leg fatigue (0–10). Heart rate (HR) and peripheral oxygen saturation (SpO 2 ) were measured at 1‐min intervals during each test. Mood valence ratings were obtained using Self‐Assessment Manikin (SAM) scale (1–9). Results Mood valence ratings were significantly higher when viewing positive (end‐exercise mean ± SEM = 7.6 ± 0.3) compared to negative IAPS images (2.4 ± 0.3, p < 0.001). Dyspnoea intensity (mean ± SEM = 5.8 ± 0.4) and dyspnoea unpleasantness (5.6 ± 0.3) when viewing negative images were significantly higher compared to positive images (4.2 ± 0.4, p = 0.004 and 3.4 ± 0.5, p = 0.003). Eighty‐five percent of participants ( n = 17) met the minimal clinically important difference (MCID) criteria for both dyspnoea intensity and unpleasantness. HR, SpO 2 and leg fatigue did not differ significantly between conditions. Conclusion These findings indicate that the negative affective state worsens dyspnoea in COPD, thereby suggesting strategies aimed at reducing the likelihood of negative mood or improving the mood may be effective in managing morbidity associated with dyspnoea in COPD.
In our society men are considered more impulsive than women, especially in the violent and sexual domain. This correlation of sex and impulsivity might trace back to enhanced male impulsivity in general or a domain specific effect of emotions on impulsivity. The evidence for sex differences in the interaction of emotional or sexual stimuli and impulsivity has been relatively inconclusive so far. In this study, we investigated the effects of various emotional stimuli on responsivity in a Go/No-Go task. Participants had to respond quickly to a visual cue and withhold their response to another visual cue, while different emotional pictures were presented in the background, including sexual stimuli, non-sexual positive stimuli and negative stimuli. Both men (N = 37) and women (N = 38) made most commission errors in the sexual condition, indicating a disinhibiting effect in both genders. On top of this, men made even more commission errors than women, specifically in the sexual condition and not in other conditions. Men rated sexual stimuli as more positive, but did not differ from women in arousal ratings and pupil dilation. These findings may partly indicate increased impulsive behavior under sexual arousal in men, most likely driven by enhanced approach motivation due to more positive value but not higher arousal of sexual stimuli. The results are consistent with the theory of evolutionarily based concealment of sexual interest in women.
The organization of brain functional networks dynamically changes with emotional stimuli, but its relationship to emotional behaviors is still unclear. In the DEAP dataset, we used the nested-spectral partition approach to identify the hierarchical segregation and integration of functional networks and investigated the dynamic transitions between connectivity states under different arousal conditions. The frontal and right posterior parietal regions were dominant for network integration whereas the bilateral temporal, left posterior parietal, and occipital regions were responsible for segregation and functional flexibility. High emotional arousal behavior was associated with stronger network integration and more stable state transitions. Crucially, the connectivity states of frontal, central, and right parietal regions were closely related to arousal ratings in individuals. Besides, we predicted the individual emotional performance based on functional connectivity activities. Our results demonstrate that brain connectivity states are closely associated with emotional behaviors and could be reliable and robust indicators for emotional arousal.
Nowadays it is seen the active development of text sentiment analysis tools, which help solve a huge range of problems, ranging from emotional tone or polarity to the nature of the analyzed text. Although significant advances have been made in sentiment analysis for languages with abundant computing resources, the problem still remains relevant for resource-poor languages, where Uzbek is one of them. The lack of complete linguistic databases and tools for this language represents a major obstacle for the scientific community.The present study aims to address this limitation by introducing a rule-based approach for sentiment analysis in the Uzbek language. In particular, the methodology uses three key vocabularies: the first includes a comprehensive list of more than 300 affixes used to generate word forms; the second focuses on cataloging exception words that defy standard morphological rules; and the third contains a list of word roots, some of the word roots are already marked with positive or negative sentiment polarity.When combined, morphological analysis with these lexical resources can effectively recognize tonal orientation in Uzbek sentences. This is especially important given the agglutinative nature of language and complex morphological structures that add layers of subtle meaning, influencing the interpretation of sentiments.
Overgeneralization of conditioned fear is associated with anxiety disorders (AD). Most results stem from studies done in adult patients, but studies with children are rare, although the median onset of anxiety disorders lies already in childhood. Thus, the goal of the present study was to examine fear learning and generalization in youth participants, aged 10-17 years, with AD (n = 39) compared to healthy controls (HC) (n = 40). A discriminative fear conditioning and generalization paradigm was used. Ratings of arousal, valence, and US expectancy (the probability of an aversive noise following each stimulus) were measured, hypothesizing that children with AD compared to HC would show heightened ratings of arousal and US expectancy, and decreased positive valence ratings, respectively, as well as overgeneralization of fear. The results indicated that children with AD rated all stimuli as more arousing and less pleasant, and demonstrated higher US expectancy ratings to all stimuli when compared to HC. Thus, rather than displaying qualitatively different generalization patterns (e.g., a linear vs. quadratic slope of the gradient), differences between groups were more quantitative (similar, but parallel shifted gradient). Therefore, overgeneralization of conditioned fear does not seem to be a general marker of anxiety disorders in children and adolescents.
Academic success in adolescence is a strong predictor of well-being and health in adulthood. A healthy lifestyle and moderate/high levels of physical activity can influence academic performance. Therefore, we aimed to assess the relationship between the physical activity levels and body image and academic performance in public school adolescents. The sample consisted of 531 secondary school students in Porto (296 girls and 235 boys) aged between 15 and 20 years. The study variables and instruments were satisfaction with body image (The Body Image Rating Scale), assessment of physical activity (International Physical Activity Questionnaire for Adolescents (IPAQ-A), assessment of academic performance (academic achievement), school motivation (Academic Scale Motivation). The statistical analysis performed was descriptive analysis, an analysis of covariance, and a logistic regression. Regarding the results obtained, although there was no association between physical activity level and academic performance, it was observed in 10th grade students that the school average was higher for those practicing group or individual sports compared to students practicing artistic expression. Regarding the level of satisfaction with body image, we found different results in both genders. Our results support the importance of an active lifestyle, with the presence of regular physical activity being an important factor in improving academic performance.
Definitions of news are increasingly fraught in today’s media environment, making audience assessments of news-ness – the degree to which something is considered news – particularly important. Drawing from literature on representation in news and news-ness, we explore how seeing news that features a similar age group affects ratings of news-ness. We also argue that relevance offers as a psychological mechanism to explain how audiences make assessments about news-ness. Using two experiments among teens and adults in the United States, our results confirm that proximity (in terms of age) of the groups represented in news affects audience evaluations of news-ness, with relevance acting as a mediator to linking proximity and news-ness. These results are replicated across issues and among American teens and adults. We discuss the importance of representation and relevance as strategic initiatives to involve people in news.
Numerous cybercriminals are active in the online realm, carrying out cyber-crimes according to predetermined and preplanned agendas. Cyberbullying, which was formerly limited to physical limits, has now expanded online as a result of technology advancements. One type of cyberbullying is denigration or insult. The cyberbullying cases are in exponential rise in social media as per the reports of Computer Emergency Team by Sri Lanka. Insulting words are changeable in dynamic and the same terminology may have numerous meanings depending on the context. Bullying cannot be defined just because a statement comprises such a term. As a result, when classifying comments, standard keyword detecting approaches are insufficient. Other languages also may have dealt with this issue by utilizing lexical databases like WordNet, which might give synonyms as well as homonyms for words. Because no adequate lexical database mainly for the English language has been built, recognizing a word like bullying is difficult. As a result, employed rules to solve the problem. Facebook comments containing profanity were gathered, outliers were eliminated, and the remaining messages were pre-processed. Five feature extraction rules were employed to assess insult in the text. Following that, used the Support Vector Machine (SVM) technique. Using an F1-score of 85%, the findings demonstrate that when compared to existing works, SVM performs better. The focus on English language cyberbully identification, which has never been addressed earlier, distinguishes this study.
The article deals with the role of the contemporary Bulgarian linguist in the complex processes of codification of literary-linguistic norms. The text is linked to the 75th anniversary of the eminent Bulgarian linguist Prof. Vladko Murdarov. The object of the analytical observations in the text are problems related to the codification of literary-language norms, to the curricula and textbooks on Bulgarian language in secondary school; to the notion of language policy, as well as to the issues of the public image of the Bulgarian language in the media space. A brief overview is given of important historical processes that have shaped the development of the Bulgarian language and its science. The author summarizes the important features that every contemporary linguist should possess - depth of linguistic knowledge in synchronous and diachronic terms, a sense of the dynamics of linguistic processes in order to be able to objectively reflect the complex processes of language development, linking them to the complex and dynamic processes of social development.
In this paper, I will explore the concept of 'yakuwarigo' in Japanese language and present text analysis of three female characters, Eboshi, Rin and Yubaba from two anime movies of Hayao Miyazaki (Princess Mononoke, Spirited Away). My aim is to demonstrate how the well-known director employs role language, particularly masculine language, to empower his female characters to take on prominent roles in a society where men traditionally hold dominance. In terms of film analysis, Miyazaki places his female characters in the public sphere, making them active participants in the storyline while consistently defying traditional Japanese feminine conventions. This study is closely tied to the field of gender linguistics and linguistic ideology from an analytical perspective, aiming to illustrate how Miyazaki's female characters diverge from linguistic norms in their dialogues.
The goal of visual word sense disambiguation is to find the image that best matches the provided description of the word's meaning. It is a challenging problem, requiring approaches that combine language and image understanding. In this paper, we present our submission to SemEval 2023 visual word sense disambiguation shared task. The proposed system integrates multimodal embeddings, learning to rank methods, and knowledge-based approaches. We build a classifier based on the CLIP model, whose results are enriched with additional information retrieved from Wikipedia and lexical databases. Our solution was ranked third in the multilingual task and won in the Persian track, one of the three language subtasks.
This paper confirms that, in English binary coordinations, left conjuncts tend to be shorter than right conjuncts, regardless of the position of the governor of the coordination. We demonstrate that this tendency becomes stronger when length differences are greater, but only when the governor is on the left or absent, not when it is on the right. We explain this effect via Dependency Length Minimization and we show that this explanation provides support for symmetrical dependency structures of coordination (where coordination is multi-headed by all conjuncts, as in Word Grammar or in enhanced Universal Dependencies, or where it single-headed by the conjunction, as in the Prague Dependency Treebank), as opposed to asymmetrical structures (where coordination is headed by the first conjunct, as in the Meaning–Text Theory or in basic Universal Dependencies).
Attention mechanisms have become a crucial aspect of deep learning, particularly in natural language processing (NLP) tasks. However, in tasks such as constituency parsing, attention mechanisms can lack the directional information needed to form sentence spans. To address this issue, we propose a Bidirectional masked and N-gram span Attention (BNA) model, which is designed by modifying the attention mechanisms to capture the explicit dependencies between each word and enhance the representation of the output span vectors. The proposed model achieves state-of-the-art performance on the Penn Treebank and Chinese Penn Treebank datasets, with F1 scores of 96.47 and 94.15, respectively. Ablation studies and analysis show that our proposed BNA model effectively captures sentence structure by contextualizing each word in a sentence through bidirectional dependencies and enhancing span representation.
Multiword expressions (MWEs) are challenging and pervasive phenomena whose idiosyncratic properties show notably at the levels of lexicon, morphology, and syntax. Thus, they should best be annotated jointly with morphosyntax. We discuss two multilingual initiatives, Universal Dependencies and PARSEME, addressing these annotation layers in cross-lingually unified ways. We compare the annotation principles of these initiatives with respect to MWEs, and we put forward a roadmap towards their gradual unification. The expected outcomes are more consistent treebanking and higher universality in modeling idiosyncrasy.
The Bitext Synonym Data - General Language includes 31,723 entries and more than 100,000 synonyms for English language. This dataset is a set of synonyms developed to augment the English version of Wordnet, a powerful open-source lexical database, released in 2005. All synonyms can be linked to Bitext Lexical Data - English (see ELRA-L0140) for lemmatization, POS and morphological information.
This paper describes the development and validation of 3D Affective Virtual environments and Event Library (AVEL) for affect induction in Virtual Reality (VR) settings with an online survey; a cost-effective method for remote stimuli validation which has not been sufficiently explored. Three virtual office-replica environments were designed to induce negative, neutral and positive valence. Each virtual environment also had several affect inducing events/objects. The environments were validated using an online survey containing videos of the virtual environments and pictures of the events/objects. They survey was conducted with 67 participants. Participants were instructed to rate their perceived levels of valence and arousal for each virtual environment (VE), and separately for each event/object. They also rated their perceived levels of presence for each VE, and they were asked how well they remembered the events/objects presented in each VE. Finally, an alexithymia questionnaire was administered at the end of the survey. User ratings were analysed and successfully validated the expected affect and presence levels of each VE and affect ratings for each event/object. Our results demonstrate the effectiveness of the online validation of VE material in affective and cognitive neuroscience and wider research settings as a good scientific practice for future affect induction VR studies.
The growing interest in harnessing natural environments to enhance mental health, including cognitive functioning and mood, has yielded encouraging results in initial studies. Given that images of nature have demonstrated similar benefits, they are frequently employed as proxies for real-world environments. To ensure precision and control, researchers often manipulate images of natural environments. The effectiveness of this approach relies on standardization of imagery and therefore inconsistency in methods and stimuli has limited the synthesis of research findings in the area. Responding to these limitations, the current paper introduces the Salford Nature Environments Database (SNED), a standardized database of natural images created to support ongoing research into the benefits of nature exposure. The SNED currently exists as the most comprehensive nature image database available, comprising 500 high-quality, standardized photographs capturing a variety of possible natural environments across the seasons. It also includes normative scores for user-rated (801 participants) characteristics of fascination, refuge and prospect, compatibility, preference, valence, arousal and approach-avoidance, as well as data on physical properties of the images, specifically luminance, contrast, entropy, CIELAB color space parameter values, and fractal dimensions. All image ratings and content detail, along with participant details, are available open access online. Researchers are encouraged to use this open access database in accordance with the specific aims and design of their study. The SNED represents a valuable resource for continued research in areas such as nature-based therapy, social prescribing and experimental approaches investigating underlying mechanisms that help explain how natural environments improve mental health and wellbeing.
No description provided.
Abstract: Parent–child attachment is robustly associated with typical patterns of emotion regulation but rarely examined in relation to changes in emotion in response to events. We studied how attachment is related to emotion reactivity to positive and negative events and to immediate and delayed emotion recovery from social exclusion. The sample (78% White, 46% girls) included 110 children (9–12 years). Children completed a story stem measure that was scored for security, avoidance, ambivalence, and disorganization. Emotion reactivity and recovery were assessed with child-reported and observer positive affect and negative affect ratings. Parents rated child temperament. More avoidant children showed dampened emotion responding (low reactivity and recovery), whereas more ambivalent children showed heightened emotion responding (high reactivity and recovery). Attachment security, disorganization, and temperament were not consistently related to emotion reactivity or recovery. The findings highlight that emotion regulation occurs in response to contextual changes and is related to attachment.
The folder contains subjective arousal ratings, eye-tracking data (fixation times) and EEG data relative to the processing of emotional body expressions presented at the end of a virtual promenade within different architectural forms. Scripts are also provided for the reproducibility of statistical data analysis. Please refer to the "Readme.txt" file for more detailed information
Filipino is a changing language that poses several challenges. The study was done to overcome the time needed to construct a corpus with it syntactic structure, treebank, as a part of the project where the researchers’ part is for Filipino Language Treebank. This project is, Open Collaboration for Developing and Using Asian Language Treebank (ALT). The objective of this study is to parse Tagalog text and determine grammar rules and part-of-speech tags using a Probabilistic Context-free Grammar. The model was trained and tested using the Asian Language Treebank Filipino dataset which includes Filipino sentences and their grammar rules. The system achieved 86.43% f-measure score in parsing the grammar rules and part-of-speech tags of Filipino sentences. The model can reliably be used to automatically create treebanks for Filipino sentence grammar.
This dataset contains EEG, ECG and audio recordings of 11 individual musicians playing emotional music on their instrument. The dataset consists of two parts: Experiment: Musicians’ self-reported ratings, audio recordings, and physiological recordings where the 11 expert musicians were asked to play at least 4~2-minute unfamiliar (non-popularly known) musical pieces. Participants were asked to play at least once one of the following emotions: happiness, sadness, relaxation, and anger. For the physiological recordings EEG, ECG, and GSR signals were recorded. Each musican’s data is denoted by MS_ followed by the order in which they were recorded. Self-report questionnaire: A self-assessment questionnaire and their answers where 11 expert musicians were asked to rate musical pieces recorded based on: Objective valence & arousal they felt the piece had; Felt valence & arousal during playing. For a more detailed explanation of the dataset, its recording procedure, and its contents, see L. Turchet, B. O'Sullivan, R. Ortner & C. Gugher (2024). Emotion Recognition of Playing Musicians from EEG, ECG, and Acoustic Signals. IEEE Transactions on Human-Machine Systems. File Listing The following files are available (each explained in more detail below): Name Format Contents EEG_ECG_data_for_each_musician mat This folder contains the raw EEG, ECG, & GSR data for all 11 musicians for each piece they played as well as a resting state recording, which was recorded while a neutral audio stimulus was played. audio_data_for_each_musician wav, JSON, csv This folder contains 3 subfolders: 1) wav_audio_files_original: This is the raw audio data recorded for each musician and includes all 56 pieces included in the paper reported above, please see below regarding rejected trials. 2) wav_audio_files (original split_into_3_parts): Here, the 56 pieces are appropriately split into 3 separate parts.3) analysis_audio_files: This folder contains the acoustic features extracted, 1714 acoustic features were extracted from each split trial. Each result is stored in.JSON, however, the collated results can be seen in all_results.csv. self_report_questionnaire pdf, xls Two files exist in this folder:1. The self-reported questionnaire given to the musicians of the questionnaire during the experiment.2. File Details EEG_ECG_data_for_each_musician These are the original raw data recordings. EEG data were recorded using a g.GAMMAcap2 by g.tec Medical Engineering, a 64-channel cap with g.SCARABEO active electrodes, with two g.GAMMAsys reference active ear clip electrodes. Two g.GAMMAbox electrode connector boxes were used to connect the active electrodes to two g.USBamp biosignal amplifiers with a sampling frequency of 256 Hz.The following 31 EEG channels were used: Fp1, Fp2, AFz, AF3, AF4, AF7, AF8, Fz, F3, F4, F7, F8, Cz, C3, C4, CP3, CP4, CP5, CP6, P1, P2, P3, P4, P5, P6, P7, P8, PO7, PO8, O1, and O2. AFz was used as a ground electrode and Cz was used as a re-reference electrode. The right-side ear clip electrode was used as a reference electrode. ECG data were recorded using a single g.GAMMAclip active electrode clip connected directly to the g.GAMMAbox, sharing the same ground electrode with the EEG cap and placed on position V4 of the subjects.GSR data was recorded using the g.GSRsensor² box which contains two small dry electrodes placed underneath the participant’s toes (due to the amount of hand movement required for playing an instrument). The g.GSRsensor² was connected directly to a g.GAMMAbox, using jumper cables connected to different ports in the g.USBAMPs to share the same reference and ground electrodes as the EEG cap. The locations of the channels and their corresponding number in the raw data is as follows: Channel Number Channel Name 1 Time series 2 AF3 3 AF4 4 AF7 5 AF8 6 CP3 7 CP4 8 CP5 9 CP6 10 P1 11 P2 12 P5 13 P6 14 P7 15 P8 16 O1 17 O2 18 Cz 19 Fp1 20 Fp2 21 F3 22 F4 23 F7 24 F8 25 C3 26 C4 27 P3 28 P4 29 PO7 30 PO8 31 AFz 32 GSR 33 ECG audio_data_for_each_musician/wav_audio_files_original This folder contains all of the raw audio files recorded during the experiment. All pieces were recorded using the software Audacity and exported as WAV files encoded with a bit depth of 32-bits and a sampling rate of 44.1 kHz. 56 of the included trials are present in this folder. audio_data_for_each_musician/wav_audio_files (original split_into_3_parts) This folder contains the above mentioned 56 raw audio pieces separated into 3 appropriately sized recordings. audio_data_for_each_musician/analysis_audio_filesThis folder contains the 1714 acoustic features from each of the 3 separated trials from 56 accepted pieces, these are denoted by MS_ followed by the order in which the musicians were recorded and the order in which the trails were split, the individual features are in JSON format. Within this folder, we have collated all results in a csv file (all_results.csv) where the columns show the intended emotion, trial name and number, followed by the names of the acoustic features. self_report_questionnaire This pdf is the questionnaire which each musician was given following each trial relating to the emotions communicated and felt during each trial. Most* questions in the questionnaire were multiple-choice and speak pretty much for themselves. The answers for which were collated and are described below. self_report_questionnaire_anwsers This.xls file contains the results of all the musicians self-reported ratings from the above described questionnaire. Column name Description Subject The subject code of the musician, denoted by MS_ recording_filename The name of the trial denoted by MS_01_ followed by the trial number. intended_emotion The intended emotion the musician was instructed to communicate. The emotions are as follows: Angry Sad Relaxed Happy They are intended to be reported by using the valence-arousal space. valence_communicated The valence rating (integer between 1 and 5), that participants were asked to objectively rate how they thought the music played would be perceived. arousal_communicated As above, but relating to arousal rather than valence. valence_felt The valence rating (integer between 1 and 5), that participants were asked to rate how they felt while playing the piece. arousal_felt As above, but relating to arousal rather than valence. instrument_played The instrument played for each trial.
OBJECTIVE: To test and initially describe a new handheld wireless ultrasound technique (TE Air) for clinical use. METHODS: In this pilot study, the new ultrasound device TE Air from Mindray was used to examine the hepatic and renal vessels of healthy volunteers for first impressions. The probe has a sector transducer with a frequency range of 1.8-4.5 MHz. The B-mode and color-coded doppler sonography (CCDS) scanning methods were used. A high-end device from the same company (Resona 9, Mindray) was used as a reference. The results were evaluated using an image rating scale ranging from 0 to 5, with 0 indicating not assessable and 5 indicating without limitations. RESULTS: Altogether, 61 participants (n = 34 female [55.7%], n = 27 male [44.3%]), age range 18-83 years, mean age 37.9±16.5 years) could be adequately studied using TE AIR and the high-end device. With one exception, the image quality score for TE Air never fell below 3 and had a mean/median scored of 4.97/5.00 for the B-mode, 4.92/5.00 for the color flow (CF) mode, and 4.89/5.00 for the pulse wave (PW) mode of the hepatic vein, 4.90/5.00 for the portal vein, 4.11/4.00 for the hepatic artery, and 4.57/5.00 for the renal segmental artery. A significant difference in the assessment of flow measurement of the hepatic artery and renal segmental arteries was found between TE AIR and the high-end device. CONCLUSIONS: TE Air represents a new dimension in point-of-care ultrasound via wireless handheld devices. Especially, its flow measurement ability offers a relevant advantage over other available handheld models. TE Air provides a formally sufficient image quality in terms of diagnostic significance.
The Dutch anaphoric possessive construction (APC), as exemplified by Tom zijn fiets ‘Tom his bike’, shows a peculiar mix of regularity and idiosyncracy. The article provides a theory-neutral description of its properties and quantitative information about its use in two treebanks, one of spoken Dutch (CGN) and one of written Dutch (Lassy Small). It argues that the APC has a right branching structure and models it in the framework of Constructional Head-driven Phrase Structure Grammar. The latter’s organization of constructions in terms of a finegrained hierarchy of phrase types is shown to provide the means to capture both what the APC has in common with other possessive constructions and what is idiosyncratic of it.
This dataset offers a multilingual corpus for studying Persian translations of Plato’s Crito in comparison with the Ancient Greek source text. It is designed for scholars and students working in fields such as translation studies, computational linguistics, and philology. This repository includes: The Ancient Greek text of Crito (Burnet edition) Five Persian student translations, word-level and sentence-level alignments Three finalized Persian translations, word-level andsentence-level alignments, including UD treebank annotations Two English translations (Jowett and Fowler), sentence-level aligned One German translation (Schleiermacher), sentence-level aligned A Greek–Persian lexical wordlist for Crito Source Texts & References Greek text: Plato. Platonis Opera, ed. John Burnet. Oxford University Press. 1903.Source: Perseus Digital Library – https://www.perseus.tufts.edu/hopper/text?doc=Perseus%3Atext%3A1999.01.0169%3Atext%3DCrito%3Asection%3D43a Student Persian translations: Produced during a Greek translation course by elementary students after a 30-hour introductory course to Ancient Greek. Each translation was made using Ancient Greek commentaries, lexicons, and parallel English and German translations available in this dataset. Alignments were created by the translators using Ugarit.These translations are available as word-level and sentence-level alignments. The word alignments are extracted from Ugarit, but they are still visualized on Ugarit. Translators’ Ugarit profiles: - Shouresh Assimi: https://ugarit.ialigner.com/userProfile.php?userid=50956- Aylar Mahmoudzadeh Sarabi: https://ugarit.ialigner.com/userProfile.php?userid=63464&tgid=9576- Nima Mohammadi: https://ugarit.ialigner.com/userProfile.php?userid=52434&tgid=9362- Kimia Nikpour: https://ugarit.ialigner.com/userProfile.php?userid=52378- Farshid Rahimi: https://ugarit.ialigner.com/userProfile.php?userid=50932&tgid=9727 Finalized Persian translations: Three finalized Persian translations, revised under the supervision of Farnoosh Shamsian.These translations are available as word-level and sentence-level alignments, accompanied by UD treebank annotation English translations: Fowler’s translation:Plato, H. N. Fowler, W. Lamb, Plato in Twelve Volumes, Vol. 1, translated by Harold North Fowler; Introduction by W.R.M. Lamb, volume 1, Harvard University Press and Wiliam Heinemann Ltd., Cambridge, MA and London, 1966.Fowler's translation on Perseus Digital Library:https://www.perseus.tufts.edu/hopper/text?doc=plat.+crito+43aAlignment: Sentence-level Jowett's translation:Plato, B. Jowett, Crito, The Internet Classics Archive, Massachusetts Institute of Technology, http://classics.mit.edu/Plato/crito.html.Alignment: Sentence-level German translation: Plato, F. Schleiermacher, Platons Werke, In der Realschulbuchhandlung, 1809Link to the text on Project Gutenberg:https://www.projekt-gutenberg.org/platon/platowr1/kriton.html Alignment: Sentence-level Dataset Contents To ensure proper rendering of Persian RTL text, files are available in both CSV and XML formats when necessary. Student translations aligned at sentence-level:Sentence-level alignments of the Ancient Greek text of Plato’s Crito with all five student Persian translations, two English translations (Jowett and Fowler), and one German translation (Schleiermacher).The sentence number is the treebank sentence ID followed by the Stephanus paginations.Files and Formats available:crito_sentence_alignments_student_translations.csvcrito_sentence_alignments_student_translations.xlsxNotes: All eight translations aligned to the same sentence boundaries as the Greek text. Student alignment data extracted from Ugarit:Word-level alignments of the student translations, which were done throughout the course by each student on Ugarit. The zip file includes five folders, each containing the Ugarit user ID of the student who did the translation and the alignment, followed by their last name in parentheses. Each folder includes the alignments of Crito done by the student, exported from Ugarit as a JSON file. The name of each JSON file is the identifier of the alignment on Ugarit.Formats available:student_alignments_ugarit.zip, containing JSON files Three finalized Persian translations aligned at sentence-level:Sentence-level alignments of the Ancient Greek text of Crito with the three finalized Persian translations under the supervision of Farnoosh Shamsian. The sentence number is the treebank sentence ID followed by the Stephanus paginations. Formats available:crito_finalized_translations_sentence_alignment.csvcrito_finalized_translations_sentence_alignment.xlsx Three finalized Persian translations aligned at word-level to the UD treebanks:Word-level alignments of the three finalized Persian translations to the Ancient Greek text, accompanied by UD Treebank annotations. These alignments are done by a single annotator according to alignment guidelines previously published on Zenodo here:Shamsian, F. (2023). Alignment Guidelines for Greek-Persian. Zenodo. https://doi.org/10.5281/ZENODO.8039931Formats available:crito_finalized_translations_treebank_word_alignment.csvcrito_finalized_translations_treebank_word_alignment.xlsxNotes: The translations are aligned to the treebanks provided by the Pedalion project, which were automatically generated and manually corrected. See more here: https://perseids-publications.github.io/pedalion-trees/ A Greek-Persian wordlist:A bilingual wordlist containing Greek–Persian lexical correspondences for every Greek word attested in Plato’s Crito.Formats available:crito_greek_persian_wordlist.csvcrito_greek_persian_wordlist.xlsx Notes This is the actively maintained and most complete release. An earlier version (only the three finalized translations) is now outdated, but still archived on Zenodo for reference here: Shamsian, F., Assimi, S., Mahmoudzadeh Sarabi, A., Mohammadi, N., Nikpour, K., & Rahimi, F. (2023). Persian Translations of Crito, Three Versions with Word-level Alignment. Zenodo. https://doi.org/10.5281/zenodo.8333681 Refrences to the consulted print sources and other Persian translations: Adam, J. (1888). Platonis Crito. Cambridge University Press. Emlyn-Jones, C. (1999). Crito. London: Bloomsbury Academic. Fowler, H. N., Lamb, W. R. M., & Shorey, P. (1966). Plato in Twelve Volumes: With an English Translation. Harvard University Press. Jowett, B. (1909). The Apology, Phaedo, and Crito of Plato, with introduction, notes and illustrations. New York P.F. Collier & son. http://archive.org/details/apologyphaedocri0000plat Plato (1938) Hekmate Soghrāt va Aflātun (M. A. Foroughi, Trans.) Tehran: Majles Publishing house. Plato (1979) Dore-ye Āsāre Aflātun (H. Lotfi, Trans.) Tehran: Kharazmi. Plato (1989) Vāpasin Ruzhāye Soghrāt (J. Jahanshahi, Trans.) Tehran: Porsesh. Plato (2013) Mohākeme-ye Soghrāt (L. Golestan, Trans.) Tehran: Markaz. Plato (2023) Kriton (I. Shafi’-Beyk, Trans.) Tehran: Ney. Schleiermacher, F. (1910). Platons Apologie und Kriton. Verlag Philipp Reclam.
The aim of this paper is to present criteria for identifying adjectival Lithuanian participles, which function as adjectives but have participial forms, and to explain how adjectival participles can be identified through corpus data analysis. The relevance of the research lies in the fact that adjectival participles are considered as adjectives in dictionaries. Moreover, the devised identification criteria may be of added value for part-of-speech tagging, teaching Lithuanian, and other related domains.The article examines the concept of adjectivization, proposing a potential classification of Lithuanian participles based on the extent of their adjectivization (i.e., fully adjectival and partially adjectival participles). It also suggests criteria for identifying adjectival participles. The criteria, categorized into grammatical, semantic, derivational, and quantitative, are distinguished based on prior research and a pilot study. The data for the pilot study was extracted from Mokomasis lietuvių kalbos vartosenos leksikonas (The Lexical Database of Lithuanian Language Usage), encompassing 200 most frequent verbs and 49 adjectives with participial forms analyzed across two Lithuanian corpora. In addition, the order of criteria application is demonstrated in the present paper. In total, 185 frequent verbs were analyzed in the corpora to identify adjectival participles based on the following parameters: (1) the inclination of the participle to form an adverb with the suffix -ai; (2) its propensity to generate comparative and superlative forms and be used with adverbs of measure/degree; and (3) its inclination to be used attributively without verbal arguments. Methodologically, a quantitative approach was applied to analyze the data (only frequent participle forms were examined, as it is hypothesized that particularly frequent participles are adjectival).The results show that the proposed criteria, along with quantitative information, aid in the identification of adjectival participles. Seventeen identified participles are classified as separate words rather than verbal forms by Bendrinės lietuvių kalbos žodynas (The Dictionary of the Standard Lithuanian Language); one identified participle is categorized as an adjective by Lietuvių kalbos žodynas (The Lithuanian Language Dictionary); one identified participle was found to be adjectival, and another participle was used as a pronoun, while eight identified participles are considered partly adjectival. Six identified participles should not be considered adjectival. Thus, it appears that additional criteria need to be applied, and the methodology requires refinement for a more accurate identification of adjectival participles.