Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
16504 papers
This article describes the first steps towards a open-source dependency treebank for Erzya based on universal dependency (UD) annotation standards. The treebank contains 610 sentences with 6661 tokens and is based on texts from a range of open-source and public domain original Erzya sources. This ensures its free availability and extensibility. Texts in the treebank are first morphologically analyzed and disambiguated after which they are annotated manually for dependency structure. In the article we present some issues in dependency syntax for Erzya and how they are analyzed in the universal-dependency framework. Preliminary statistics are given for dependency parsing of Erzya, along with points of interest for future research.
We present a progress report of the Turkish Treebank concentrating on various aspects of its design and implementation. In addition to a review of the corpus compilation process and the design of the annotation scheme, we describe the details of various pre-processing stages and the computer-assisted annotation process.
This paper describes our system (HIT-SCIR) submitted to the CoNLL 2018 shared task on Multilingual Parsing from Raw Text to Universal Dependencies. We base our submission on Stanford's winning system for the CoNLL 2017 shared task and make two effective extensions: 1) incorporating deep contextualized word embeddings into both the part of speech tagger and parser; 2) ensembling parsers trained with different initialization. We also explore different ways of concatenating treebanks for further improvements. Experimental results on the development data show the effectiveness of our methods. In the final evaluation, our system was ranked first according to LAS (75.84%) and outperformed the other systems by a large margin.
Implicit discourse relation recognition is a challenging task as the relation prediction without explicit connectives in discourse parsing needs understanding of text spans and cannot be easily derived from surface features from the input sentence pairs. Thus, properly representing the text is very crucial to this task. In this paper, we propose a model augmented with different grained text representations, including character, subword, word, sentence, and sentence pair levels. The proposed deeper model is evaluated on the benchmark treebank and achieves state-of-the-art accuracy with greater than 48% in 11-way and F1 score greater than 50% in 4-way classifications for the first time according to our best knowledge.
Larysa Kolibaba PhD in Philology, Senior Research Scientist of the Department of Grammar and Scientific Terminology, Institute of the Ukrainian Language of National Academy of Sciences of Ukraine 4 Hrushevskyi St., Kyiv 01001, Ukraine Е-mail: kolibaba.lm@meta.ua Heading: Researches Language: Ukrainian Abstract: In this article the problem of fixing of morphological forms of nouns in the Ukrainian dictionaries of different time and its […]
We introduce the syntactic scaffold, an approach to incorporating syntactic information into semantic tasks. Syntactic scaffolds avoid expensive syntactic processing at runtime, only making use of a treebank during training, through a multitask objective. We improve over strong baselines on PropBank semantics, frame semantics, and coreference resolution, achieving competitive performance on all three tasks. them he After encouraging them, told goodbye and left for STIMULATE _EMOTION
Version 1 of the Late Latin Charter Treebank (LLCT1). Early medieval Latin documentary texts with morphological and syntactic annotation. Ancient Language Dependency Treebank compatible linguistic annotation, Prague style treebank format (PML). See full description in Korkiakangas & Lassila, 2013, "Abbreviations, fragmentary words, formulaic language: treebanking medieval charter material". Will be replaced by LLCT2 in 2018 (version 2).
In this paper, we describe the annotation and development of Telugu treebank following the Universal Dependencies framework. We manually annotated 1328 sentences from a Telugu grammar textbook and the treebank is freely available from Universal Dependencies version 2.1.1 In this paper, we discuss some language specific annotation issues and decisions; and report preliminary experiments with POS tagging and dependency parsing. To the best of our knowledge, this is the first freely accessible and open dependency treebank for Telugu.
Preview this article: The added value of diachronic treebanks for historical linguistics, Page 1 of 1 < Previous page | Next page > /docserver/preview/fulltext/dia.00004.eck-1.gif
In this paper we describe the extensions we made to an existing treebank query application (GrETEL). These extensions address user needs expressed by multiple linguistic researchers and include (1) facilities for uploading one’s own data and metadata in GrETEL; (2) conversion and cleaning modules for uploading data in the CHAT format; (3) new facilities for analysing the results of the treebank queries in terms of data, metadata and combinations of them. These extensions have been made available in a new version (Version 4) of GrETEL.
We present PAWS, a multi-lingual parallel treebank with coreference annotation. It consists of English texts from the Wall Street Journal translated into Czech, Russian and Polish. In addition, the texts are syntactically parsed and word-aligned. PAWS is based on PCEDT 2.0 and continues the tradition of multilingual treebanks with coreference annotation. The paper focuses on the coreference annotation in PAWS and its language-specific differences. PAWS offers linguistic material that can be further leveraged in cross-lingual studies, especially on coreference.
The present paper shows how the current Universal Dependency treebanks can be used for language typology studies and can reveal structural syntactic features of languages. Two methods, one existing method and one newly proposed method, based on dependency treebanks as typological measurements, are presented and tested in order to assess both the coherence of the underlying syntactic data and the validity of the methods themselves. The results show that both methods are valid for positioning a language in the typological continuum, although they probably reveal different typological features of languages.
The article presents a quantitative analysis of some syntactic dependency properties in Czech. A dependency frame is introduced as a linguistic unit and its characteristics are investigated. In particular, a ranked frequencies of dependency frames are observed and modelled and a relationship between particular syntactic functions and the number of dependency frames is examined. For the analysis, the Czech Universal Dependency Treebank is used.
We present a novel abstractive summarization framework that draws on the recent development of a treebank for the Abstract Meaning Representation (AMR). In this framework, the source text is parsed to a set of AMR graphs, the graphs are transformed into a summary graph, and then text is generated from the summary graph. We focus on the graph-to-graph transformation that reduces the source semantic graph into a summary graph, making use of an existing AMR parser and assuming the eventual availability of an AMR-to-text generator. The framework is data-driven, trainable, and not specifically designed for a particular domain. Experiments on gold-standard AMR annotations and system parses show promising results. Code is available at: https://github.com/summarization
Introduction. The article explores the impact of various types of verbal representation of ethnic stereotypes in the framework of a polyethnical academic community, i.e. educational environment in modern international university. Although the educational process with subjects of different cultural backgrounds plays a crucial role in conveying world views of representatives of different cultures, the research on the linguistic representation of stereotyped views on representatives of other nationalities has not been conducted yet. This aspect determines the relevance of the study. The aim of the research is to compare the impact levels of purely linguistic and speech ways of verbalising heterostereotypes by ways of employing relevant linguistic data for academic purposes during foreign language classes. Materials and Methods. The first stage of the experiment resulted in preparation of the linguistic corpus for the research: by means of comprehensive vocabulary research the lexical database with ethnonyms or ethnonym-based adjectives was compiled. To reveal the potential of their usage in the education processes, the participants were offered the preliminary and final surveys held as free associatio n experiment. Results. The influential potential for purely linguistic and speech ways of representing national stereotypes was compared to find out if they relate to the descriptors and scripts revealed through analysis of phraseological units and national anecdote respectively, while the latter was marked as a more efficient way of delivering ethnic stereotypes. The conclusions based on the analysis of the data obtained were drawn on how to use relevant linguistic material for academic purposes in order to appropriately develop attitudes to other ethnic groups. Discussion and Conclusions. The conducted research revealed more significant impact degree for ethnic anecdotes against investigation of lexical-phraseological units containing ethnonyms or ethnonym-based adjectives. It was illustrated by collection and further analysis of verbal reactions provided by students of non-linguistic departments of the modern University who took part in the preliminary and final stages which were in line with the beginning and end of the academic term accordingly. The portraits of typical national representatives made by the students at the completion of the course which included sessions on studying dictionary extracts and national anecdotes, to a greater extent conformed with the stereotypes delivered by ethnic anecdotes than the linguistic corpus of lexical-phraseological units. The research results may be considered during the development of the curriculum for foreign language courses in international universities with polyethnical academic environment.
Activities can increase quality of life for residents with dementia, however determining which activities are high in quality is often subjective. In this study, trained researchers observed 22 residents in common areas of a memory care unit (10 males, 12 female, all consented for research). Assessments occurred in 15-minute sessions across multiple days (totaling 7000 minutes). Each minute involved co-observation of staff interactions (Quality Interaction Scale, Dean, et al., 1993) and residents’ positive affect (Philadelphia Geriatric Center Affect Rating Scale, Lawton, et al., 1996). Observers noted types of activities underway. Z-scores indicated proportionally higher positive affect in residents during preplanned activities, compared to non-facilitated/unplanned activities (z = -3.09, p <.001). Compared to residents’ positive affect during “no activity”, positive affect was proportionally highest (p <.001) during music therapy (z = -23.43) and motor activity (z = -13.67), and lowest (p = n.s.) during dance performances (z = -1.07) and art/crafts (z = 1.83). Compared to positive staff interactions during “no activity”, positive staff interactions was proportionally highest (p <.001) during motor activities (z = -12.74) and music therapy (z = -11.86), and lowest (p = n.s.) during cognitive activities (z = -0.30) and music presentations (z = -0.17). Commonalities in quality activities included residents being able to see, engage, and move about if they wanted, staff considering residents’ autonomy and staff using active efforts to converse with residents. Staff can observe affect in residents to evaluate engagement during activities, and adjust delivery and interaction frequency where needed.
Det är lättare att minnas emotionellt laddad information. Även om minnet ofta försämras vid normalt åldrande, kvarstår den emotionella förstärkningseffekten. Studier antyder att unga vuxna minns bättre negativt än positivt material, men att minnespreferensen för negativt material minskar eller t.o.m. ersätts av en preferens för positivt material med åldern. I avhandlingen undersöktes hur unga och äldre friska vuxna personer minns ord med olika emotionellt värde (positiv/negativ) och emotionell intensitet. Även underliggande mekanismer i hjärnan hos äldre vuxna utforskades. Avhandlingen visade att det inte fanns skillnader mellan hur unga (21-35 år) och äldre (50-79 år) vuxna minns emotionella ord. Här sågs alltså ingen skillnad i minnespreferens för negativa eller positiva ord eller för ord med olika stark emotionell intensitet mellan åldersgrupperna. I avhandlingen identifierades samband mellan minne för positiva ord och gråsubstansvolym i hjärnans bakre delar hos friska äldre vuxna. Bättre minne för positiva ord hade samband med lägre integritet hos vitsubstansbanor i hjärnan, vilket antyder att en kompensationsmekanism kunde spela en roll. Avhandlingen visade även att språkliga och kulturella faktorer har större betydelse än ålder och kön för bedömningen av emotionella egenskaper hos ord. ------------------------------------- On helpompi muistaa tunnesisältöistä materiaalia. Vaikka muisti usein heikkenee normaalissa ikääntymisessä, tunnesisällön muistia vahvistava vaikutus pysyy. Tutkimukset osoittavat, että nuoret aikuiset muistavat kielteistä materiaalia myönteistä paremmin, mutta että kielteisen materiaalin suosiminen vähenee tai jopa vaihtuu myönteisen materiaalin suosimiseksi ikääntyessä. Väitöskirjassa tutkittiin miten nuoret ja ikääntyvät terveet aikuiset muistavat erityyppisiä tunnesisältöisiä sanoja. Selvitettiin myös tunnemuistin aivostollista perustaa ikääntyvillä aikuisilla. Väitöskirja osoitti, ettei nuorten (21-35 v.) ja ikääntyvien (50-79 v.) aikuisten välillä ollut eroja tunnesisältöisten sanojen muistamisessa. Tässä ei siis havaittu ikään liittyvää eroa kielteisten tai myönteisten sanojen tai eri tunneintensiteetillä varustettujen sanojen suosimisessa. Väitöskirjassa havaittiin yhteys myönteisten sanojen muistamisen ja aivojen takaosien harmaan aineen volyymin välillä terveillä ikääntyvillä aikuisilla. Mitä paremmin myönteisiä sanoja muistettiin, sen matalampi valkean aineen radastojen integriteetti oli normaalisti ikääntyvillä, mikä voi viitata korvaavan mekanismin vaikutukseen. Väitöskirja osoitti myös, että kielellisillä ja kulttuurillisilla tekijöillä on suurempi merkitys kuin iällä tai sukupuolella sanojen tunnesisällön arvioinnissa.
In situations of real threat, showing a fear reaction makes sense, thus, increasing the chance to survive. The question is, how could anybody differentiate between a real and an apparent threat? Here, the slogan counts “better safe than sorry”, meaning that it is better to shy away once too often from nothing than once too little from a real threat. Furthermore, in a complex environment it is adaptive to generalize from one threatening situation or stimulus to another similar situation/stimulus. But, the danger hereby is to generalize in a maladaptive manner involving as it is to strong and/or fear too often “harmless” (safety) situations/stimuli, as it is known to be a criterion of anxiety disorders (AD). Fear conditioning and fear generalization paradigms are well suited to investigate fear learning processes. It is remarkable that despite increasing interest in this topic there is only little research on fear generalization. Especially, most research on human fear conditioning and its generalization has focused on adults, whereas only little is known about these processes in children, even though AD is typically developing during childhood. To address this knowledge gap, four experiments were conducted, in which a discriminative fear conditioning and generalization paradigm was used. In the first two experiments, developmental aspects of fear learning and generalization were of special interest. Therefore, in the first experiment 267 children and 285 adults were compared in the differential fear conditioning paradigm and generalization test. Skin conductance responses (SCRs) and ratings of valence and arousal were obtained to indicate fear learning. Both groups displayed robust and similar differential conditioning on subjective and physiological levels. However, children showed heightened fear generalization compared to adults as indexed by higher arousal ratings and SCRs to the generalization stimuli. Results indicate overgeneralization of conditioned fear as a developmental correlate of fear learning. The developmental change from a shallow to a steeper generalization gradient is likely related to the maturation of brain structures that modulate efficient discrimination between threatening and (ambiguous) safety cues. The question hereby is, at which developmental stage fear generalization gradients of children adapt to the gradients of adults. Following up on this question, in a second experiment, developmental changes in fear conditioning and fear generalization between children and adolescents were investigated. According to experiment 1 and previous studies in children, which showed changes in fear learning with increasing age, it was assumed that older children were better at discriminating threat and safety stimuli. Therefore, 396 healthy participants (aged 8 to 12 years) were examined with the fear conditioning and generalization paradigm. Again, ratings of valence, arousal, and SCRs were obtained. SCRs indicated differences in fear generalization with best fear discrimination in 12-year-old children suggesting that the age of 12 years seems to play an important role, since generalization gradients were similar to that of adults. These age differences were seen in boys and girls, but best discrimination was found in 12-year-old boys, indicating different development of generalization gradients according to sex. This result fits nicely with the fact that the prevalence of AD is higher in women than in men. In a third study, it was supposed that the developmental trajectory from increased trait anxiety in childhood to manifest AD could be mediated by abnormal fear conditioning and generalization processes. To this end, 394 children aged 8 to 12 years with different scores in trait anxiety were compared with each other. Results provided evidence that children with high trait anxiety showed stronger responses to threat cues and impaired safety signal learning contingent on awareness as indicated by arousal at acquisition. Furthermore, analyses revealed that children with high trait anxiety showed overall higher arousal ratings at generalization. Contrary to what was expected, high trait anxious children did not show significantly more fear generalization than children with low trait anxiety. However, high-trait-anxious (HA) participants showed a trend for a more linear gradient, whereas moderate-trait-anxious (MA) and low-trait-anxious (LA) participants showed more quadratic gradients according to arousal. Additionally, after controlling for age, sex and negative life experience, SCR to the safety stimulus predicted the trait anxiety level of children suggesting that impaired safety signal learning may be a risk factor for the development of AD. Results provide hints that frontal maturation could develop differently according to trait anxiety resulting in different stimuli discrimination. Thus, in a fourth experiment, 40 typically developing volunteers aged 10 to 18 years were screened for trait anxiety and investigated with the differential fear conditioning and generalization paradigm in the scanner. Functional magnetic resonance imaging (fMRI) were used to identify the neural mechanisms of fear learning and fear generalization investigating differences in this neural mechanism according to trait anxiety, developmental aspects and sex. At acquisition, HA participants showed reduced activation in frontal brain regions, but at generalization, HA participants showed an increase in these frontal regions with stronger linear increase in activation with similarity to CS+ in HA when compared to LA participants. This indicates that there is a hyper-regulation in adolescents to compensate the higher difficulties at generalization in form of a compensatory mechanism, which decompensates with adulthood and/or may be collapsed in manifest AD. Additionally, significant developmental effects were found: the older the subjects the stronger the hippocampus and frontal activation with resemblance to CS+, which could explain the overgeneralization of younger children. Furthermore, there were differences according to sex: males showed stronger activation with resemblance to CS+ in the hippocampus and frontal regions when compared to females fitting again nicely with the observation that prevalence rates for AD are higher for females than males. In sum, the studies suggest that investigating developmental aspects of (maladaptive) overgeneralization may lead to better understanding of the mechanisms of manifest anxiety disorders, which could result in development and provision of prevention strategies. Although, there is need for further investigations, the present work gives some first hints for such approaches.
espanolLa interpretacion de espanol y castellano ha tenido casi siempre una tendencia sinonimica total a lo largo de su historia. Sin embargo, existen razones historicas y linguisticas para no considerarlo asi. Por un lado, la expansion del Imperio espanol parece ser la causa fundamental para el uso del primero de los terminos mientras que Norma linguistica sevillana serviria como prueba irrefutable de la validez del segundo, de una manera sistemica, y a pesar de que los propios sevillanistas normalmente se habrian manifestado en contra. Se presentan entonces una serie de argumentos contrapuestos que hacen especial hincapie en el marco normativo e historico del idioma y su descripcion geolectal en la actualidad. Con ello se pretende esclarecer diferencias entre los dos conceptos y, de paso, justificar tambien su significado como denominacion compuesta, espanol castellano. NB: Este articulo se basa en el texto “Del castellano al espanol y viceversa”, incluido en mi tesis doctoral (2017) y se redacta de acuerdo con “Opciones linguisticas avanzadas en clase ELE” (2016). EnglishThe interpretation of Spanish and Castilian has been almost meant to be synonymic in absolute terms in history. However, there exist linguistic and historic reasons in order not to state this. On the one hand, the expansion of the Spanish Empire seems to be the paramount cause for the use of the first term whereas Sevillian Linguistic Norm would prove successful in validating the second of them, in a systemic sense, and in spite of the fact that the selfsame sevillianists would have regularly claimed the opposite. Contrasted arguments are presented then, which put special emphasis on the historic and normative framework about this language and its geolectal description at present day. With this, it is intended to clarify some differences between the two concepts and, via that, to also justify their meaning as a compound definition, Castilian Spanish. NB: this article is based on the text “Del castellano al espanol y viceversa”, included in my doctoral thesis (2017) and it has been composed according to “Opciones linguisticas avanzadas en clase ELE” (2016).
Fibromyalgia syndrome (FMS), a common chronic pain condition, is often incompletely treated by conventional medical therapies. It can cause disability, psychological distress, work-related absenteeism, increased use of healthcare resources, and result in the inability to carry out the tasks of daily living. The purpose of this quantitative, correlational study was to investigate the potential influence of laughter on affect and pain in individuals with FMS. Laughter produces beneficial effects on acute pain and on chronic pain in general and has been found to improve temporary affective states, but there have been no studies testing the effects of laughter on the pain and affect of fibromyalgia patients. Informing this study were the gate control and neuromatrix theories of pain, as well as the dynamic model of affect theory. The research questions addressed whether laughter frequency is associated with affect and or with perceived chronic pain levels in these individuals. Forty-one adult fibromyalgia patients documented all laughter episodes daily and assessed their pain and affective states 3 times per day for 14 days. Hierarchical regressions revealed that increased overall laughter frequency was significantly associated with decreases in overall pain and increases in overall positive affect but was not associated with measures of negative affect. Also, morning laughter frequency was predictive of increased afternoon and evening positive affect ratings, as well as with decreased afternoon pain ratings, but was not significantly associated with evening pain ratings. The knowledge gained from these results may have positive social change implications at the individual level, within those individuals' larger social networks, and within the research and medical communities.
This study aimed to address the following questions regarding the emotional experience of Dialectical Behavior Therapy clients with Borderline Personality Disorder: 1) How do positive and negative emotions change in therapy? 2) Does the severity of clients’ symptoms relate to affect? 3) Is affect related to clients’ perceptions of therapeutic alliance? 4) How are clients’ and therapists’ affect related? To test these questions, positive and negative affect ratings were collected from clients (N=77) and therapists (N=25) at the start and end of session. These ratings were tested in relation to alliance and severity ratings using Hierarchical Linear Modeling. Results indicated that clients’ positive affect increased while negative affect decreased from the start to the end of session. This pattern was mirrored over the course of treatment, but only the increase in positive affect was statistically significant over that time period. Severity was significantly related to affect, but in an unexpected direction (higher ratings of emotion dysregulation were associated with slight decreases in negative emotion and emotion lability). Additionally, clients’ positive emotion significantly predicted therapeutic alliance ratings, and therapist positive affect was significantly, positively related to clients’ positive emotion. These results indicate that client affect appears to change in treatment and may be related to severity, alliance, and therapist affect. Further exploration is needed to clarify these complex relationships given the differences between positive and negative affect and the surprising direction of the association between negative affect and emotion dysregulation.
Demand-withdraw is an ineffective communication pattern frequently experienced by distressed couples. Therapists often attempt to address this pattern by helping partners understand and regulate the emotions that underlie these behaviors. To date, there is a lack of research focusing on the emotional experiences underlying the demand-withdraw pattern of interaction in couples. Related lines of research focus on emotional arousal and the expression of hard and soft emotions, but this research does not specifically investigate demand-withdraw interactions. The purpose of this study is to identify what emotions underlie demanding behavior in both men and women during marital demand-withdraw conflict interactions. Six couples were chosen from a five-year longitudinal randomized clinical trial that compared Integrative Behavioral Couple Therapy (IBCT) and Traditional Behavioral Couple Therapy (TBCT). Researchers viewed 10-minute pre-treatment problem-solving interactions to observe the demand-withdraw pattern in vivo among couples seeking therapy. The Behavioral Affective Rating Scale (BARS) was used to code the emotions observed during the interactions. The results indicated that the types of emotions varied not only depending on who initiated the problem-solving interaction (e.g., wife topic-husband topic) but also between the different couples, and when comparing gender. Anxiety (#2) and aggression (#4) were in the top four most commonly observed emotions for husbands, while they were two of the least observed emotions for wives. Moreover, frustration and hurt were the two most observed emotions for wives, while they were the least observed emotions for husbands.
Customers’ opinions on social network platforms are known to influence peer behaviour (Bai, 2011; Eirinaki, Pisal, &amp; Singh, 2012). Customers are also known to be more engaged in sharing their experiences by writing online reviews and recommendations that may be useful to others (Cantallops &amp; Salvi, 2014; Tang &amp; Guo, 2015; Xu &amp; Li, 2016). Actually, user-generated content (UGC) on social network platforms has emerged as an important source for understanding and managing consumers’ expectations, particularly using automated and semi-automated knowledge extraction techniques from text such as text mining and sentiment analysis (Zhang, Zeng, Li, Wang, &amp; Zuo, 2009). This research analyses dimensions of online customer engagement and associated concepts in customers’ reviews through (i) a global sentiment analysis using positive, neutral and negative sentiments and (ii) a topic-sentiment analysis to capture latent topics in online reviews. Furthermore, it examines what influences customers to contribute their online reviews, beyond the features of each focal company or brand. The research methodology is based on a text mining approach, using the MeaningCloud tool. The study focuses on Yelp.com reviews and includes a random sample of 15,000 unique reviews of restaurants, hotels and nightlife entertainment in eleven cities in the USA. An innovative customer engagement dictionary is created, based on previously validated scales using known dimensions of engagement, experience, emotions and brand advocacy, and extended using WordNet 2.1 lexical database. The research findings reveal a high impact of the engagement cognitive processing dimension and hedonic experience on customers’ review endeavour. The study results further indicate that customers seem to be more engaged in positively advocating a company/brand than the contrary. The findings will help social network managers to reinforce their platforms.
Freedom-of-movement in thought (the degree to which thought is constrained in its variety as opposed to being free to change) can be empirically dissociated from other well-studied dimensions of thought such as its task-unrelatedness in everyday life setting, but has yet to be studied in a controlled experimental environment. While there are several proposed mechanisms by which thought can become constrained (and therefore less freely moving), none have bene explored empirically. The present study set out to uncover which constructs associated with thought’s task-relatedness were also related to its freedom-of-movement and to test potential mechanisms of constraint. Motivation was within-subjects through a variable-value time-sensitive task and administered experience sampling probes asking participants to self-report the level of freely-moving thought, task-unrelated thought, deliberate control, arousal, and valence they were experiencing. Electrodermal activity and pupillometry were used as an index of physiological arousal in addition to self-reports. When participants were more highly motivated they reported having more constrained and more task-related thoughts, and having greater control over their thoughts. Control fully mediated motivation’s impact on freedom-of-movement of thought, but only partially mediated motivation’s impact on task-unrelated thought. Neither self-report or physiological measures of arousal were impacted by the manipulation, but high levels of both task-unrelated and freely-moving thought were associated with high ratings or self-reported arousal, higher pupillary responses to stimuli and smaller average pupil size, with freely-moving thought being uniquely associated with reduced skin conductance. Additionally, high freely-moving thought ratings were uniquely associated with slower responses, while high task-unrelated thought ratings were uniquely associated with low valence ratings. Overall, these findings support and further extend the previously identified dissociation between task-unrelatedness and freedom-of-movement as two separable dimensions of thought. They indicate that a person’s degree of control over their own thoughts is a crucial determinant of the content and especially the dynamics of that thought, but further work needs to be done to explore what nonconscious factors constrain thought movement.
A number of firms in Northern Europe and especially in Denmark are owned by private foundations similarly to what would have been the case if the Ford Foundation had owned a majority of the shares in Ford Motor Company. Foundation-owned companies appear to perform surprisingly well in terms of profitability and growth despite lacking governance mechanisms like profit incentives or takeover threats. Given their non-profit ownership, they might be expected to behave more responsibly towards stakeholders such as employees or customers (Hansmann 1980), but so far there has been little empirical evidence to support this hypothesis. This paper presents new research on the reputation and responsibility of foundation-owned companies. In a panel of large Danish companies 2001-2011 we find that foundation-owned firms have better reputations and are regarded as more socially responsible in corporate image ratings. Secondary evidence on labour market behaviour is consistent with these findings. Using matched employer-employee data we show that foundation-owned companies are more stable employers, pay their employees better and keep them for longer. Altogether, the evidence indicates that foundation-ownership is associated with more responsible business behaviour towards employees.
Listeners rate the speech of boys and girls as young as four years old as sounding gendered: boys are rated as sounding boy-like and girls as girl-like (Perry et al., J. Acoust. Soc. Am. [2001]). Recent research found that the extent to which boys’ speech sounds boy-like is correlated with measures of their gender identity and expression (Li et al., J. Phonetics [2016], Munson et al., J. Acoust. Soc. Am. [2015]). Munson et al. found that boys with a diagnosis of gender identity disorder [GID] were rated as sounding less boy-like than boys without GID. Munson et al.’s experiment used only a small number of girls’ productions as filler items. The current study examined listeners’ ratings of the gender typicality of speech of boys with GID and both boys and girls without GID. Significant differences in sex-typicality ratings were found between the two groups of boys. Boys with GID elicited ratings intermediate to those for boys and girls without GID. However, the differences between boys with and without GID were much smaller than those in Munson et al., suggesting that the sex distribution in the stimulus set can affect ratings of the sex typicality of children’s voices.
Users review about an app is a crucial component for open mobile application market, such as the AppStore and the Google play. Analyzing these reviews can reveal user's sentiment towards a feature in the app. There exist several analytical tools to summarize user reviews and extract meaningful sense out of them. However, these tools are still limited in terms of expressiveness and accurately classifying the reviews into more than a positive and a negative review. There is a need to get more insights from user app reviews and direct it to future app development. In this paper, we present our result of analyzing user reviews of 20 food journaling and health tracking apps. We gathered and analyzed reviews per app and classified them into three distinct categories using the sentiment treebank with recursive neural tensor network. We then analyzed the vocabulary frequency per category using the Gensim implementation of Word2Vec model. The analysis result clustered the reviews into good, bad and ugly feature reviews. Different usage patterns were detected from users review. We identified major reasons why users express a certain sentiment towards an app and learned how users' satisfaction or complaints was related to a specific feature. This research could be a guideline for app developers to follow when developing an app to refrain from adopting techniques that might demotivate (hinder) the application use or adopt those perceived positively by the users.
The article deals with the variant terms in normative aspect codified in Ukrainian art lexicography of the 21st century. Dictionary codification of variant terms indicates changing in the language and deliberate influence of the society on the development of terminological norm. Variation is a existence form of objects of the surrounding reality, in particular, of scientific concepts, which defines the laws of their function and interaction. The choice of sources is due to the fact that the selected dictionaries are represented modern art knowledge. Dictionaries play a significant role in the normalization of language, the spread of linguistic norms, and therefore they are a grateful and relevant material for the analysis of variation in the Ukrainian art terminology. The article focuses on the importance of the scientific philological study of art terminology – the field of knowledge, which is rapidly developing in modern conditions, acquiring new meanings and forms. The variant terms of the art terminology, codified in Ukrainian special vocabulary, are analyzed. Three types of variant terms, phonemic, derivational and morphological-phonemic units, are fixed in the Ukrainian art terminology. It was found out that among the reasons for the occurrence of phonemic variant terms of the analyzed terminology tends to facilitate articulation of the learned term; the appearance of derivation of variant terms is conditioned by the presence of various derivative models in the Ukrainian language and the search for forms of terms that correspond most closely to modern productive models of term derivation; functioning of morphological-phonemic variant terms is explained by different degrees of grammatical adaptation of foreign-language art terms. It also traces the effect of an analogy inherent to all three of the varieties mentioned. In general, the article discusses the essence of the problem of terminological variation as one of the most relevant processes in the regulation and standardization of the Ukrainian art terminology.
This era, in which we currently stand, is an era of public opinion and mass information. People from all around the globe are joined together through various information junctions to create a global community, where one thing from the far east reaches to the people of the far west within seconds. Nothing is hidden, everything and anything can be scrutinized to its core and through these global criticisms and mass discussions of gigantic magnitude, we have reached to the pinnacle of correct decisions and better choices. These pseudo social groups and data junctions have bombarded our society so much that they now hold the forelock of our opinions and sentiments, ergo, we reach out to these groups to achieve a better outcome. But, all this enormous data and all these opinions cannot be researched by a single person, hence, comes the need of sentiment analysis. In this paper we’ll try to accomplish this by creating a system that will enable us to fetch tweets from twitter and use those tweets against a lexical database which will create a training set and then compare it with the pre-fetched tweets. Through this we will be able to assign a polarity to all the tweets by means of which we can address them as negative, positive or neutral and this is the very foundation of sentiment analysis, so subtle yet so magnificent.
OBJECTIVE: To evaluate variables of tobacco health warnings associated with their emotional impact, the perception of smoking risks and the perceived effectiveness to avoid tobacco use. MATERIALS AND METHODS: Teenagers (151) and adults (168) evaluated 27 tobacco health warnings selected from the sets used on tobacco packages in Argentina and in other countries. A standardized affective rating-scale system and a structured questionnaire measured respectively the emotional impact (hedonic valence and emotional arousal), and the cognitive-behavioral attributions. The correlation between emotional and cognitive-behavioral evaluations was analyzed by age, sex, education level, smoker status,stage of quitting and susceptibility of non-smokers teenagers. RESULTS: Strong significant correlations between cognitivebehavioral and emotional assessments were observed. The warnings depicting graphic images of tobacco-related injuries and suffering were considered more valuable for tobacco. control, helping quitting and preventing initiation. CONCLUSIONS: Using graphic images with high emotional arousal is recommended for both adults and teenagers.
Recurrent Neural Networks (RNNs) play a major role in the field of sequential\nlearning, and have outperformed traditional algorithms on many benchmarks.\nTraining deep RNNs still remains a challenge, and most of the state-of-the-art\nmodels are structured with a transition depth of 2-4 layers. Recurrent Highway\nNetworks (RHNs) were introduced in order to tackle this issue. These have\nachieved state-of-the-art performance on a few benchmarks using a depth of 10\nlayers. However, the performance of this architecture suffers from a\nbottleneck, and ceases to improve when an attempt is made to add more layers.\nIn this work, we analyze the causes for this, and postulate that the main\nsource is the way that the information flows through time. We introduce a novel\nand simple variation for the RHN cell, called Highway State Gating (HSG), which\nallows adding more layers, while continuing to improve performance. By using a\ngating mechanism for the state, we allow the net to "choose" whether to pass\ninformation directly through time, or to gate it. This mechanism also allows\nthe gradient to back-propagate directly through time and, therefore, results in\na slightly faster convergence. We use the Penn Treebank (PTB) dataset as a\nplatform for empirical proof of concept. Empirical results show that the\nimprovement due to Highway State Gating is for all depths, and as the depth\nincreases, the improvement also increases.\n
We live in a society where the large majority of the population has a camera-equipped smartphone. In addition, hard drives and cloud storage are getting cheaper and cheaper, leading to a tremendous growth in stored personal photos. Unlike photo collections captured by a digital camera, which typically are pre-processed by the user who organizes them into event-related folders, smartphone pictures are automatically stored in the cloud. As a consequence, photo collections captured by a smartphone are highly unstructured and because smartphones are ubiquitous, they present a larger variability compared to pictures captured by a digital camera. To solve the need of organizing large smartphone photo collections automatically, we propose here a new methodology for hierarchical photo organization into topics and topic-related categories. Our approach successfully estimates latent topics in the pictures by applying probabilistic Latent Semantic Analysis, and automatically assigns a name to each topic by relying on a lexical database. Topic-related categories are then estimated by using a set of topic-specific Convolutional Neuronal Networks. To validate our approach, we ensemble and make public a large dataset of more than 8,000 smartphone pictures from 10 persons. Experimental results demonstrate better user satisfaction with respect to state of the art solutions in terms of organization.
We live in a society where the large majority of the population has a camera-equipped smartphone. In addition, hard drives and cloud storage are getting cheaper and cheaper, leading to a tremendous growth in stored personal photos. Unlike photo collections captured by a digital camera, which typically are pre-processed by the user who organizes them into event-related folders, smartphone pictures are automatically stored in the cloud. As a consequence, photo collections captured by a smartphone are highly unstructured and because smartphones are ubiquitous, they present a larger variability compared to pictures captured by a digital camera. To solve the need of organizing large smartphone photo collections automatically, we propose here a new methodology for hierarchical photo organization into topics and topic-related categories. Our approach successfully estimates latent topics in the pictures by applying probabilistic Latent Semantic Analysis, and automatically assigns a name to each topic by relying on a lexical database. Topic-related categories are then estimated by using a set of topic-specific Convolutional Neuronal Networks. To validate our approach, we ensemble and make public a large dataset of more than 8,000 smartphone pictures from 10 persons. Experimental results demonstrate better user satisfaction with respect to state of the art solutions in terms of organization.
Abstract The proposal presented in this study seeks to properly represent natural language to ontologies and vice-versa. Therefore, the semi-automatic creation of a lexical database in Brazilian Portuguese containing morphological, syntactic, and semantic information that can be read by machines was proposed, allowing the link between structured and unstructured data and its integration into an information retrieval model to improve precision. The results obtained demonstrated that the methodology can be used in the risco financeiro (financial risk) domain in Portuguese for the construction of an ontology and the lexical-semantic database and the proposal of a semantic information retrieval model. In order to evaluate the performance of the proposed model, documents containing the main definitions of the financial risk domain were selected and indexed with and without semantic annotation. To enable the comparison between the approaches, two databases were created based on the texts with the semantic annotations to represent the semantic search. The first one represents the traditional search and the second contained the index built based on the texts with the semantic annotations to represent the semantic search. The evaluation of the proposal was based on recall and precision. The queries submitted to the model showed that the semantic search outperforms the traditional search and validates the methodology used. Although more complex, the procedure proposed can be used in all kinds of domains.
This paper describes our system (SLT-Interactions) for the CoNLL 2018 shared task: Multilingual Parsing from Raw Text to Universal Dependencies. Our system performs three main tasks: word segmentation (only for few treebanks), POS tagging and parsing. While segmentation is learned separately, we use neural stacking for joint learning of POS tagging and parsing tasks. For all the tasks, we employ simple neural network architectures that rely on long short-term memory (LSTM) networks for learning task-dependent features. At the basis of our parser, we use an arc-standard algorithm with Swap action for general non-projective parsing. Additionally, we use neural stacking as a knowledge transfer mechanism for cross-domain parsing of low resource domains. Our system shows substantial gains against the UDPipe baseline, with an average improvement of 4.18% in LAS across all languages. Overall, we are placed at the 12 th position on the official test sets.
Tree-structured neural network architectures for sentence encoding draw inspiration from the approach to semantic composition generally seen in formal linguistics, and have shown empirical improvements over comparable sequence models by doing so. Moreover, adding multiplicative interaction terms to the composition functions in these models can yield significant further improvements. However, existing compositional approaches that adopt such a powerful composition function scale poorly, with parameter counts exploding as model dimension or vocabulary size grows. We introduce the Lifted Matrix-Space model, which uses a global transformation to map vector word embeddings to matrices, which can then be composed via an operation based on matrix-matrix multiplication. Its composition function effectively transmits a larger number of activations across layers with relatively few model parameters. We evaluate our model on the Stanford NLI corpus, the Multi-Genre NLI corpus, and the Stanford Sentiment Treebank and find that it consistently outperforms TreeLSTM
In this paper we present the linguistic databases developed during our 8-year lexicographic research on the Modern Greek Standard (MGS) verbal system. Apart from the intermediate databases presented, the main products are (a) a new conjugation system of 385 paradigmatic models, which allows for the automatic generation of all verbal lexical morphemes and monolexical forms (b) a statistically established database of 151,536 distinctive verb-final grapheme sequences which allow for the automatic tagging of all monolexical verbal tokens without the traditional intervention of any built-in lexicon, and (c) a linear Iemmatisation morphophonological rule system accessed on the basis of the distinctive grapheme sequences identified.
Detecting lexical entailment plays a fundamental role in a variety of natural language processing tasks and is key to language understanding. Unsupervised methods still play an important role due to the lack of coverage of lexical databases in some domains and languages. Most of the previous approaches were either based on statistical hypothesis of specific entailment relations or tried to encode word relations in low-dimensional vector embeddings. This thesis builds upon one of the few approaches which intrinsically model entailment in a vector space. We then further generalize this model by introducing an alternative, distributional representations for words which harnesses tools from optimal transport to define distance or entailment measures between such representations. We evaluated the models on hypernymy detection where our distributional estimate significantly improves over the underlying model and even outperforms state-of-the-art on some datasets.
Syntactic parsing plays a crucial role in improving the quality of natural language processing tasks. Although there have been several research projects on syntactic parsing in Vietnamese, the parsing quality has been far inferior than those reported in major languages, such as English and Chinese. In this work, we evaluated representative constituency parsing models on a Vietnamese Treebank to look for the most suitable parsing method for Vietnamese. We then combined the advantages of automatic and manual analysis to investigate errors produced by the experimented parsers and find the reasons for them. Our analysis focused on three possible sources of parsing errors, namely limited training data, part-of-speech (POS) tagging errors, and ambiguous constructions. As a result, we found that the last two sources, which frequently appear in Vietnamese text, significantly attributed to the poor performance of Vietnamese parsing.
Shi, Huang, and Lee (2017) obtained state-of-the-art results for English and Chinese dependency parsing by combining dynamic-programming implementations of transition-based dependency parsers with a minimal set of bidirectional LSTM features. However, their results were limited to projective parsing. In this paper, we extend their approach to support non-projectivity by providing the first practical implementation of the MH_4 algorithm, an $O(n^4)$ mildly nonprojective dynamic-programming parser with very high coverage on non-projective treebanks. To make MH_4 compatible with minimal transition-based feature sets, we introduce a transition-based interpretation of it in which parser items are mapped to sequences of transitions. We thus obtain the first implementation of global decoding for non-projective transition-based parsing, and demonstrate empirically that it is more effective than its projective counterpart in parsing a number of highly non-projective languages
It is no secret that people often use taboo words when speaking about persons and objects in their environment. Taboo words are charged with emotion and have observable impact on the listener as well as the speaker. The purpose of this study was to determine whether taboo words were quantitatively more offensive when used in combination with a proper name versus being used with a non-human object. We found that using taboo words to describe proper names does not cause a significant effect; however, we found that participants rated certain categories of taboo words as more offensive than other categories. In a second experiment, taboo words did affect ratings and memory for proper names and non-human objects.
The statistical parsing of morphologically rich languages is hindered by the inability of parsers to collect solid statistics because of the large number of word types in such languages. There are however two separate but connected problems, reducing data sparsity of known words and handling rare and unknown words. Methods for tackling one problem may inadvertently negatively impact methods to handle the other. We perform a tightly controlled set of experiments to reduce data sparsity through class-based representations in combination with unknown word signatures with two PCFG-LA parsers that handle rare and unknown words differently on the German TiGer treebank. We demonstrate that methods that have improved results for other languages do not transfer directly to German, and that we can obtain better results using a simplistic model rather than a more generalized model for rare and unknown word handling.
The focus of the current study was on idiom comprehension in younger and older adults. Due to inconsistent results in previous studies, it is unclear whether older adults may have problems understanding idioms. For the current study, I used a sentence-to-word matching task presented on an iPad with software that recorded participants’ response time and accuracy. Participants also completed a familiarity task where they rated idioms on how frequently these phrases were encountered. I predicted that older adults would have more difficulty comprehending idioms because of the context in which the idioms were embedded and the timed nature of the task. I also predicted that both age groups would rate the idioms as highly familiar because we purposefully selected these types of expressions. With respect to the sentence-to-word matching task, results showed that although older adults were slower overall, both younger and older adults showed faster response times and greater accuracy for idiomatic targets following idiomatically-biased contexts than for literal targets following literally-biased contexts. With respect to the familiarity ratings task, results showed that both age groups were very familiar with the idioms. These findings suggest that older adults are able to successfully use context to understand familiar ambiguous idioms and that they do not have difficulty comprehending idioms in a cognitively demanding timed task.
Abstract People remember events and materials better when these are congruent with their mood at retrieval; this is known as the mood-congruent memory bias. This effect is largest when the materials are self-referential and this is known as the self-reference effect. We present two word rating studies, to create a list of self-referential valenced words that may be used as stimuli to investigate the influence of valence on cognitive processing in depressive ruminators. Words selected from the Affective Norms for English Words pool were rated by an unselected sample for self-referentiality (Study 1) and validated with ratings provided by depressive ruminators. As hypothesized, depressive ruminators rated negative words as more self-referential than an unselected sample. Using this list, valence differentiated performance between depressive ruminators and healthy controls in a working memory updating task. We thus created a list of self-referential valenced words matched on factors that influence word processing.
This study aims at exploring new norms as to the textual additions in parentheses (=TAiPs) in the translation of a Quranic text as writer-oriented devices of textuality. Coding for this sort of information could be useful in establishing an impact on any decision-making process on the TL version; such TAiPs can give a translated text of the Quran unity and purpose and distinguish it from a disconnected sequence of sentences. Six small-sized chapters of the Quran were selected as a research sample including a number of four handred forty two (442) TAiPs. Two writer-oriented kinds of textuality were found: cohesivity at the levels of grammar and lexis to be in form of recurrence, reference, substitution, ellipsis and conjunction; and relationality by coherence and intentionality to be in form of reiteration, collocation, connotation, evocation and interpretation. The study is a detailed analysis of such a severely criticized yet officially approved English interpretation of the Quran as the Hilali and Khan Translation (=HKT) against a predetermined set of text-linguistic norms. The strength or weakness of TAiPs as to how they might alleviate or aggravate the TL version is eventually identified for sake of improvement.
The first edition of one of the most important and mysterious novels of the 20th century appeared more than fifty years ago. Despite the passage of time The Master and Margarita still enjoys popularity; it also intrigues and inspires. Until now five Polish translations of Bulgakov’s novel have appeared. It is known that the interpretation of the original might be expressed in the form of many potential texts that are communicatively equivalent. There is no doubt that it is the translator who plays a vital role in any translation; her/his personality, life experience, knowledge, skills, and also the times s/he lives in regulate the target text. That is why, no matter how many times a text is translated, the final product will always be different. Taking this into consideration, the author will compare the three Polish translations of Bulgakov’s Master and Margarita, paying attention to the diachronic perspective as far as linguistic norms are concerned, the modernity of language, and the way the anthroponyms are expressed.
The repertoire of forms of address can be considered as one of the determinants of the discourse genre, which makes it possible to capture its evolution and cultural variations. From such comparative, intra- and intercultural perspective, adopting an interactive approach in the analysis of political discourse, we will look at the practice of addressing one another in the French and Polish politicalmedia discourse. While in both languages the linguistic norm recommends the use of the polite forms of address in official situations, the cases of the use of the familiar pronoun tu / ty in media interactions between politicians are not rare at all. Whether it is an informal talk of politicians caught by the media, a television pre-election debate, or a meeting of the heads of state, addressing the other person by the familiar forms is a manifestation of a deliberate blurring of the boundaries between the front-stage and backstage in political discourse in order to create the impression of intimacy andequality between the interlocutors.