Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
In werkwoordsgroepen met een vervangende infinitief (IPP) wordt in het Nederlands de keuze van het hulpwerkwoord van de voltooide tijd doorgaans bepaald door de IPP, zoals in heeft kunnen komen, maar de keuze kan ook door het hoofdwerkwoord bepaald worden, zoals in is kunnen komen. Gebruik makend van een aantal treebanks en corpora (CGN, Lassy, SoNaR) hebben we onderzocht welke IPP’s deze alternantie vertonen. Dat blijken er naast kunnen nog minstens 13 andere te zijn. Voor de twee meest frequente (moeten en kunnen) hebben we vervolgens nagegaan wat de verhouding is van de voorkomens met hebben en zijn in die gevallen waarin de alternantie mogelijk is. Daarbij is gebleken dat in gemiddeld 80% van de gevallen de keuze van het hulpwerkwoord door de IPP wordt bepaald. Een belangrijke factor bij de keuze is de aard van het hoofdwerkwoord.
Abstract We present a work in progress aimed at extracting translation pairs of source and target dependency treelets to be used in a dependency-based machine translation system. We introduce a novel unsupervised method for parallel tree segmentation based on Gibbs sampling. Using the data from a Czech-English parallel treebank, we show that the procedure converges to a dictionary containing reasonably sized treelets; in some cases, the segmentation seems to have interesting linguistic interpretations.
It is difficult to square the rhetoric about the current “crisis” in the humanities with the abundant, if anecdotal, evidence that Greco-Roman antiquity continues to thrive in the popular imagination. As I am writing this, Mary Beard's new history of Rome is flying off the shelves; general interest magazines publish articles on Greek papyri; the first translation of Homer's Iliad by a woman has appeared to wide acclaim; the challenge of teaching ancient Greek made it to the op-ed pages of The New York Times; a remake of the film Ben-Hur is scheduled for release this summer; a traveling exhibition of large-scale Hellenistic bronzes has become a “must see” show of the season; productions of Greek tragedies and their adaptations continue to be a staple of professional and amateur theater; and television programs abound on ancient topics ranging from Cleopatra to the Colosseum.2 Of course, this preoccupation with the past has a negative side as well, since even the modern attempt to mythologize Zenobia as an Arab queen who resisted Roman power was not enough to save her city Palymra from those in Syria who were hell-bent on erasing any signs of what they deemed to be unorthodox. But even such wanton acts of destruction, which seek to obliterate history, only provide further proof that the past is still very much alive in the present.3That said, there are different ways to assess the health of a field than by measuring popular interest in the objects of its study.4 These signs of robust interest–of a fascination fueled perhaps by the way in which Greek and Roman culture is simultaneously familiar and foreign to us–do not tell the whole story. If we turn instead to data usefully amassed by the Humanities Indicators of the American Academy of Arts and Sciences, and by other professional sources, we get a somewhat different picture at the institutional level–small (though relatively steady) numbers of students majoring in classics, respectable enrollments in Greek and Latin (though modest by comparison with many modern languages), and some retrenchment in faculty hiring (though it is not across-the-board and is offset by hiring in other schools and colleges).5Even more striking, and encouraging, is the fact that, as the number of individuals specializing in the field has shrunk, more students than ever before are encountering Greece and Rome through courses on “classics in translation.” A staple of undergraduate general education programs (whether distributional or core requirements) and popular as electives, these courses explore such topics as “Classical Mythology,”“Women in Antiquity,” “Sport and Spectacle in the Ancient World,” “Ancient Religion,” “Greek and Roman Drama,” and “Cinema and the Classics”–to name just a few. Rather than “dumbing down” the field, as some critics have claimed, and being harbingers of further decline, these courses have succeeded in educating a whole new generation of citizens, hardly an unworthy goal. They have also helped to recruit new majors who had not encountered this material before college. And they have even supplied a modest pipeline into the profession, as some of those latecomers to the field, upon graduation, make up for gaps in their linguistic training by enrolling in post-baccalaureate programs, yet another creative adaptation by which the field prepares students for entry into doctoral programs and scholarly and teaching careers.The visibility of antiquity in the curriculum testifies to the resilience of the field in the face of “crisis”–or, rather, “crises.” Greco-Roman studies has long been recognized as the canary in the coal mine of the humanities, having faced early on some of the pressures that the other humanities would encounter only later. In the late nineteenth and early twentieth centuries, the field lost its curricular hegemony, as American colleges and universities jettisoned Latin as a requirement for admission or graduation. Then, as private schools, particularly Catholic ones, made Latin optional or dropped it altogether, one important pipeline for college majors dried up. Later, as the quintessential home of “dead white males,” the field was at the epicenter of the culture wars.6 And, now, in a climate of economic anxiety, vocationalism, and concern with financial return on educational investment, it is again vulnerable. Rather than circling the wagons, the field has confronted these challenges in creative ways. The curricular engagement noted above was one of these strategies. In fact, in a reversal of the usual model whereby research influences what is taught in the classroom, this curriculum also became a powerful driver (though by no means the only one) of exciting new research agendas that focus on contemporary issues where the past has something to teach us.And so, if ancient Greco-Roman culture is alive and well in the popular imagination and in the general curriculum, the most important evidence of its vitality must nevertheless be sought in the quality of current research. While the past several decades may have seen no grand paradigm shift,7 it is clear that our understanding of the past has been dramatically enhanced–and in some cases radically altered–by new evidence, new methods, and new questions. As befits a scholarly field whose history began to be written even in antiquity, it is not surprising that there are periodic moments of taking stock. The year 2000 occasioned several, including Classics in Progress, a volume of essays by British scholars that was published for our sister society, the British Academy.8 This special issue of Dædalus was inspired by a different sort of milestone, the important work of the American Academy's Commission on the Humanities and Social Sciences. The idea for this issue started to come into view at the same time that the Commission was preparing its report, The Heart of the Matter; and the appearance of this issue coincides roughly with the publication of the Commission's follow-up report, which documents the extensive activities that have taken place over the past two years.9 There could be no better time to focus on the oldest of the humanities fields, Greco-Roman studies, and to assess (in the words of this volume's title) “what is new about the old.”10Taken together, the essays in this volume exemplify some of the most important recent developments in Greco-Roman studies. Here I would single out only four. The first is, paradoxically, the persistence of the old amidst the new–the continued focus on the text, whether literary or documentary, and hence the continued importance of philology and the traditional specialisms necessary for recovering and recuperating this category of evidence, such as palaeography, textual criticism, and linguistics. It is sometimes assumed that the vagaries of transmission have left us all that we will ever have of ancient literature–a minute percentage of the total production, to be sure, but more than any one person could read in many lifetimes. But new material regularly turns up, whether in a manuscript miscatalogued in a monastic library, or in a “quotation fragment” (the work of one author cited by another), or, more commonly, on a scrap of papyrus recovered from the dry and preservative sands of Egypt.11 Indeed, one scholar estimates that “Over the last hundred years, one literary papyrus has been published, on average, every ten days; the agglomeration provides, for Greek literature at least, a small new renaissance.”12 (For a recent discovery that has attracted much attention, see the elegant translation by Rachel Hadas of the so-called “Brothers Poem” by Sappho in the box on page 40.)13These discoveries not only enlarge our store of ancient literature, but also enable us to restore what we already have, to recognize previously unknown connections among works, and, on occasion, to rewrite history, literary or otherwise. Meanwhile, extant texts regularly require philological attention. To take just one example: new editions of authors are needed not only to incorporate the new discoveries noted above, but also to take into account several phenomena, only recently understood. One is contaminatio, the fact that most family trees of manuscripts (stemmata codicum) are complicated by horizontal transmission (the cross-fertilization of distinct traditions, when a copyist relying mainly on one manuscript nevertheless incorporates readings from another with a different lineage). Another is even more basic: the that in an where texts were or from the In other there may be no one And just as new editions the new and studies provide their and on the of the research. In fact, a is which not just the words upon a but also the of the text, including the of the ancient (the papyrus and and its not only for textual criticism, but also for ancient in the perhaps the most since it has been for over Greco-Roman studies has up dramatically in of its and This is sometimes as the of other But this which the of a more complicated Greco-Roman studies had been even to this their in the other humanities, not only scholars of and literature but also ancient and In fact, most of these their to the of antiquity, In the and early twentieth these became from their and started to different The was that scholars of Greco-Roman antiquity as a and, over became more from developments in the that they had but that had in different several decades as a if in that Greco-Roman studies was not much a single as a field, and scholars started to take out with their the work of ancient literary and began to be by the and of those (For an elegant see the box on page where of a from traditional philological also being by contemporary literary such as and And of this was a since scholars of the ancient in with their and made to particularly in such as the history of and In an even more scholars who were these began also to as individuals the sort of that had previously in the of their or the as a literary scholars the texts they were ancient history and history a and the same these scholars also upon other that had their of the field, such as and and And through they began to in such as and (the having had a particularly and in the recent of the all of this scholarly no one or has even for a and a of The has been that a field seen by some as more has become much more have an of the at work in the very of the This has to the ancient material and sometimes with a and a of It has also to the that not only past scholars but also they to the evidence they and the they there is of the of in in Greco-Roman studies is the most recent and perhaps the most the new of A of this the for in our of the such as and have been for a long But these have been by other powerful for is from that had been in the of or is us to ancient and and the of such as and us to and and ancient and of some of these and their in the box on page are not just to objects or but to tell a story. The data recovered in this way an that it even to old to and to rewrite to these developments is the important by Greco-Roman studies has been with of with the from to and from the to the the field was an early the of what has come to be as humanities, and it has been a to that field ever one has to evidence, as the of texts and has made research on a previously and has up whole new of But at another not only to evidence but also powerful for ranging from of to the of Greek and Latin texts (the linguistic of every in a and is the of the The for has been dramatically not just by new but also as a of the new noted Greek and Roman at the of are taking of their to of ancient and understanding of the Ancient such topics as and are also being And that the literature of the Hellenistic is in the scholars are their to the the Greek literature of the Roman and the literature of early the the and of the of for are as traditional and as and are seen to be over and of and that not with the the focus on Greece and Rome has way to studies of the and the ancient that recognize the of their at different And even where there is evidence of history for those who work in the Greco-Roman field to explore that one culture or The current interest in or is an of this as is the of a new field, ancient studies, which as its this sort of of of and Greco-Roman studies is being the of as the of a or material is to be a not only of the and in which it was and but also of other and have the of the Greco-Roman in all its is no just the of the As Mary and have is a that in that us and the of the and The by Classics are the by our from and at the same time by our to and by its to The of Classics is not only to or the ancient is also to and our to that And to that one Greco-Roman studies into the square and to the of to and to the past is to our contemporary persistence of the to new and the new of antiquity, and the of developments in Greco-Roman studies over the past several decades are on in the essays that this a are in just for the of the field, I must that many important are from this of and was the could it and the were the to their as they the same these essays are not general or of the of research. While most their work in the of recent they their essays to be studies that new in and, in some in new the is this volume from literature to and material ancient history, and, the institutional in which Greco-Roman studies are Of course, this the among these and also among the essays which a of and This is all the more since the not with one another or in other ways. But this only to the of this as noted of the field, the of of and the of The that are to something that the to the articles could not to out some of these connections and also to a since these when read come to a about “what is new about the that the on texts is of the field, the first essays in this volume the past several have left their on literary including not the and the or In to readings of current also a wide of including the of the noted and, its as the of ancient the and in which texts were and as and and more of and name just a on Greek literature, that category and its scholarly have been as the traditional has the of some of these different To take one example: to texts their and on the other studies to the of texts and about their as her Greek the in which the these two is perhaps most a of the that Greek at a the when of and were also for the of to these in contemporary is about but it is also about the and of as the power to in the as well as to the that one to in ways that are difficult to by on Latin literature, its with its Greek the literature of a Rome had The of this and the understanding of it has been a staple of But of have way to an of the creative of that were at work in one literature to this further by linguistic on Roman about their a of the could to linguistic translation Greek texts into and whose for Greek culture into that could be as a over an text, but it may have to the of any from two essays in different ways the of texts has into the of Greco-Roman studies. The by this on the of the Greek and Roman classics, the recent from a model that a whose be through by the ancient texts as for this and a horizontal one in which these texts are that from and encountering and of these As a on and on two different to in which the the of the in this it a of this of essays on literature, turns to one of had of the idea in Roman antiquity, interest is in the contemporary of which has made Greco-Roman texts to students and the While translation studies has recently as its her focus is not on or criticism, but on as befits one who has just published her translation of the not much a scholarly as a a and the of the that are there in of her of literature, the volume a to more in antiquity than the as the of the to its the current the of ancient and the way that is that is, or and modern with its its powerful its interest in and its linguistic to the that ancient modern and that ancient in the way they take on that have out of contemporary And there are recent signs of the may be its on the field, of may be more to ancient and, at a time of some are that texts of the past a place where one again about some of the traditional issues of in a more In ancient texts one again to see the for the A is the of where there has been a creative engagement the old and the The was a one in antiquity most the view that is not an and contemporary are and with that to by the that the ancient will continue to us the challenges we and that they will also teach contemporary to to those and in ways that we two essays our from ancient literature and to and material In a way that is familiar from of ancient the of Greco-Roman has been the of history, which on the modern and history has also to its to philology and the engagement with to these challenges are familiar from the of literature in this One is to focus on on and ways in which Greco-Roman has continues to and and and Another to the in their this is an in which of by contemporary their focus on of and, an important to there is to the and text, which is to the in literary studies. by a from the which about the with of and and by also on material objects and but of a different the written that an important for research. These texts on and and most of from two The first is a new interest in these were texts and of make it to the of whereby material of writing and writing have come to into the of the of writing in ancient The a and on the in which the written were and what that about different in their and that the two are the that and are This material focus a in from the literary and philological of a generation history in a have from being only in the of a new of Sappho to to who was and has in the of the of the ancient in their and from This is, the of But if it were the of an that would be no essays turn our to ancient In recent years, has traditional and history to also and has from in to the of and including and and to such topics as and and on one of these the of and on the sort of evidence that has as the city of in which for a long time and has different of and and with one another the they were being and by and and they were in which were to the and transmission of and only the of this but even its a as when an is as a its in having by become It is that about not the late antiquity, when and and all other of the most of the of this is in a the of the of as the of the by a very different category of evidence, not just textual and but also a of ancient history, the of that had not in past that Rome was an and its was of a of from in to the of and that the of and other But if and the the for no we on the of evidence, that climate also that the of had been for much of the In the the which has as was through the the very that the same in the of the of what has as a in and again in the to a in that had the And a the of started in in and the Roman over the The of the was not as the of any one but instead to a of that was to climate and and that in a of the These the the of two essays that explore in ancient history, and on the of the They two different of ancient model that Greece and Rome as the that since they were in history, and the which is in its and to the of The have and for two hundred and years, with the model taking in the and the the But as evidence and are than ever the is in the the that to most began not in Greece and but with the of in the more than ten or the in of modern more than one hundred or of the But if the model most of the history, the model has its much of what the and the that is, much of The authors an way of ancient history, which is and and first is the the of the first when of at roughly the same time in different and much evidence of The is the of Rome and for but they had very different and their be only by The and the of two for research in which and work But to this, to new evidence, methods, and and recognize that the ancient was much ancient history much our made last two essays in this volume return to a that was at the of this the institutional and professional of Greco-Roman studies. But the now, is on the to curriculum and the of from its in American education in the of the of universities at that including the of whose from of into an of the an to contemporary where more have a than ever and where studies. a for the in this by that the of the field, the way it different of and is to the as a But these different are in one curriculum, the of and to the of a different that the are the same challenges that faced decades that the make the for research by of the past through popular and which a and make our teaching a is also to having Greco-Roman studies as a But is the power of research and teaching be by upon in this but to which the from to a our to as a from the Academy's Humanities Indicators that the humanities to the general is not by scholars in this and other humanities But this the humanities for the for the of research ways to this One is to which is by no means for this more important is to come up with new for Greco-Roman studies in a one which not the with of the but to including in other it for such to and make to is that Greco-Roman studies in a to up not only to different and but also to of and One traditional name for the field, the fact that there are many other and than those of Greece and institutional with scholars of and of to work with the of in and Greek and is to the sort of ancient history that and and is with the out of the field noted While not all may about the or of some of these not only as a to this but also as a to the essays that
The neural networks associated with socio-affective (empathy, compassion) and socio-cognitive processes (mentalizing/Theory of Mind) have been well-characterized over the last years. The goal of the present talk is twofold: (1) To explore the separability of these functions during online social understanding on a subjective, behavioral and on a neural level and (2) to investigate the embedding of the related neural substrates in large-scale task-free neural networks. To this end, we acquired resting state as well as behavioral and neuroimaging data (fMRI) during a social video task in a large sample of participants (N = 178). The videos were short autobiographical narrations of emotionally negative and neutral events that allowed for asking Theory of Mind questions about the thoughts of the narrators and factual reasoning questions about the content of the stories, thereby allowing for independent assessment of socio-affective and socio-cognitive processing. Linking the phenomenological with the neural level, participants reported increased negative affect after emotional stories, which covaried with activity strength in the meta-analytically defined “empathy network”, but not with activity in the “Theory of Mind network”. Vice versa, performance in answering the Theory of Mind questions correlated with “Theory of Mind network”, but not “empathy network” activity. Interestingly, neither behavioral markers of social affect and mentalizing (i.e. emotional valence ratings and Theory of Mind performance) nor activity in the two respective neural networks correlated with each other. Furthermore, resting state functional connectivity to task activation based seed regions for empathy and Theory of Mind yielded distinct networks that strongly overlapped with the respective task activations and correspond to the well-described default mode network (Theory of Mind seeds) and the salience or central executive network (empathy). The data strongly argue for dissociable and independent socio-affective and -cognitive functions that are embedded in large-scale task-unrelated neural circuits.
Sentiment lexicons with valence-arousal ratings are useful resources for the development of dimensional sentiment applications. In order to solve the significant lack of Chinese valence and arousal lexicons, the objective of the DSAW is to automatically acquire the valence-arousal ratings of Chinese affective words. In this task, we develop a novel approach that integrate word embeddings into a graph-based model with K-Nearest Neighbor to identify both valence and arousal dimensions. We also propose to use character embeddings to represent unseen words, which is a major challenge in collecting large corpora. The evaluation results demonstrate that our system is effective in dimensional sentiment analysis for Chinese words with 0.847 and 1.281 mean absolute error (MAE) for valence and arousal respectively.
The present study compared the effects of: (a) PETTLEP imagery (e.g. imaging in the environment), (b) prior-observation (i.e. observing prior to imaging), and (c) traditional imagery (e.g. imaging sat in a quiet room) on the ease and vividness of external visual imagery (EVI), internal visual imagery (IVI), and kinaesthetic imagery (KI) of movements. Fifty-two participants (28 female, 24 male, Mage = 19.60 years, SD = 1.59) imaged the movements described in the Vividness of Movement Imagery Questionnaire-2 under the three conditions in a counterbalanced order. Vividness and ease of imaging ratings were recorded for each movement. A repeated measure MANOVA revealed that ease and vividness ratings for EVI, IVI, and KI were higher during the PETTLEP imagery condition compared to the traditional imagery condition, and vividness of EVI was higher during the observation imagery condition compared to traditional imagery. Findings indicate that incorporating PETTLEP elements into the imagery instructions leads to easier and more vivid movement EVI, IVI, and KI imagery.
PURPOSE: The focus of this study was to examine the influence of fundamental frequency (F0) and vocal tract length (VTL) modifications on speaker gender recognition in cochlear implant (CI) recipients for different stimulus types. METHOD: Single words and sentences were manipulated using isolated or combined F0 and VTL cues. Using an 11-point rating scale, CI recipients and listeners with normal hearing rated the maleness/femaleness of the corresponding voice. RESULTS: Speaker gender ratings for combined F0 and VTL modifications were similar across all stimulus types in both CI recipients and listeners with normal hearing, although the CI recipients showed a somewhat larger ambiguity. In contrast to listeners with normal hearing, F0-VTL and F0-only modifications revealed similar ratings in the CI recipients when using words as stimuli. However, when sentences were used, a difference was found between F0-VTL-based and F0-based ratings. Modifying VTL cues alone did not affect ratings in the CI group. CONCLUSIONS: Whereas speaker gender ratings by listeners with normal hearing relied on combined VTL and F0 cues, CI recipients made only limited use of VTL cues, which might be one reason behind problems with identifying the speaker on the basis of voice. However, use of the voice cues depended on stimulus type, with the greater information in sentences allowing a more detailed analysis than single words in both listener groups.
It is well recognised in psychology that music has affective connotations and that musical stimuli can modify affective states. The aim of this study was to assess the affective connotations of 120 fifteen-second musical excerpts, covering both modern musical genres such as pop, rock, jazz, rap/R&B and electronic music (5 x N = 20), and classical music ( N = 20). Expert judges used predetermined criteria to select excerpts with positive or negative valence that induced high arousal or low arousal. The excerpts were assessed by 50 undergraduate students (25 women) from different academic departments, aged between 18 and 28 years ( M = 21.46 years, SD = 1.85). They listened to all 120 fragments and rated them with respect to six dimensions: valence, arousal, dominance, origin, subjective significance and imageability. Analyses showed that ratings were reliable, with high split-half correlations and Cronbach’s alpha estimates. We did not identify any gender differences concerning affective reactions to the music. Some music genre specificity was found for all measures, and initial music preference appeared to shape affective ratings. The results presented here will be of interest to researchers working on musical perception and the influence of music on affective outcomes and emotional regulation.
We train a language-universal dependency parser on a multilingual collection of treebanks. The parsing model uses multilingual word embeddings alongside learned and specified typological information, enabling generalization based on linguistic universals and based on typological similarities. We evaluate our parser's performance on languages in the training set as well as on the unsupervised scenario where the target language has no trees in the training data, and find that multilingual training outperforms standard supervised training on a single language, and that generalization to unseen languages is competitive with existing model-transfer approaches.
The Prague Czech-English Dependency Treebank 2.0 Coref (PCEDT 2.0 Coref) is a parallel treebank building upon the original PCEDT 2.0 release and enriching it with the extended manual annotation of coreference, as well as with an improved automatic annotation of the coreferential expression alignment.
This article attempts to place dependency annotation options on a solid theoretical and applied footing. By verifying the validity of some basic choices of the current dependency reference framework, Universal Dependencies (UD), in a perspective of general annotation principles, we show how some choices can lead to inconsistencies and discontinuities, partly due to UD's alternation between syntax and semantics. For some constructions, we propose better suited alternative structures with a clear-cut distinction of syntax and semantics. We propose a classification of conception-oriented, annotatororiented, and finally, treebank end-useroriented considerations to be used in the creation of new annotation schemes.
In this paper, a German verb resource for verb-centered sentiment inference is introduced and evaluated. Our model specifies verb polarity frames that capture the polarity effects on the fillers of the verb's arguments given a sentence with that verb frame. Verb signatures and selectional restrictions are also part of the model. An algorithm to apply the verb resource to treebank sentences and the results of our first evaluation are discussed.
Interpersonal forgiveness is a burgeoning area of research in psychology and has been linked to lower levels of depression and perseverative cognitive states such as rumination.As much of the extant research employs self-reported assessments of forgiveness, the aims of the present work are to test a novel operational definition of forgiveness using behavioral outcomes from economic games-specifically, the Ultimatum Game (UG) and Dictator Game (DG)-and to explore how such behavior corresponds with phasic heart rate (HR), heart rate variability (HRV), individual differences, and psychosocial variables.Participants (n = 89; age M = 19 years; 46% female) were instrumented with continuous electrocardiogram and were seated for a 5minute resting baseline, a pre-post adaptation of the DG and UG with digital opponents, randomization to forgiveness or rumination imagery, and a 5-minute recovery period.Participants reported affective ratings as a manipulation check as well as questionnaires on state and trait forgiveness, hostility, and rumination.Forgiving behavior was operationalized as more generous monetary offers to previously unfair, provoking opponents (positive value for post-minus pre-manipulation DG offer).As hypothesized, individuals who imagined forgiving previously unfair, provoking opponents showed more behavioral forgiveness in their return offers and reported less negative affect
Using a multiple study format, the current research sought to examine how the combination of facial and vocal cues would affect assessments for short-term and long-term mating of Caucasian female faces of varying attractiveness levels. It was hypothesized for study 1 that a high-pitched voice would be matched with an attractive face and that a low pitched voice would be matched with an unattractive face. The results from study 1 were consistent with the hypothesis, finding that participants did match a high-pitched voice with an attractive female face and a low-pitched voice with an unattractive face. Study 2 sought to determine how combinations of neutral faces of varying attractiveness level and voice pitches would affect ratings of evolutionarily significant traits. It was hypothesized that when faces and voices were congruent that ratings would differ on a variety of traits as compared to when the faces and voices are incongruent (based off of results from study 1 of matching). The results indicated that the traits that were most affected by these combinations for both the attractive and unattractive face when being evaluated as both short and long-term mates were warmth, masculinity, femininity, nurturance, friendliness, and lovingness. Study 3 hypothesized that a smiling expression would alter the perceptions of traits from study 2. The hypotheses were only partially supported, finding some traits were impacted by expression type. These results suggest that people rely more on bimodal cues associated with facial attractiveness and vocal pitch more than cues that apparent from a smile.
The present study assessed if children would present different information in drawings of emotion eliciting stimuli when drawing for an adult or a child audience. This question is important as children’s drawings of emotive topics are used in interview and diagnostic settings often without reference to whether children alter their graphic communication depending on the audience of their drawings. Ninety‐five 6‐year‐olds (54 boys and41 girls) were allocated to three groups: the reference group, the child audience group and the adult audience group. The reference group were not informed of an audience and the child and adult audience groups received audience appropriate instructions. All children completed a drawing session where they first drew a neutral uncharacterised figure, followed by drawings of a sad and a happy figure in counterbalanced order. Affect ratings towards the drawn topics were taken immediately after completion of each drawing. \nThe findings demonstrated that children did consider who would be viewing their drawings when communicating affective information and included different features for different audiences within their drawings. The study showed that children altered positive and negative affective elements of their drawings depending upon whether they were drawing for a peer or an adult audience highlighting the need to understand the role of the audience when interpreting children’s drawings for emotional information about the child artists. It is posited that further systematic research is required to continue to examine communicative aspects of children’s drawings.
espanolEl ingles tiene cada vez mas variantes geograficas, y se utiliza cada vez mas en comunicaciones internacionales sobre plataformas digitales. Basado en un metodo de investigacion secundaria, se recogen las conclusiones a partir de una revista de literatura multidisciplinar sobre la evolucion del paradigma de aprendizaje del ingles incluyendo la evolucion socio-linguistica del idioma, del paradigma del aprendizaje y de las competencias comunicativas asociadas. Los resultados de esta revista de literatura muestran que el soporte vehicular de expresion en ingles se esta convirtiendo cada vez mas en las redes en linea, lo que supone que la competencia linguistica tiende a incluir el dominio de la semiotica propia a la web y soportes digitales de comunicacion. Por otro lado, el hablante de ingles debe adquirir la conciencia de la relatividad de su propia cultura y aceptar la divergencia de normas linguisticas. Para concluir, se propone unas adaptaciones a la definicion de las competencias genericas listadas en el Marco Europeo de Referencia de Idiomas para tomar en cuenta las caracteristicas del ingles como idioma internacional y usado en la web. EnglishThe English language is diverging into ever stronger regional variants and is used increasingly in international communications over digital platforms. Based on secondary research, conclusions are drawn from a multidisciplinary literature review on the evolution of the paradigm of the English language, including the socio-linguistic evolution of the language, the learning paradigm and associated communicative competences. Results of this literature review show that the vehicular support of expression of English is ever more that of online networks, implying that mastery of the language goes hand in hand with that of web-specific semiotics and digital media supports. Additionally, the speaker of English must acquire consciousness of the relativity of one's own culture and accept divergence from linguistic norms. To conclude, adaptations are proposed to the definition of the generic competences listed in the Common European Framework of Languages, to take into account the characteristics of English as an international language used over the web and digital networks.
Natural language parsing is known to potentially produce a high number of syntactic interpretations for a sentence. Some of them may contain multiword expressions (MWEs) and achieving them faster than compositional alternatives proved efficient in symbolic parsing (see below). We propose to apply this strategy to symbolic LTAG (Lexicalized Tree Adjoining Grammar) parsing using an architecture adaptable to probabilistic parsing. We are particularly interested in LTAGs because, according to (Abeille and Schabes 1989), they show several advantages with respect to parsing MWEs. Firstly, unification constraints on feature structures attached to tree nodes allow one to naturally express dependencies between arguments at different depths in the elementary trees (as in NP 0 vider DET sac 'to express one's secret thoughts', where the determiner DET embedded in the direct object must agree in person and number with the subject NP 0). Secondly, the so-called extended domain of locality offers a natural framework for representing two different kinds of discontinuities. Namely, discontinuities coming from the internal structure of a MWE are directly visible in elementary trees and are handled in parsing mostly by substitution. Disconti-nuities coming from insertion of modifiers (e.g. a bunch of NP, a whole bunch of NP) are invisible in elementary trees but are handled in parsing by adjunction. Consider the sentence in example (1). (1) Acid rains in Ghana are equally grim. When it is being scanned by a left-to-right parser, two competing interpretations are syntactically valid for the first 4 words. One of them considers rains as a verb whose subject is acid while, according to the other, rains is the head noun of the NN compound acid rains. Our objective is to propose a parsing strategy which would promote the latter interpretation due the fact that it contains a known MWE. More precisely, the parser should: (i) trivially, admit only grammar-compliant analyses of a sentence, (ii) achieve MWE-oriented interpretations more rapidly than potential compositional interpretations, (iii) eliminate no grammar-compliant interpretations. Note that all these conditions could rather easily be met for sentence (1) in a pre-processing-based approach in which potential MWEs are identified prior to parsing and conflated into word-with-spaces tokens. Such an approach might however lead to a parsing failure in the case of sentence (2) if the two initial tokens are wrongly merged into a nominal compound in the pre-parsing step. In order to avoid errors of this kind, MWE identification and parsing should be performed jointly. (2) Hunger strikes the civilians since 2001. Seminal works, such as (Finkel and Manning 2009, Green et al. 2011, 2013, Constant et al. 2013), show that the results of probabilis-tic MWE identification and/or parsing are improved when both tasks are performed simultaneously. (Wehrli et al. 2010) point out that such an improvement (also within further parsing-based applications, e.g. machine translation) occurs in symbolic parsing (here: in a Chomskian grammar-based approach) when the knowledge about a potential occurrence of MWEs guides the parsing process. Our goal is to apply a similar strategy to the one in (Wehrli et al. 2010), i.e. to systematically promote MWE-oriented interpretations, within LTAG parsing 1 We additionally wish to design the parser architecture in such a way that corpus-based probabilities about MWE contexts can be 1 The parsing algorithm should of course abstract away from the way the input LTAG grammar was obtained (manually crafted, generated from a metagram-mar, or learned from a treebank).
This paper discusses the process of creating corpora of the sign languages used in Finland, Finnish Sign Language (FinSL) and Finland-Swedish Sign Language (FinSSL). It describes the process of getting informants and data, editing and storing the data, the general principles of annotation, and the creation of a web-based lexical database, the FinSL Signbank, developed on the basis of the NGT Signbank, which is a branch of the Auslan Signbank. The corpus project of Finland’s Sign Languages (CFINSL) started in 2014 at the Sign Language Centre of the University of Jyväskylä. Its aim is to collect conversations and narrations from 80 FinSL users and 20 FinSSL users who are living in different parts of Finland. The participants are filmed in signing sessions led by a native signer in the Audio-visual Research Centre at the University of Jyväskylä. The edited material is stored in the storage service provided by the CSC – IT Center for Science, and the metadata will be saved into CMDI metadata. Every informant is asked to sign a consent form where they state for what kinds of purposes their signing can be used. The corpus data are annotated using the ELAN tool. At the moment, annotations are created on the levels of glosses and translation.
The Czech Legal Text Treebank (CLTT) is a collection of 1133 manually annotated dependency trees. CLTT consists of two legal documents: The Accounting Act (563/1991 Coll., as amended) and Decree on Double-entry Accounting for undertakers (500/2002 Coll., as amended).
India is a country with 22 officially recognized languages and 17 of these have WordNets, a crucial resource.Web browser based interfaces are available for these WordNets, but are not suited for mobile devices which deters people from effectively using this resource.We present our initial work on developing mobile applications and browser extensions to access WordNets for Indian Languages.Our contribution is two fold: (1) We develop mobile applications for the Android, iOS and Windows Phone OS platforms for Hindi, Marathi and Sanskrit WordNets which allow users to search for words and obtain more information along with their translations in English and other Indian languages.(2) We also develop browser extensions for English, Hindi, Marathi, and Sanskrit WordNets, for both Mozilla Firefox, and Google Chrome.We believe that such applications can be quite helpful in a classroom scenario, where students would be able to access the WordNets as dictionaries as well as lexical knowledge bases.This can help in overcoming the language barrier along with furthering language understanding.
While many traditional studies on semantic relatedness utilize the lexical databases, such as WordNet or Wikitionary, the recent word embedding learning approaches demonstrate their abilities to capture syntactic and semantic information, and outperform the lexicon-based methods. However, word senses are not disambiguated in the training phase of both Word2Vec and GloVe, two famous word embedding algorithms, and the path length between any two senses of words in lexical databases cannot reflect their true semantic relatedness. In this paper, a novel approach that linearly combines Word2Vec and GloVe with the lexical database WordNet is proposed for measuring semantic relatedness. The experiments show that the simple method outperforms the state-of-the-art model SensEmbed.
This paper introduces our project for developing Asian Language Treebank (ALT). The ALT project aims to advance the state-of-the-art Asian natural language processing (NLP) techniques through the open collaboration for developing and using ALT. The project is a joint effort of six institutes for making a parallel treebank for seven languages: English, Indonesian, Japanese, Khmer, Malay, Myanmar, and Vietnamese. The process of building ALT began with sampling about 20,000 sentences from English Wikinews, and then these sentences were translated into the other six languages. ALT will have word segmentation, part-of-speech tags, syntactic analysis annotations, together with word alignment links among these languages.
Cross-linguistically consistent annotation is necessary for sound comparative evaluation and cross-lingual learning experiments. It is also useful for multilingual system development and comparative linguistic studies. Universal Dependencies is an open community effort to create cross-linguistically consistent treebank annotation for many languages within a dependency-based lexicalist framework. In this paper, we describe v1 of the universal guidelines, the underlying design principles, and the currently available treebanks for 33 languages.
In this study, we have developed an Indonesian WordNet through four main phases: synonym set extraction (synset) as the smallest entity of lexical database from a natural language, semantic relation establishment between synsets (hypernym-hyponym and holonym-meronym), gloss extraction for synset collection, and the visual editor creation. The Semi-automatic term refers to the three initial phases which are automatically done using a number of machine learning approaches, while using visual editor to collaboratively complement the results collected from the previous phases. A number of raw data used on synset acquisition, semantic relations and glosses come from Kamus Besar Bahasa Indonesia (Great Dictionary of the Indonesian Language, abbreviated as KBBI) and Tesaurus Bahasa Indonesia (Indonesian Language Thesaurus), large collection of web pages from search engines, Wikipedia, and even Princeton WordNet for mapping purpose. This study shows that the proposed system successfully achieve 37,485 synsets, 24,256 hypernym-hyponym relations, 11,044 holonym-meronym relations and 6,520 gloss synsets. Similar approach is believed to accelerate lexical database development like WordNet for other languages.
The paper introduces the DeriNet lexical database, which includes more than 969,000 Czech words interconnected by 718,000 links corresponding to derivational relations (relations between a base word and a word derived from it). Derivational relations were identified by semi-automatic procedures and manual annotation. As the DeriNet network is fully compatible with a large inflectional dictionary of Czech (MorfFlex CZ), it can be used as a resource for an integrating approach to derivational and inflectional morphology of Czech both in linguistic research and in natural language processing.
One of the characteristics of writing in Modern Standard Arabic (MSA) is that the commonly used orthography is mostly consonantal and does not provide full vocalization of the text. It sometimes includes optional diacritical marks (henceforth, diacritics or vowels). Arabic script consists of two classes of symbols: letters and diacritics. Letters comprise long vowels such as A, y, w as well as consonants. Diacritics on the other hand comprise short vowels, gemination markers, nunation markers, as well as other markers (such as hamza, the glottal stop which appears in conjunction with a small number of letters, dots on letters, elongation and emphatic markers) which in all, if present, render a more or less exact precise reading of a word. In this study, we are mostly addressing three types of diacritical marks: short vowels, nunation, and shadda (gemination). Diacritics are extremely useful for text readability and understanding. Their absence in Arabic text adds another layer of lexical and morphological ambiguity. Naturally occurring Arabic text has some percentage of these diacritics present depending on genre and domain. For instance, religious text such as the Quran is fully diacritized to minimize chances of reciting it incorrectly. So are children's educational texts. Classical poetry tends to be diacritized as well. However, news text and other genre are sparsely diacritized (e.g., around 1.5% of tokens in the United Nations Arabic corpus bear at least one diacritic (Diab et al., 2007)). In general, building models to assign diacritics to each letter in a word requires a large amount of annotated training corpora covering different topics and domains to overcome the sparseness problem. The currently available diacritized MSA corpora are generally limited to the newswire genres (those distributed by the LDC) or religion related texts such as Quran or the Tashkeela corpus. In this paper we present a pilot study where we annotate a sample of non-diacritized text extracted from five different text genres. We explore different annotation strategies where we present the data to the annotator in three modes: basic (only forms with no diacritics), intermediate (basic forms–POS tags), and advanced (a list of forms that is automatically diacritized). We show the impact of the annotation strategy on the annotation quality. It has been noted in the literature that complete diacritization is not necessary for readability Hermena et al. (2015) as well as for NLP applications, in fact, (Diab et al., 2007) show that full diacritization has a detrimental effect on SMT. Hence, we are interested in discovering the optimal level of diacritization. Accordingly, we explore different levels of diacritization. In this work, we limit our study to two diacritization schemes: FULL and MIN. For FULL, all diacritics are explicitly specified for every word. For MIN, we explore what is a minimum and optimal number of diacritics that needs to be added in order to disambiguate a given word in context and make a sentence easily readable and unambiguous for any NLP application. We conducted several experiments on a set of sentences that we extracted from five corpora covering different genres. We selected three corpora from the currently available Arabic Treebanks from the Linguistic Data Consortium (LDC). These corpora were chosen because they are fully diacritized and had undergone significant quality control, which will allow us to evaluate the anno tation accuracy as well as our annotators understanding of the task. We select a total of 16,770 words from these corpora for annotation. Three native Arabic annotators with good linguistic background annotated the corpora samples. Diab et al. (2007), define six different diacritization schemes that are inspired by the observation of the relevant naturally occurring diacritics in different texts. We adopt the FULL diacritization scheme, in which all the diacritics should be specified in a word. Annotators were asked to fully diacritize each word. The text genres were annotated following the different strategies: - Basic: In this mode, we ask for annotation of words where all diacritics are absent, including the naturally occurring ones. The words are presented in a raw tokenized format to the annotators in context. - Intermediate: In this mode, we provide the annotator with words along with their POS information. The intuition behind adding POS is to help the annotator disambiguate a word by narrowing down on the diacritization possibilities. - Advanced: In this mode, the annotation task is formulated as a selection task instead of an editng task. Annotators are provided with a list of automatically diacritized candidates and are asked to choose the correct one, if it appears in the list. Otherwise, if they are not satisfied with the given candidates, they can manually edit the word and add the correct diacritics. This technique is designed in order to reduce annotation time and especially reduce annotator workload. For each word, we generate a list of vowelized candidates using MADAMIRA (Pasha et al., 2014). MADAMIRA is able to achieve a lemmatization accuracy 99.2% and a diacritization accuracy of 86.3%. We present the annotator with the top three candidates suggested by MADAMIRA, when possible. Otherwise, only the available candidates are provided. We also provided annotators with detailed guidelines, describing our diacritization scheme and specifying how to add diacritics for each annotation strategy. We described the annotation procedure and specified how to deal with borderline cases. We also provided in the guidelines many annotated examples to illustrate the various rules and exceptions. In order to determine the most optimized annotation setup for the annotators, in terms of speed and efficiency, we test the results obtained following the three annotation strategies. These annotations are all conducted for the FULL scheme. We first calculated the number of words annotated per hour, for each annotator and in each mode. As expected, following the Advanced mode, our three annotators could annotate an average of 618.93 words per hour which is double those annotated in the Basic mode (only 302.14 words). Adding POS tags to the Basic forms, as in the Intermediate mode, does not accelerate the process much. Only − 90 more words are diacritized per hour compared to the basic mode. Then, we evaluated the Inter-Annotator Agreement (IAA) to quantify the extent to which independent annotators agree on the diacritics chosen for each word. For every text genre, two annotators were asked to annotate independently a sample of 100 words. We measured the IAA between two annotators by averaging WER (Word Error Rate) over all pairs of words. The higher the WER between two annotations, the lower their agreement. The results obtained show clearly that the Advanced mode is the best strategy to adopt for this diacritization task. It is the less confusing method on all text genres (with WER between 1.56 and 5.58). We also conducted a preliminary study for a minimum diacritization scheme. This is a diacritization scheme that encodes the most relevant differentiating diacritics to reduce confusability among words that look the same (homographs) when undiacritized but have different readings. Our hypothesis in MIN is that there is an optimal level of diacritization to render a text unambiguous for processing and enhance its readability. We showed the difficulty in defining such a scheme and how subjective this task can be. Acknowledgement This publication was made possible by grant NPRP-6-1020-1-199 from the Qatar National Research Fund (a member of the Qatar Foundation).
This paper describes our two discourse parsers (i.e., English discourse parser and Chinese discourse parser) for submission to CoNLL-2016 shared task on Shallow Discourse Parsing. For English discourse parser, we build two separate argument extractors for single sentence (SS) case, and adopt a convolutional neural network for Non-Explicit sense classification based on (Wang and Lan, 2015b)'s work. As for Chinese discourse parser, we build a pipeline system following the annotation procedure of Chinese Discourse Treebank in Our English discourse parser achieves better performance than the best system of CoNLL-2015 and the Chinese discourse parser achieves encouraging results. Our two parsers both rank second on the blind datasets.
The purpose of this study was to describe the increase in the ability of understanding, responsibility and discipline fifth grade students in learning social studies using models of Make A Match at SDN 03 Lengayang, South Coastal District. This research is a classroom action research. This research was conducted in two cycles, each cycle consisting of two meetings and one final exam cycle. The subjects were fifth grade students at SDN 03 Lengayang, South Coastal District, totaling 33 students. The research instrument used was the observation sheet affective student assessment, teacher observation sheet teaching activities, students' test results, field notes, and cameras. Based on the research capability of understanding the students in the first cycle of 39.40% and 87,88% in the second cycle with an average of 68,33 in the first cycle an increase in cycle II with an average of 80. The observation sheet in response to the student affective ratings Affective Aspects realm Responsibility 48,49% in the first cycle and the second cycle of 81.81% with average cycle I 65,15 an increase in cycle II 81,43. Then Domains Affective Aspects of Discipline in the first cycle of 45,45% and 90,90% in the second cycle to cycle I average 62,5 an increase in cycle II with an average of 84.47. This means learning social studies using a model of Make A Match can improve student learning outcomes in primary school class V 03 Lengayang, South Coastal District. Keywords: Learning Outcomes, IPS, MAKE A MATCH
Les relatives en dont ont fait l’objet de plusieurs études théoriques (Godard 1988) mais ont moins été étudiées d’un point de vue empirique (Gapany 2004). C. Blanche-Benveniste (1995) notait leur rareté dans des corpus oraux, où ils sont principalement compléments de certains verbes. Notre objectif est d’observer le comportement des relatives en dont en français d’un point de vue empirique. Pour ce faire, nous avons en premier lieu étudié les relatives dans deux corpus, le French Treebank (corpus écrit) et le Corpus de Français Parlé Parisien des années 2000 (corpus oral). Ces études nous ont permis de confirmer en partie les travaux de C. Blanche-Benveniste (1995): l’usage de dont en français écrit diffère effectivement de son usage en français oral, mais dont est loin d’être réservé à l’oral à quelques expressions figées. Nous avons pu aussi confirmer l’étude de Godard (1980): les dont complément de nom (extractions hors de SN) ne sont pas marginaux, et les plus frequents sont ceux où le nom est sujet, ce qui considéré comme agrammatical par certaines théories linguistiques comme agrammaticales. De ce point de vue, les corpus nous ont permis de vérifier que les extractions hors du SN sujet ne posent pas de problèmes en français. Leur usage est même plutôt répandu, et inclut les extractions hors de « vrais sujets » de verbes transitifs, contrairement aux prédictions des hypothèses de la grammaire générative. En grammaire générative en effet, les relatives comme les autres structures à extraction sont soumises à des contraintes d’îles (Ross 1957). Selon la contrainte de l’îlot nominal, il est plus difficile d’extraire un complément de nom qu’un complément de verbe. La contrainte de l’îlot sujet exclut quant à elle l’extraction du complément du sujet, qui n’est possible que lorsque le sujet est derivé (passif, moyen…) ou argument interne (Chomsky 2008). Or, nous constatons que les extractions hors de SN sujet sont plus fréquentes dans les corpus que les extractions hors de SN objet (en ce qui concerne les relatives en dont). Ce même fait est un argument en faveur de la théorie de la localité de Gibson, qui prévoit que l’extraction hors de SN sujet est non seulement bien formée mais moins coûteuse que celle hors de SN objet. D’autres observations sur les relatives en dont des deux corpus semblent corroborer cette théorie, comme par exemple la plus grande fréquence des sujets pronominaux ou inversés dans les relatives en dont dans lesquelles dont correspond à un complément du verbe. Dans un deuxième temps, nous avons mené une étude expérimentale de jugement d’acceptabilité afin de comparer les jugements de locuteurs natifs face à des relatives en dont correspondant à une extraction hors du SN sujet d’une part et à une extraction hors de SN objet d’autre part (la forme du sujet dans ces dernières, qui joue un rôle dans la théorie de la localité de Gibson, a été testée également). Les résultats de cette expérience sont eux aussi en accord avec les prévisions de la théorie de la localité de Gibson, les jugements étant significativement supérieurs pour les extractions hors de SN sujet par rapport aux extractions hors de SN objet.
In this paper, we present an overview of some issues related to the use of Big Data in the area of Linguistics that have been debated in workshops and conferences in the last two years. We also consider some requirements that "big" linguistic databases should have in order to tackle some of these issues; finally, we discuss a set of possible interactive visualization approaches of large datasets that may have an impact in this research field.
• Picture rating study: 57 younger and 48 older adults rated 48 pictures each from the same four emotional categories: arousing positive (12 pictures), arousing negative (12 pictures), non-arousing positive (12 pictures), non-arousing negative (12 pictures). There were two different sets of pictures; half of the participants saw Set 1 and the rest saw Set 2 (meaning we have ratings for 96 pictures in total, with each picture having been rated by at least 24 people). Both sets contained equal numbers of pictures in each emotional category and mean arousal and valence were equated across sets for each of the four categories. Participants made their ratings on the self-assessment manikin, just as in Bradley, Lang, and Cuthbert (2008). Lang, P.J., Bradley, M.M., & Cuthbert, B.N. (2008). International affective picture system (IAPS): Affective ratings of pictures and instruction manual. Technical Report A-8. University of Florida, Gainesville, FL.<br><br>
The International Affective Picture System (IAPS) is a picture set used by researchers to select pictures that have been prerated on valence. Researchers rely on the ratings in the IAPS to accurately reflect the degree to which the pictures elicit affective responses. Here we show that this may not always be a safe assumption. More specifically, the scale used to measure valence in the IAPS ranges from positive to negative, implying that positive and negative feelings are end-points of the same construct. This makes interpretation of midpoint, or neutral ratings, especially problematic because it is impossible to tell whether these ratings are the result of neutral, or of mixed feelings. In other words, neutral ratings may not be as neutral as researchers assume them to be. Investigating this, in this work we show that pictures that seem neutral according to the valence ratings in the IAPS indeed vary in levels of ambivalence they elicit. Furthermore, the experience of ambivalence in response to these pictures is predictive of the arousal that people report feeling when viewing these pictures. These findings are of particular importance because neutrality differs from ambivalence in its specific psychological consequences, and by relying on seemingly neutral valance ratings, researchers may unwillingly introduce these consequences into their research design, undermining their level of experimental control. (PsycINFO Database Record
This work elaborates the semi-semantic part of speech annotation guidelines for the URDU.KON-TB treebank: an annotated corpus. A hierarchical annotation scheme was designed to label the part of speech and then applied on the corpus. This raw corpus was collected from the Urdu Wikipedia and the Jang newspaper and then annotated with the proposed semi-semantic part of speech labels. The corpus contains text of local & international news, social stories, sports, culture, finance, religion, traveling, etc. This exercise finally contributed a part of speech annotation to the URDU.KON-TB treebank. Twenty-two main part of speech categories are divided into subcategories, which conclude the morphological, and semantical information encoded in it. This article reports the annotation guidelines in major; however, it also briefs the development of the URDU.KON-TB treebank, which includes the raw corpus collection, designing & employment of annotation scheme and finally, its statistical evaluation and results. The guidelines presented as follows, will be useful for linguistic community to annotate the sentences not only for the national language Urdu but for the other indigenous languages like Punjab, Sindhi, Pashto, etc., as well.
Affective computing is a very important issue. An increasing amount of research focused on representing affective states as continuous numerical values on multiple dimensions. Such as the emotional space, which is about the valence and arousal. Due to the affective dimension representation can be useful to sentiment analysis, building dimensional sentiment resources with valence-arousal ratings are very important. Therefore, this study proposes a method to automatically obtain the valence-arousal ratings of affective words. Experiment results using the evaluation metrics to get the error rates about the mean absolute error and pearson correlation coefficient.
This paper reports on the development of a French FrameNet, within the ASFALDA project. While the first phase of the project focused on the development of a French set of frames and corresponding lexicon (Candito et al., 2014), this paper concentrates on the subsequent corpus annotation phase, which focused on four notional domains (commercial transactions, cognitive stances, causality and verbal communication). Given full coverage is not reachable for a relatively " new " FrameNet project, we advocate that focusing on specific notional domains allowed us to obtain full lexical coverage for the frames of these domains, while partially reflecting word sense ambiguities. Furthermore, as frames and roles were annotated on two French Treebanks (the French Treebank (Abeillé and Barrier, 2004) and the Sequoia Treebank (Candito and Seddah, 2012), we were able to extract a syntactico-semantic lexicon from the annotated frames. In the resource's current status, there are 98 frames, 662 frame-evoking words, 872 senses, and about 13000 annotated frames, with their semantic roles assigned to portions of text. The French FrameNet is freely available at alpage.inria.fr/asfalda.
The semantic similarity measures are designed to compare terms that belong to the same ontology. Many of these are based on a graph structure, such as the well-known lexical database for the English language, named WordNet, which groups the words into sets of synonyms called synsets. Each synset represents a unique vertex of the WordNet semantic graph, through which is possible to get information about the relations between the different synsets. The literature shows several ways to determine the similarity between words or sentences through WordNet (e.g., by measuring the distance among the words, by counting the number of edges between the correspondent synsets), but almost all of them do not take into account the peculiar aspects of the used dataset. In some contexts this strategy could lead toward bad results, because it considers only the relationship between vertexes of the WordNet semantic graph, without giving them a different weight based on the synsets frequency within the considered datasets. In other words, common synsets and rare synsets are valued equally. This could create problems in some applications, such as those of recommender systems, where WordNet is exploited to evaluate the semantic similarity between the textual descriptions of the items positively evaluated by the users, and the descriptions of the other ones not evaluated yet. In this context, we need to identify the user preferences as best as possible, and not taking into account the synsets frequency, we risk to not recommend certain items to the users, since the semantic similarity generated by the most common synsets present in the description of other items could prevail. This work faces this problem, by introducing a novel criterion of evaluation of the similarity between words (and sentences) that exploits the WordNet semantic graph, adding to it the weight information of the synsets. The effectiveness of the proposed strategy is verified in the recommender systems context, where the recommendations are generated on the basis of the semantic similarity between the items stored in the user profiles, and the items not evaluated yet.
The historical development and the linguistic triggering environments for Oberfeld formation in German subordinate clauses represent long-standing research questions in Germanic Linguistics, dating at least as far back as Jacob Grimm's famous Deutsche Grammatik. The present corpus study traces this historical development back to the 17th century. The study is based on three text corpora. For contemporary German, two syntactically annotated newspaper corpora were consulted: the TüBa-D/Z and TüPP-D/Z treebanks,1 which provide linguistic annotations for articles published in the daily newspaper die tageszeitung (taz). For diachronic data, the corpus collection Deutsches Textarchiv (DTA)2 was utilized. The DTA contains texts ranging from 1610 to 1900. All three corpus collections are part of the Common Language Resources and Technology Infrastructure (CLARIN) initiative3 and will be developed further as part of the CLARIN research infrastructure. The study demonstrates the added value that annotated corpora can provide for in-depth studies in historical syntax. At the same time it showcases the added value of interoperable language resources for linguistic investigations that require access to and analysis of multiple linguistic resources.
Chinese word segmentation and Part-of-speech (POS) tagging have been studied for decades. However, most of the previous works mainly focus on pipeline method which will lead to error propagation. In order to make word segmentation and POS tagging jointly in one model, in this paper, we propose an effective neural network model to improve the accuracy of the segmentation and tagging. Our model works based on the hierarchical Long Short-Term Memory (LSTM) and trained jointly in one objective function. What's more, to better utilizing the transition features between tags, we further introduce the transition matrix which can help to search the best tagging sequence. Experiment on Chinese Treebank shows that our model achieves competitive accuracy on word segmentation and POS tagging.
We propose a method for improving the dependency parsing of complex sentences. This method assumes segmentation of input sentences into clauses and does not require to re-train a parser of one's choice. We represent a sentence clause structure using clause charts that provide a layer of embedding for each clause in the sentence. Then we formulate a parsing strategy as a two-stage process where (i) coordinated and subordinated clauses of the sentence are parsed separately with respect to the sentence clause chart and (ii) their dependency trees become subtrees of the final tree of the sentence. The object language is Czech and the parser used is a maximum spanning tree parser trained on the Prague Dependency Treebank. We have achieved an average 0.97% improvement in the unlabeled attachment score. Although the method has been designed for the dependency parsing of Czech, it is useful for other parsing techniques and languages.
El acto de destrucción de la lengua que lleva cabo uno de los sujetos de la escritura de Altazor abre el espacio y el tiempo para el surgimiento de nuevos significantes. Para este sujeto, el uso sistemático de la lengua -cumpliendo la norma lingüística- impide la representación (aparición de imágenes) y referencia de correlatos radicalmente nuevos. Este sujeto es discontinuo y coexiste con otros, en especial, con un sujeto voluntarioso, que sigue aferrándose a la tradicional concepción del mundo, fundada en la trascendencia divina. Las operaciones sobre la lengua de este poeta altazoriano se realizan, sobre todo, como utilización paródica de ella, como demolición intencional de la lengua, como transformación de las ruinas idiomáticas en significantes, como uso alegórico de la lengua (en el sentido sugerido por W. Benjamin). Este sujeto altazoriano propone orientar y sostener sentimentalmente la constitución de nuevas imágenes (en el sentido de los universales fantásticos de Vico). La poesía se hace, así, acontecimiento (Ereignis, según lo nombra Heidegger), actividad cuya plenitud se produce en su consumación y consunción, en la mostración de su temporalidad fundante. Pero la energía del poeta no alcanza para la continuidad del acontecimiento poético -en el caso de que sea posible-, para la urgencia y necesidad de su aparición. The act oflanguage destruction carried out by one ofthe subjects in the writing of "Altazor", opens both space and time to the birth ofnew signifiers. For such subject, the systematic use oflanguage sticking to the linguistic norm, hinders the representation (the creation o.fimages and the reference to radically new co-relatives. Such subject is discontinuous and coexists with others, in particular, with a willful subject that keeps its allegiance to the traditional world view, based on divine transcendence. The language operations ofthis altazorean poet are effected, above all, as parody; as transformation ofidiomatic debris into signifiers; and as an allegorical use of language (in the sense suggested by Walter Benjamín). The altarzorean subject propases to orientate and sustain sentimentally the creation of new images (in the sense of the "fantastic universals" of Vico). In this way, poetry becomes 'event' ("Ereignis", as named by Heidegger) an activity whosefullnes is achieved through its realization and consumption, in the exhibition of its founding tenporality.
INTRODUCTIONThe Strategic Integrated Management Seminar (SIMS) course is mandatory for every senior student in the school of business at a mid-size private university in the northeastern United States. The course allows students to integrate their accumulated knowledge and apply this knowledge to issues from a strategic perspective. It examines a firm from the position of top level management, focusing on the role of the general manager in formulating and implementing corporate and business level strategy. Strategic issues of an entire athletic (hereon, footwear company or company) and industry are analyzed. Students are expected to draw their accumulated knowledge of the functional areas of their majors into a homogenous team effort. Each individual student works on developing his/her ability to analyze information, draw logical conclusions, and offer sound supporting evidence for their arguments in written form and classroom discussions. The course is highly interactive with students taking the lead and the professors sharing knowledge and offering supplementary support.The SIMS course uses the Business Strategy Game (BSG) simulation to enable students to experience a top management team perspective in running a and experiencing competitive conditions in the athletic industry. In the BSG, students compete in teams (each team constitutes a company, hereon team/company will be synonymous) within a global arena that encompasses four regions - Europe-Africa, North America, Asia-Pacific, and Latin America (The Business Strategy Game, 2016). They compete against teams in their individual classes and compare/contrast data with teams/companies worldwide. Each competes head-to-head against companies run by other teams in the course, hence competition plays an important role in the experience. Each sells its brand of to retailers worldwide and to individuals buying online at the company's website.Competing in the BSG requires a series of complex decisions by the students, taking into account the team's strategy for their and the competitive conditions in the industry and the strategies of their competitors. The simulation allows for numerous decisions for each round, requiring students to choose which decisions are most important to implement their strategy and which areas of the business must receive attention in order for their firm to be its most competitive. Beyond overall strategy (corporate, competitive) are several key functional areas for decision making. Decision areas in operations include capacity planning (either adding to existing plants or building new plants in new geographic locations), production quality decisions for the athletic footwear, plant operations efficiency, and labor decisions. Footwear must be shipped to distribution centers around the world and students must choose where it is best to manufacture the and where to ship taking into consideration demand, shipping costs, tariffs, and exchange rates. Marketing decisions include pricing the product in a wholesale and a retail environment, advertising and use or non-use of celebrity endorsements, rebates, and incentives to retailers. Financial decisions include funding the capital structure of the firm using debt, equity, and/or cash. Dividend payouts and stock repurchases may be used by the companies.The simulation has students take control of an athletic that has been in operation for ten years. Teams make in total eight years of decisions (years 11 - 18), approximately one per week. Each decision rollover represents one year and includes many decisions within the decision. The simulation evaluates team performance based on five investor expectation performance targets: Earnings Per Share (EPS), Return on Equity (ROE), credit rating, image rating (a combination of market share and shoe quality), and stock price. Each measure of performance is equally weighted at 20% of the total score (The Business Strategy Game, 2016). …
This paper aims at filling the gap between the accuracy of Italian and English constituency parsing: firstly, we adapt the Bllip parser, i.e., the most accurate constituency parser for English, also known as Charniak parser, for Italian and trained it on the Turin University Treebank (TUT). Secondly, we design a parse reranker based on Support Vector Machines using tree kernels, where the latter can effectively generalize syntactic patterns, requiring little training data for training the model. We show that our approach outperforms the state of the art achieved by the Berkeley parser, improving it from 84.54 to 86.81 in labeled F1.