Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
This paper formalizes a sound extension of dynamic oracles to global training, in the frame of transition-based dependency parsers. By dispensing with the precomputation of references, this extension widens the training strategies that can be entertained for such parsers; we show this by revisiting two standard training procedures, early-update and max-violation, to correct some of their search space sampling biases. Experimentally, on the SPMRL treebanks, this improvement increases the similarity between the train and test distributions and yields performance improvements up to 0.7 UAS, without any computation overhead.
This study attempts to determine the common features and differences between the Latin language of the inscriptions of Aquincum, Salona, Aquileia and the provincial countries of Pannonia Inferior, Dalmatia and Venetia et Histria, compared with each other and the rest of the Latin speaking provinces of the Roman empire, and we intend to demonstrate whether a regional dialect area over the Alps–Danube–Adria region of the Roman empire existed, a hypothesis suggested by József Herman. For our research, we use all relevant linguistic data from the Computerized Historical Linguistic Database of Latin Inscriptions of the Imperial Age. We will examine the relative distribution of diverse types of non-standard data found in the inscriptions, contrasting the linguistic phenomena of an earlier period with a later stage of Vulgar Latin. The focus of our analysis will be on the changes in the vowel system and the grammatical cases between the two chronological periods within each of the three examined cities. If we succeed in identifying similar tendencies in the Vulgar Latin of these three cities, the shared linguistic phenomena may suggest the existence of a regional variant of Latin in the Alps–Danube–Adria region.
<h1> TaPvex: A Tagged and Phrased word2vec Model </h1> Benjamin J. Radford <br> September 29, 2017 <h2> Summary </h2> <i>TaPvex</i> is a trained word2vec model of part-of-speech-tagged and named-entity-tagged words and phrases. The model was trained on a large corpus of English language news text from the early 2010s. Words have been tagged using Stanford CoreNLP to include <a href="https://nlp.stanford.edu/software/CRF-NER.shtml">named entities</a> (NER) and <a href="https://www.ling.upenn.edu/courses/Fall_2003/ling001/penn_treebank_pos.html">Penn Treebank</a> parts-of-speech (POS). Tagged words have been concatenated into n-gram phrases.<br> The model contains 1.17 million unique words and phrases. Word vectors are of size 150. <h2> Use </h2> All three files (TaPvex, TaPvex.syn0.npy, TaPvex.syn1.npy) must be located in the same directory. The file can be opened with the <a href="https://radimrehurek.com/gensim/models/word2vec.html"><i>gensim</i></a> Python module using: <pre> from gensim.models import Word2Vec model = Word2Vec.load("/path/to/model/TaPvex") </pre> <h2> Examples </h2> Tokens are of the form: <pre>[WORD]:[NER]:[POS]</pre> Phrases are of the form: <pre>[WORD]:[NER]:[POS]_[WORD]:[NER]:[POS]</pre> Example tokens include: <pre> BUSH:O:NN BUSHES:O:NNS BUSH:PERSON:NNP GEORGE:PERSON:NNP_BUSH:PERSON:NNP GEORGE:PERSON:NNP_W:PERSON:NNP_BUSH:PERSON:NNP NEW:O:JJ NEW:LOCATION:NNP_YORK:LOCATION:NNP NEW:ORGANIZATION:NNP_YORK:ORGANIZATION:NNP_TIMES:ORGANIZATION:NNP </pre>
Dans cette thèse, nous décrivons la création du French FrameNet (FFN), une ressource de type FrameNet pour le français créée à partir du FrameNet de l’anglais (Baker et al., 1998) et de deux corpus arborés: le French Treebank (Abeillé et al., 2003) et le Sequoia Treebank (Candito et Seddah, 2012). La ressource séminale, le FrameNet de l’anglais, constitue un modèle d’annotation sémantique de situations prototypiques et de leurs participants. Elle propose à la fois:a) un ensemble structuré de situations prototypiques, appelées cadres, associées à des caractérisations sémantiques des participants impliqués (les rôles);b) un lexique de déclencheurs, les lexèmes évoquant ces cadres;c) un ensemble d’annotations en cadres pour l’anglais. Pour créer le FFN, nous avons suivi une approche «par domaine notionnel»: nous avons défini quatre «domaines» centrés chacun autour d’une notion (cause, communication langagière, position cognitive ou transaction commerciale), que nous avons travaillé à couvrir exhaustivement à la fois pour la définition des cadres sémantiques, la définition du lexique, et l’annotation en corpus. Cette stratégie permet de garantir une plus grande cohérence dans la structuration en cadres sémantiques, tout en abordant la polysémie au sein d’un domaine et entre les domaines. De plus, nous avons annoté les cadres de nos domaines sur du texte continu, sans sélection d’occurrences: nous préservons ainsi la distribution des caractéristiques lexicales et syntaxiques de l’évocation des cadres dans notre corpus. à l’heure actuelle, le FFN comporte 105 cadres et 873 déclencheurs distincts, qui donnent lieu à 1109 paires déclencheur-cadre distinctes, c’est-à-dire 1109 sens. Le corpus annoté compte au total 16167 annotations de cadres de nos domaines et de leurs rôles. La thèse commence par resituer le modèle FrameNet dans un contexte théorique plus large. Nous justifions ensuite le choix de nous appuyer sur cette ressource et motivons notre méthodologie en domaines notionnels. Nous explicitons pour le FFN certaines notions définies pour le FrameNet de l’anglais que nous avons jugées trop floues pour être appliquées de manière cohérente. Nous introduisons en particulier des critères plus directement syntaxiques pour la définition du périmètre lexical d’un cadre, ainsi que pour la distinction entre rôles noyaux et non-noyaux.Nous décrivons ensuite la création du FFN: d’abord, la délimitation de la structure de cadres utilisée pour le FFN, et la création de leur lexique. Nous présentons alors de manière approfondie le domaine notionnel des positions cognitives, qui englobe les cadres portant sur le degré de certitude d’un être doué de conscience sur une proposition. Puis, nous présentons notre méthodologie d’annotation du corpus en cadres et en rôles. à cette occasion, nous passons en revue certains phénomènes linguistiques qu’il nous a fallu traiter pour obtenir une annotation cohérente; c’est par exemple le cas des constructions à attribut de l’objet.Enfin, nous présentons des données quantitatives sur le FFN tel qu’il est à ce jour et sur son évaluation. Nous terminons sur des perspectives de travaux d’amélioration et d’exploitation de la ressource créée.
Proof theory began in the 1920's as a part of Hilbert's program, which aimed to secure the foundations of mathematics by modeling infinitary mathematics with formal axiomatic systems and proving those systems consistent using restricted, finitary means. The program thus viewed mathematics as a system of reasoning with precise linguistic norms, governed by rules that can be described and studied in concrete terms. Today such a viewpoint has applications in mathematics, computer science, and the philosophy of mathematics.
The use of semantically related words in new vocabulary lists is common practice in second language textbooks. However, research has suggested that organizing new foreign language vocabulary in semantic sets may cause interference and even slow acquisition as compared to other organizational methods (Erten & Tekin 2008). When it comes to assessing the frequency of this semantic organization of new vocabulary in foreign language textbooks, relatively little research has been done, especially with regard to Spanish textbooks. López-Jiménez's study (2012) on vocabulary in Spanish textbooks found that semantic organization was present in six of twelve textbooks surveyed, but the definition and measurement of semantic organization were not clear. Therefore, the present study seeks to improve this measurement based on a definition incorporating WordNet's online lexical database to calculate the semantic distance (using the Wu-Palmer method) between the words in each set. For this purpose, three textbooks used in the lower- and intermediate-level Spanish courses at Iowa State University were selected for analysis. To assess the levels of semantic relatedness within each set, the semantic distance calculations were then compared to a threshold determined by performing the same calculations on word sets that had been determined by the existing literature to be semantically related or unrelated. The results suggest that improvement can be made in the organization of words in each textbook and in the sample as a whole. Furthermore, a more precise definition is presented for determining semantic relatedness among vocabulary in foreign language textbooks. These results will be instrumental in helping to better inform the choice of materials for Spanish classrooms in the future.\nReferences:Erten, I. H., & Tekin, M. (2008). Effects on vocabulary acquisition of presenting new words in semantic sets versus semantically unrelated sets. System, 36, 407-422.López-Jiménez, M. D. (2012). A Critical Analysis of the Vocabulary in L2 Spanish Textbooks. Porta Linguarum, 21, 163-181.
Data acquired by space borne systems is subjected to inherent deviations in geometric and radiometric aspects in terms of true representation of observed earth features. Significantly, high resolution data is more sensitive towards its radiometric accuracies and its quality. Raw data is preprocessed incorporating all corrections which are modeled using prelaunch test data. However, it is necessary to estimate post launch performance and its characterization for understanding the deviations in the orbit phase from the lab characterization to produce real world objects accurately in terms of their positional and spectral characteristics. In this paper the emphasis is made w.r.t the early orbit characterization of the Resourcesat-2A, LISS4 multispectral data products which were operationalized during February 2017. The Radiometric characterization is performed through in flight LED Calibration, Non illumination area imaging, global Radiometric Sites, cross calibration with the contemplating sensors. MTF is estimated through edge based artificial targets and vicarious calibration exercises. The Geometric calibration exercises were carried out by estimating positional accuracy consistency through longer paths, different homogeneous and heterogeneous terrains. Assurance of time series data consistency and systematic coverage is estimated after path lock, to provide contiguous data for the support of applications based on larger extents. All the calibrations exercises were carried out in iterative mode and finally achieved the targeted specifications. Over all Image rating is estimated by NIIRS.
In this paper, we propose a general methodology for designing semantic role/relation system. Based on this methodology, we establish a succinct semantic relation system for consecutive predicative constituents for Chinese, which includes serial verb construction, discourse construction, and other constructions describing serial events. This semantic relation system has 13 middle-level classes and 24 fine-grained sub-classes in contrast to conventional complex classification schemes and meets the uniqueness and completeness criteria of semantic relation identification. We conduct experiments on our system by training four annotators in 1 h to label 200 sentences extracted from Sinica Treebank and HIT-CDTB. With the help of our predesigned feature-based decision tree and a connective markers checklist, the annotators attain a 73% consistency with the reference standard annotation and substantial agreement by Cohen’s kappa coefficient for middle-level labeling. By analyzing the labeling error types, we slightly revise our classification scheme and propose six methods to improve the classification and labeling system, hoping to achieve even better agreement in the future.
Every 10 years there seems to be a new wave of dissertations on Dutch verb clusters.The special properties of the Dutch verbal end group have been keeping Dutch linguists employed since at least the 50s, and are still leading to new insights into language syntax and language variation.In the mid '90s the interest was mainly in novel theoretical syntactic analyses of the phenomenon, and in the mid '00s several dissertations came out on the wide range of factors that allow the word order variation to occur.Now, in the mid '10s, another round of dissertations is set to appear or have already appeared, including Liesbeth Augustinus's study on cluster formation.The current wave appears to be all about describing and studying the scope of the variation that occurs when people use verb clusters, both in dialects and in standard Dutch.The range of variation appears to be broader than what was previously thought.The likely reason behind this direction of study is the increased availability of Dutch linguistic resources, such as digital dialect atlases and large language corpora, making it easier to study large amounts of empirical data.Liesbeth Augustinus's study, though mostly focusing on standard Dutch, exemplifies this approach.The study is corpus-based throughout, featuring language usage data extracted from a treebank (a corpus of Dutch, annotated with syntactic trees).Furthermore, the study not only discusses standard verb clusters, it also addresses various (though not all) less frequent verb cluster constructions.Clusters with te-infinitives for example, and 'cluster creepers' where the verb cluster is interrupted by non-verbal materials, and Infinitivus Pro Participio (IPP) constructions involving uncommon IPP verbs, where a verb that takes an infinitive appears in participial form instead of as an infinitive.In this way, a wide range of verb cluster variation is addressed and incorporated into the theoretical discussion.A quick look at the table of contents shows that this is quite an interdisciplinary affair.On the one hand there are more theoretical chapters, involving both transformational grammar, the framework in which most previous work has been conducted, as well as monostratal grammarmainly Head-driven Phrase Structure Grammar (HPSG), the formal framework that Liesbeth Augustinus uses.
Background: Anhedonia has long been associated with schizophrenia (SZ); however, the true nature of this deficit remains elusive. Given the role of hedonic capacity within the larger motivational framework, we sought to examine reward responsiveness (RR) and reward expectancy (RE) across a spectrum of motivation deficits in SZ (Study 1). Further, we sought to better understand the relationship between hedonic capacity and specific facets of the motivational system (Study 2). Methods: In study 1, RR and RE were assessed using the self-report Temporal Experience of Pleasure Scale (TEPS) in a sample of 72 SZ patients and 74 healthy controls. In study 2, 99 healthy undergraduate students completed the TEPS as well as objective measures of RR, RE, reward valuation, effort valuation, and goal-directed decision-making using the International Affective Picture System (IAPS), Cued Reinforcement Reaction Time (CRRT) task, Kirby Delay Discounting (DD) task, Virtual Reality Progressive Ratio (ViPR) task, and the Multitasking in the City Test (MCT), respectively. Further, the Schizotypal Personality Questionnaire (SPQ) was administered to characterize subclinical schizotypal traits. In both studies, the Apathy Evaluation Scale (AES) was used to characterize participants into low, moderate, and high amotivation groups. Results: In both studies, a multivariate analysis of variance revealed a main effect of amotivation such that participants with high levels of amotivation reported significantly lower levels of RR and RE compared to those at low and moderate levels (Study 1: F(4, 280) = 2.962, P =.02, η2 =.041; Study 2: F(4, 170) = 4.453, P =.002, η2 =.095). In Study 1, an interaction effect revealed that patients with moderate levels of amotivation endorsed significantly higher levels of RE compared to healthy controls at the same level, and to patients at both low and high levels of amotivation (F(2) = 2.674, P =.007). In Study 2, correlational analyses revealed that both RR (r =.32, P =.002) and RE (r =.38, P <.001) were correlated with IAPS pleasantness ratings. Further, RE was correlated with IAPS arousal ratings on the IAPS (r =.33, P =.001) and the ViPR task (r = −.25, P =.032). RE (r = −.37, P <.001) and IAPS pleasantness (r = −.31, P =.003) and arousal (r = −.34, P =.001) ratings were also correlated with the negative subscale of the SPQ. Conclusion: Overall, the results of both studies suggest that impairments in RR and RE emerge exclusively in individuals with high levels of motivation deficits, regardless of diagnosis. Further, Study 1 illustrates the complex relationship between self-reported RE, amotivation, and diagnosis. Correlational analyses in Study 2 suggest that emotional arousal and cost–benefit analyses are related to the evaluation of prospective rewards on the TEPS. Going forward, utilizing both subjective and objective measures of hedonic capacity may serve to further our understanding of the nuances of motivation and reward system impairments in SZ.
Connections play a crucial role in neural network (NN) learning because they determine how information flows in NNs. Suitable connection mechanisms may extensively enlarge the learning capability and reduce the negative effect of gradient problems. In this paper, a new delay connection is proposed for Long Short-Term Memory (LSTM) unit to develop a more sophisticated recurrent unit, called Delay Connected LSTM (DCLSTM). The proposed delay connection brings two main merits to DCLSTM with introducing no extra parameters. First, it allows the output of the DCLSTM unit to maintain LSTM, which is absent in the LSTM unit. Second, the proposed delay connection helps to bridge the error signals to previous time steps and allows it to be back-propagated across several layers without vanishing too quickly. To evaluate the performance of the proposed delay connections, the DCLSTM model with and without peephole connections was compared with four state-of-the-art recurrent model on two sequence classification tasks. DCLSTM model outperformed the other models with higher accuracy and F1[Formula: see text]score. Furthermore, the networks with multiple stacked DCLSTM layers and the standard LSTM layer were evaluated on Penn Treebank (PTB) language modeling. The DCLSTM model achieved lower perplexity (PPL)/bit-per-character (BPC) than the standard LSTM model. The experiments demonstrate that the learning of the DCLSTM models is more stable and efficient.
Abstract The paper deals with the field of Czech corpus linguistics and represents one of various current studies analysing text coherence through language interactions. It presents a corpusbased analysis of grammatical coreference and sentence information structure (in terms of contextual boundness) in Czech. It focuses on examining the interaction of these two language phenomena and observes where they meet to participate in text structuring. Specifically, the paper analyses contextually bound and non-bound sentence items and examines whether (and how often) they are involved in relations of grammatical coreference in Czech newspaper articles. The analysis is carried out on the language data of the Prague Dependency Treebank (PDT) containing 3,165 Czech texts. The results of the analysis are helpful in automatic text annotation - the paper presents how (or to what extent) the annotation of grammatical coreference may be used in automatic (pre-)annotation of sentence information structure in Czech. It demonstrates how accurately we may (automatically) assume the value of contextual boundness for the antecedent and anaphor (as the two participants of a grammatical coreference relation). The results of the paper demonstrate that the anaphor of grammatical coreference is automatically predictable - it is a non-contrastive contextually bound sentence item in 99.18% of cases. On the other hand, the value of contextual boundness of the antecedent is not so easy to estimate (according to the PDT, the antecedent is contextually non-bound in 37% of cases, non-contrastive contextually bound in 50% and contrastive contextually bound in 13% of cases).
This paper examines the mobilisation of linguistic ideologies as a form of dissent from dominant discourses of identity in contemporary Middle Eastern media. As part of my broader doctoral research on non-government Jordanian radio today, it takes a linguistical anthropological perspective focused on the notion of indexicality: the non-referential meanings that are invoked contingently in language use, and thus articulate links to broader social and cultural ideologies, including stereotypes of identity categories such as gender and geographic origins.I examine two case studies in which speakers problematise and reframe such stereotypes. The first involves the indexical mechanism of implicature, whereby a talk show caller mounts a challenge to dominant discourses of urban linguistic refinement through the ironic use of a ‘sanitised’ pronunciation of a local Jordanian dish (ča‘āčīl / ka‘ākīl). The second, from a programme in honour of a Jordanian pilot executed by the Islamic State (IS) in Syria, exhibits the performance of an evaluative stance towards Jordanian military activity as a form of patriotic nationalism, through the use of the [g] pronunciation of the sound /q/ (qāf) by a female broadcaster – a usage that defies gendered linguistic norms Jordanian radio, which require female speakers to use the [ʔ] (glottal stop) pronunciation instead.While these contingent uses of implicature and stance form challenges to certain dominant discourses, they are nevertheless ambiguous in that they draw on other problematic ideologies, including localist linguistic ‘authenticity’ and patriotic Jordanian nationalism. Thus, while details of language use provide important potential for dissent, this paper also problematises this potential – asking whether (1) subversive linguistic practices always need to draw on other dominant discourses in order to be meaningful, and (2) whether such references necessarily make dissent compromised or illegitimate.
You have accessJournal of UrologyProstate Cancer: Advanced (including Drug Therapy) II1 Apr 2017PD37-06 A NEW ERA: AUTOMATED EXTRACTION OF DETAILED PROSTATE CANCER INFORMATION FROM NARRATIVELY WRITTEN HEALTH RECORDS. PIONEER WORK FROM A EUROPEAN TERTIARY CARE CENTER Sami-Ramzi Leyh-Bannurah, Zhe Tian, Pierre Karakiewicz, Dirk Pehrke, Hartwig Huland, Markus Graefen, and Lars Budäus Sami-Ramzi Leyh-BannurahSami-Ramzi Leyh-Bannurah More articles by this author, Zhe TianZhe Tian More articles by this author, Pierre KarakiewiczPierre Karakiewicz More articles by this author, Dirk PehrkeDirk Pehrke More articles by this author, Hartwig HulandHartwig Huland More articles by this author, Markus GraefenMarkus Graefen More articles by this author, and Lars BudäusLars Budäus More articles by this author View All Author Informationhttps://doi.org/10.1016/j.juro.2017.02.1568AboutPDF ToolsAdd to favoritesDownload CitationsTrack CitationsPermissionsReprints ShareFacebookTwitterLinked InEmail INTRODUCTION AND OBJECTIVES Detailed pathological information are necessary for follow-up analyses and new prediction tool development. Most institutional databases harbour a lack of either quantity or quality of such data. Expensive manpower needs to be manually invested for this purpose in daily clinical practice. Natural language (NLP) processing presents immense potential to automate the information-gathering process in the field of urology. To propose and validate a novel tool to extract specific detailed pathological information from written health records that contain continuous, narrative text in an automated and precise way. METHODS Overall, 1500 postoperative narrative pathology reports of patients undergoing radical prostatectomy in 2015 were analyzed. Of these, 750 reports were randomly selected as training data and the remaining 750 reports were used as validation data. Using domain knowledge from clinical experts and Stanford treebank parser for German, rule based extraction algorithms were created for pathological staging, Gleason percentages, prostate dimension and laterality of the tumor by iterative review of misclassified reports until there are no longer any misclassified reports in the training data. Each number found in the reports was also verified by clinical experts. RESULTS By applying the developed NLP system on the validation data, we assed the accuracy of each information extracted. The NLP derived accuracy for pT-stage, pN-stage, Gleason percentage, prostate weight, prostate volume and laterality of the tumour were 100%, 100%, 100%,100% and 97%, respectively. CONCLUSIONS We developed a novel NLP method that could extract detailed pathological information from a narrative, written pathological report with very high accuracy. This automated method can be implemented with the aim to greatly increase the efficiency and accessibility of research data in academic centers. © 2017FiguresReferencesRelatedDetails Volume 197Issue 4SApril 2017Page: e676 Advertisement Copyright & Permissions© 2017MetricsAuthor Information Sami-Ramzi Leyh-Bannurah More articles by this author Zhe Tian More articles by this author Pierre Karakiewicz More articles by this author Dirk Pehrke More articles by this author Hartwig Huland More articles by this author Markus Graefen More articles by this author Lars Budäus More articles by this author Expand All Advertisement Advertisement PDF downloadLoading...
Vavák’s Memoirs, written in 1770–1816, can serve by dint of their significance and their scope to describe the development of Czech at that time. The article evaluates the content and language modifications in their modern edition and follows and suggests the use of their language material in grammar books, dictionaries, lexical databases and text corpuses.
This paper reports on a suite of experiments that evaluates how the linguistic granularity of part-of-speech tagsets impacts the performance of tagging and syntactic dependency parsing. Our results show that parsing accuracy can be significantly improved by introducing more finegrained morphological information in the tagset, even if tagger accuracy is compromised. Our taggers and parsers are trained and tested using the annotations of the Norwegian Dependency Treebank.
Borderline Personality Disorder (BPD) is characterized by unstable mood states, chaotic interpersonal relationships, and behavioral dysregulation in the form of selfinjurious acts that results in notable functional impairment. Emotion dysregulation, marked by strong shifts in emotional states away from baseline levels across subjective and physiological substrates, is believed to reflect one mechanism in the relationship between BPD and functional impairment. However, it remains unclear whether emotion dysregulation represents a general tendency to experience both positive and negative emotions keenly, or to specifically be sensitized to negative mood states. The present study examined the relationship between BPD symptoms and emotion dysregulation across neutral, negative, and positive valenced emotional states in a sample of twenty-two community dwelling adults with histories of psychiatric disorders. Emotion dysregulation was measured via subjective affect ratings and pupillary responses that index sympathetic nervous system reactivity when participants recalled neutral, stressful, and pleasant events that occurred during the prior 3 months. Results and clinical implications are discussed.
From sociolinguistic assumptions about issues involving linguistic variations that have direct implications on language teaching and learning, the paper falls on the research line about language and social practices. The objective is to analyse the linguistic attitudes that teachers of Portuguese who hold university degree have towards the norm they and their students use. As far as theoretical support is concerned, the research takes as reference the distinction between the concepts of polite norm and standard norm; the concept of sociolinguistic norm that allows the systematisation of linguistic behaviour, assessment of the linguistic behaviour and convergence of processes of linguistic change. Methodologically, as instrument of data collection, it was used a questionnaire designed on the basis of three categories: uses of Portuguese; acceptance of Portuguese; and assessment of informants’ and their students’ level of competence. The results related to the uses of Portuguese show that there are differences between European and Mozambican Portuguese, but these differences are more expressed at semantic and phonological level than at syntactic and morphosyntactic level. As far as acceptance of Portuguese is concerned, the assessment is positive, in a way that they agree with the need of adapting European Portuguese to the Mozambican context through inclusion, in teaching, of the structures already consolidated in speaking. In relation to linguistic competences, the informants assess their competence as being excellent and their students’ as being weak, reflecting that linguistic competence is only limited to the mastery of grammatical structure but not to a simultaneous development of discursive, sociolinguistic and strategic abilities in language usage.Keywords: linguistic norm, linguistic variation, social assessment.
Word segmentation is a basic problem in natural language processing. With the languages having the complex writing system like the Khmer language in Southern of Vietnam, this problem really very intractable, posing the significant challenges. Although there are some experts in Vietnam as well as international having deeply researched this problem, there are still no reasonable results meeting the demand, in particular, no treated thoroughly the ambiguous phenomenon, in the process of Khmer language processing so far. This paper present a solution based on the syllable division into component clusters using two syllable models proposed, thereby building a Khmer syllable database, is still not actually available. This method using a lexical database updated from the online Khmer dictionaries and some supported dictionaries serving role of training data and complementary linguistic characteristics. Each component cluster is labelled and located by the first and last letter to identify entirety a syllable. This approach is workable and the test results achieve high accuracy, eliminate the ambiguity, contribute to solving the problem of word segmentation and applying efficiency in Khmer language processing.
Despite the frequent use of sketch maps in assessing environmental knowledge, it remains unclear how and to what degree familiarity impacts sketch map content. In the present study, we assess whether different levels of familiarity relate to differences in the content and spatial accuracy of environmental knowledge depicted in sketch maps drawn for the purpose of route instructions. To this end, we conduct a real-world wayfinding study with 91 participants, all of whom have to walk along a pre-defined route of approximately 2.3km length. Prior to the walk, we collect self-report familiarity ratings from participants for both a set of 15 landmarks and a set of areas we define as hexagons along the route. Once participants finished walking the route, they were asked to sketch a map of the route, specifically a sketch that would enable a person who had never walked the route to follow it. We found that participants unfamiliar with the areas along the route sketched fewer features than familiar people did. Contrary to our expectations, however, we found that landmarks were sketched or not regardless of participants' level of familiarity with the landmarks. We were also surprised that the level of familiarity was not correlated to the accuracy of the sketched order of features along the route, of the position of sketched features in relation to the route, nor to the metric locational accuracy of feature placement on the sketches. These results lead us to conclude that different aspects of feature salience influence whether the features are included on sketch maps, independent of familiarity. They also point to the influence of task context on the content of sketch maps, again independent of familiarity. We propose further studies to more fully explore these ideas.
It is no secret that people often use taboo words when speaking about persons and objects in their environment. Taboo words are charged with emotion and have observable impact on the listener as well as the speaker. The purpose of this study was to determine whether taboo words were quantitatively more offensive when used in combination with a proper name versus being used with a non-human object. We found that using taboo words to describe proper names does not cause a significant effect; however, we found that participants rated certain categories of taboo words as more offensive than other categories. In a second experiment, taboo words did affect ratings and memory for proper names and non-human objects.
This article proposes an ontology design pattern leading knowledge providers to represent knowledge in more normalized, precise and inter-related ways, hence in ways that help automatic matching and exploitation of knowledge from different sources. This pattern is also a knowledge sharing best practice that is domain and language independent. It can be used as a criteria for measuring the quality of an ontology. This pattern is: "using binary relation types directly derived from concept types, especially role types or process types". The article explains this pattern and relates it to other ones, thereby illustrating ways to organize such patterns. It also provides a top-level ontology for generating relation types from concept types, e.g., those from lexical ontologies such as those derived from the WordNet lexical database. This generation and categorization helps normalizing knowledge, reduces having to introduce new relation types and helps keeping all the types organized.
Translation: Bryan Vit, Beatrix Busse and Ruth Möhlig-Falke This article attempts to sketch by example how discussions about English language norms have developed from the late 16th cen- tury until today. These complex discussions are closely related to the processes of standardisation and codification of English. They reflect the changing social norms that are shaped in the course of the 18th and 19th centuries as a consequence of industrialisation and urbanisation, as well as through the emergence of the British Empire on the one hand, and the growing economic and political importance of the United States on the other. While the discussion of language norms in the 18th and early 19th century is largely normative and prescriptive, the late 19th and 20th century sees the emergence and development of a descriptive tradition focused on linguistic diversity mainly in academic discourse, which is further influenced by linguistic anthropology and sociolinguistics since the mid-20th century. Today, public discourse about language norms is still frequently prescriptive, which is reflected for instance in the debates about politically correct language use or a fixed linguistic norm in edu- cation, as well as in discussions about the alleged decline of the English language due to its growing role as international lingua franca and global language.
The role of orthographic neighbors (e.g. bank—tank) in word processing has been discussed in many experimental studies. However, these studies have been conducted on a limited pool of languages, and many important questions are still unresolved. After creating a lexical database StimulStat that contains various neighborhood parameters for Russian, we conducted the rst experiment with substitution neighbors in Russian. We used lexical decision task with priming, and manipulated the following factors: whether the prime is more or less frequent than the target, whether the prime is a nominative singular (primary) form or an oblique form, and whether the substituted letter is word- nal or in the middle of the word. The results sug- gest that noun forms undergo morphological decomposition at a very early stage and shed new light on the process of activating candidates during lexical. The results also have practical signi cance because it is well known that spelling errors are in uenced by neighborhood e ects
This study aims to achieve qualitative research on the linguistic consciousness of the internet. This paper reviewed works of linguistic regarding norm and consciousness, and took note of social tax that is negative impact on the writer’s ethos. Then internet replies at the internet news media were collected according to the procedures. 1,500 replies analysed by linguistic norms were recorded in Excel and 29,380 replies on 1,500 replies were investigated to observe responses to errors. 161 replies in response to replies containing errors were categorized three types: ‘simple proofreading’, ‘making fun of error’, ‘negative evaluation’. As a result, it was confirmed that there was a self-purification of the using non-normative language in the Internet language ecosystem. The public saw the non-normative language and reacted differently. The public has shown a variety of responses to non - normative languages, and when negative evaluations have been excessive, they have manifested repulsion. In Korean language education, in order to cope with this phenomenon, it is necessary to expand the outline of peer assessment and to make good assessment of peer as education contents. Finally, there must be a large-scale study of non-normative language that bear more social taxes and expressions that do not.
Este artigo apresenta uma discussao sobre a variacao linguistica, mostrando as tres normas que coexistem no portugues do Brasil: norma-padrao, normas cultas e normas populares. Faz uma breve analise do conceito de norma, surgido no seio do Estruturalismo, a partir de trabalhos de Eugenio Coseriu (1980), antes de evidenciar a visao de diferentes pesquisadores da area da Sociolinguistica sobre essa tematica, a saber Bagno (2001; 2003; 2009), Faraco (2002; 2008), Lucchesi (2001; 2002) e Bortoni-Ricardo (2009). ABSTRACT: This article presents a discussion about linguistic variation, showing the three norms that coexist in brasilian portuguese: standard norm, cultured norm and popular norms. Give a brief review of the concept of norm, emerged in the framework of Structuralism from the works of Eugenio Coseriu (1980), before evidencing the vision of different researchers in the area of Sociolinguistic on this subject, named Bagno (2001; 2003; 2009), Faraco (2002; 2008), Lucchesi (2001; 2002) e Bortoni-Ricardo (2009). KEYWORDS: Linguistic variation, linguistic norms, standard norm.
The current study presents a direct comparison of the level of association of ingroup favoritism and outgroup hostility with opposition to multiculturalism policies in New Zealand. With both predictors operationalized as affect ratings of warmth and anger across separate models, ingroup favoritism and outgroup hostility were independently associated with European New Zealanders’ (N = 10,869) opposition to both resource-specific and symbolic policies. Furthermore, ingroup favoritism was more strongly associated with opposition to resource-specific policies which represent high realistic threat (compared with symbolic policies). In contrast, outgroup hostility was more consistently associated with both policy domains.
This article describes a pilot project of automatic morphosyntactic analysis system development and some results of our work. The approach, developed and adopted in this research, is caused by peculiarities of Arabic morphology and syntax, and implies parsing of morphosyntactic structures (both morphology and syntax) instead of traditional tokenization and division of language into morphology and syntax, which seems considerably artificial for many widespread phenomena of Arabic and leads to problems with parsing. Some existing Arabic corpora are discussed, some of them being even treebanks, but none of them having a common reliable underlying uniform formal model (a formal grammar of any kind) in public domain. The methodology of this project is described, and techniques used in the pilot study are discussed, including the software technologies adopted and developed. Current results are described with examples of morphosyntactic structures, immediate constituent classes (with information on dependencies), and code snippets of the grammar module.
Solving the issue of intuition activity dynamics opens up possibilities in the management of intuition, with consideration of factors affecting the temporal characteristics of its functioning. This is particularly important in dangerous professions, where the price for failure is high, such as the activities of Interior Ministry members. Subject of our research is features of intuition activity dynamics in police officers from different substructures with consideration for gender. For our research purposes, we used a set of methods (both theoretical and practical) that includes: generalization, theoretical analysis of potential factors of dynamics of intuition activity; the interpretation of the experimental procedure with affective pictures relevant to the dangerous part of activity of police officer; authors’ software stimuli (modified for police officers as the test subjects), database International Affective Picture System” (IAPS); descriptive statistics, binomial statistical test, “Angular transformation of Fisher” statistical test. For majority of groups in men, the maximum of efficiency is observed during 21–30 attempts (for joined sample is 53.43%, p = 0.001 against 50% by binomial test). monotonic increase of efficiency during previous attempts possibly indicates that during anticipation of dangerous situations, relevant to the real activity of police officers, the training for inclusion into a task is quite important. Women police officers in general regardless to the number of trial, anticipate dangerous situations significantly more effectively (p = 0.036 by “Angular transformation of Fisher” test). Our conclusions are as follows: (i) pictures with content relevant to professional activity of police officers, more effectively anticipated during 21–30 trials; (ii) gender affects to the dynamics of anticipation of dangerous situations; (iii) in general, women police officers anticipate the dangerous contented pictures more effectively than men. Interior Ministry members; intuition activity; anticipation’ visual stimuli dangerous content; IAPS database Bem D., Tressoldi P., Rabeyron T., Duggan M. Feeling the Future: A Meta-Analysis of 90 Experiments on the Anomalous Anticipation of Random Future Events. F1000Research 4 (2015): 1188. DOI: 10.12688/f1000research.7177.1. Bem D.J. Feeling the Future: Experimental Evidence for Anomalous Retroactive Influences on Cognition and Affect. Journal of Personality and Social 100 (2011): 407–425. Binhi V.N. Principles of Electromagnetic Biophysics Moscow: FIZMATLIT Publisher, 2011. (In Russian). Geodakyan V.A. Evolutionary theory of sex. Priroda [The Nature] 8 1991: 60–69. (In Russian)/ Grigoriev P.E., Vasilieva I.V. Signal Role of Emotions in Activity Actualizing Intuition. Space and Time 4 (2015): 292–299. (In Russian). Grigoriev P.E., Vasileva I.V. The Dependence of Effectiveness of Affective Colored Images Predicting on Basic Needs Satisfaction. Space and Time 3 (2015): 350–358. (In Russian). Grigoriev P., Vasileva I.V., Ignatov A.N. The Possible Role of Intuition in Predicting Emotiogenic and Criminogenic Situations. All-Russian Scientific-Practical Conference Countering Extremism and Terrorism in the Crimean Federal District: Theory and Practice (Simferopol, Krasnodar University of the Crimean Branch of the Ministry of Internal Affairs of Russia, 8 Oct. 2015) Simferopol, 2015, pp. 234–238. (In Russian). Grigoriev P., Vasilieva I.V., Ignatov A.N. Individually-Psychological Correlates of Features of Dangerous Situations Intuitive Prediction in Members of Interior. Bulletin of Krasnodar University of Ministry of Interior 1 (2016): 196–201. (In Russian). Jung C.G. Psychic Energy Moscow: Akademichesky proekt Publisher, Foundation ‘Mir’ Publisher, 2010. (In Russian). Lang P.J., Bradley M.M., Cuthbert B.N. International Affective Picture System (IAPS): Affective Ratings of Pictures and Instruction Manual. Technical Report A-8. Gainesville, FL: University of Florida, 2008. Li A.G. On the Matter of Method for Studying Some Unusual Phenomena of Human Psyche. Parapsychology in the USSR 2 (1991): 34–38. (In Russian). Li A.G. Statistical Approaches to Processing and Interpreting Results of Experiments to Identify a Person's Abilities for Extrasensory Perception. Parapsychology in the USSR 1 (1992): 23–30. (In Russian). Naftulin D.H., Ware J.E., Donnelly F.A. The Doctor Fox Lecture: a Paradigm of Educational Seduction. Journal of Medical Education 48 (July 1973): 630–635. Rosenthal R., Jacobson L. Pygmalion in the Classroom. New York: Irvington, 1992. Schmidt H. Observation of a Psychokinetic Effect under Highly Controlled Conditions. Journal of Parapsychology 57 (1993): 351–372. Schmidt H. The Strange Properties of Psychokinesis. Journal of Scientific Exploration 1.2 (1987): 103–118. Smith M.D. The of the 'Psi-conducive' Experimenter.': Personality, Attitudes towards Psi, and Personal Psi Experience. Journal of Parapsychology 67 (2003): 117–128. Smith M.D. The Role of the Experimenter in Parapsychological Research. Journal of Consciousness Studies 10 (2003): 69–84. Vasilieva I.V. Intuitive Mechanisms in Self-regulation Structure in Extreme Situations in Representatives of Dangerous Professions Tyumen: Pechatnik Publisher, 2010. (In Russian). Vasilieva I.V. Questionnaire for Studying Intuition Parameters in Self-regulation Structure in Extreme Situations in Representatives of Dangerous Professions. Bulletin of Tyumen State University 5 (2010): 148–154. (In Russian). Vasilieva I., Grigoriev P.E. Feedback Effect in Activity Actualizing Intuition. Innovative Projects and Programs in Education 5 (2014): 52–60. (In Russian). Vasilieva I.V., Grigoriev P.E. Intuition: From the Contradictory of Theoretical Explanations to the Methodology of Evidence-based Empirical Research. Bulletin of Tyumen State University. Pedagogy. 9 (2014): 189–195. (In Russian) Vasilieva I.V., Grigoriev P.E., Taratukhin A.A. Computer Maintenance of Intuition Studies. Proceedings of the 3rd All-Russian Conference on Psychological Diagnosis Modern Psychodiagnostics in Russia. Overcoming the Crisis (Chelyabinsk, 9–11 Sep. 2015). Chelyabinsk: Research Centre of South Ural State University Publisher, 2015, volume 1, pp. 45–48. (In Russian). Vasilieva I.V., Zaeva M.A., Grigoriev P.E. Relation between Gender and Features of Intuition in Young People. Proceedings of the Scientific Conference Ananiev Readings – 2015: Fundamental Problems of Psychology (St. Petersburg, October 20-22, 2015). St. Petersburg: St. Petersburg State University Publisher, Skifiya-print Publisher, 2015, pp. 12–13. (In Russian). Grigoriev, P. G., I. V. Vasilieva, and A. N. Ignatov. Dynamics of Intuitive Activity for Predicting Socially Dangerous Situations in Interior Ministry Members. Space and Time 1 (2017): 275–283. (In Russian). Fixed network address 2226-7271provr_st1-27.2017.103.
The lexicon of air transport has received the influence of terms from other modalities of locomotion, of sciences such as meteorology and ornithology and it has also incorporated loan words from different languages. Although it has been the object of lexicological study in different languages, there are few studies that address the appropriation of this lexicon by non-specialists and no dictionary exists that accounts for the non-technical uses of such lexical units. This article presents the lexicographic decisions adopted to design El lxico espaol del vuelo, a proposal that attempts to record the daily use of Spanish aeronautical words, giving an account of their diatopic, diachronic and diaphasic variation. Specifically, it describes both the selection of materials and the conformation of the lexical database as well as the text macrostructure, the criteria that operated in the lemmas selection and ordering and, finally, the microstructure of the lexicographic articles.
This study investigated how potential customers (N = 28) respond to two types of electronic word-of-mouth (eWOM) regarding the same product. The study simulated reality by having participants read either mainly negative comments from an independent discussion forum (n = 14) or mainly positive comments from a marketer's website (n=14). The results showed that the participants' valence ratings were positive after reading eWOM on the marketer's website and negative after reading eWOM on the independent forum. Although this seems obvious, it is interesting that even though the comments on the independent forum were not considered trustworthy or expert, reading these comments negatively influenced the product image. Participants who read the independent forum rated the product image significantly lower than participants who read the marketer's website. After watching commercial videos, both groups rated the product image higher; however, the difference between the groups remained significant. The results suggest that the emotions evoked by eWOM play a key role in product image. A practical implication for companies may be purchasing targeted advertising on discussion forums to manage potential customers' negative affective reactions.
Some French discussion forums devoted to pregnancy or medically assisted procreation show that commenters—all of them women—may sometimes use words such as brybry, gygy, fofos, zozos, or zhom to speak about the embryo, gynecologist, follicles, sperm, or their male companions.This article investigates how these commenters conceive of the use of such a vocabulary in relation to what they may consider the linguistic norm as well as in relation to the community that they may identify with. The folk linguistics approach allows us to identify several perspectives: on the one hand, this vocabulary, which has not yet incorporated the system of linguistic standards, may still be perceived as normal or even as prescriptive insofar as the mode of communication (the discussion forum) used by commenters is concerned; on the other hand, this same vocabulary can sometimes be considered morally blameworthy given the perceived rectitude of the standard language; finally, it can be thought of as an important element that ensures the cohesion of a discursive community, which is defined here by the mode of communication, but also and especially by gender and the bodily experience, real or desired, of pregnancy.
The peculiarities of the system of teaching Chinese students Russian as a foreign language are brought about by the specificity of the Chinese language. The article substantiates the system of methods and approaches facilitating efficient acquisition of a foreign language and the culture of its speakers; the author believes that the systemic approach (allowing appropriation of linguistic norms as a system embracing deferent levels) and the process-focused approach (allowing formation of a variable communicative field of work with the text) are the leading ones. The author pays special attention to teaching reading and working with the text. The author also discusses some of the didactic techniques that can help to train students in text comprehension, and the techniques for checking understanding of the text. The author believes that the main accent in the process of teaching Russian and speech culture to Chinese students should be oriented towards enrichment of vocabulary, consolidation of morphology and syntactical constructions characteristic of certain situations of communication, as well as towards development of speech culture and acquisition of the rules of speech etiquette. The author offers a short plan of a lesson as an example of such work.
Data-driven syntactic parsers are usually trained, tested and developed on web-news data. Little has been done to evaluate them on literary genres of different ages, which are still low-resource varieties in terms of syntactic annotation. In this paper, I will describe methodology and results concerning first experiments in testing two different kinds of dependency parsers on aesthetic writings by Schiller and F. Schlegel. First, I trained and tested the parsers on de-ud1.2, a treebank collecting German web-news texts. Second, I manually annotated excerpts by the two authors with syntactic metadata. Third, I tested the parsers on these excerpts, after training them on de-ud1.2
Due to the level of abstraction and subjectivity, the teaching-learning process of the opposition indefinido-perfecto in the indicative mode is a complex content to teach in in the Spanish classrooms as a foreign language. For Italian speakers who learn Spanish in the sociocultural context where the language is spoken, the structural similarities between their mother language and the target language hinder the learning of these verbal tenses, because the perception of minimum distance allows the commutation of the linguistic systems of both languages. The previous bring about the excessive use of the transfer, the fossilization of errors and the consequent stagnation of the Interlingua. The insufficiencies that these students present, in particular, those of the elementary level, influence in the production of texts, both oral and written, with correction and in accordance with the linguistic norms of the sociocultural context where he/she learns the language. These reasons motivate the analysis of some theoretical-methodological bases on the particularities of the teaching-learning process of the opposition indefinido-perfecto in the indicative mode in the Spanish language for Italian students.
Norms are essential to the human condition. Whether in the guise of tradition, culture, canon or rules, norms are therefore central to studies in the humanities. This book focuses on Russian language culture of the post-revolutionary and post-Soviet periods, times when norms — linguistic and otherwise — have been eagerly debated, challenged, broken and redefined. Exploring the intersections between linguistic authority and creative response, an international team of scholars examines different realms of linguistic practice (literary fiction, internet slang, literary criticism and aesthetics, writers’ blogs, linguistic play) and various arenas for “talk about talk” (the classroom, blogs, the media, or the courtroom). By combining various approaches and disciplines — linguistics, literary criticism, new media studies — the book as a whole explores the multiplicity of meanings that are accorded to the notion of linguistic norms in the Russian community. The result is both a broad and a detailed picture of important trends in modern Russian language culture.
The peculiarities of the system of teaching Chinese students Russian as a foreign language are caused by features of Chinese. The article substantiates the system of methods and approaches facilitating efficient acquisition of a foreign language and the culture of its speakers; the author believes that the systemic approach (allowing appropriation of linguistic norms as a system at deferent levels) and the process-focused approach (allowing formation of a communicative field in different variants of work with the text) are the leading ones. The author also considers certain didactic techniques which make it possible to reach a high level of formation of a foreign language communicative competence, for example, the use of tongue twisters as didactic means. The author believes that the main accent in teaching Chinese students Russian and the standard of speech should focus on vocabulary enrichment, grammatical and syntactic constructions typical of certain situations, and also development of speech culture and speech etiquette. The author offers a plan of one lesson as an example of such work.
Due to the level of abstraction and subjectivity, the teaching-learning process of the opposition indefinido-perfecto in the indicative mode is a complex content to teach in in the Spanish classrooms as a foreign language. For Italian speakers who learn Spanish in the sociocultural context where the language is spoken, the structural similarities between their mother language and the target language hinder the learning of these verbal tenses, because the perception of minimum distance allows the commutation of the linguistic systems of both languages. The previous bring about the excessive use of the transfer, the fossilization of errors and the consequent stagnation of the Interlingua. The insufficiencies that these students present, in particular, those of the elementary level, influence in the production of texts, both oral and written, with correction and in accordance with the linguistic norms of the sociocultural context where he/she learns the language. These reasons motivate the analysis of some theoretical-methodological bases on the particularities of the teaching-learning process of the opposition indefinido-perfecto in the indicative mode in the Spanish language for Italian students.
Norms are essential to the human condition. Whether in the guise of tradition, culture, canon or rules, norms are therefore central to studies in the humanities. This book focuses on Russian language culture of the post-revolutionary and post-Soviet periods, times when norms — linguistic and otherwise — have been eagerly debated, challenged, broken and redefined. Exploring the intersections between linguistic authority and creative response, an international team of scholars examines different realms of linguistic practice (literary fiction, internet slang, literary criticism and aesthetics, writers’ blogs, linguistic play) and various arenas for “talk about talk” (the classroom, blogs, the media, or the courtroom). By combining various approaches and disciplines — linguistics, literary criticism, new media studies — the book as a whole explores the multiplicity of meanings that are accorded to the notion of linguistic norms in the Russian community. The result is both a broad and a detailed picture of important trends in modern Russian language culture.
The aim of this paper is to identify the influence degree among the main rating agencies and the other variables that affect rating changes for sub-sovereign entities in Germany, Austria, Belgium, France, Italy and Spain, using a total of 32 territorial entities between 1996 and 2012. Due to the shortage of European sub-sovereigns with more than 2 ratings, we estimated six binary probit regressions as a combination of 3 rating agencies two to two. We conclude that Fitch is the most influential agency on the other two rating agencies, but Standard and Poor's is the leader. There are other relevant
Norms are essential to the human condition. Whether in the guise of tradition, culture, canon or rules, norms are therefore central to studies in the humanities. This book focuses on Russian language culture of the post-revolutionary and post-Soviet periods, times when norms — linguistic and otherwise — have been eagerly debated, challenged, broken and redefined. Exploring the intersections between linguistic authority and creative response, an international team of scholars examines different realms of linguistic practice (literary fiction, internet slang, literary criticism and aesthetics, writers’ blogs, linguistic play) and various arenas for “talk about talk” (the classroom, blogs, the media, or the courtroom). By combining various approaches and disciplines — linguistics, literary criticism, new media studies — the book as a whole explores the multiplicity of meanings that are accorded to the notion of linguistic norms in the Russian community. The result is both a broad and a detailed picture of important trends in modern Russian language culture.
In this paper we focus on collocations, which have been studied in computational linguistics since they constitute a key factor when processing natural languages. For instance, they usually represent a challenge in automatic translation because the association of two terms is not easily computed. We proposed that the parser should be provided with a lexical database in order to make more effective the identification of collocations during the parsing process. We assessed this claim by using a corpus of 6’000 sentences retrieved from the British magazine The Economist Espresso. The corpus was parsed twice, first with the collocation detection component turned on and then with it turned off, and to make the comparison the Fips tagger was used. The results showed an improvement of the quality when the parser has access to collocation knowledge.
Norms are essential to the human condition. Whether in the guise of tradition, culture, canon or rules, norms are therefore central to studies in the humanities. This book focuses on Russian language culture of the post-revolutionary and post-Soviet periods, times when norms — linguistic and otherwise — have been eagerly debated, challenged, broken and redefined. Exploring the intersections between linguistic authority and creative response, an international team of scholars examines different realms of linguistic practice (literary fiction, internet slang, literary criticism and aesthetics, writers’ blogs, linguistic play) and various arenas for “talk about talk” (the classroom, blogs, the media, or the courtroom). By combining various approaches and disciplines — linguistics, literary criticism, new media studies — the book as a whole explores the multiplicity of meanings that are accorded to the notion of linguistic norms in the Russian community. The result is both a broad and a detailed picture of important trends in modern Russian language culture.
This chapter concerns the largely ignored phenomenon of inner speech within religious groups. In three major parts, the first considers some linguistic, philosophical and theological approaches used to frame the phenomenon in the past, the second exemplifies and compares inner speech in, Sikh devotional focus on divine words, Zen Buddhist philosophical concern with silence and human being and, one contemporary Christian context concerning prayer. The third part asks how inner speech might relate to the interpersonal dynamics of religious group membership. One contextual background to the phenomenology of inner speech concerns the cultural shift from spoken to silent reading. Sikhism strongly advocates inner speech for devotional goals. Sikhism is both a radically corporate endeavour of a religious community and a profoundly individual pursuit of union with the divine. The sociological interest of this Zen case lies precisely in the attempt at a serious deconstruction of linguistic norms and 'structures'.
This paper discusses the use of ‘by’-phrases in Impersonal Passives in Icelandic. It has been claimed in the literature on Icelandic syntax that ‘by’-phrases that express the agent are not very good or even ungrammatical in Impersonal Passives. The pa-per shows that this point of view oversimplifies the facts because various examples of this pattern can be found in natural data and these examples do not seem to reflect mistakes in linguistic performance. We discuss examples from the Icelandic treebank and from the web and we suggest that ‘by’-phrases are more likely to be used in Impersonal Passives if they involve new information and/or if they are heavy. One of the conclusions of the article is that large and well annotated corpora are important for linguistic research that focuses on rare constructions.
Previous research has shown that good looks, particularly being deemed as attractive or competent-looking, can provide an electoral advantage. There is also evidence to support the notion that more dominant looks are associated with military success as cadets are more likely to rise in the ranks early in their career if they are more dominant-looking. To date, there has been little research into the effect of looks on political leadership success in a non democratic setting. This project explores the effect of facial attractiveness and dominance on the political success of leaders after leading a successful coup d’état. We examine a comprehensive set of coup d'états from 1946 to 2013. Attractiveness and dominance ratings are created via surveys, as in previous research, but with a novel way to control for the potential bias arising from respondent characteristics. Defining political success as taking executive power, longer time-to-office exit, and avoiding constraint on executive power, we find that both dominant and attractive facial features provide distinct advantages for leaders.
The article describes the process of retaining the recessive units in the correctness publications which have been the sources of the codified norm for the last hundred years. This includes the following forms: interesa, czochrze, w Prusiech. Such elements marked i.a. as rare, former, out of use or outdated belonged in the given period to the linguistic norm of some speakers of Polish. Therefore linguists made attempts to retain them in the codified norm at least for some time so that they could serve as evidence of their former correctness. They were used with qualifiers which provided information about potential question concerning the topicality of the recessive elements. Such units met different fates. They were no longer provided or perceived as wrong in the succeeding dictionaries that were examined. They often functioned as recessive until contemporaneity or even returned as equal variants. A detailed typology of these relations was discussed in the article.
ForFun is a database of linguistic forms and their syntactic functions built with the use of the multi-layer annotated corpora of Czech, the Prague Dependency Treebanks. The purpose of the Prague Database of Forms and Functions (ForFun) is to help the linguists to study the form-function relation, which we assume to be one of the principal tasks of both theoretical linguistics and natural language processing. A prototypical question to be asked is What purposes does a preposition 'po' serve for or What are the linguistic means in the sentence that can express the meaning 'a destination of an action'?. There are almost 1500 distinct forms (besides the 'po' preposition) and 65 distinct functions (besides the 'destination').
Dependency parsing is considered as the state of the art technology for a better information extraction methodology in Natural Language Processing. With the ever-growing need for linguistic analysis for different languages, the demand for multilingual dependency parsing has increased dramatically. In this research work we studied a novel collection of treebanks with homogeneous syntactic dependency annotation [1] for six languages along with other recent techniques in this area. We investigated the possibility of adding new languages in this module and successfully added universal Bangla dependency annotation. Additionally, we combined simple and complex feature representations to improve parsing output.
This chapter provides an overview of language policy and the media by reviewing the state of the art, both in terms of literature and in terms of research. It outlines key terms and their uses, and explains the types of language policy and media. The chapter also provides an overview of disciplinary perspectives on language policy and the media, with a particular focus on the evolution from traditional news to new and social media. It reviews research within a range of national and globalized contexts, and discusses the core areas of status, corpus and acquisition planning. The chapter examines relationship between language policy and the media in two overarching areas: in the chronological transition from nation-states to globalization, and in status, corpus, and acquisition planning. Language policy concerns the production and enforcement of linguistic norms; the reality of policymaking and its implementation is much more complex than such a simplistic label implies.