Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
16504 papers
<p><em>Abstrak - </em><strong>Penelitian ini berjudul “Norma dan Eksploitasi Tipe Semantis Properti Fisik Ajektiva pada Frasa nomina <em>‘Eye’</em></strong><strong><em> </em></strong><strong>d</strong><strong>alam COCA</strong><strong>”</strong><strong>.</strong><strong> </strong><strong>Bahasannya mengenai klasifikasi adjekitva terhadap nomina’eye’pada tipe semantis <em>PHYSICAL PROPERTY</em> di 50 frekuensi tertinggi dan 50 frekuensi terendah dari 500 frekuensi frasa nomina ‘eye’.</strong><strong> </strong><strong>Metode yang digunakan berdasarkan t</strong><strong>eori Creswell </strong><strong>dalam</strong><strong> </strong><strong><em>Research Design Qualitative, quantitative and Mixed Methods Approaches </em></strong><strong>(2009) </strong><strong> Hasil temuan dianalisis melalui teori yang diciptakann Hanks, ‘</strong><strong>L</strong><strong>exical Analysis Norms and Exploitations’ untuk melihat dinamika ekspoitasi frasa nomina <em>‘eye’</em></strong><strong><em></em></strong></p><p><strong><em> </em></strong></p><p><strong><em>Kata Kunci</em></strong><em>: Frasa nomina ‘Eye’ dalam COCA, Teori Hanks, Teori Creswell, </em></p><p> </p><p>Abstract - <strong>This research entitled </strong><strong>“</strong><strong>Norma dan Eksploitasi Tipe Semantis Properti Fisik Ajektiva pada Frasa nomina ‘Eye’</strong><strong> d</strong><strong>alam COCA</strong><strong>”</strong><strong>. The study is on the classification of semantic type for PHYSICAL PROPERTY in adjective of ‘eye’ as noun. The Method used is basing on Creswell theory in</strong><strong> Research Design Qualitative, quantitative and Mixed Methods Approaches (2009)</strong><strong>. The finding is analysed through Hanks’ theory ‘ Lexical Analysis Norms and exploitations’, to see the dynamic exploitation of noun Phrase’eye’.</strong><strong></strong></p><p> </p><p><strong><em>Keywords</em></strong><em>: Noun Phrase ‘Eye’ in COCA, Hanks Theory, Creswell Theory.</em><em></em></p>
The International Affective Picture System (IAPS; Lang, Bradley, & Cuthbert, 2008) is a stimulus database that is frequently used to investigate various aspects of emotional processing. Despite its extensive use, selecting IAPS stimuli for a research project is not usually done according to an established strategy, but rather is tailored to individual studies. Here we propose a standard, replicable method for stimulus selection based on cluster analysis, which re-creates the group structure that is most likely to have produced the valence arousal, and dominance norms associated with the IAPS images. Our method includes screening the database for outliers, identifying a suitable clustering solution, and then extracting the desired number of stimuli on the basis of their level of certainty of belonging to the cluster they were assigned to. Our method preserves statistical power in studies by maximizing the likelihood that the stimuli belong to the cluster structure fitted to them, and by filtering stimuli according to their certainty of cluster membership. In addition, although our cluster-based method is illustrated using the IAPS, it can be extended to other stimulus databases.
This paper studies the two approaches to translation which have been the subject of debate for a long timethe literal approach and the free approach. It presents the theoretical and conceptual framework of literal translation by illuminating the studies of a number of researchers with different attitudes to this phenomenon. The paper shows that there is no consensus among researchers on whether literal translation or free translation should be considered primarily an approach to translation. The study also analyzes the notion of literalism and distinguishes between the main types of literalisms (etymological, semantic, lexical and grammatical) by illustrating examples. The paper uses general scientific methods such as observation, analysis and synthesis, as well as descriptive, classification, generalization and selection methods. The author proves that literalism often causes the distortion of meaning and violates the norms of the target language. Therefore, the translator should carefully select equivalents to prevent the artificiality of translation.
The Quality Department of the French National Space Agency (CNES, Centre National d’Études Spatiales) wishes to design a writing guide based on the real and regular writing of requirements. As a first step in this project, the present article proposes a linguistic analysis of requirements written in French by CNES engineers. One of our goals is to determine to what extent they conform to several rules laid down in two existing Controlled Natural Languages (CNLs), namely the Simplified Technical English developed by the AeroSpace and Defense Industries Association of Europe and the Guide for Writing Requirements proposed by the International Council on Systems Engineering. Indeed, although CNES engineers are not obliged to follow any controlled language in their writing of requirements, we believe that language regularities are likely to emerge from this task, mainly due to the writers’ experience. We are seeking to identify these regularities in order to use them as a basis for a new CNL for the writing of requirements. The issue is approached using natural language processing tools to identify sentences that do not comply with the rules or contain specific linguistic phenomena. We further review these sentences to understand why the recommendations cannot (or should not) always be applied when specifying large-scale projects.
This study investigates the performance of 22 monolingual and 54 bilingual children with and without specific language impairment (SLI), in a nonword repetition (NWRT) and a sentence repetition task (SRT). Both tasks were constructed according to the principles for LITMUS tools (Language Impairment Testing in Multilingual Settings) developed within COST Action IS0804, and incorporated phonological or syntactic structures that are linguistically complex and have been shown to be difficult for children with SLI across languages. For phonology these are in particular (non)words containing consonant clusters. In morphosyntax complexity has been attributed to factors such as embedding and/or syntactic movement. Tasks focusing on such structures are expected to identify SLI in bilinguals across language combinations. This is notoriously difficult because structures that are problematic for typically developing bilinguals (BiTDs) and monolingual children with SLI (MoSLI) often overlap. We show that the NWRT and the SRT are reliable tools for identification of SLI in bilingual contexts. However, interpretation of the performance of bilingual children depends on background information as provided by parental questionnaires. In order to evaluate the accuracy of our tasks we recruited children in ordinary kindergartens or schools and in Speech Language Therapy centers and verified their status with a battery of standardized language tests, assessing bilingual children in both their languages. We consider a bilingual child language impaired if she shows impairments in two language domains in both her languages. For assessment we used tests normed for monolinguals (with one exception) and adjusted the norms for bilingualism and for language dominance. This procedure established the following groups: 10 typical monolinguals (MoTD), 12 MoSLI, 46 BiTD and 8 bilingual children with SLI (BiSLI). Our results show that both tasks target relevant structures: monolingual children are classified with 100% accuracy. Crucially, both our tasks distinguish BiTDs from MoSLIs and BiTDs from BiSLIs. The NWRT shows high accuracy and only minimal influence of language dominance. The SRT can be scored as “identical repetition” or as “target structure”, the latter aiming for scoring the mastery of a syntactic structure, ignoring lexical and specific case or gender errors. Focusing on the latter measure, we
Teesid: Käesolev artikkel käsitleb uusklassikalist luulet ehk luulet, mis tärkab humanistliku hariduse pinnalt ja on loodud nn klassikalistes keeltes ehk vanakreeka ja ladina keeles. Artikli esimene pool toob välja paar üldist probleemi varauusaja poeetika käsitlemises nii Eestis kui mujal. Teises osas esitatakse alternatiivina mõned näited (autoriteks G. Krüger, H. Vogelmann, L. Luden, O. Hermelin ja H. Bartholin) Tartu ja Tallinna uusklassikalisest luulest värsstõlkes koos poeetika analüüsidega, avalikkusele tundmata luuletuste puhul esitatakse ka originaaltekstid. SUMMARYThis article discusses poetry in classical languages (Humanist Greek and Neo-Latin) belonging to the classical literary tradition while focusing on poetry from Tallinn and Tartu from the sixteenth and seventeenth centuries. It does not aim to present an overview of this tradition in Estonia (already an object of numerous studies), but rather to discuss some general problems connected to such studies—both in Europe and Estonia—and to show some alternative (or complementary) analyses of neo-classical poetics, together with verse translations and texts that are not easily available or are unknown to the scholars.The discussion of neo-classical poetry in Estonia finds problems in a detachment from poetics and the consequent discrepancies. Firstly, although scholarly treatises stress the value of casual poetry (forming the most eminent part of Estonian Neo-Latin and Humanist Greek poetry), the same treatises present this poetry from the viewpoint of its social background, focusing more on the authors and events than the poetic form. For example, in the Anthology of Tartu casual poetry and the corpus of Neo-Latin poetry from Tartu, texts are presented according to genre, which is defined only according to the classification of social events (epithalamia, epicedia, congratulations for rectorate, disputations, etc). Secondly, in most cases (the anthology, re-editions), this poetry is presented to readers as prose translations. As in the case of ancient Greek and Roman poetry, the established norm in Estonia is verse translation. Translating poetry into prose, therefore, signals that these works are not to be considered poetry. Thirdly, commentaries on this poetry tend to list lexical parallels with authors from classical antiquity without distinguishing actual quotations from the usage of poetic formulae while simultaneously (mostly) ignoring the impact of pagan and Christian texts from late antiquity and renaissance and humanist literature.One alternative is to present Neo-Latin and Humanist Greek poetry as verse translations and focus more on discussing poetic devices and the impact of its contemporary poetry. Therefore, the second part of this article presents five poems as translations of verse and a subsequent analysis of their poetics.The first example is from a manuscript in the Tallinn City Archives and represents the earliest collection of neo-classical poetry, containing one Latin and five Greek poems belonging to the epistolary poem genre. Its author, Gregor Krüger Mesylanus (a latinized Greek translation of the name of his birth-town Mittenwalde, near Berlin), worked as a priest in Reval after his studies in Wittenberg during the time of Ph. Melanchthon (which explains Krüger‘s chosen poetic form). The Greek cycle is regarded thematically as variations on the same subject of the author‘s longing for home and his unhappiness with the jealousy and hostility of his fellow citizens in Reval. His choice of meter is influenced by Latin poetry, the initial long elegy balanced by four shorter poems of different meters (iambic and choriambic patterns). The final poem of the Greek cycle (Enviless Moon) is presented together with a metrical translation and analysis to demonstrate how sonorous patterns orchestrate the thematic development of the poem: the author‘s wish to be like the moon, who receives its light from the brighter sun, but remains still happy and grateful to God for his own gift and ability to bring a smaller light to others.The second example analyzes the structure and poetic motives of a metrical translation of a Greek Pindaric Ode by Heinrich Vogelmann from 1633. The paper’s author also examines the European tradition of The second example analyzes the structure and poetic motives of a metrical translation of a Greek Pindaric Ode by Heinrich Vogelmann from 1633. The paper’s author also examines the European tradition of such odes (including more than sixty examples from 1548 until 2004). The third example discusses two alternative translations and additional translation possibilities of a recently discovered anagrammatic poem by Lorenz Luden. The fourth and fifth examples are congratulatory poems addressed to Andreas Borg for the publication of his disputation on civil liberty (in 1697). A Latin congratulatory poem by Olaus Hermelin is an example of politically engaged poetry, which addresses not the student but the subject of his disputation and contemporary political situation (the revolt of Estonian nobility against the Swedish king, who had recaptured donated lands, and the exile of its leader, Johann Reinhold Patkul). The Greek poem by H. Bartholin refers to the arts of Muses to demonstrate the changes in poetical representations of university studies: by the end of the seventeenth century the motives of the dancing and singing, flowery Muses is replaced with the stress of the toil in the stadium and the labyrinth of Muses.This article discusses poetry in classical languages (Humanist Greek and Neo-Latin) belonging to the classical literary tradition while focusing on poetry from Tallinn and Tartu from the sixteenth and seventeenth centuries. It does not aim to present an overview of this tradition in Estonia (already an object of numerous studies), but rather to discuss some general problems connected to such studies—both in Europe and Estonia—and to show some alternative (or complementary) analyses of neo-classical poetics, together with verse translations and texts that are not easily available or are unknown to the scholars.The discussion of neo-classical poetry in Estonia finds problems in a detachment from poetics and the consequent discrepancies. Firstly, although scholarly treatises stress the value of casual poetry (forming the most eminent part of Estonian Neo-Latin and Humanist Greek poetry), the same treatises present this poetry from the viewpoint of its social background, focusing more on the authors and events than the poetic form. For example, in the Anthology of Tartu casual poetry and the corpus of Neo-Latin poetry from Tartu, texts are presented according to genre, which is defined only according to the classification of social events (epithalamia, epicedia, congratulations for rectorate, disputations, etc). Secondly, in most cases (the anthology, re-editions), this poetry is presented to readers as prose translations. As in the case of ancient Greek and Roman poetry, the established norm in Estonia is verse translation. Translating poetry into prose, therefore, signals that these works are not to be considered poetry. Thirdly, commentaries on this poetry tend to list lexical parallels with authors from classical antiquity without distinguishing actual quotations from the usage of poetic formulae while simultaneously (mostly) ignoring the impact of pagan and Christian texts from late antiquity and renaissance and humanist literature. One alternative is to present Neo-Latin and Humanist Greek poetry as verse translations and focus more on discussing poetic devices and the impact of its contemporary poetry. Therefore, the second part of this article presents five poems as translations of verse and a subsequent analysis of their poetics. The first example is from a manuscript in the Tallinn City Archives and represents the earliest collection of neo-classical poetry, containing one Latin and five Greek poems belonging to the epistolary poem genre. Its author, Gregor Krüger Mesylanus (a latinized Greek translation of the name of his birth-town Mittenwalde, near Berlin), worked as a priest in Reval after his studies in Wittenberg during the time of Ph. Melanchthon (which explains Krüger‘s chosen poetic form). The Greek cycle is regarded thematically as variations on the same subject of the author‘s longing for home and his unhappiness with the jealousy and hostility of his fellow citizens in Reval. His choice of meter is influenced by Latin poetry, the initial long elegy balanced by four shorter poems of different meters (iambic and choriambic patterns). The final poem of the Greek cycle (Enviless Moon) is presented together with a metrical translation and analysis to demonstrate how sonorous patterns orchestrate the thematic development of the poem: the author‘s wish to be like the moon, who receives its light from the brighter sun, but remains still happy and grateful to God for his own gift and ability to bring a smaller light to others. The second example analyzes the structure and poetic motives of a metrical translation of a Greek Pindaric Ode by Heinrich Vogelmann from 1633. The paper’s author also examines the European tradition of This article discusses poetry in classical languages (Humanist Greek and Neo-Latin) belonging to the classical literary tradition while focusing on poetry from Tallinn and Tartu from the sixteenth and seventeenth centuries. It does not aim to present an overview of this tradition in Estonia (already an object of numerous studies), but rather to discuss some general problems connected to such studies—both in Europe and Estonia—and to show some alternative (or complementary) analyses of neo-classical poetics, together with verse translations and texts that are not easily available or are unknown to the scholars.The discussion of neo-classical poetry in Estonia finds problems in a detachment from poetics and the consequent discrepancies. Firstly, although scholarly
In this article, we introduce an explicit count-based strategy to build word space models with syntactic contexts (dependencies). A filtering method is defined to reduce explicit word-context vectors. This traditional strategy is compared with a neural embedding (predictive) model also based on syntactic dependencies. The comparison was performed using the same parsed corpus for both models. Besides, the dependency-based methods are also compared with bag-of-words strategies, both count-based and predictive ones. The results show that our traditional count-based model with syntactic dependencies outperforms other strategies, including dependency-based embeddings, but just for the tasks focused on discovering similarity between words with the same function (i.e. near-synonyms).
The article is devoted to a lexical analysis of the New Testament translation into Russian performed by an established statespersonKonstantin Pobedonostsev (1827Pobedonostsev ( -1907) ) at the beginning of the 20 th century. The lexical particularities of it have been revealed by means of its lexical comparison with the Synodal translation and the Church-Slavonic liturgical version. According to academic interpretation of the data collected the author stated that Konstantin Pobedonostsev's translationpreservesmore resemblance to the Church-Slavonic liturgical version, as there are 185 wordsinhis work, which were not found in the Synodal translation, but were used in Church-Slavonic liturgical version. The major part of the words is registered in lexicographical sources and reflects language norms of Pobedonostsev's lifetime. However, the smaller part is registered in the Russian Language National Corpus as being presented in the texts of the 1820-1920 period. The vocabulary specificity of the translation version under study lies in vast references to the Church-Slavonic liturgical version word pool. This fact is explained by stylistic preferences of the Gospel translators as well as by the target addressee image (people who are well informed about the orthodox liturgical tradition). In conclusion the author states that in his translation Konstantin Pobedonostsev never took words from the Church-Slavonic texts without prolonged meditation, the replacement cases seem to be an intentional choice of the words that were frequently used in the speech of his epoch.
online materials cited were simply unfindable. The index runs to almost forty pages of quadruple columns—but what is included? One third of my randomly selected “test”NLs were absent (Le Bouscat, Germignan, Le Taillan, Le Haillan, Sabres,Auros). Parentis-en-Born is listed s.v.“Born” but not under “p.” Souesmes was missing, while Solliès-Toucas figured only as Solliès—part of an enumeration of terms of ensoleillement, most of which are absent from the index. Finally, the key question for a volume of this scope: why is there no electronic edition, effortlessly searchable? In the end, the Trésor’s hybrid approach can never stray far from linguistic analysis. Brunet the geographer stands on the shoulders of linguists who have unearthed and identified (défriché, déchiffré) the etyma underlying his concept-based categories, without which this volume would not have been possible. University of Hawaii, Ma -noa Kathryn Klingebiel Cannone, Belinda, et Christian Doumet. Dictionnaire des mots manquants. Vincennes: Thierry Marchaisse, 2016. ISBN 978-2-36280-094-8. Pp. 216. Embracing the notion of the lexical gap, this book takes a rigorous and philosophical approach to populating the patchy and often inconsistent lexical landscape of the French language. It is a literary dictionary of sorts, which employs a method of semantic triangulation to visualize and articulate a series of lexical gaps: a dominant keyword serves as the theme of each association, while the two other members serve to delimit the nature of their relationship, the result of which is sometimes one or many novel terms the author(s) have cobbled together from preexisting French morphemes, for example entre-deux-pouvoir-vouloir pveux (peux+veux) to describe someone’s inability to do something because, in fact, they never wanted to; deuilparent -enfant im-père (in+père) to describe a father who has lost his only child. Composed of fifty-nine entries alphabetized by keyword and penned by forty-four expert users of the language (contemporary French authors, poets, philosophers, translators, and language and/or literature professors), the text reads like an edited volume of short stories. Some authors employ an academic style, presenting a collection of historical facts and offering well-paved lines of reasoning for the reader to follow on his guided semantic exploration (e.g., langaige françoys-interprétationspolitique ), whereas others ruminate more indirectly, instead telling a story: setting a scene, describing its players, and letting the unnamed concept emerge from the background all on its own (e.g., envers-visage-occiput). Regardless of their approach, the goal of these authors is not neologism for neologism’s sake, but rather the pursuit of enhanced expression, one so clear and unmistakable that it is made possible only by the kind of lexical precision one might attain after a series of rigorous exercises in both semantic and morphological permutation. This volume not only fills countless 272 FRENCH REVIEW 91.2 Reviews 273 lexical gaps in the French language with its small set of well-motivated innovations, but paves the way for other contemporary users to take action in situations of expressive lacunae by legitimizing a varied yet thoughtful innovation process accessible to all.As such, its value is all but limitless: for native users of French, it is a documentation of language agility—a testament to all the ways the language could neatly package recurring concepts out of familiar building blocks, but for arcane reasons does not. For second-language learners, it is a documentation of language fragility—a testament to just a few of the language’s idiosyncrasies, with the larger lesson that being a successful language user involves much more than knowing how to assemble familiar chunks of meaning into words that logically ought to exist. For this reason, this book is simultaneously a unique resource for experienced French writers to challenge and diversify their lexicon,and a semantic guidebook for second-language learners building their awareness and written expression one case study at a time. Bibliophiles and Francophiles alike will delight in the impressive artistry presented for reaching deep into the French lexicon and its sociohistorical norms for the sake of engineering one’s own mot juste. University of South...
Dyslexia has been claimed to be causally related to deficits in visuo-spatial attention. In particular, inefficient shifting of visual attention during spatial cueing paradigms is assumed to be associated with problems in graphemic parsing during sublexical reading. The current study investigated visuo-spatial attention performance in an exogenous cueing paradigm in a large sample (N = 191) of third and fourth graders with different reading and spelling profiles (controls, isolated reading deficit, isolated spelling deficit, combined deficit in reading and spelling). Once individual variability in reaction times was taken into account by means of z-transformation, a cueing deficit (i.e. no significant difference between valid and invalid trials) was found for children with combined deficits in reading and spelling. However, poor readers without spelling problems showed a cueing effect comparable to controls, but exhibited a particularly strong right-over-left advantage (position effec)
In this paper, extensive experiments are conducted to study the impact of features of different categories, in isolation and gradually in an incremental manner, on Arabic Person name recognition. We present an integrated system that employs the rule-based approach with the machine learning (ML)-based approach in order to develop a consolidated hybrid system. Our feature space is comprised of language-independent and language-specific features. The explored features are naturally grouped under six categories: Person named entity tags predicted by the rule-based component, word-level features, POS features, morphological features, gazetteer features, and other contextual features. As decision tree algorithm has proved comparatively higher efficiency as a classifier in current state-of-the-art hybrid Named Entity Recognition for Arabic, it is adopted in this study as the ML technique utilized by the hybrid system. Therefore, the experiments are focused on two dimensions: the standard dataset used and the set of selected features. A number of standard datasets are used for the training and testing of the hybrid system, including ACE (2003–2004) and ANERcorp. The experimental analysis indicates that both language-independent and language-specific features play an important role in overcoming the challenges posed by Arabic language and have demonstrated critical impact on optimizing the performance of the hybrid system.
Reviewed by: Corpus Stylistics as Contextual Prosodic Theory and Subtext by Bill Louw, Marija Milojkovic Feng (Robin) Wang (bio) and Philippe Humblé (bio) Bill Louw and Marija Milojkovic. Corpus Stylistics as Contextual Prosodic Theory and Subtext. John Benjamins Publishing Company, 2016. xix + 419 pp. $149. The term corpus stylistics, usually regarded as a near-synonym for stylometry, stylometrics, statistical stylistics, or stylogenetics, is closely related to statistics and corpus linguistics. Despite an increasing number of studies in the field, people still do not attain a clear line of demarcation between corpus linguistics and corpus stylistics. Corpus linguists are typically concerned with “repeated occurrences, generalizations and the description of typical patterns,” while corpus stylistic studies relate to “deviations from linguistic norms that account for the artistic effects of a particular text” (Mahlberg, “Corpus Stylistic Perspective” 19). However, more needs to be known about what new perspectives corpus linguistics can offer to the depiction of stylistic devices and the interpretation of stylistic values. Under these circumstances, Bill Louw and Marija Milojkovic’s Corpus Stylistics as Contextual Prosodic Theory and Subtext is instructive and worthy of reading, for it offers valuable perspectives for interdisciplinary investigations. This volume comprises two parts: the first part (Chapters 1–6) is devoted to the theoretical construction of Contextual Prosodic Theory (CPT), and the second part (Chapters 7–12) applies CPT to literary criticism, translation studies, and foreign language teaching. Chapter 1 revisits the proposal on “language and literature integration” in foreign language teaching. Louw dissolves the doubts from language teachers about “integration” by sufficiently discussing lexical syllabus design and progressive delexicalization. Having critically reviewed different theoretical perspectives on collocation, the authors argue in [End Page 550] Chapter 2 that one objective characteristic of literary devices is that they will demonstrate some evidence of relexicalization through collocation. Chapter 3 focuses on the theoretical interpretation of semantic prosody. Semantic prosody, according to Louw, is the “consistent aura of meaning with which a form is imbued by its collocations” (80). In Chapter 4, the author expounds that data-driven reading will produce a class of negotiator distinct from the intuitive counterparts. Chapter 5 affirms the role of collocation in terms of predicting and grading the potential success of all humorous contexts of situation as well as composition. Moreover, the interaction between collocation and events in the external world is capable of isolating humorous situations that are “waiting to happen” (132). Chapter 6 introduces subtext, a core concept of CPT, and proceeds to explore what these deviations from logical semantic prosody (subtext) can tell us about an author’s text. The second part (Chapters 7–12) is written by Milojkovic and adapts CPT to other disciplines. In this sense, the volume can be considered as a necessary reference for a consortium of scholars. In order to test the applicability and universality of CPT, Milojkovic applies CPT to Slavic languages, namely, Russian and Serbian. Based on a synthesis of the theoretical tools of CPT (i.e., collocation, semantic prosody, and subtext), Milojkovic analyzes the logical construction of literary worlds as well as a hitherto uncharted domain in corpus stylistics: authorial intention, that is, whether the author sincerely means what he or she writes. Chapter 8 reveals the subtext of “in the * of” in a translated poem of Pushkin as a picture of action verging on conflict, which inspires Milojkovic to probe into whether this is an incompatible grammatical pattern to express Pushkin’s call for resignation. Methodologically, the application of CPT in translation studies enriches the theoretical toolkit of corpus-based translation studies. Chapter 9 distinguishes inspired writing from banality by evaluating the deviation from the reference corpus. Chapter 10 puts forward the hypothesis that inspired writing will differ from uninspired in the density of its subtextual and prosodic clashes, and that the clashes themselves will be indicative of the presence of inspiration (274). In order to test this hypothesis, Milojkovic, in Chapter 10, contacts several poets to elicit clear-cut cases of inspired writing. The final two chapters, concerning applications for foreign language teaching, pertain to time-honored pedagogical stylistics. Chapter 11 is a piece [End Page 551] of classroom corpus stylistics research with a twofold purpose: empirically, to verify Louw...
Language comprehension involves the simultaneous processing of information at the phonological, syntactic, and lexical level. We track these three distinct streams of information in the brain by using stochastic measures derived from computational language models to detect neural correlates of phoneme, part-of-speech, and word processing in an fMRI experiment. Probabilistic language models have proven to be useful tools for studying how language is processed as a sequence of symbols unfolding in time. Conditional probabilities between sequences of words are at the basis of probabilistic measures such as surprisal and perplexity which have been successfully used as predictors of several behavioural and neural correlates of sentence processing. Here we computed perplexity from sequences of words and their parts of speech, and their phonemic transcriptions. Brain activity time-locked to each word is regressed on the three model-derived measures. We observe that the brain keeps track of t)
Plagiarism takes place when we use any person’s work without giving due acknowledgment. There are several fields where the text similarity is involved like web document retrieval, information mining, and searching related articles. Several approaches have been introduced for detecting plagiarism in the text documents based on the syntactic structure of the text, string similarity, fingerprinting, semantic meaning underlying the text, etc. The basic limitation of plagiarism detection systems these days is that they fail to detect tough cases of plagiarism. The proposed plagiarism detection approach is the hybrid of semantic and syntactic similarity between the text documents. This novel approach exploits linguistic information sources non-linearly using the lexical database for finding the relatedness between text documents. The proposed approach uses semantic knowledge to perform cognitive-inspired computing. The framework is capable of detecting intelligent plagiarism cases like a verbatim copy, paraphrasing, rewording in a sentence, and sentence transformation. The approach has been evaluated on the standard PAN-PC-11 dataset. The experiments show that our technique has outperformed other strong baseline techniques in terms of precision, recall, F-measure, and plagiarism detection (PlagDet) score. (PsycINFO Database Record (c) 2018 APA, all rights reserved)
Cognitive mechanisms for sign language lexical access are fairly unknown. This study investigated whether phonological similarity facilitates lexical retrieval in sign languages using measures from a new lexical database for American Sign Language. Additionally, it aimed to determine which similarity metric best fits the present data in order to inform theories of how phonological similarity is constructed within the lexicon and to aid in the operationalization of phonological similarity in sign language. Sign repetition latencies and accuracy were obtained when native signers were asked to reproduce a sign displayed on a computer screen. Results indicated that, as predicted, phonological similarity facilitated repetition latencies and accuracy as long as there were no strict constraints on the type of sublexical features that overlapped. The data converged to suggest that one similarity measure, MaxD, defined as the overlap of any 4 sublexical features, likely best represents mechanisms of phonological similarity in the mental lexicon. Together, these data suggest that lexical access in sign language is facilitated by phonologically similar lexical representations in memory and the optimal operationalization is defined as liberal constraints on overlap of 4 out of 5 sublexical features—similar to the majority of extant definitions in the literature. (PsycINFO Database Record (c) 2018 APA, all rights reserved)
The CELEX lexical database (Baayen, Piepenbrock & van Rijn 1995) was developed in the 1990s, providing a database of the syntactic, morphological, phonological and orthographic forms of between 50,000 and 125,000 words of Dutch, English and German. This database was used as the basis for the development of the PolyLex lexicons, which included syntactic, morphological and phonological information for around 3,000 words of Dutch, English and German. Orthographic information was subsequently added in the PolyOrth project. The PolyOrth project was based on the assumption that the underlying, lexical phonological forms could be used to derive the surface orthographic forms by means of a combination of phoneme-grapheme mappings and sets of autonomous spelling rules for each language. One of the complications encountered during the project was the fact that the phonological forms in CELEX were not always genuinely underlying forms which made deriving the orthographic forms tricky. This paper discusses the nature and status of underlying phonological forms, their relation to orthography and the issues of finding this information in databases. (PsycINFO Database Record (c) 2018 APA, all rights reserved)
In this study we present a novel set of discrimination-based indicators of language processing derived from Naive Discriminative Learning () theory. We compare the effectiveness of these new measures with classical lexical-distributional measures—in particular, frequency counts and form similarity measures—to predict lexical decision latencies when a complete morphological segmentation of masked primes is or is not possible. Data derive from a re-analysis of a large subset of decision latencies from the English Lexicon Project, as well as from the results of two new masked priming studies. Results demonstrate the superiority of discrimination-based predictors over lexical-distributional predictors alone, across both the simple and primed lexical decision tasks. Comparable priming after masked and type primes, across two experiments, fails to support early obligatory segmentation into morphemes as predicted by the morpho-orthographic account of reading. Results fit well with theory, wh)
Measuring the similarity between two sentences is often difficult due to their small lexical overlap. Instead of focusing on the sets of features in two given sentences between which we must measure similarity, we propose a sentence similarity method that considers two types of constraints that must be satisfied by all pairs of sentences in a given corpus. Namely, (a) if two sentences share many features in common, then it is likely that the remaining features in each sentence are also related, and (b) if two sentences contain many related features, then those two sentences are themselves similar. The two constraints are utilized in an iterative bootstrapping procedure that simultaneously updates both word and sentence similarity scores. Experimental results on SemEval 2015 Task 2 dataset show that the proposed iterative approach for measuring sentence semantic similarity is significantly better than the non-iterative counterparts. [ABSTRACT FROM AUTHOR], Copyright of PLoS ONE is the )
Grammatical words represent the part of grammar that can be most directly contrasted with the lexicon. Aphasiological studies, linguistic theories and psycholinguistic studies suggest that their processing is operated at different stages in speech production. Models of sentence production propose that at the formulation stage, lexical words are processed at the functional level while grammatical words are processed at a later positional level. In this study we consider proposals made by linguistic theories and psycholinguistic models to derive two predictions for the processing of grammatical words compared to lexical words. First, based on the assumption that grammatical words are less crucial for communication and therefore paid less attention to, it is predicted that they show shorter articulation times and/or higher error rates than lexical words. Second, based on the assumption that grammatical words differ from lexical words in being dependent on a lexical host, it is hypothesiz)
Previous research has mainly considered the impact of tone-language experience on ability to discriminate linguistic pitch, but proficient bilingual listening requires differential processing of sound variation in each language context. Here, we ask whether Mandarin-English bilinguals, for whom pitch indicates word distinctions in one language but not the other, can process pitch differently in a Mandarin context vs. an English context. Across three eye-tracked word-learning experiments, results indicated that tone-intonation bilinguals process tone in accordance with the language context. In Experiment 1, 51 Mandarin-English bilinguals and 26 English speakers without tone experience were taught Mandarin-compatible novel words with tones. Mandarin-English bilinguals out-performed English speakers, and, for bilinguals, overall accuracy was correlated with Mandarin dominance. Experiment 2 taught 24 Mandarin-English bilinguals and 25 English speakers novel words with Mandarin-like tones,)
This is the first study to examine the effect of phonetic contexts on children’s lexical tone production. Mandarin tones in disyllabic words produced by forty-four 2- to 6-year-old children and twelve mothers were low-pass filtered to eliminate lexical information. Native Mandarin-speaking adults categorized the tones based on the pitch information in the filtered stimuli. All mothers’ tones were categorized with ceiling accuracy. Counter to the findings in most previous studies on children’s tone acquisition and the prevailing assumption in models of speech development that children acquire suprasegmental features much earlier than segmental features, this study found that children as old as six years of age have not mastered the production of Mandarin tones. Children’s tones were judged with significantly lower accuracy than mothers’ productions. Tone accuracy improved, while cross subject variability in tone accuracy decreased, with age. Children’s tone accuracy was affected by the)
Word recognition includes the activation of a range of syntactic and semantic knowledge that is relevant to language interpretation and reference. Here we explored whether or not the number of arguments a verb takes impinges negatively on verb processing time. In this study, three experiments compared the dynamics of spoken word recognition for verbs with different preferred argument structure. Listeners’ eye movements were recorded as they searched an array of pictures in response to hearing a verb. Results were similar in all the experiments. The time to identify the referent increased as a function of the number of arguments, above and beyond any effects of label appropriateness (and other controlled variables, such as letter, phoneme and syllable length, phonological neighborhood, oral and written lexical frequencies, imageability and rated age of acquisition). The findings indicate that the number of arguments a verb takes, influences referent identification during spoken word re)
Sentence reading involves multiple linguistic operations including processing of lexical and compositional semantics, and determining structural and grammatical relationships among words. Previous studies on Indo-European languages have associated left anterior temporal lobe (aTL) and left interior frontal gyrus (IFG) with reading sentences compared to reading unstructured word lists. To examine whether these brain regions are also involved in reading a typologically distinct language with limited morphosyntax and lack of agreement between sentential arguments, an FMRI study was conducted to compare passive reading of Chinese sentences, unstructured word lists and disconnected character lists that are created by only changing the order of an identical set of characters. Similar to previous findings from other languages, stronger activation was found in mainly left-lateralized anterior temporal regions (including aTL) for reading sentences compared to unstructured word and character li)
Though metaphoric language comprehension has previously been investigated with event-related potentials, little attention has been devoted to extending this research from the monolingual to the bilingual context. In the current study, late proficient unbalanced Polish (L1)–English (L2) bilinguals performed a semantic decision task to novel metaphoric, conventional metaphoric, literal, and anomalous word pairs presented in L1 and L2. The results showed more pronounced P200 amplitudes to L2 than L1, which can be accounted for by differences in the subjective frequency of the native and non-native lexical items. Within the early N400 time window (300–400 ms), L2 word dyads evoked delayed and attenuated amplitudes relative to L1 word pairs, possibly indicating extended lexical search during foreign language processing, and weaker semantic interconnectivity for L2 compared to L1 words within the memory system. The effect of utterance type was observed within the late N400 time window (400–)
Objectives: The present study explored tone perception ability in school age Mandarin-speaking children with otitis media with effusion (OME) in noisy listening environments. The study investigated the interaction effects of noise, tone type, age, and hearing status on monaural tone perception, and assessed the application of a hierarchical clustering algorithm for profiling hearing impairment in children with OME. Methods: Forty-one children with normal hearing and normal middle ear status and 84 children with OME with or without hearing loss participated in this study. The children with OME were further divided into two subgroups based on their severity and pattern of hearing loss using a hierarchical clustering algorithm. Monaural tone recognition was measured using a picture-identification test format incorporating six sets of monosyllabic words conveying four lexical tones under speech spectrum noise, with the signal-to-noise ratio (SNR) conditions ranging from -9 to -21 dB. Re)
Biomedical knowledge claims are often expressed as hypotheses, speculations, or opinions, rather than explicit facts (propositions). Much biomedical text mining has focused on extracting propositions from biomedical literature. One such system is SemRep, which extracts propositional content in the form of subject-predicate-object triples called predications. In this study, we investigated the feasibility of assessing the factuality level of SemRep predications to provide more nuanced distinctions between predications for downstream applications. We annotated semantic predications extracted from 500 PubMed abstracts with seven factuality values (, , , , , , and ). We extended a rule-based, compositional approach that uses lexical and syntactic information to predict factuality levels. We compared this approach to a supervised machine learning method that uses a rich feature set based on the annotated corpus. Our results indicate that the compositional approach is more effective than th)
Using a wireless single channel EEG device, we investigated the feasibility of using short-term frontal EEG as a means to evaluate the dynamic changes of mental workload. Frontal EEG signals were recorded from twenty healthy subjects performing four cognitive and motor tasks, including arithmetic operation, finger tapping, mental rotation and lexical decision task. Our findings revealed that theta activity is the common EEG feature that increases with difficulty across four tasks. Meanwhile, with a short-time analysis window, the level of mental workload could be classified from EEG features with 65%–75% accuracy across subjects using a SVM model. These findings suggest that frontal EEG could be used for evaluating the dynamic changes of mental workload. [ABSTRACT FROM AUTHOR], Copyright of PLoS ONE is the property of Public Library of Science and its content may not be copied or emailed to multiple sites or posted to a listserv without the copyright holder's express written permissio)
Readers’ eye movements were recorded to examine the role of character positional frequency on Chinese lexical acquisition during reading and its possible modulation by word spacing. In Experiment 1, three types of pseudowords were constructed based on each character’s positional frequency, providing congruent, incongruent, and no positional word segmentation information. Each pseudoword was embedded into two sets of sentences, for the learning and the test phases. In the learning phase, half the participants read sentences in word-spaced format, and half in unspaced format. In the test phase, all participants read sentences in unspaced format. The results showed an inhibitory effect of character positional frequency upon the efficiency of word learning when processing incongruent pseudowords both in the learning and test phase, and also showed facilitatory effect of word spacing in the learning phase, but not at test. Most importantly, these two characteristics exerted independent inf)
The ability of Baboons (papio papio) to distinguish between English words and nonwords has been modeled using a deep learning convolutional network model that simulates a ventral pathway in which lexical representations of different granularity develop. However, given that pigeons (columba livia), whose brain morphology is drastically different, can also be trained to distinguish between English words and nonwords, it appears that a less species-specific learning algorithm may be required to explain this behavior. Accordingly, we examined whether the learning model of Rescorla and Wagner, which has proved to be amazingly fruitful in understanding animal and human learning could account for these data. We show that a discrimination learning network using gradient orientation features as input units and word and nonword units as outputs succeeds in predicting baboon lexical decision behavior—including key lexical similarity effects and the ups and downs in accuracy as learning unfolds—w)
Congenital amusia is a lifelong disorder of fine-grained pitch processing in music and speech. However, it remains unclear whether amusia is a pitch-specific deficit, or whether it affects frequency/spectral processing more broadly, such as the perception of formant frequency in vowels, apart from pitch. In this study, in order to illuminate the scope of the deficits, we compared the performance of 15 Cantonese-speaking amusics and 15 matched controls on the categorical perception of sound continua in four stimulus contexts: lexical tone, pure tone, vowel, and voice onset time (VOT). Whereas lexical tone, pure tone and vowel continua rely on frequency/spectral processing, the VOT continuum depends on duration/temporal processing. We found that the amusic participants performed similarly to controls in all stimulus contexts in the identification, in terms of the across-category boundary location and boundary width. However, the amusic participants performed systematically worse than co)
In tonal languages, such as Mandarin Chinese, the pitch contour of vowels discriminates lexical meaning, which is not the case in non-tonal languages such as German. Recent data provide evidence that pitch processing is influenced by language experience. However, there are still many open questions concerning the representation of such phonological and language-related differences at the level of the auditory cortex (AC). Using magnetoencephalography (MEG), we recorded transient and sustained auditory evoked fields (AEF) in native Chinese and German speakers to investigate language related phonological and semantic aspects in the processing of acoustic stimuli. AEF were elicited by spoken meaningful and meaningless syllables, by vowels, and by a French horn tone. Speech sounds were recorded from a native speaker and showed frequency-modulations according to the pitch-contours of Mandarin. The sustained field (SF) evoked by natural speech signals was significantly larger for Chinese th)
Recent studies have shown that concurrent physical activity enhances learning a completely unfamiliar L2 vocabulary as compared to learning it in a static condition. In this paper we report a study whose aim is twofold: to test for possible positive effects of physical activity when L2 learning has already reached some level of proficiency, and to test whether the assumed better performance when engaged in physical activity is limited to the linguistic level probed at training (i.e. L2 vocabulary tested by means of a Word-Picture Verification task), or whether it extends also to the sentence level (which was tested by means of a Sentence Semantic Judgment Task). The results show that Chinese speakers with basic knowledge of English benefited from physical activity while learning a set of new words. Furthermore, their better performance emerged also at the sentential level, as shown by their performance in a Semantic Judgment task. Finally, an interesting temporal asymmetry between the)
Previous work suggests that, when attended, pictures may be processed more readily than words. The current study extends this research to assess potential differences in processing between these stimulus types when they are actively ignored. In a dual-task paradigm, facilitated recognition for previously ignored words was found provided that they appeared frequently with an attended target. When adapting the same paradigm here, previously unattended pictures were recognized at high rates regardless of how they were paired with items during the primary task, whereas unattended words were later recognized at higher rates only if they had previously been aligned with primary task targets. Implicit learning effects obtained by aligning unattended items with attended task-targets may apply only to conceptually abstract stimulus types, such as words. Pictures, on the other hand, may maintain direct access to semantic information, and are therefore processed more readily than words, even whe)
This present study investigated the link between speech-in-speech perception capacities and four executive function components: response suppression, inhibitory control, switching and working memory. We constructed a cross-modal semantic priming paradigm using a written target word and a spoken prime word, implemented in one of two concurrent auditory sentences (cocktail party situation). The prime and target were semantically related or unrelated. Participants had to perform a lexical decision task on visual target words and simultaneously listen to only one of two pronounced sentences. The attention of the participant was manipulated: The prime was in the pronounced sentence listened to by the participant or in the ignored one. In addition, we evaluate the executive function abilities of participants (switching cost, inhibitory-control cost and response-suppression cost) and their working memory span. Correlation analyses were performed between the executive and priming measurements)
To date there is no software that directly connects the linguistic analysis of a conversation to a network program. Networks programs are able to extract statistical information from data basis with information about systems of interacting elements. Language has also been conceived and studied as a complex system. However, most proposals do not analyze language according to linguistic theory, but use instead computational systems that should save time at the price of leaving aside many crucial aspects for linguistic theory. Some approaches to network studies on language do apply precise linguistic analyses, made by a linguist. The problem until now has been the lack of interface between the analysis of a sentence and its integration into the network that could be managed by a linguist and that could save the analysis of any language. Previous works have used old software that was not created for these purposes and that often produced problems with some idiosyncrasies of the target lan)
Sound units play a pivotal role in cognitive models of auditory comprehension. The general consensus is that during perception listeners break down speech into auditory words and subsequently phones. Indeed, cognitive speech recognition is typically taken to be computationally intractable without phones. Here we present a computational model trained on 20 hours of conversational speech that recognizes word meanings within the range of human performance (model 25%, native speakers 20–44%), without making use of phone or word form representations. Our model also generates successfully predictions about the speed and accuracy of human auditory comprehension. At the heart of the model is a ‘wide’ yet sparse two-layer artificial neural network with some hundred thousand input units representing summaries of changes in acoustic frequency bands, and proxies for lexical meanings as output units. We believe that our model holds promise for resolving longstanding theoretical problems surroundin)
In this paper we explore the results of a large-scale online game called ‘the Great Language Game’, in which people listen to an audio speech sample and make a forced-choice guess about the identity of the language from 2 or more alternatives. The data include 15 million guesses from 400 audio recordings of 78 languages. We investigate which languages are confused for which in the game, and if this correlates with the similarities that linguists identify between languages. This includes shared lexical items, similar sound inventories and established historical relationships. Our findings are, as expected, that players are more likely to confuse two languages that are objectively more similar. We also investigate factors that may affect players’ ability to accurately select the target language, such as how many people speak the language, how often the language is mentioned in written materials and the economic power of the target language community. We see that non-linguistic factors a)
We used a computational linguistic approach, exploiting machine learning techniques, to examine the letters written by King George III during mentally healthy and apparently mentally ill periods of his life. The aims of the study were: first, to establish the existence of alterations in the King’s written language at the onset of his first manic episode; and secondly to identify salient sources of variation contributing to the changes. Effects on language were sought in two control conditions (politically stressful vs. politically tranquil periods and seasonal variation). We found clear differences in the letter corpus, across a range of different features, in association with the onset of mental derangement, which were driven by a combination of linguistic and information theory features that appeared to be specific to the contrast between acute mania and mental stability. The paucity of existing data relevant to changes in written language in the presence of acute mania suggests tha)
Authorship attribution is to identify the most likely author of a given sample among a set of candidate known authors. It can be not only applied to discover the original author of plain text, such as novels, blogs, emails, posts etc., but also used to identify source code programmers. Authorship attribution of source code is required in diverse applications, ranging from malicious code tracking to solving authorship dispute or software plagiarism detection. This paper aims to propose a new method to identify the programmer of Java source code samples with a higher accuracy. To this end, it first introduces back propagation (BP) neural network based on particle swarm optimization (PSO) into authorship attribution of source code. It begins by computing a set of defined feature metrics, including lexical and layout metrics, structure and syntax metrics, totally 19 dimensions. Then these metrics are input to neural network for supervised learning, the weights of which are output by PSO a)
The present study investigated interactions between cognitive processes and finger actions called “kusho,” meaning “air-writing” in Japanese. Kanji-culture individuals often employ kusho behavior in which they move their fingers as a substitute for a pen to write mostly done when they are trying to recall the shape of a Kanji character or the spelling of an English word. To further examine the visualization role of kusho behavior on cognitive processing, we conducted a Kanji construction task in which a stimulus (i.e., sub-parts to be constructed) was simultaneously presented. In addition, we conducted a Kanji vocabulary test to reveal the relation between the kusho benefit and vocabulary size. The experiment provided two sets of novel findings. First, executing kusho behavior improved task performance (correct responses) as long as the participants watched their finger movements while solving the task. This result supports the idea that visual feedback of kusho behavior helps cogniti)
In the field of word recognition and reading, it is commonly assumed that frequently repeated words create more accessible memory traces than infrequently repeated words, thus capturing the word-frequency effect. Nevertheless, recent research has shown that a seemingly related factor, contextual diversity (defined as the number of different contexts [e.g., films] in which a word appears), is a better predictor than word-frequency in word recognition and sentence reading experiments. Recent research has shown that contextual diversity plays an important role when learning new words in a laboratory setting with adult readers. In the current experiment, we directly manipulated contextual diversity in a very ecological scenario: at school, when Grade 3 children were learning words in the classroom. The new words appeared in different contexts/topics (high-contextual diversity) or only in one of them (low-contextual diversity). Results showed that words encountered in different contexts we)
The experiments reported here used “Reversed-Interior” (RI) primes (e.g., cetupmor-COMPUTER) in three different masked priming paradigms in order to test between different models of orthographic coding/visual word recognition. The results of Experiment 1, using a standard masked priming methodology, showed no evidence of priming from RI primes, in contrast to the predictions of the Bayesian Reader and LTRS models. By contrast, Experiment 2, using a sandwich priming methodology, showed significant priming from RI primes, in contrast to the predictions of open bigram models, which predict that there should be no orthographic similarity between these primes and their targets. Similar results were obtained in Experiment 3, using a masked prime same-different task. The results of all three experiments are most consistent with the predictions derived from simulations of the Spatial-coding model. [ABSTRACT FROM AUTHOR], Copyright of PLoS ONE is the property of Public Library of Science and i)
Automatic extraction of protein-protein interaction (PPI) pairs from biomedical literature is a widely examined task in biological information extraction. Currently, many kernel based approaches such as linear kernel, tree kernel, graph kernel and combination of multiple kernels has achieved promising results in PPI task. However, most of these kernel methods fail to capture the semantic relation information between two entities. In this paper, we present a special type of tree kernel for PPI extraction which exploits both syntactic (structural) and semantic vectors information known as Distributed Smoothed Tree kernel (DSTK). DSTK comprises of distributed trees with syntactic information along with distributional semantic vectors representing semantic information of the sentences or phrases. To generate robust machine learning model composition of feature based kernel and DSTK were combined using ensemble support vector machine (SVM). Five different corpora (AIMed, BioInfer, HPRD50, )
In spite of decades of theorizing, the origins of Zipf’s law remain elusive. I propose that a Zipfian distribution straightforwardly follows from the interaction of syntax (word classes differing in class size) and semantics (words having to be sufficiently specific to be distinctive and sufficiently general to be reusable). These factors are independently motivated and well-established ingredients of a natural-language system. Using a computational model, it is shown that neither of these ingredients suffices to produce a Zipfian distribution on its own and that the results deviate from the Zipfian ideal only in the same way as natural language itself does. [ABSTRACT FROM AUTHOR], Copyright of PLoS ONE is the property of Public Library of Science and its content may not be copied or emailed to multiple sites or posted to a listserv without the copyright holder's express written permission. However, users may print, download, or email articles for individual use. This abstract may be )
Background: Most of earlier studies in the field of literature-based discovery have adopted Swanson's ABC model that links pieces of knowledge entailed in disjoint literatures. However, the issue concerning their practicability remains to be solved since most of them did not deal with the context surrounding the discovered associations and usually not accompanied with clinical confirmation. In this study, we aim to propose a method that expands and elaborates the existing hypothesis by advanced text mining techniques for capturing contexts. We extend ABC model to allow for multiple B terms with various biological types. Results: We were able to concretize a specific, metabolite-related hypothesis with abundant contextual information by using the proposed method. Starting from explaining the relationship between lactosylceramide and arterial stiffness, the hypothesis was extended to suggest a potential pathway consisting of lactosylceramide, nitric oxide, malondialdehyde, and arteria)
Despite the ongoing growth in the number of published randomized controlled trials (RCTs) and increased quality assessment of RCTs, the association between the quality and characteristics in the text has not been sufficiently studied. We are interested in a specific question: what kind of sentences is a good indicator of high quality RCTs? To help researchers to efficiently screen articles worth reading, this study aims 1) to quantify the linguistic features of articles and 2) to build a document assessment model to evaluate quality of RCTs using only the abstract. All RCTs that were conducted in Japan in 2010 as original articles were included in the analysis. Data were independently assessed by two reviewers using a risk-of-bias tool. Three aspects of linguistic style were quantitatively measured, and a document model was constructed to evaluate the RCTs. A total of 302 RCTs were selected for quality assessment. Of these, 255 articles were assessed as high quality and 47 as low qual)
Infants preferentially discriminate between speech tokens that cross native category boundaries prior to acquiring a large receptive vocabulary, implying a major role for unsupervised distributional learning strategies in phoneme acquisition in the first year of life. Multiple sources of between-speaker variability contribute to children’s language input and thus complicate the problem of distributional learning. Adults resolve this type of indexical variability by adjusting their speech processing for individual speakers. For infants to handle indexical variation in the same way, they must be sensitive to both linguistic and indexical cues. To assess infants’ sensitivity to and relative weighting of indexical and linguistic cues, we familiarized 12-month-old infants to tokens of a vowel produced by one speaker, and tested their listening preference to trials containing a vowel category change produced by the same speaker (linguistic information), and the same vowel category produced )
Background: To facilitate informed consent, consent forms should use language below the grade eight level. Research Ethics Boards (REBs) provide consent form templates to facilitate this goal. Templates with inappropriate language could promote consent forms that participants find difficult to understand. However, a linguistic analysis of templates is lacking. Methods: We reviewed the websites of 124 REBs for their templates. These included English language medical school REBs in Australia/New Zealand (n = 23), Canada (n = 14), South Africa (n = 8), the United Kingdom (n = 34), and a geographically-stratified sample from the United States (n = 45). Template language was analyzed using Coh-Metrix linguistic software (v.3.0, Memphis, USA). We evaluated the proportion of REBs with five key linguistic outcomes at or below grade eight. Additionally, we compared quantitative readability to the REBs’ own readability standards. To determine if the template’s country of origin or the presenc)
In decision making, similarity measure and distance between two objects are crucial to be able to determine the relationship between those objects. Many researchers have received much attention for their research on this subject. In this study, we propose two novel similarity measures between hesitant fuzzy linguistic term sets (HFLTSs). In addition, two extensions of Technique for Order of Preference by Similarity to Ideal Solution (TOPSIS) are proposed in the hesitant fuzzy linguistic environments. Furthermore, an example of an application concerning traditional Chinese medical diagnosis and an MCDM problem have been given to illustrate the applicability and validation of these similarity measures of HFLTSs. Furthermore, the results of examples demonstrate that the Dice and Jaccard similarity measures are more reasonable than the cosine similarity measure with respect to HFLTSs. [ABSTRACT FROM AUTHOR], Copyright of PLoS ONE is the property of Public Library of Science and its conten)
We present a new open source software tool called BEASTling, designed to simplify the preparation of Bayesian phylogenetic analyses of linguistic data using the BEAST 2 platform. BEASTling transforms comparatively short and human-readable configuration files into the XML files used by BEAST to specify analyses. By taking advantage of Creative Commons-licensed data from the Glottolog language catalog, BEASTling allows the user to conveniently filter datasets using names for recognised language families, to impose monophyly constraints so that inferred language trees are backward compatible with Glottolog classifications, or to assign geographic location data to languages for phylogeographic analyses. Support for the emerging cross-linguistic linked data format (CLDF) permits easy incorporation of data published in cross-linguistic linked databases into analyses. BEASTling is intended to make the power of Bayesian analysis more accessible to historical linguists without strong programmi)