Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
This chapter provides an overview of the contribution the Prague School of Linguistics has made to the study of linguistic norm. Departing from a functionalist approach to the analysis of language, the Prague School shaped and theorized the influential notions of language culture (jazykova kultura, Sprachkultur) and language cultivation (Sprachpflege) at a time when, in Western Europe, prescriptive approaches to language were considered unworthy of scientific attention. Distancing themselves from "purist" and "school-grammar" conceptions of language maintenance and based on the case of Czech language culture, the representatives of the Prague School advocated for the cultivation of the written (standard) language in functional terms in order to achieve a "stable" yet "elastic" concept of norm.
Universal Dependencies is an open community effort to create cross-linguistically consistent treebank annotation for many languages within a dependency-based lexicalist framework. The annotation consists in a linguistically motivated word segmentation; a morphological layer comprising lemmas, universal part-of-speech tags, and standardized morphological features; and a syntactic layer focusing on syntactic relations between predicates, arguments and modifiers. In this paper, we describe version 2 of the guidelines (UD v2), discuss the major changes from UD v1 to UD v2, and give an overview of the currently available treebanks for 90 languages.
Cupping therapy has recently gained public attention and is widely used in many regions. Some patients are resistant to being treated with cupping therapy, as visually unpleasant marks on the skin may elicit negative reactions. This study aimed to identify the cognitive and emotional components of cupping therapy. Twenty-five healthy volunteers were presented with emotionally evocative visual stimuli representing fear, disgust, happiness, neutral emotion, and cupping, along with control images. Participants evaluated the valence and arousal level of each stimulus. Before the experiment, they completed the Fear of Pain Questionnaire-III. In two-dimensional affective space, emotional arousal increases as hedonic valence ratings become increasingly pleasant or unpleasant. Cupping therapy images were more unpleasant and more arousing than the control images. Cluster analysis showed that the response to cupping therapy images had emotional characteristics similar to those for fear images. Individuals with a greater fear of pain rated cupping therapy images as more unpleasant and more arousing. Psychophysical analysis showed that individuals experienced unpleasant and aroused emotional states in response to the cupping therapy images. Our findings suggest that cupping therapy might be associated with unpleasant-defensive motivation and motivational activation. Determining the emotional components of cupping therapy would help clinicians and researchers to understand the intrinsic effects of cupping therapy.
As the number, size, and complexity of building construction projects increase, code compliance checking becomes more challenging because of the time-consuming, costly, and error-prone nature of a manual checking process. A fully automated code compliance checking would be desirable in facilitating a more efficient, cost effective, and human error-proof code checking. Such automation requires automated information extraction from building designs and building codes, and automated information transformation to a format that allows automated reasoning. Natural language processing (NLP) is an important technology to support such automated processing of building codes, because building codes are represented in natural language texts. Part-of-speech (POS) tagging, as an important basis of NLP tasks, must have a high performance to ensure the quality of the automated processing of building codes in such a compliance checking system. However, no systematic testing of existing POS taggers on domain specific building codes data have been performed. To address this gap, the authors analyzed the performance of seven state-of-the-at POS taggers on tagging building codes and compared their results to a manually-labeled gold standard. The authors aim to: (1) find the best performing tagger in terms of accuracy, and (2) identify common sources of errors. In providing the POS tags, the authors used the Penn Treebank tagset, which is a widely used tagset with a proper balance between conciseness and information richness. An average accuracy of 88.80% was found on the testing data. The Standford coreNLP tagger outperformed the other taggers in the experiment. Common sources of errors were identified to be: (1) word ambiguity, (2) rare words, and (3) unique meaning of common English words in the construction context. The found result of machine taggers on building codes calls for performance improvement, such as error-fixing transformational rules and machine taggers that are trained on building codes.
Noun phrases convey key information in communication and are of interest in NLP tasks. A base NP is defined as the headword and left-hand side modifiers of a noun phrase. In this thesis, we identify base NPs in Universal Dependencies treebanks in English and French using an RNN architecture.The data of this thesis consist of three multi-layered treebanks in which each sentence is annotated in both constituency and dependency formalisms. To build our training data, we find base NPs in the constituency layers and project them onto the dependency layer by labeling corresponding tokens. For input features, we devised 18 configurations of features available in UD annotation. We train RNN models with LSTM and GRU cells with different numbers of epochs on these configurations of features.Tested on monolingual and bilingual test sets, our models delivered satisfactory token-based F1 scores (92.70% on English, 94.87% on French, 94.29% on bilingual test set). The most predicative configuration of features is found out to be pos_dep_parent_child_morph, which covers 1) dependency relations between the current token, its syntactic head, its leftmost and rightmost syntactic dependents; 2) PoS tags of these tokens; and 3) morphological features of the current token.
In order to extract the semantic and grammatical information of sentences more effectively, this paper proposes a sentence sentiment classification method based on Self-supervised and Self-attention mechanism (SS-SAtt-BiLSTM). In this method, BiLSTM network is used to extract the feature of text context relationship, and self-supervised (SS) learning mode is introduced into the supervised sentence representation model. The sentence itself is used as the label data information of current words, and an improved self-attention mechanism (SA) is used to calculate the attention weight of each moment. The experimental results of MR and Stanford sentient treebank (sst-5) data sets show that this method reduces the dependence on tagged data, and the improved self-attention mechanism enables the model to learn more key features of sentences and improve the classification performance.
Semantic Role Labelling (SRL) is the process of automatically finding the semantic roles of terms in a sentence. It is an essential task towards creating a machine-meaningful representation of textual information. One public linguistic resource commonly used for this task is the FrameNet Project. FrameNet is a human and machine-readable lexical database containing a considerable number of annotated sentences, those annotations link sentence fragments to semantic frames. However, while the annotations across all the documents covered in the dataset link to most of the frames, a large group of frames lack annotations in the documents pointing to them. In this paper, we present a data augmentation method for FrameNet documents that increases by over 13% the total number of annotations. Our approach relies on lexical, syntactic, and semantic aspects of the sentences to provide additional annotations. We evaluate the proposed augmentation method by comparing the performance of a state-of-the-art semantic-role-labelling system, trained using a dataset with and without augmentation.
Introduction. High-quality language education in technical universities requires its interdisciplinary relation to the content of highly specialised subjects corresponding to the training programmes aimed at instructing the future specialists. Educational materials in a foreign language are highly productive if they emphasise the terminology and professional vocabulary authentic to the current state of the scientific field. The aim of the study presented in the article was to assess the validity of the lexical material delivered in the course “English for Business Communication”, to determine the selection criteria for this vocabulary as well as the methods for its assimilation and practical application. Methodology and research methods. The applied corpus software enabled to obtain quantitative indicators of the distribution of foreign-language business vocabulary in the given training course. The lexical material being currently offered to students and the professional thesaurus identified via linguistic databases was compared with the use of comparative analysis and synthesis. Results and scientific novelty. The lexical units (terms, set expressions), which are the most active in the business sphere, were identified on the basis of its frequency. The authors established the correlation between them and educational vocabulary, both from the perspective of its integration into the course without block concentration throughout the course of university training, and from the perspective of the variety of methods used to practice this vocabulary. It is concluded that the applied educational material needs to be substantially adjusted. The vocabulary does not completely reflect the realities of the business communication sphere and the distribution of active vocational vocabulary regulated by methodological guidelines does not entirely contribute to its strong assimilation. According to the authors, the necessary changes to the approaches and methods for selecting and compiling lexical material and to the methodology for designing a foreign language course should be made on the basis of integrating pedagogical and linguistic knowledge, in particular, the methodology of teaching foreign languages and the corpus linguistics. Practical significance. The ways of integrating corpus programs in the process of developing the content of language disciplines, which are part of the main educational program of technical universities, are demonstrated as one of the methods to increase the effectiveness of teaching foreign languages to students of non-linguistic specialties.
Language users and learners are sensitive to distributional information in their environment, which enables them to extract regularities that occur in the language input that they are exposed to. This process is referred to as statistical learning. While the statistical learning phonotactic literature thoroughly investigates the learning of overall phonotactics in specific languages, little is known about cases where different phonological systems coexist within a single language. The Japanese lexicon is generally classified into four lexical strata according to the etymological status of each word (Itô & Mester, 1995, 1999, 2001). Although each stratum includes the internal phonological similarity in the Japanese language as a whole, there are also distinctive phonological properties. A recent study suggests that language users should be able to learn phonotactics of each sublexicon based on the same kind of statistical probabilities that computers analyse from language users’ accumulated lexicons (Morita, 2018). This thesis examines whether second-language (L2) learners can learn the loanword phonotactics/phonology of Japanese through experience of using and/or passive exposure to Japanese lexical stratification. Using two loanword phonological regularities (categorical and gradient rules) as a case study, two fully-crossed perceptual experiments involving English- speaking learners of Japanese, native speakers of Japanese, and English-speaking monolinguals are presented. The first experiment explores listeners’ phonotactic/phonological knowledge of nativised loanwords in Japanese using a well-formedness task which shows the adaptation of English final consonants in monosyllabic words. Listeners judge whether the pronunciation they hear is how the word would be pronounced if it was a Japanese word, rating how confident they are on a scale of 1-5. This study shows that L2 learners learn categorical rules, but not gradient patterns. This study also confirms that loanword phonotactics and overall phonotactics make separate contributions to perceived well-formedness. L2 learners access and make use of the sublexicon-specific probabilities of Japanese during the task. The second perceptual experiment is designed to support the findings in the first experiment, by testing for discrimination of non-native consonantal contrasts. Even under high memory demand, L2 learners show the ability to discriminate non-native consonantal contrasts (i.e., CVCV/CVCCV) effectively enough to support findings in the first experiment. These results suggest that L2 learners can implicitly detect the statistical structure of a language’s sublexicon phonology over the course of acquiring a natural language. However, while native speakers of Japanese learn a gradient rule, L2 learners of Japanese do not. A potential explanation for the differences in gradient rule learning is that the vocabulary size of the target language might play a crucial role. This remains an open question. In addition, the present work provides a basis for future investigation into whether L2 learners of Japanese, whose native language is other than English, are able to learn Japanese loanword phonotactics/phonology. L1 English-L2 Japanese speakers might gain advantage in perceiving the English input which inevitably overlaps with the phonological form of the host language.
The goal of this special issue of Critical Multilingualism Studies “National Standards – Local Varieties: A Cross-Linguistic Discussion on Regional Variation in L2 Studies” is to incite a conversation on how topics such as linguistic norms and variation, dominant practices, ideologies, identities, and politics surrounding languages are discussed from a view outside of the dominant centers of linguistic norms.
Because of its focus on the past and on historical languages, the classics is a discipline that is particularly interested in translations and text alignment. Starting from a diachronic perspective, this contribution demonstrates how issues related to text alignment, present since antiquity, can be approached from a different angle and with entirely new opportunities thank to tools and methods developed in the field of digital humanities. By comparing examples from antiquity (e.g. Origen’s Hexapla from the third century CE) with modern projects based on treebanking and dependency grammar (e.g. the Ancient Greek and Latin Dependency Treebank [AGLDT] as part of the Perseus Digital Library from Tufts University), we shall present some new approaches and their potentials. In doing so, we shall also examine what status English has in these projects and how the different languages involved in each of them interact with English and/or with each other.
We study the effect of rich supertag features in greedy transition-based dependency parsing. While previous studies have shown that sparse boolean features representing the 1-best supertag of a word can improve parsing accuracy, we show that we can get further improvements by adding a continuous vector representation of the entire supertag distribution for a word. In this way, we achieve the best results for greedy transition-based parsing with supertag features with $88.6\%$ LAS and $90.9\%$ UASon the English Penn Treebank converted to Stanford Dependencies.
This is an introduction to the proposed theme, in which the importance of sociolinguistic studies for the teaching, acquisition and learning of languages is emphasized. In addition, each text of the material is presented, starting with interviews with significant and current representatives of the variation sociolinguistics (Francisco Moreno Fernández and Juan Manuel Hernández Campoy) from the Hispanic and Anglo-Saxon spheres, respectively; then, it discusses the ten articles that deal with the theme from two perspectives: linguistic attitudes and beliefs of speakers and linguistic norms and policies. Finally, the reviews of two books related to the Special issue are commented: The Routledge handbook of Spanish as a heritage language, edited by Kim Potowsky, 2018, New York, Routledge publisher, and La trastienda de la enseñanza de lenguas extranjeras, by Francisco García Marcos, 2018, from the Interlingua collection of Editora Comares de Granada / Spain. The presentation is an invitation to readers to enjoy reading the Special issue.
Context: Parkinson’s disease (PD) is a neurodegenerative disease caused by degeneration of the dopaminesynthesizing cells of the mesostriatal-mesocortical neuronal pathway,which affects motor pathway in basal ganglia (BG). Neuropsychological studies showed that degeneration of dopamine neuroreceptor also affects nigrostriatal and mesocortical limbic system which is associated with emotional processing in PD. However, very few studies have identified deficit in selective attention in patients with PD patients except in patients with PD-MCI (PD-Mild Cognitive Impairment) or PD-D (PD-Dementia). Thus, the present study examined the effect of emotion on attentional processing in PD and matched control. Emotional flanker task was designed by using pictures selected from the International Affective Picture System (IAPS) based on their normative valence ratings. Results revealed that attentional processing of emotional images were slower in PD patients in comparison to matched healthy control.
The paper deals with the problem of the studying of dialectal personality phenomenon, in particular the researching of the Eastern Polissian dialectal speakers’ metalingual consciousness. The purpose of the study is to analyze the Eastern Polissian dialectal speakers’ metatextual utterances about the peculiarities of the dialectal language at different language levels in the context of scientific researches. Sources of the research is classified dialectal material collected by the field method and recorded by the author in 76 villages of Chernihiv and Sumy oblasts during 2009–2019, using a special questionnaire devoted to the study of verb vocabulary semantic variation, and recordings of dialectal speakers’ texts-stories. Metatextual utterances were recorded on occasion. Dialectal material of lexicographic and linguogeographical sources, which contain information about the studied linguistic phenomena, are also used. Based on the analyzed dialectal material it is concluded that dialectal speakers evaluate the dialectal language in compliance / inconsistency with the established norm with a focus on their own dialect, other dialects, which language, according to dialectal speakers’ opinion, is close to the Ukrainian literary language, and the Ukrainian literary language. It is found out that dialectal speakers are aware mainly the phonetic (accent variation, substitution of unstressed [o] through [a], equivalents */o/ in stressed closed syllable, hardening of [p]) and lexical features, ignoring word-formation, morphology and syntax. It is stated that in metatextual utterances dialectal speakers first of all pay attention to the functioning of dinphthongs / monophthongs in dialects and to the substitution of unstressed [o] through [a] - the phenomenon of akannia, which, according to researchers’ opinion, is the edge of the Eastern Slavic akannia area and has no influence of the Ukrainian literary language. It is determined that most of the data of the Eastern Polissian dialectal speakers’ metalingual consciousness confirms by the author’s recorded dialectal material and information from various scientific researches.
Croatian accentual norm is in a constant state of flux. Its stability is impeded, first of all, by two mutually intertwined forces: the nature of the accentual norm, which belongs to speech (dynamic dimension, individual realisation), and the disagreement amongst linguists as to what to record and prescribe (in constant interaction between the stress accent and pitch accent systems). The modern accentual norm is obtained from non-orthoepical manuals, i.e. grammar books, dictionaries, handbooks (which further complicates the clarification of the orthoepical reality). We will conduct a comparative analysis of the approach, in modern handbooks, to accent alternations in morphology, falling accent in non-initial syllables in word formation, post-tonic length, uncertainties regarding lexical stress, etc. Grammar books and dictionaries approach the open questions in different ways and this paper gives an overview of the (systematic and non-systematic) solutions offered by linguists today, with the aim of presenting the dynamics of the codified norm (which carries the label of being “conservative” and “hidebound”). The changes in the modern norm are compared then to usus occurrences, illustrated by a narrower speech corpus – the speech of actors. In their orthoepical research, linguists resort to the speech of radio and television presenters, linguists in specialised radio and television programmes, students of the Croatian language or phonetics, Croatian language teachers, etc., and, more recently, to the speech of actors reading audio books (MP3 files are available at www.lektire.skole.hr). Presenters, teachers and actors have always been perceived as quintessential competent speakers of the standard language, so close observation of their speech as one of the steps in the process of describing and prescribing is the basis of every orthoepical research. Since the modern speech/pronunciation (e-lektira, audio versions of school reading list books available online) has still not been analysed and valorised linguistically/orthoepically, and since it is available to those learning and listening to speech values in this type of material, the paper turns to the corpus with the intention of determining the basic features of pronunciation. Prose texts whose pronunciation has been analysed are those written in or translated into the standard language. Special attention has been given to accent (stress placement and stress shift) and to the prosodic word. Specific pronunciation traits (especially those related to the accentual norm) have been compared to those prescribed in handbooks. Finally, the accentual traits acknowledged by the modern conception of accentual norm and codification were clarified as well as those that are systematically ignored in modern prescription.
1. Introduction One of the grounds of changing contemporary literature in Iran is translation or interpretation of foreign literary works which studying that can express the necessities, the creativities and the harms of this category well. Sometimes Contemporary poets have conformed or idiomatically domesticated the component parts of original context to literary and cultural requirements of Persian speaking community by displacing, increasing and decreasing them. The present article is based on expressing domestication methods in moving from "extra - system" (original context) to "system" (target context) and understanding the poets success or unsuccess in translating foreign literary works to Persian poem according to Yuri Lotman's theory and here we just study poetic translations of an fable entitled "The Fox and the Crow" by Lafontaine as example. 2. Methodology This research has been performed by a descriptive method and content analysis & the relationship quality among the culture elements with the context and extra – context background of the original work compared to its translations has been expressed by Lotman's semiotic theory of culture. It has been tried to show the creativities and deficiencies of the new texts against the original text while referring to different methods and ways of domestication and brief comparison of contemporary poets translations from "The Crow and the Fox" story by Lafontaine. 3. Discussion Nasim Shomal, Iraj Mirza, Nayyere Saeidi and Habib Yaghmayi have versified "The Crow and Fox" story chronologically. These versifiers mainly aim to put unfamiliar culture elements in native culture forms according to the present times aesthetics culture. Iranian culture system hasn't accepted the entered content from extra – system in the same existing form and has made it similar to fixed norms of its memory and since literal type of the source text is didactic and adjustable with those norms has organized system and traditional poem form because traditional form is an element that has been located in the center of semiospher of Iran's literal culture and any disagreement with that has been considered a riot. The reality of focusing on the type of expression is the result of a one – to – one correlation between the level of expression and content and the effect of expression on content (Lotman & Ospenski, 1390: 52). Decreasing, increasing, changing & moving that these poets have performed, show socio cultural features of that age and any way in most cases in order to guarantee the necessary consistency, are directly associated with the orientation to the past; for example, Iraj Mirza has changed a western story in to the form of classic Persian poems by choosing traditional measure and form, using ancient words and grammar and direct conclusion at the end of poetry. This approach can be considered one of the harms of domestication; because translation that is a means to introduce new ideas and methods will become a tool for maintaining traditional taste. Certainly in these versified translations, didactic &critical aspect has been dominated the satire aspect of the original work. Iraj Mirza has made the text critical by changing the symbol (salamander instead of phoenix) and Nasim Shomal has changed not only the tone & style of a foreign work, but also its direction and aim and has created a text with his own intended critical purpose by decreasing and increasing the elements, lexical reviews and speeding up the language. Nasim Shomal poem it at the service of society not ethics and it can be said that the intended meanings of the original author have been dominated by the author's ideas and beliefs and the target language words. In Nayyere Saeidi's poem Intertextuality & musical and Gnostic terms and word selection related to Iran's culture sign system have made the text strong but insisting on bringing same synonyms in the manner of some ancient Persian works has added redundancies to his poem. We can also observe encountering two different cultures in La Fontaine and Nayyere Saeidi's treatment with fox trickery and flattery. Each of them considers himself far from this forbidden phenomenon by a different method; La Fontaine does that by mock in and scorning and Iranian poet by adding adjectives that can be interpreted ethically and permit a connection with the focus of Iranian traditions. Habib Yaghmayi has selected fluent words in his translation and cared about the intimacy of poem language and word flexibility phonetically. Although Yaghmayi tries to define every innovation in terms of traditional Persian literature, language naturalness along with the application of old word makes his poem like a city in which old buildings are next to new buildings These applications are an effort for transferring language from "diachronic" limitation to "synchronic" range; also the existence of movement and cinematic pictures have shown his work newer and more modern. In Yaghmayi's method, the present traditional relationship between expression and content that can be seen in other translations of the story, is not known as the only possible relationship; although this works expression affects content more than other translations. The expression type is not naturalized except in some rare examples and utilizes dramatic art. This method of extra – system has been absorbed and the poet invalidates the last translations of this story by a new movement with creativity so that Habib Yaghmayi's poem is replaced by Iraj Mirza work in textbooks. Reticence and word choice or selectivity in words has also been from old & bright favorite traditions of Iranian and is located in the center of semiospher & the texts which have had this feature like Yaghmayi translation, have had the most life & longevity in Iranian culture. This work is the only translation which has become strong in dispute with other systems and domestication in interaction with creative methods of imagery has resulted in its richness; since besides employing the elements of culture which are located in the center of semiospher, hasn’t paid any attention to the marginal elements of the culture which are being forgotten. His way of expression is a combination of modernity & tradition. Yaghmayi has used Nezami's poem that itself is very organized and its mixture with the original text has created a more valuable text. To compare the ways of addressing "The Crow and the Fox" story by the translators shows gradual development of domestication & poets interaction in moving from extra – system to system. 4. Conclusion The results of the study show that the poets have conformed the texts with literal traditions and Islamic – Iranian culture symbols in this process and this domestication has been performed in three systems including lingual, literal and mental in images, thoughts, expression methods and language processes. Decreasing, increasing, changing and moving that these poets have performed, express social and cultural features of its own age. Adding measure& rhyme has been performed for preserving the spirit of Persian classical poetry and its traditional form and didactic and critical aspect has been dominated the satire aspect of the original work. In versified translation of Habib Yaghmayi, the movement and Cinematic pictures and language simplicity and intimacy have shown the work newer and more modern. This work is the only translation that has become strong in dispute with other systems and domestication in interaction with creative methods of imagery has resulted in its richness Yaghmayi poetry has more validity & stability in the memory of the culture of Iranian society; since besides employing the elements of culture which are located in the center of semiospher, has paid no attention to marginal elements of culture which are being forgotten. His expression style is a combination of modernity & tradition. His imajic &dramatic expression has been observed from extra – system & has combined traditional values located in the center of culture system with that. Also analyzing domestication techniques & aesthetics elements in this research show that the poetries have become more complete in order of composition time & Habib Yaghmayi translation is the result of the perfection of poetry language towards more brevity. It seems that at the rest of this research direction, we can also address analyzing domestication in the story works obtained from translation.
The article discusses the possibility of using author’s dictionaries as sources for replenishing the general language explanatory dictionary of the Ukrainian language. By the beginning of the 20th century, Ukrainian author’s lexicography acquired clear features of a separate vocabulary direction. Index words, concordances, complete and differential dictionaries, monolingual and bilingual have been published. In addition to dictionaries in a book format, electronic dictionaries, online dictionaries are now being created. Оne of the tasks of the writer’s dictionary is to describe the individual characteristics of the author’s language, in a broader sense, to popularize the linguistic personality. On the other hand, the material from the author’s dictionary serves as a factual basis for updating the register, semantic, stylistic characteristics, illustrative material of general language explanatory dictionaries. A number of researchers emphasize the need for a close relationship between explanatory lexicography and the author’s one for the successful development of national vocabulary work. “Dictionary of the language in Hryhorii Kvitka-Osnovianenko’s works” in 3 volumes occupies a special place among the author’s dictionaries for such characteristics as structure, volume. It includes all words, phraseology contained in a six-volume edition of the writer’s works, in other publications, in archival works. Kvitka-Osnovianenko is the founder of the prose genre in Ukrainian literature. As evidenced by the materials of the “Dictionary of the Ukrainian language” in 11 vol., its compilers turned to Kvitka-Osnovianenko’s works to fill in the vocabulary zones. Nevertheless, the materials of the “Dictionary of the language in Hryhorii Kvitka-Osnovianenko’s works” contain language units that may be of interest to compilers of large explanatory dictionaries. These are deminitives that are characteristic of the language of many genres of Ukrainian literature, as well as phraseological units. An example of the author’s non-fiction lexicography is the dictionary of the language of the publicist I.M. Dziuba, which is currently being developed. In Soviet times, the works of a well-known literary critic and publicist were not used as literary sources for the selection of words or illustrative material in explanatory lexicography. The materials of the language dictionary of I.M. Dziuba represent a significant resource for replenishing the vocabulary of the explanatory dictionary, for filling in the structural zones of the dictionary entry. The presented specific examples of word usage by the two authors correspond to lexical, word-formation norms. Further scientific and lexicological study of such linguistic facts will show the possibility of their involvement in explanatory lexicography. Keywords: author’s lexicography, general language explanatory dictionary, vocabulary register, H.F. Kvitka-Osnovianenko, language of works by I.M. Dziuba.
The lexical units connected with religion in Uzbek are borrowed from Arabic. In the process of assimilation, phonemes that do not exist in the Uzbek language are replaced by similar alternatives, and this is the norm for the speakers of the Uzbek language. However, in the speeches of religious preachers, we see that Arabic phonemes are pronounced in accordance with the rules of the Arabic language, not Uzbek alternatives. This is a feature peculiar only to the religious texts.
For those looking to deepen their understanding of the letter to the Colossians (and Philemon, in the case of Beale’s commentary), the scholarly choices are looking bright. As the following review will evidence, studies in Colossians are on the rise. Members of the academic world are no strangers to the authors of the following volumes. G. K. Beale holds the J. Gresham Machen chair of New Testament at Westminster Theological Seminary, Glenside, PA; Scot McKnight holds the Julius R. Mantey chair of New Testament at Northern Seminary, Lisle, IL; and Joel White is university lecturer in New Testament at Freie Theologische Hochschule, Giessen, Germany.Given the nature of commentaries and the amount of material being surveyed and reviewed, the following engagement will be somewhat limited in its attention to particulars. Instead, I want to focus on the layout of each commentary and some of the distinctives one can glean from each volume. I will begin with Beale’s commentary because it is the only one of the three that deals with two books. The book begins with the typical introduction of addressing questions concerning authorship, dating, location, occasion, and outline/argument. The exegesis follows a straightforward outline, splitting the letter into four sections: the letter opening (1:1–2), the letter thanksgiving (1:3–23), the letter body (1:24–4:6), and the letter closing (4:7–18). Philemon is analyzed similarly with four major sections: the letter opening (vv. 1–3), the introductory thanksgiving and prayer (vv. 4–7), the letter body (vv. 8–21), and the letter closing (vv. 22–25). The textual analysis is followed by five excurses. These sections address the difficulties of establishing Pauline and non-Pauline authorship, the criteria for discerning OT allusions, interpreting “Christ among the Gentiles,” the OT background of “the uncircumcision of your flesh,” and the master-slave relationship.McKnight’s commentary opens similarly to Beale’s. He addresses authorship, opponents, date and imprisonment, Paul’s theology in Colossians, and the structure of the book. McKnight’s structural analysis concludes that the book is divided into four parts: introduction (1:1–2:5), doctrinal correction (2:6–3:4), practical exhortation (3:5–4:6), and conclusion (4:7–18). One fundamental difference in McKnight’s approach is his preference for non-epistographical designations for each section. As he states, he is less confident of rhetorical conclusions concerning the letter’s message and prefers to split the letter according to the pastoral and theological emphases.White’s commentary also begins with the standard introductory matters as the previous commentaries, with one small difference: he devotes a small section to the textual integrity of the text, which examines the manuscript evidence of Colossians. This section devoted to textual witnesses is helpful for those who want to find this information in one place. The commentary proceeds with its textual analysis. White opts for a threefold structure: introduction (1:1–2:5), letter body/corpus (2:6–4:6), and letter closing (4:7–18). Each of the volumes concludes with extensive bibliographies that have various arrangements but are all equally helpful.The benefits of each volume are as long as the volumes themselves. I will limit my comments to a few advantages of each commentary. Beale’s treatment of the text is less a word-for-word exegesis, and more focus is given to individual thought units (paragraphs) and how they relate to what precedes and what follows. Consider his treatment of Col 1:9–14, where Beale notes the lexical parallels with 1:3–8 and the ensuing literary chiasm (or concentric structuring) between the two sections. If one is familiar with Beale’s other projects, particularly his attention to the OT in the NT, they will recognize the emphasis on OT allusions and their function within Colossians. Beale has placed helpful charts throughout the commentary to illustrate the parallels between passages and their respective antecedents (see the comparison of Col 2:2–3 with Dan 2 [Theod] and Prov 2:3–6 on p. 156). One benefit to those who may not know Greek is Beale’s translations of difficult words/phrases and “additional notes” at the end of structurally pivotal points in the letter. These additional notes give parsing information for difficult Greek terms and their use in the rest of the NT and extracanonical literature. Another interesting feature of Beal’s commentary is his use of the NASB, which is a departure from the norm of the BECNT series.The advantages of McKnight’s commentary align with his differing perspectives on theological frameworks. McKnight begins his theological exploration of Paul not through the lens of soteriology but rather through the lens of the Christological mystery. McKnight argues that Paul’s theology begins with a “missional, ecclesial theology” tied to a given location. In so doing, McKnight highlights how Paul’s writing is concerned with the mission of cosmic reconciliation in forming a group set apart for Christian fellowship (emphasis McKnight’s, p. 50). Another aspect of McKnight’s commentary that is commendable is his handling of prior scholarship. As he notes, some of his teachers and professional predecessors (Harris, Dunn, Moo, Pao, and Campbell) have written commentaries on Colossians. He pays tribute to their work in his introduction and throughout the commentary.The contribution of White’s commentary is its footnotes. White’s primary interlocutors are his fellow Germans. Read side-by-side with the commentaries of Beale and McKnight allows the reader to grasp scholarship from both sides of the Atlantic. White also does not burden the reader with excessive amounts of footnotes. In places where the secondary literature is extensive, White cuts through the burden for the interpreter. Consider the exegesis of Col 1:24. Here, interpreters have been vexed by how properly to understand Paul’s role in completing “what is lacking with regard to the sufferings of the Messiah.” White does his due diligence in explaining the relationship between Paul’s apostolic calling and the Isaianic Servant traditions but in a clear and condensed manner. Beale initially provides eight pages of exegesis followed by three more pages in his “additional notes” section while McKnight provides an exegetical section and a six-page excursus on the topic. Both of these latter interpreters have copious numbers of footnotes for further research. One volume is not necessarily to be preferred, but each volume offers different approaches and attention to the problems of interpreting this verse.In conclusion, each volume can be recommended for its distinctives. The three authors have given us research that models charitable engagement with interlocutors, various grammatical and historical gems, and implementation of new theological frameworks for reading this important letter. While no one will agree with all conclusions, reading the breadth of these commentaries will do well to provide exegetically significant insights for students, pastors, and scholars.
This paper presents the foundations, procedures, tests and first results of a dependency treebank of the Spanish Sign Language (LSE). Dependency syntax offers many advantages over other alternatives for the systematic and exhaustive syntactic analysis of a corpus. Nevertheless, the visual modality that is characteristic of sign languages poses unique challenges for their syntactic analysis, among which the most prominent is the simultaneity of expression: both hands, face and other non-manual components. Taking into account these and other particularities of sign languages, the paper explores the main difficulties faced when one tries to apply some usual categories and relations from the syntactic analysis of spoken and written languages to LSE.
Individuals with autism spectrum disorders have difficulty understanding verbal and non-verbal cues, and display atypical gaze behaviour during social interactions. The aim of this study was to examine differences among groups of individuals with high, medium, and low levels of autistic traits with regard to their gaze behaviour and their ability to assess peers’ social status. 54 university students who completed the short Autism Quotient (AQ10) were eye-tracked as they watched six 20-second video clips of individuals (“targets”) involved in a group decision-making task. The specific experimental instruction to the participants was to "think about who you would want to work with on a subsequent task". The video clips included moments of debate, humour, interruptions, and cross talk, simulating natural, everyday social interactions. Fixations were labelled by region of interest (body, face, or eyes). Participants then completed the Dominance and Prestige Peer Rating Scales, which asked them to rate the video targets in terms of status, prestige, and dominance. High-scorers on the AQ10 (i.e., those with more autistic traits) did not differ from the low- and medium-scorers in the status, prestige, and dominance ratings they gave the video targets. Unlike the low- and medium-scorers, high-scorers attended to the body of high dominance targets significantly more than they attended to the low and medium dominance targets, suggesting high-scorers found the high dominance target far more compelling than the medium and low dominance targets. In all other cases, high-scorers did not differ from low- and medium-scorers in either their ability to evaluate social status or in gaze behaviour. This suggests that deficits exhibited by individuals with autistic traits in reading social cues may be reduced in tasks probing certain social skills abilities. The results are discussed in terms of their implications towards the Theory of Mind, Weak Central Coherence, and Social Motivation theories of autism.
Abstract In this article we provide a practical demonstration of how syntactically annotated corpora (treebanks), particularly the English Historical Parsed Corpora Series, can be used to investigate research questions with a diachronic depth and synchronic breadth that would not otherwise be possible. The phenomenon under investigation is split coordination, in which two parts of a conjoined constituent appear separated in the clause (e.g., and this is where my aunt lives and my uncle ). It affects every type of coordinated constituent (subject/object DPs, predicate and attributive ADJPs, ADVPs, PPs and DP objects of P) in Old English (OE); and it, or a superficially similar construction, occurs continuously throughout the attested period from approximately 800 to the present day. Despite its synchronic range and diachronic persistence, split coordination has received surprisingly little attention in the diachronic literature, with the exception of Perez Lorido’s (2009) limited study of split subjects in eight OE texts. Its modern counterpart is most frequently analysed as Bare Argument Ellipsis (BAE). Although the OE and Present-Day English constructions appear superficially similar, we show that not all of the OE data is amenable to a BAE analysis. We bring to bear different types of evidence (structural, discourse/performance effects, rate of change, etc.) to argue that split coordination in fact represents two different constructions, one of which remains stable over time while the other is lost in the post-Middle English period.
The present paper aims to broaden the current understanding of students’ misconception of scientific terminology by identifying the gaps between Arabic and English scientific terminologies and between everyday language and scientific language. The paper compares the polysemy, prototypes, and motivating factors of English energy with those of Arabic طَاقَة (ṭāqa), with more focus on students’ prior knowledge. The study employs Lakoff’s (1987) idealized cognitive models and Rosch’s (1975) prototype theory to reveal the radial members of both categories, i.e., energy and طَاقَة (ṭāqa), and to explain the kinds of cognitive mechanisms that motivate the extension as well as understanding of the meanings of these terms. To this end, the study uses several English and Arabic dictionaries, lexical databases and corpora. This is to explore all the meanings, prototypes and motivating factors of the terms under investigation. The results show that the terms energy and طَاقَة (ṭāqa) overlap in prototypical meanings and motivating factors but differ in less prototypical and peripheral meanings. English and Arabic learners may then face similar issues in learning scientific concepts due to the difference between their pre-existing knowledge and scientific language.
Many NLP applications, such as biomedical data and technical support, have\n10-100 million tokens of in-domain data and limited computational resources for\nlearning from it. How should we train a language model in this scenario? Most\nlanguage modeling research considers either a small dataset with a closed\nvocabulary (like the standard 1 million token Penn Treebank), or the whole web\nwith byte-pair encoding. We show that for our target setting in English,\ninitialising and freezing input embeddings using in-domain data can improve\nlanguage model performance by providing a useful representation of rare words,\nand this pattern holds across several different domains. In the process, we\nshow that the standard convention of tying input and output embeddings does not\nimprove perplexity when initializing with embeddings trained on in-domain data.\n
Abstract The current investigation examined the development of second language (L2) intensifier use in spoken Spanish over a 6-week immersion program in Madrid ( n = 45). Native Spanish speakers from Madrid ( n = 10) served as a comparison group to represent the local ambient input or sociopragmatic norm to which L2 learners were exposed. Data were extracted from semi-structured interviews. Results exposed different developmental trends over the program for intensifier frequency, intensifier lexical diversity, and intensifier collocations. While learners already had a strong sense of which intensifiers were most frequent in Spanish and how to use them in appropriate linguistic environments at the beginning of the program, the immersion program had positive impacts on the development of intensifier frequency and intensifier lexical diversity. The findings also highlighted different intensifier frequency developmental trends among learners, which collectively suggested that learners adjusted to the sociopragmatic norm of intensifier use in Madrid over the immersion experience.
O presente artigo, resultante da pesquisa de mestrado, apresenta a extração automatizada de palavras-chaves da tragédia grega Héracles, de Eurípides, associada à anotação morfossintática em árvore (treebanking) com a finalidade de determinar temas e aspectos formais da obra. Dessa forma, primeiramente Héracles fora comparada estatisticamente com as outras dezoito obras de Eurípides, sem lematização. Posteriormente, foram selecionados trechos contendo as palavras-chaves com a finalidade de analisá-las morfossintaticamente em árvore. Desse modo, os resultados evidenciaram quatro temas centrais na tragédia como os personagens, os laços familiares, a caracterização de Héracles por meio de suas armas e trabalhos, e, por fim, sua loucura.
In this welcome study of the second-century Apocryphal Acts of Paul and Thecla (ATh for Acts of Thecla), J. D. McLarty is particularly interested in questions of emotion and identity. The exploration of themes such as gender, class, and citizenship is driven by narrative analysis and comparison with one of the five complete Greek novels, Chariton's Callirhoe. Revised from McLarty's 2011 PhD thesis, Thecla's Devotion leads the reader carefully through both ATh and Callirhoe, showing their striking emotive similarities and differences. McLarty argues that, despite being a work of Christian fiction, ATh provides a helpful example of how early ‘Christians in one part of the eastern Empire constructed an identity for themselves’ (p. 232). In Chapter 1, McLarty examines the context of ATh, discussing questions of composition, origin, and date. She accepts the view that the Thecla episode was originally included in the larger narrative of the Acts of Paul but was later disseminated independently with the development of the cult of St. Thecla (already present in the fourth century). It is likely that the text of ATh originated in ‘south central Asia Minor’ (p. 5) and is dated to the ‘mid-to-late second century’ (p. 7). Concerning whether the author of the text was male or female, McLarty concludes, ‘even if one wishes to argue for at least some female contribution to the narrative in the form of oral legend, these contributions have been absorbed into a masculine literary culture’ (p. 9). However, the readers and hearers of the text were likely mixed in both gender and class. Rather than a primarily oral development, McLarty sees the composition as reflecting a ‘predominantly literary milieu’ (p. 17), which both imitated and reworked traditions from the Acts of the Apostles. The world of ATh was a changing one, with a growing interest in the ethics of the individual and self-control. This focus is clearest in relation to discussions of marriage and continence contemporary to ATh. McLarty is, therefore, interested in the intersection of pagan ideas of self-control, reflected in Callirhoe, with a Christian view that contains a distinct ‘spiritual dimension to the control of the passions’ (p. 23). Story, in contrast to pure discourse, provides a unique opportunity to affect the emotions of the reader, but may also provide insight into the author's worldview. This study is broken into two parts. The first part analyses the plot of ATh in comparison to Callirhoe. In Chapter 2, McLarty outlines her methodology, focused especially on the concept of ‘ “affect” – the emotive atmosphere created by the construction of plot’ (p. 28). The next chapter provides a diachronic study through an extended discussion of the plots of Callirhoe (Book 1) and ATh. With both narratives, the points of interest are (1) the teleology of the plot, (2) the loci of tension, (3) and causation. While Callirhoe anticipates a safe return home, ATh ends with all of Thecla's ties to family and home broken, leaving her in ‘a plot space that is without definition’ (p. 89) as she goes about evangelizing. Plot tension presents a challenge to the reader's assumption that all will end well for the protagonists. This tension also provides Thecla with an opportunity to mature through perseverance, a process that culminates in self-baptism. While the concept of ‘Chance’ (τύχη) – sometimes personified as a deity – plays an important role in the causation of Callirhoe, it does not appear in ATh. Instead, God appears in the narrative through deus ex machina (in the theater of Iconium), in order to save the heroine. Human causation in ATh is largely confined to the lack of emotional control, with Paul being the main exception. Chapters 4 and 5 are devoted to a synchronic study, consisting of an analysis of time (Chapter 4), space, and place (Chapter 5). While the author of Callirhoe uses many technical literary features for emotive affect, ATh is often simplified and provides additional moral exhortation. McLarty discusses how the idea of place is subverted in ATh: Tombs and prisons become places of Christian devotion, and the stadium is used for baptism, rather than death. While McLarty claims that ‘Thecla is not portrayed as overturning male authority’ (p. 119), she does transgress boundaries of gender and class. Thecla is found denying her betrothed and, therefore, the civic value of marriage; she roams the streets unaccompanied and ultimately develops into a wandering ‘ “type” of Christian philosopher’ (p. 124). The second part of this study examines character and characterization as the locus of expressed emotion in the narratives. From here, ‘one can therefore glean useful information about the role of emotion in the ATh author's own community’ (p. 132). McLarty proposes a methodology that combines ancient and modern theories of character and compares ATh with Callirhoe as its ‘genre model’ (though she is clear that it is not the only model). Chapters 7 and 8 explore characterization in Callirhoe (Chapter 7) and ATh (Chapter 8), each divided by female and male characters in the narratives. Characters are described in terms of social class and emotional expression, which reveals that, in these narratives, emotion (produced from desire) is reserved mostly for the upper class. Callirhoe also shows the ‘potential destructiveness of uncontrolled masculine emotions’ (p. 166). At times the heroine is depicted as more masculine or ‘warlike’ (p. 167), while the hero is feminized. However, the conclusion of the narrative provides a return home, both physically and emotionally, with both characters fulfilling their culturally defined gender and societal roles. ATh, on the other hand, contrasts and subverts many of these expectations. Although Thecla is understood to be overtaken by passion for Paul, not unlike to the heroine of Callirhoe, it becomes clear to the reader that her true desire is for the gospel he preaches. McLarty concludes that Thecla's character remains that of the ‘perpetual παρθένος [unmarried girl]’ (p. 219). In spite of this, the narrative ends with Thecla reaching maturity through baptism and possessing ‘a masculine control of her emotions’ (p. 219). Again, the contrast with Callirhoe is apparent when the reader understands that Thecla does not return home but remains independent and isolated. This combination of celibacy and isolation subverts many of the second century assumptions about gender and class – namely, that young women (especially of high status) should marry and let their husbands protect them from falling victim to eros (desire). While Paul does not subvert gender norms as such, he does challenge many social categories. ATh portrays Paul as an opponent to the leading men of the city, who are fighting for Thecla's affection. However, ‘Paul does not take part in this contest, an indication that the honour system of the Christian community is different from that of the wider pagan society’ (p. 220). McLarty's final chapter provides a useful summary of emotion and identity in ATh, while also returning to some motifs that she hints at in the beginning of the study. In particular, she discusses the Christian narrative's interaction with pagan philosophy (especially Stoic and Cynic). ATh presents Paul as ‘like a philosopher in his mastery of the passions’ and as a wise man ‘countering the kinds of argument [against Christians] advanced by Celsus’ (p. 227). In contrast to these philosophies, the ‘mastery of the passions’ achieved by both Paul and Thecla comes ‘through Christ’ (p. 232). There is much to admire in McLarty's approach to ATh. Her detailed lexical analysis of both Callirhoe (Book 1) and ATh is sure to provide a useful guide for readers. McLarty also draws from a broad range of classical and Christian texts as comparanda to this apocryphal book. Although she makes a strong case for the predominantly Graeco–Roman background of ATh, one might wonder whether comparison with a Jewish novel like Joseph and Aseneth could further shed light on character and emotion. It is not clear that the bibliography has been fully updated from her 2011 thesis. For example, one will not find interaction with Susan Hylan's A Modest Apostle: Thecla and the History of Women in the Early Church (Oxford: OUP, 2015), who argues that Thecla does not depart from modesty, contrary to what McLarty claims (p. 170). Finally, it is possible that McLarty's emphasis on Thecla's isolation overshadows certain hints at her accumulation of followers. For this reason, McLarty does not mention the ‘band of young men and maidens’ (ATh p. 40), who accompany Thecla in her final return to Paul. This does not negate the overall point about Thecla's subversive character but may indicate some desire in the second century to follow the heroine's example. In the end, McLarty's captivating prose and persuasive arguments ensure that this work is an important contribution to the study of ATh.
This paper explores the possibility of improving the performance of\nspecialized parsers for pre-modern Slavic by training them on data from\ndifferent related varieties. Because of their linguistic heterogeneity,\npre-modern Slavic varieties are treated as low-resource historical languages,\nwhereby cross-dialectal treebank data may be exploited to overcome data\nscarcity and attempt the training of a variety-agnostic parser. Previous\nexperiments on early Slavic dependency parsing are discussed, particularly with\nregard to their ability to tackle different orthographic, regional and\nstylistic features. A generic pre-modern Slavic parser and two specialized\nparsers -- one for East Slavic and one for South Slavic -- are trained using\njPTDP (Nguyen & Verspoor 2018), a neural network model for joint part-of-speech\n(POS) tagging and dependency parsing which had shown promising results on a\nnumber of Universal Dependency (UD) treebanks, including Old Church Slavonic\n(OCS). With these experiments, a new state of the art is obtained for both OCS\n(83.79\\% unlabelled attachment score (UAS) and 78.43\\% labelled attachement\nscore (LAS)) and Old East Slavic (OES) (85.7\\% UAS and 80.16\\% LAS).\n
Neste artigo, buscamos descrever as concepcoes de lingua e norma linguistica que sao veiculadas em gramaticas produzidas pela Real Academia de Espanola (RAE) – instituicao fundada na Espanha, no inicio do seculo XVIII, cuja missao principal e a “defesa da unidade da lingua”. Foram feitas analises de excertos de dois manuais da instituicao, a saber: Esbozo de una nueva gramatica de la lengua espanola (1982) e Nueva gramatica de la lengua espanola - Manual (2010). Na analise, buscamos identificar de qual concepcao de Norma Linguistica o discurso normativo da RAE mais se aproxima, observando tanto os capitulos introdutorios quanto um capitulo com descricoes propriamente ditas. Apresentaremos, aqui, os resultados obtidos ao longo do projeto, que mostraram uma aproximacao maior a visao mais prescritiva do termo norma, apesar de todo o projeto de pan-hispanismo que teve como evento importante a publicacao dessas duas obras. ABSTRACT: In this paper, we describe how concepts such as ‘language’ and ‘linguistic norm’ are presented on the grammars produced by Real Academia Espanola (RAE) – an institution founded in Spain in the 18th century who claims to have the mission of “defending the unity of the language”. We made an analysis of two books published by RAE: Esbozo de una nueva gramatica de la lengua espanola (1982) and Nueva gramatica de la lengua espanola - Manual (2010). In the analysis, we seek to identify from which conception of Linguistic Norm the normative discourse of RAE comes closest and we have done that by studying the introductory chapters and a chapter with grammatical content from both publications. Here, we present the results of this research, which showed that the content of both of RAE’s books has more closeness to a more prescriptivist vision of said concepts, even though they are a direct result of the whole pan-hispanic politics promoted by the academy. KEYWORDS: Linguistic norm; Grammaticography; Real Academia Espanola; Spanish.
This article offers a review of the most relevant general dictionaries of Spanish. In doing so, we consider briefly the relationship between linguistic norms, standardization and dictionaries, highlighting the concept of "general dictionary" as most relevant to the issue of normativity. Thereafter, we take a look at the historical origins and emergence of Spanish general lexicography. Finally, we pay attention to more recent developments in the field as there are the lexicographic implications and consequences of pluricentricity which implies a new mode of codification as well as new developments due to digitalization.
Fully data-driven, deep learning-based models are usually designed as language-independent and have been shown to be successful for many natural language processing tasks. However, when the studied language is low-resourced and the amount of training data is insufficient, these models can benefit from the integration of natural language grammar-based information. We propose two approaches to dependency parsing especially for languages with restricted amount of training data. Our first approach combines a state-of-the-art deep learning-based parser with a rule-based approach and the second one incorporates morphological information into the parser. In the rule-based approach, the parsing decisions made by the rules are encoded and concatenated with the vector representations of the input words as additional information to the deep network. The morphology-based approach proposes different methods to include the morphological structure of words into the parser network. Experiments are conducted on the IMST-UD Treebank and the results suggest that integration of explicit knowledge about the target language to a neural parser through a rule-based parsing system and morphological analysis leads to more accurate annotations and hence, increases the parsing performance in terms of attachment scores. The proposed methods are developed for Turkish, but can be adapted to other languages as well.
Universal Dependencies conversion of the Late Latin Charter Treebank 2
Abstract Consumerism is an inherent feature of a modern consumer-minded society which enhances in some people both a hunger for collecting and a more serious desire developing into oniomania (shopping mania, shopaholism), kleptomania or pathological hoarding (syllogomania, Diogenes syndrome, Plyushkin’s syndrome). The paper proposes an interdisciplinary approach to the problem of pathological hoarding of unnecessary things and domestic animals by tenants of condominiums in Russian cities. Socio-legal prerequisites for this psychosocial disease still insufficiently studied in the country are also analyzed in the paper. The data of the Federal State Statistics Service of the Russian Federation and various legal acts were examined. In the course of the study, methods of quality content analysis and visual sociology were used to analyze cases of pathological hoarding highlighted in Russian digital media in recent years. The Clutter-Hoarding Scale and Clutter Image Rating Scale were used to interpret photos of cluttered Russian flats in condos. In conclusion, recommendations are given on improving state policy and the legislation of the Russian Federation.
Deep learning has promoted remarkable progress in various tasks while the effort devoted to these hand-crafting neural networks has motivated so-called neural architecture search (NAS) to discover them automatically. Recent aging evolution (AE) automatic search algorithm turns to discard the oldest model in population and finds image classifiers beyond manual design. However, it achieves a low speed of convergence. A nonaging evolution (NAE) algorithm tends to neglect the worst architecture in population to accelerate the search process whereas it obtains a lower performance compared with AE. To address this issue, in this letter, we propose to use an optimized evolution algorithm for recurrent NAS (EvoRNAS) by setting a probability ϵ to remove the worst or oldest model in population alternatively, which can balance the performance and time length. Besides, parameter sharing mechanism is introduced in our approach due to the heavy cost of evaluating the candidate models in both AE and NAE. Furthermore, we train the sharing parameters only once instead of many epochs like ENAS, which makes the evaluation of candidate models faster. On Penn Treebank, we first explore different ϵ in EvoRNAS and find the best value suited for the learning task, which is also better than AE and NAE. Second, the optimal cell found by EvoRNAS can achieve state-of-the-art performance within only 0.6 GPU hours, which is 20 × and 40 × faster than ENAS and DARTS. Moreover, the transferability of the learned architecture to WikiText-2 also shows strong performance compared with ENAS or DARTS.
RST-based discourse parsing is an important NLP task with numerous downstream applications, such as summarization, machine translation and opinion mining. In this paper, we demonstrate a simple, yet highly accurate discourse parser, incorporating recent contextual language models. Our parser establishes the new state-of-the-art (SOTA) performance for predicting structure and nuclearity on two key RST datasets, RST-DT and Instr-DT. We further demonstrate that pretraining our parser on the recently available large-scale "silver-standard" discourse treebank MEGA-DT provides even larger performance benefits, suggesting a novel and promising research direction in the field of discourse analysis.
Many language teachers use Information and Communications Technology (ICT) in their classrooms to create tasks, quizzes, or polls with general online learning platforms. Few teachers have experience, however, of incorporating online corpus tools in their teaching or assessment practices. This paper will explore how autonomous learning can be fostered by gradually introducing freely available lexical databases, online collocation dictionaries, pronunciation guides, concordancers, N-gram extractors, and other text analysis tools for vocabulary building, skills practice, or self-checking. Tasks used with English as a Foreign Language (EFL) undergraduates and teacher trainees on a Master’s Teaching English as a Foreign Language (MA TEFL) course will be presented. I will also explain why having some familiarity with linguistics research can enable teachers to use these applications more meaningfully.
Abstract The paper presents the application of non-specialized lexical database and semantic metrics on transcripts of co-design protocols. Three different and previously analyzed design protocols of co-creative sessions in the field of packaging design, carried out with different supporting tools, are used as test-bench to highlight the potential of this approach. The results show that metrics about the Information Content and the Similarity maps with sufficient precision the differences between ICT- and non-ICT-supported sessions so that it is possible to envision future refinement of the approach.
Temperature scaling has been widely used as an effective approach to control the smoothness of a distribution, which helps the model performance in various tasks. Current practices to apply temperature scaling assume either a fixed, or a manually-crafted dynamically changing schedule. However, our studies indicate that the individual optimal trajectory for each class can change with the context. To this end, we propose contextual temperature, a generalized approach that learns an optimal temperature trajectory for each vocabulary over the context. Experimental results confirm that the proposed method significantly improves state-of-the-art language models, achieving a perplexity of 55.31 and 62.89 on the test set of Penn Treebank and WikiText-2, respectively. In-depth analyses show that the behaviour of the learned temperature schedules varies dramatically by vocabulary, and that the optimal schedules help in controlling the uncertainties. These evidences further justify the need for the proposed method and its advantages over fixed temperature schedules.
In this Letter, the authors introduce a novel approach to learn representations for sentence‐level paraphrase identification (PI) using BERT and ten natural language processing tasks. Their method trains an unsupervised model called BERT with two different tasks to detect whether two sentences are in paraphrase relation or not. Unlike conventional BERT, which fine tunes the target task such as PI to pre‐trained BERT, twice fine‐tuning deep neural networks first fine tune each task (e.g. general language understanding evaluation tasks, question answering, and paraphrase adversaries from word scrambling task) and second fine tune target PI task. As a result, the multi‐fine‐tuned BERT model outperformed the fine‐tuned model only with Microsoft Research Paraphrase Corpus (MRPC), which is paraphrase data, except for one case of Stanford Sentiment Treebank ‐ 2 (SST‐2). Multi‐task fine‐tuning is a simple idea but experimentally powerful. Experiments show that fine‐tuning just PI tasks to the BERT already gives enough performance, but additionally, fine‐tuning similar tasks can affect performance (3.4% point absolute improvement) and be competitive with the state‐of‐the‐art systems.
abstract Reaction times for a translation recognition study are reported where novice to expert English–ASL bilinguals rejected English translation distractors for ASL signs that were related to the correct translations through phonology, semantics, or both form and meaning (diagrammatic iconicity). Imageability ratings of concepts impacted performance in all conditions; when imageability was high, participants showed interference for phonologically related distractors, and when imageability was low participants showed interference for semantically related distractors, regardless of proficiency. For diagrammatically related distractors high imageability caused interference in experts, but low imageability caused interference in novices. These patterns suggest that imageability and diagrammaticity interact with proficiency – experts process diagrammatic related distractors phonologically, but novices process them semantically. This implies that motivated signs are dependent on the entrenchment of language systematicity; rather than decreasing their impact on language processing as proficiency grows, they build on the original benefit conferred by iconic mappings.
Pain is evolutionarily hardwired to signal potential danger and threat. It has been proposed that altered pain-related associative learning processes, i.e., emotional or fear conditioning, might contribute to the development and maintenance of chronic pain. Pain in or near the face plays a special role in pain perception and processing, especially with regard to increased pain-related fear and unpleasantness. However, differences in pain-related learning mechanisms between the face and other body parts have not yet been investigated. Here, we examined body-site specific differences in associative emotional conditioning using electrical stimuli applied to the face and the hand. Acquisition, extinction, and reinstatement of cue-pain associations were assessed in a 2-day emotional conditioning paradigm using a within-subject design. Data of 34 healthy subjects revealed higher fear of face pain as compared to hand pain. During acquisition, face pain (as compared to hand pain) led to a steeper increase in pain-related negative emotions in response to conditioned stimuli (CS) as assessed using valence ratings. While no significant differences between both conditions were observed during the extinction phase, a reinstatement effect for face but not for hand pain was revealed on the descriptive level and contingency awareness was higher for face pain compared to hand pain. Our results indicate a stronger propensity to acquire cue-pain-associations for face compared to hand pain, which might also be reinstated more easily. These differences in learning and resultant pain-related emotions might play an important role in the chronification and high prevalence of chronic facial pain and stress the evolutionary significance of pain in the head and face.
We propose a novel constituency parsing model that casts the parsing problem into a series of pointing tasks. Specifically, our model estimates the likelihood of a span being a legitimate tree constituent via the pointing score corresponding to the boundary words of the span. Our parsing model supports efficient top-down decoding and our learning objective is able to enforce structural consistency without resorting to the expensive CKY inference. The experiments on the standard English Penn Treebank parsing task show that our method achieves 92.78 F1 without using pre-trained models, which is higher than all the existing methods with similar time complexity. Using pre-trained BERT, our model achieves 95.48 F1, which is competitive with the state-of-theart while being faster. Our approach also establishes new state-of-the-art in Basque and Swedish in the SPMRL shared tasks on multilingual constituency parsing.