Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
The PROIEL Treebank is a dependency treebank with morphosyntactic and information-structure annotation. It includes texts in several ancient Indo-European languages and is freely available under a Creative Commons Attribution-NonCommercial-ShareAlike 3.0 License.
The object of considerations focuses on problems related with the translation of semantic innovations of idioms. The first part of the article introduces the concepts of idioms, idiomatic standards and phraseological innovations. Phraseological connections or idioms are multi-word units that have a relatively stable form and meaning. This meaning is not a combination of the meanings of separate words that form the unit. A norm is accepted by a community and defined as a standard set of rules pertaining to the use and occurrence of idioms. The most common phraseological innovations are demetaphorisation, semantic contamination, desemantisation and idiom accumulation in the text. The effect of innovations is the fusion of the above mentioned factors with a text making it unique and one of its kind. Innovations expand and determine utterances. What is more, they are added value expressed by a pun. Translation of idioms may cause many problems, especially when they appear in the form and meaning different from the standard one. The main strategies in the translation of innovations are literal translation, using a different idiom that reflects the meaning of the original text, paraphrasing or creating new lexical units. Examples of semantic innovations and their translation strategies have been derived from the novel by Gunter Grass and its Polish translation.
Cette thèse s’inscrit dans le cadre général de la linguistique descriptive et s’intéresse à l’espagnol familier d’usage au Chili, à travers les affects exprimés par les locuteurs chiliens. Notre hypothèse consiste dans le fait que l’espagnol du Chili est défini sur la base des affects, et que l’expression de ceux-ci, au niveau linguistique, est observable dans les phénomènes les plus divers, pouvant aller d’un simple morphème à une structure syntagmatique complexe. Afin de concevoir l’expression linguistique des affects d’un point de vue plus large, cette recherche vise l’étude de quatre phénomènes clairement différenciés: le suffixe –it dans le cadre de la communauté linguistique chilienne, la « paronomase orientée », le défigement phraséologique dans le domaine des locutions verbales et adverbiales et la particule illocutoire poh. Notre méthodologie consiste en l’extraction d’un nombre d’exemples conséquents pour illustrer les phénomènes choisis à partir de 103 numéros du journal chilien La Cuarta de 2010 et 2011. Parallèlement nous considérons un nombre significatif d’exemples issus de l’enregistrement de conversations menées auprès de 58 locuteurs chiliens pendant 9 heures 11 minutes et 33 secondes.Les quatre phénomènes de langue visés ont été étudiés sous le concept d’affectivité proposé par Bally (1965 [1913]), qui distingue les principes d’intensité et de valeur (1951 [1909]).L’étude de l’affectivité du suffixe –it dans le contexte chilien a révélé que l’affectivité est une propriété qui opère sur deux dimensions. Elle est, d’un côté, intrinsèque à un signe linguistique – base lexicale, suffixe –it et tout autre élément linguistique convergeant dans le discours – et, d’un autre côté, extrinsèque, car relatée à des éléments culturels ou idéologiques, et plus généralement à tout ce qui est extralinguistique. Dans les deux dimensions, l’affectivité du locuteur fonctionne comme élément de fusion et donne son vrai sens au signe linguistique. Concernant la « paronomase orientée » (terme que nous avons proposé), nous avons constaté que l’affectivité constitue la caractéristique principale de ce type de paronomase, et que son usage repose notamment sur la figure de la plaisanterie, cela étant dû à des caractéristiques communes qui se sont révélées être: la fonction ludique, l’intentionnalité comique et l’effet de surprise.La paronomase orientée, qui s’investit sur les plans phonétique, morphologique, lexical et sémantique, constitue pour nous une opération linguistique dérivative, où deux lexies simples ou complexes, qui se substituent l’une à l’autre au sein d’un énoncé, partagent des propriétés phonétiques, alors que leur contenu sémantique diverge.En liaison avec le défigement des locutions verbales et adverbiales, il s’avère que le défigement phraséologique est une activité linguistique naturelle faisant appel à la relation figurative constante entre les mots, qui permet de créer d’autres manières d’exprimer et d’actualiser les usages de termes déjà existants.L’évolution des phrases figées passe par un processus de transformation souhaité par les locuteurs au détriment des normes phraséologiques ou syntaxiques données. Cependant, l’intérêt de la création néologique intervient quand le locuteur a la possibilité de créer et de recréer des structures linguistiques nouvelles.Quant à la particule illocutoire poh, particule notamment orale qui provient du connecteur pues (Oroz: 1966), nous avons réussi à montrer que poh aide à la progression de la communication, favorisant la cohérence et la cohésion entre l'énoncé et le texte. De plus, par le truchement de poh, un accord consensuel entre le locuteur et l’interlocuteur s’établit et, à partir de cet accord, des fonctions pragmatico-affectives nouvelles surgissent.
Reviews 277 vernacular languages, for example by teaching malhun—centuries-old popular poetry sung and written in the vernacular language found in the Maghreb—in the schools, since it represents remembrance and cultural traditions. Examining language in media such as cinema, theater, and songs, Chachou points to the absence of Algerian Arabic and Berber in newspapers even though those languages are present on national television. She notes the strong position of standard Arabic on the public radio while local radio stations increasingly broadcast Algerian Arabic or l’arabe médian, a variety situated between the standard and the vernacular. Algerian Arabic and Berber dominate in music, although Rai folk music uses borrowings from French and Algerian Arabic and code-switches between the two. The second half of the book deals with language in newspaper advertising, covering theory (for example, Bourdieu’s notion that language and values reflect power relationships, ideologies, and norms), linguistic strategies used by advertisers (“Algerianisms,” monolingual and bilingual signs, and the predominance of standard Arabic during Ramadan and other Muslim holidays), and English and Italian borrowings. Examples are followed by brief data analysis. Overall, this book contains a great deal of useful information, but its goals, which cover a wide range of contexts, may be overly ambitious and at times not particularly original. The chapters on advertising are the book’s strength, as they offer a rich set of data. The data, however, could benefit from deeper analysis. Nonetheless, this volume provides a valuable contribution to the field of Maghrebi sociolinguistics due to its empirical approach. Manhattan College Samira Hassa Faucher, Marie. L’enfance des mots: l’étymologie vagabonde. Paris: Silène, 2013. ISBN 978-2-913947-13-9. Pp. 159. 18 a. This book of “etymology” is neither a learned study nor a dictionary, but rather what Faucher calls a search for “le berceau des mots,” original meanings, as a route to understanding unexpected connections in the French lexicon. In an attitude of wideeyed, playful enjoyment, she has gathered evidence of lexical families, some of them surprisingly widespread. The preface by poet Henri Gougaud sets the stage, reminding us that all words have stories to tell, as he praises the author’s vagabondage in the forest of words. The book includes twenty chapters, ending with a summary of chapter contents (139–48), a word index (149–55), and a bibliography (157) of three titles: the Robert étymologique (RE), the Robert historique (RH), and, curiously, a student’s etymological dictionary dating from 1909. Unfortunately, Faucher has taken as her reporting scheme and model the“présentation synthétique”of RE (devised by Picoche as an antidote to strictly alphabetical presentation such as found in RH). Inspired by the RE, which shows coussin ( having been consulted, only two of the fifteen pairs such as coussin/cuisse listed under“Did you know that...?”prove to demonstrate the kind of etymological connection this reviewer understands by the term“vient de.” Forêt does come from Latin FORESTIS (SILVA); and divan is correctly identified— although one cannot accept that “divan vient de douane,” since, at best, each is an extension of the original Turkish meaning of ‘conseil politique, salle de conseil.’In the body of the text, Faucher’s efforts are generally more successful, her phrasing more careful, for example, “outil vient d’USITILIUM, qui vient d’UTI, USUS” (67). The etymologist is happy to see an accusative case as etymon and happier yet to see that the noun in –IUM is shown as deriving from the verb. Faucher’s presentation of the twenty or so words in the family of VEN RE (44–45) is clean and clear, as are many others of the 250 words presented.Yet her reported understanding of the etymological relationships between chosen pairs of words remains open to question, and her basic premise (“sens premier”=“sens propre”[12]) appears indefensible, since—as Faucher herself admits—the etymological origins of words may be far from their current meaning. In following her search to“laver les mots pour leur rendre leur netteté,”she has nonetheless succeeded in finding a gracious way to link members of many lexical families and, in so doing, to illuminate their interrelationships. Caveat...
Abstract The tension in the sixteenth century between Christian institutions, viewed as corrupt, and the Scriptures, taken to be the unsullied source of spiritual renewal, gives rise to an energetic biblical erudition intended to transmit this clarity--an effort that ironically ends up obscuring that supposedly limpid source. Spinoza subjects the Book to norms of reason that distinguish sharply between the intended meaning (sensus) and truth (veritas). The truth of the text, for Erasmus, is either above it (in a philosophy of the revelation) or below (in lexical, historical and grammatical knowledge). In the seventeenth century, the logic of Port-Royal sunders the text into pre-existent concepts or ideas, and signifiants, or phonetic signs that are meaningless in themselves. The distinction between metaphor (and musicality) and strict meaning (stripped of rhetoric and affect) makes translation a battlefield between beauty and truth. The remainder of the chapter gives a detailed account and assessment of the two major seventeenth-century biblical translators: Sacy, whose focus is the intended meaning of the Author (the spirit of God); and Simon, who treats the text as an object. The text itself must first be "established," and then its obscurities elucidated by compared versions and an abundant critical apparatus.
У статті проаналізовано норми усного наукового спонтанного висловлення у зв’язку з психолінгвальними чинниками формування спонтанного мовного потоку. При цьому актуалізовано ідеї, погляди, підходи Олександра Опанасовича Потебні щодо пояснення специфіки ословлення думки, формування судження, поняття. Відзначено діалектичність взаємозв’язку символічних і слабких мовних норм, що становлять якісну ознаку усної спонтанної мови загалом. (The article analyzes the scientific standards of spontaneous verbal expression because of psycholinguistic factors in the formation of spontaneous speech stream. The analyzed material showed that spontaneous verbal text is formed gradually, in portions, through associative connections. Talking can do pause and return to have said, repeating the expression of certain components and adjust its content and the way verbalization sense, control logic, composite, syntactic and lexical and phraseological correctness, connectivity, logical adequacy. Because of this oral scientific markers in spontaneous utterance – a significant factor in the organization of information flow language. Elements of the so-called redundancy and elements of hesitation – a situational caused by non-normative terms written practice elements that prove the spontaneity of the speech stream, psychological and emotional stress speaking in terms of formal scientific communication, which requires a correct, clear, logical design attitudes, opinions senders. Such phenomena study began relatively recently. In order to make sweeping generalizations, need comprehensive, typological studies of spontaneous oral language in different genre and discursive conditions. This will clarify regulatory features of this genre and stylistic variety. Modified arguments and ideas, opinions, attitudes Olexandr Potebnia to explain specific verbalization opinion forming judgments, concepts. Attention to this theoretic material due to the fact that this year marks the philological community 180 years from the birth of the philosopher and Slavist. His work always in the circle of attention of philologists and still in need of updating, the promotion is in Ukrainian linguistics, for the scientist made his time Kharkiv philological school. Noted dialectical relationship and weaknesses of symbolic language rules that make quality a sign of spontaneous oral language in general. More analysis of symbolic language communication standards regulations. Much attention is given to the weak representation of language norms in oral scientific spontaneous utterance.)
Implementation intention (IMP) has recently been highlighted as an effective emotion regulatory strategy. Most studies examining the effectiveness of IMPs to regulate emotion have relied on self-report measures of emotional change. In two studies we employed electrodermal activity (EDA) and heart rate (HR) in addition to arousal ratings (AR) to assess the impact of an IMP on emotional responses. In Study 1, 60 participants viewed neutral and two types of negative pictures (weapon vs. non-weapon) under the IMP "If I see a weapon, then I will stay calm and relaxed!" or no self-regulatory instructions (Control). In Study 2, additionally to the Control and IMP conditions, participants completed the picture rating task either under goal intention (GI) to stay calm and relaxed or warning instructions highlighting that some pictures contain weapons. In both studies, participants showed lower EDA, reduced HR deceleration and lower AR to the weapon pictures compared to the non-weapon pictures. In Study 2, the IMP was associated with lower EDA compared to the GI condition for the weapon pictures, but not compared to the weapon pictures in the Warning condition. ARs were lower for IMP compared to GI and Warning conditions for the weapon pictures.
Subject matter of the research: cognitive communicative discourse properties of a bilingual. The objective was to characterize a bilingual’s capacities of forming the types of knowledge and skills during several stages of formation of his/her competence. Methodology: in the article there were used as a methodological basis the cognitive and anthropocentric principles, competence approach, focusing on the study of the ways of formation and representation of knowledge, skills, and competencies of man in his cognitive activity. Methods of work: a conceptual analysis, dialingual analysis (identification and description of the types of interference), the competence analysis, discourse method. Results: 1) the author studied the bilingual personality as one of the types of language personality; 2) the author described in a cognitive aspect types of knowledge a bilingual acquires at various stages of formation of his/her competence; 3) identified the types of knowledge phonetic, lexical, grammatical (linguistic) and cultural at the first and second stages of a bilingual’s competence formation arising from a personality’s insufficient competence in the second language and culture of the nation of the target language; 4) clarified a bilingual’s cases of misunderstandings the connotative meanings of words and idioms in the second language; 5) proved the necessity of forming the communicative pragmatic competence, recognizing the speakers’ intentions, for regulating the communication. In conclusion it should be noted that the issue of competence building of a bilingual in the second language is one of the problems that is not completely solved. The main way to solve this issue is to take into account the need to master not only the knowledge (cognitive aspect), but discursive properties of a speech in the second language (knowledge of the speaker’s intentions, mastering the pragmatic norms, pragmatic value, the ability to model frames, to conceptualize the notion (cognitive and discursive aspects).
Cet article presente une experience d’enquete sur les cooccurrents des lemmes latins designant les femmes, en particulier ego, femina et uxor, dans les actes bourguignons des IXe-XIe siecles (reunis dans la base des donnees des CBMA- Chartae Burgundiae Medii Aevi) et dans les tomes 97 a 165 de la Patrologie latine, qui regroupent, grosso modo, des textes latins datant des IXe, Xe et XIe siecles. La recherche a fait ressortir que, dans une societe emaillee par des rapports de subordination, certaines normes d’agencement lexical relevent de la categorisation des personnes, des determinations par le genre et par le statut social.
The aim of this study was to analyse the psycholinguistic variables of the attributes and concepts involved in the recall of a concept. One hundred and twenty adults (18–40 years old) participated. A lexical recall task was administered by presenting a successive list of defining attributes. Forty concepts from different semantic categories were used. The attributes were obtained empirically from local Semantic Features Production Norms. The influence of the characteristics and attributes of the concepts on the number of participants who accessed the name of the concept and the correct guess trend was analysed. Significant values for Age of Acquisition, Presence of Distinctive Attributes and Presence of Taxonomic Attributes were observed. Results show that concepts which are acquired earliest are more easily recalled; presenting taxonomic categories narrows the search and the presence of distinctive attributes allow differentiating between such concepts within a category.
En):The paper, a case study in modern French usage and style, deals with the question of the linguistic norm in Muriel Barbery's novel, The Elegance of the Hedgehog (2006).The point is to show how the self-educated main character, Renée Michel, concierge in an elegant Parisian block of flats, refutes the stereotyped image of her social position through her immense culture and her refined usage of French.The study gives a detailed survey of lexical and grammatical forms of usage in various registers of speech in Renée's linguistic performance and concludes with the idea that Renée's deep respect for the norms of standard French is a special form of revolt against the intellectual negligence characterizing the rich inhabitants always ready to show their superiority over her.
In this article, we explore the feasibility of extracting suitable and unsuitable food items for particular health conditions from natural language text. We refer to this task as conditional healthiness classification. For that purpose, we annotate a corpus extracted from forum entries of a food-related website. We identify different relation types that hold between food items and health conditions going beyond a binary distinction of suitability and unsuitability and devise various supervised classifiers using different types of features. We examine the impact of different task-specific resources, such as a healthiness lexicon that lists the healthiness status of a food item and a sentiment lexicon. Moreover, we also consider task-specific linguistic features that disambiguate a context in which mentions of a food item and a health condition co-occur and compare them with standard features using bag of words, part-of-speech information and syntactic parses. We also investigate in how far individual food items and health conditions correlate with specific relation types and try to harness this information for classification.
The aim of the study is to identify the relationship between notions of derivational deductibility and lexical usage of derived words in the linguistic consciousness and their actual functioning in speech. Suffixed nouns with the meaning “person” formed by productive word-formation models of quality adjectives served as language material. The study was performed on the basis of the derivatives obtained by the linguistic experiment with Russian native speakers. We also controlled the frequency of use of these words in the speech, which was obtained with the help of Google search engine. As a result of the research we have identified several patterns. Firstly, the data about the functioning of such words obtained in different experimental conditions correlated. Secondly, the nouns with the meaning of “person” significantly exceed the lexical norm for the diversity of their use in speech (in quantitative and qualitative indicators), which was reflected in codified dictionaries.
ABSTRACT The research is entitled Discourse Analysis of Martin Luther King Jr.’s speech, “ I Have a Dream ” ( Addressed to the March on Washington ). It is an attempt to analyzeand explain the discourse analysis norms. The objective of this research is to identify, classify, and analyze the discourse analysis norms in Martin Luther King Jr’s Speech. The writer conducts this research by using descriptive method. In collecting data, the speech and its history were taken from the book American speech and other relevant source from internet as a source data. The data analysis was based on by Juez (2008:20) and supportedby De Beaugrande and Dressler (1986:8) and Aarts and Aarts (1982: 4).These discourse analysis norms were applied in the Martin Luther King Jr’s Speech. The theory consists of seven norms, they are: cohesion (pronoun, substitusion, ellipsis, conjuction, and lexical), coherence (mark coherence and unmark coherence), intentionality, acceptability, informatifity, situationality, and intertextuality. The result of this research shows that in the cohesion norm, there are 131 pronouns, 52 subtitusion, 1 ellipsis, 90 conjuction and 16 lexical. In coherence norm, there are 46 mark coherence and there is no unmark coherence. There are also others norms like intentionality, acceptability, informatifity, situationality, and intertextuality. This speech contains the seven norms sugested by Juez (2008:20). Keywords: Discourse Analysis, Speech, Martin L.King Jr., Norms.
Abstract This chapter concludes that Negritude allowed black poets to be full participants of the “aesthetic regime.” This “aesthetic regime” is a lyric regime, and the poetry of Negritude establishes itself solidly as a text-based (rather than oral) movement. Negritude poets not only adopted typographic innovations introduced by other writers; they also developed their own way of harnessing the resistant force that the printed word harbors in its material being. The poets of Negritude in this sense raced textuality. They drew on the complex specificities of the irracialization under modern capitalism to exert pressure on thematic, lexical prosodic, typographical, and rhetorical norms. Moreover, the Negritude poem offers the promise of an identity that can be performed but will never resolve into essence, the promise of an identity that acts like a resistant force of “materiality as it plays itself out in/as the work of art.”
There is an obvious interest in capturing general trends in the structure of phonological systems of the world's extant languages.These may hint at overall design properties of human language, which in turn may have origins in basic human cognitive properties or characteristics inherited from the earliest human language(s).One tool that can be used to study such trends is a broadly-based crosslinguistic database on phonological systems.Four of the principal challenges to providing this will be discussed in this paper, which describes the thinking behind the compilation of the LAPSyD database and draws some comparisons with other somewhat similar projects, such as PHOIBLE, SAPhon and Segrer's African consonant inventory database.
This paper presents an overview of the linguistic analyses developed in the MULTILIT project and the processing of the oral and written texts collected. The project investigates the language abilities of multilingual children and adolescents, in particular, those who have Turkish and/or Kurdish as a mother tongue. A further aim of the project is to examine from a psycholinguistic and sociolinguistic perspective the extent to which competence in academic registers is achieved on the basis of the languages spoken by the children, including the language(s) spoken at the home, the language of the country of residence and the first foreign language. To be able to examine these questions using corpus linguistic parameters, we created categories of analysis in MULTILIT. The data collection comprises texts from bilingual and monolingual children and adolescents in Germany in their first language Turkish, their second language German und their foreign language English. Pupils aged between nine and twenty years of age produced monologue oral and written texts in the two genres of narrative and discursive. On the basis of these samples, we examine linguistic features such as lexical expression (lexical density, lexical diversity), syntactic complexity (syntactic and discursive packaging) as well as phonology in the oral texts and orthography in the written texts, with the aim of investigating the pupils’ growing mastery of these features in academic and informal registers. To this end the raw data have been transcribed by the use of transcription conventions developed especially for the needs of the MULTILIT data. They are based on the commonly used HIAT and GAT transcription conventions and supplemented with conventions that provide additional information such as features at the graphic level. The categories of analysis comprise a large number of linguistic categories such as word classes, syntax, noun phrase complexity, complex verbal morphology, direct speech and text structures. We also annotate errors and norm deviations at a wide range of levels (orthographic, morphological, lexical, syntactic and textual). In view of the different language systems, these criteria are considered separately for all languages investigated in the project.
The article deals with the factors defining the creative potential of child speech. Special significance is assigned to heuristic mechanisms responsible for the system’s sharp “nose”, typical of a child at the stage of self-learning a language (in the period of pre-school ontogenesis). The article substantiates the idea about a close relationship between the compensatory function (compensating for the lexical deficit) and the conventional game function of child speech innovations. The experimental nature of child speech discloses the spontaneous quick wit of the child by potential semantic filling of “ready-made” and “invented” words, including the ability “to think by means of imagery analogy”. Object standards that lie at the basis of intentional or unintentional metaphors are characterized in the light of child mentality (“personification of everything”, dominants of personal meaning, etc.). In the situation of “ignoring” the norm, intuition, as a heuristic vector of child linguistic mentality, “prompts” them a non-standard solution. The article describes the strategy of word manipulation in the child’s communication with grown-ups as one of the early forms of manifestation of intention to language games. Literal interpretation of phraseological units is analyzed as the language game resource. The article presents a fragment of the vocabulary of “aphoristic literalisms” of child speech. Transition of child heuristics into a “conscious” state is considered to be the basic principle of organization of training verbal creativity.
Linguistic resources are very important to any natural language processing task. Unfortunately, the manual construction of these resources is laborious and time-consuming. The use of annotated corpora as a knowledge database might be a solution to a fast construction of a grammar for a given language. In this paper, we present our method to automatically induce a syntactic grammar from an Arabic annotated corpus (The Penn Arabic TreeBank), a probabilistic context free grammar in our case. To construct our resource, we first induce context free rules from the annotated corpus trees as a first step and then we calculate a specific probability for each induced rule. Finally, we present and discuss the obtained grammar.
Recent evidence suggests that grammatical aspect can bias how individuals perceive criminal intentionality during discourse comprehension. Given that criminal intentionality is a common criterion for legal definitions (e.g., first-degree murder), the present study explored whether grammatical aspect may also impact legal judgments. In a series of four experiments participants were provided with a legal definition and a description of a crime in which the grammatical aspect of provocation and murder events were manipulated. Participants were asked to make a decision (first- vs. second-degree murder) and then indicate factors that impacted their decision. Findings suggest that legal judgments can be affected by grammatical aspect but the most robust effects were limited to temporal dynamics (i.e., imperfective aspect results in more murder actions than perfective aspect), which may in turn influence other representational systems (i.e., number of murder actions positively predicts perce)
The names of people, locations, and organisations play a central role in language, and named entity recognition (NER) has been widely studied, and successfully incorporated, into natural language processing (NLP) applications. The most common variant of NER involves identifying and classifying proper noun mentions of these and miscellaneous entities as linear spans in text. Unfortunately, this version of NER is no closer to a detailed treatment of named entities than chunking is to a full syntactic analysis. NER, so construed, reflects neither the syntactic nor semantic structure of NE mentions, and provides insufficient categorical distinctions to represent that structure. Representing this nested structure, where a mention may contain mention(s) of other entities, is critical for applications such as coreference resolution. The lack of this structure creates spurious ambiguity in the linear approximation. Research in NER has been shaped by the size and detail of the available annotated corpora. The existing structured named entity corpora are either small, in specialist domains, or in languages other than English. This thesis presents our Nested Named Entity (NNE) corpus of named entities and numerical and temporal expressions, taken from the WSJ portion of the Penn Treebank (PTB, Marcus et al., 1993). We use the BBN Pronoun Coreference and Entity Type Corpus (Weischedel and Brunstein, 2005a) as our basis, manually annotating it with a principled, fine-grained, nested annotation scheme and detailed annotation guidelines. The corpus comprises over 279,000 entities over 49,211 sentences (1,173,000 words), including 118,495 top-level entities. Our annotations were designed using twelve high-level principles that guided the development of the annotation scheme and difficult decisions for annotators. We also monitored the semantic grammar that was being induced during annotation, seeking to identify and reinforce common patterns to maintain consistent, parsimonious annotations. The result is a scheme of 118 hierarchical fine-grained entity types and nesting rules, covering all capitalised mentions of entities, and numerical and temporal expressions. Unlike many corpora, we have developed detailed guidelines, including extensive discussion of the edge cases, in an ongoing dialogue with our annotators which is critical for consistency and reproducibility. We annotated independently from the PTB bracketing, allowing annotators to choose spans which were inconsistent with the PTB conventions and errors, and only refer back to it to resolve genuine ambiguity consistently. We merged our NNE with the PTB, requiring some systematic and one-off changes to both annotations. This allows the NNE corpus to complement other PTB resources, such as PropBank, and inform PTB-derived corpora for other formalisms, such as CCG and HPSG. We compare this corpus against BBN. We consider several approaches to integrating the PTB and NNE annotations, which affect the sparsity of grammar rules and visibility of syntactic and NE structure. We explore their impact on parsing the NNE and merged variants using the Berkeley parser (Petrov et al., 2006), which performs surprisingly well without specialised NER features. We experiment with flattening the NNE annotations into linear NER variants with stacked categories, and explore the ability of a maximum entropy and a CRF NER system to reproduce them. The CRF performs substantially better, but is infeasible to train on the enormous stacked category sets. The flattened output of the Berkeley parser are almost competitive with the CRF. Our results demonstrate that the NNE corpus is feasible for statistical models to reproduce. We invite researchers to explore new, richer models of (joint) parsing and NER on this complex and challenging task. Our nested named entity corpus will improve a wide range of NLP tasks, such as coreference resolution and question answering, allowing automated systems to understand and exploit the true structure of named entities.
Question classification module of a Question Answering System plays a very important role in identifying and providing results according to the user expectations. There are different types of methods involved during classification that can be applied to all kinds of domain like machine learning or using lexical database with its own advantages and disadvantages. Identifying the relevant approach for question classification for a specific domain is one of the foremost tasks. A study on different levels of questions including Blooms taxonomy and Costa taxonomy made our work to focus more on different categories of questions. To overcome these issues, we employ a question classifier using Register Linear (RL) models for a specific domain. The Register Linear (RL) Classification Model classifies the complex questions in a linear manner where each input is assigned to only one class. The RL classification model identifies the role of semantic provided in the input space which is divided into decision regions with the decision surfaces to be of linear functions of input x (sentence) for different set of classes. Initially, the Register Linear model identifies the role of semantics in a sentence and with these roles being identified, statistical relations between the concepts in the sentence are derived that produces a probability distribution over different set of classes. With these classifications, the exact answer type is identified that helps to find the answer. Our model gives better results in terms of execution time (time taken to categorize the queries), classification accuracy and result analyzing efficiency.
Units of the Russian word family with the root -gordoriginate in the Proto-Slavic etymological word family * gbrdb, which has reflexes gord/ gardin East Slavic languages, grd-, grd-, hrdin South Slavic and hrd-, hord-, hard-, gardin West Slavic. These units express the concept pride almost in all Slavic languages (with the exception of Slovenian). In Russian, Ukrainian, Czech, Sorbian, Bulgarian and Macedonian units of etymological word family * gbrdb are the main expression means of the concept pride. In the Slovak and Belarusian languages there are also units with other roots (Belarusian gonar-, Slovakianpych-) in the nucleus of the lexical semantic field pride. In the western group of South Slavic languages (Serbian, Slovenian) and Polish the basic units expressing the concept pride are lexemes of a word family with root ponos-. The semantic field of Proto-Slavic etymological word family * gbrdb includes four semantic centers: 1) pride and related concepts (obstinacy, insolence, scorn, impudence, abuse); 2) social status, which includes words with meanings 'grandeur', 'majestic', 'nobility', 'loftiness', 'hero', 'glory', 'fame', 'importance', the Russian dialect ritual wedding lexis and Old-Russian units with meanings 'austerity', which contain negative connotations. The units of the following two semantic centers are registered in all Slavic language groups: 3) common and aesthetic evaluation: a) common negative and negative aesthetic evaluation: units expressing common and aesthetic negative evaluation with meanings 'ugly', 'hideous', 'bad' are registered in South Slavic, Old Slovak, Old Russian and Upper Sorbian languages; b) positive aesthetic evaluation; such units are recorded in West Slavic and East Slavic languages; 4) exceeding norms in size and strength; the units of the etymological word family * gbrdb expressing meanings 'big', 'heavy' exist in Serbian and Upper Sorbian languages and Russian dialects. So, first and second semantic centers are common for all Slavic languages, while third and fourth are local. The units of the third center are registered in South and West Slavic languages. The fourth center is characteristic for South Slavic languages mostly. A.A. Kretov and L. Kralik suggested two well-founded etymology versions of Russian gord-. Both explain its semantics 'pride/proud/to be proud' as an extension of the meaning 'having high social status'. The main divergence of the two versions applies to the primary meaning of Proto-Slavic *gbrdb. A.A.Kretov supposes that Proto-Slavic *gbrdb primary meaning is 'dimensional height', while L. Kralik considers Proto Slavic *gbrdb as derived from Indo-European *gwher 'hot' and believes that its primary meaning is 'furious'. Detailed Slavic lexis analysis and the structure of the etymological word family semantic field corroborates Kretov's version and shows that the meaning 'pride/proud/to be proud' of Proto-Slavic *gbrdb is related genetically to meanings 'high social level' and 'dimensional height'.
Increasingly audacious steps in advertising are made to affect the customer and to encourage them to buy the advertised goods. Advertising is highly important in gaining a foothold in the business environment. Usually, the advertising texts fail to meet the norms of the standard Lithuanian language. The aim of this article is to compare the language of the advertising booklets of two pharmacies.The linguistic analysis of the advertising booklets of Camelia and Euro Pharmacy for March 2014 showed that in terms of language errors the booklets of the two pharmacies were similar, and the character of the errors was identical in both cases. The advertising booklets of both pharmacies contained lexical, syntactic, morphological, and logical errors. The advertising booklet of the Camelia pharmacy presents 121 items, which advertising descriptions contain 55.3% of language errors. The advertising booklet of the EuroPharmacy presents advertising descriptions of 103 items, where language errors comprise 57.2%. The majority of the errors detected in the advertising booklets of the two pharmacies are lexical (Camelia – 33.8%, and Euro Pharmacy – 37.3%) or syntactic (Camelia – 27.9%, and Euro Pharmacy – 37.3%). Both publications contain nearly equal numbers of lexical errors (Camelia – 17.6%, and Euro Pharmacy – 18.7%). The greatest difference was observed in the number of morphological errors (Camelia – 20.7%, and Euro Pharmacy – 5.7%).In addition to that, the name of the Camelia pharmacy is in conflict with the norms of both Lithuanian and Latin languages.
Metaphorical expressions very often involve words referring to physical entities and experiences. Yet, figures of speech such as metaphors are not intended to be understood literally, word-by-word. We used event-related brain potentials (ERPs) to determine whether metaphorical expressions are processed more like physical or more like abstract expressions. To this end, novel adjective-noun word pairs were presented visually in three conditions: (1) Physical, easy to experience with the senses (e.g., "printed schedule"); (2) Abstract, difficult to experience with the senses (e.g., "conditional schedule"); and (3) novel Metaphorical, expressions with a physical adjective, but a figurative meaning (e.g., "thin schedule"). We replicated the N400 lexical concreteness effect for concrete vs. abstract adjectives. In order to increase the sensitivity of the concreteness manipulation on the expressions, we divided each condition into high and low groups according to rated concreteness. Mirroring the adjective result, we observed a N400 concreteness effect at the noun for physical expressions with high concreteness ratings vs. abstract expressions with low concreteness ratings, even though the nouns per se did not differ in lexical concreteness. Paradoxically, the N400 to nouns in the metaphorical expressions was indistinguishable from that to nouns in the literal abstract expressions, but only for the more concrete subgroup of metaphors; the N400 to the less concrete subgroup of metaphors patterned with that to nouns in the literal concrete expressions. In sum, we not only find evidence for conceptual concreteness separable from lexical concreteness but also that the processing of metaphorical expressions is not driven strictly by either lexical or conceptual concreteness.
This reserach presents SentIta, a Sentiment lexicon for the Italian language, and Doxa, a prototype that, interacting with the lexical database, applies a set of linguistic rules for the Documentlevel Opinionated teXt Analysis. Details about the dictionary population, the semantic analysis of texts written in natural language and the evaluation of the tools will be provided in the paper.
The paper is about the modern nicknames of the residents of Perm Land of the 20th-early 21st centuries. Russian as well as Komi-Permyak and Tatar anthroponomy which is motivated by the lexemes of thematic group “Animals” is analyzed. The lexicalsemantic groups which are topical for Perm nicknames are determined (“Names of the animals and their kinds”, “Names of the regions of the animals”, “Names of the young animals”, “Nicknames of the animals”, “Names and nicknames of the animals which are the personages of the folklore / author’s works”, “Names of the groups of the animals”, “Onomatopoeia (imitations of the sounds producing by the animals”, “Words to beckon the animals”). Types and kinds of the nicknames are identified: 1) from the point of view of spontaneity / consciousness of appearance: natural and artificial nominations which are proposed to distinguish from the standpoint of language (in the linguistic opposition “norm” / “usage”) and from the standpoint of onomastics (in the extra-linguistic opposition “own” / “alien”); 2) by the subject of the nomination: “auto-nicknames” and “allo-nicknames”; 3) by the Extension: individual (family; social-group, at-school, youth, teacher’s, nicknames of adults) and collective (family, family-tribal, “family-group”; socialgroup; territorial). The factors which determine language tools of derivation of the nicknames of this group are revealed: language which is a source of the nomination, motivation of nomination (structural, phonetic, lexical, semantic, multiple), evaluation, social conditions, territorial conditions, temporal conditions. Structurally, the central (not formally distorted), transitional (derivative) and peripheral (occasional, barbarism) lexemes of this group are identified. Multiplicity of characteristics and properties of the animals which motivate the nicknames is noted. The perspectives of research of the modern Perm nicknames in lingualcultural and cognitive aspects are defined.
The article deals with psychological peculiarities of a language person forming in the process of foreign texts translation and interpretation. Foreign language competence comprises the level of the person’s professionalism and skills within the scope of the competence. It is stated that the processes of foreign texts translation and interpretation are characterized by the translator’s subjective vision of the language norms, as well as his intuition and lexical sensitivity, active participation in creative communication with the author. It is established that during a foreign text translation language conscience and the language personality of the reader are formed and his own language field and his picture of the world are developed.
Chinese Lexicon Project (Sze et al., 2014) summarized lexical decision response data of 2,500 Chinese characters. The original analysis has showed that the newest character frequency norm accounts the most variance of reaction times. The variance of these response data are analyzed in terms of the character frequency, strokes, and structures in use of the norms from Taiwan. First of all, simplified characters ranked as high frequency have greater performance on reaction times, but these characters ranked as lower frequency in traditional scripts have overestimated response points. Secondly many simplified characters are transformed from complex to simple, and strokes are substantial discrepancy between Chinese scripts. Finally, stuructre of character takes a large proportion of variance in the response data. The covariance of structure and character frequency also shows a significant trend. Our current work reveals some critical thinkings on using mega-data as the approach to study Chinese character processing.
ABSTRACT This skripsiwas made as a requirement to obtain bachelor Degree in English in Sam Ratulangi University. This research is entitled “Discourse analysis of King George VI Speech”With God’s Help, We shall Prevail”(First Radio Address, Britania. September 3, 1939)”. It is an attempt to analyze and explain the discourse analysis norms in King George VI speech. There are three steps to finish this research. First step is preparation,the writer reads some books about language, linguistics, and discourse analysis to find out the relevant theories.Second step is data collection, the writer finds King George VI speech and reads it for several times to have a deep understanding. Third step is data analysis,The data arecollected, identified, classified and analyzed. The method used in this research is taken from Alba-Juez (2008:20) and supported the theory by De Beaugrande and Dressler (1986:8) and Aarts and Aarts (1982: 4). The theory consists of seven norms, they are: Cohesion: pronoun, substitution, ellipsis, conjunction, lexical, Coherence: mark coherence, and unmark coherence, Intentionality, Acceptability, Informatifity, Situationality, and Intertextuality. The result of this research shows that in cohesion there are 50 pronouns, 7 subtitusion, no elipsis, 34 conjuction, 10 lexical. In this speech, there are 21 mark coherence and there is no unmark coherence. There are also the norms like intentionality is focused on user or producer by expressing a disappointment and sadness. Acceptabilityhas a generally acceptable meaning, according to the history of the text of this speech is received, the king get a warm welcome from the people and members of the royal. Informatifity is can provide full information,can be known through historical conditions occurred. Situationalityhas a relationship with the surrounding circumstances, refers to the situation of war. And intertextualityrefer to the agreement as a protector of the independence of Poland. Keywords: discourse analysis, speech, seven norms, King George VI
Alzheimer's disease (AD) is a neurodegenerative disorder characterized by progressive memory impairment and the presence of amyloid plaques and neurofibrillary tangles. The associated neuropathology originates in brain areas responsible for olfaction, which makes olfactory tasks potentially useful for assessing AD. The strongest genetic risk factor for AD is the apolipoprotein E (ApoE) ɛ4 allele that has been associated with increased cognitive and olfactory deficits. While individuals carrying one ɛ4 allele of the ApoE gene are at increased risk for AD relative to non-carriers, those with two copies of the ɛ4 allele demonstrate an even higher risk for developing AD. Furthermore, homozygous ApoE ɛ4/4 individuals diagnosed with AD are known to have heightened amyloid burden and a more rapid rate of cognitive decline relative to heterozygous ɛ3/4 ApoE carriers. All of these factors suggest there are differences in severity and progression of AD as a function of possessing one versus two ɛ4 alleles. The current study investigated olfactory functioning in homozygous ɛ4/4 older adults diagnosed with probable AD. Compared to demographically matched ɛ3/3 and ɛ3/4 individuals, ɛ4/4 individuals showed deficits in odor identification and remote odor memory as measured by odor familiarity ratings. The current findings suggest that these particular domains of olfactory functioning may be more impaired in AD ɛ4/4 homozygotes compared to ɛ3/4 heterozygotes and ɛ3/3 homozygotes. These deficits give insight into how the presence of two ɛ4 alleles may differentially affect the progression of AD and suggest the usefulness of odor tasks in detecting those at risk for AD.
This paper proposes a novel approach to sentiment analysis that leverages work in sociology on symbolic interactionism. The proposed approach uses Affect Control Theory (ACT) to analyze readers' sentiment towards factual (objective) content and towards its entities (subject and object). ACT is a theory of affective reasoning that uses empirically derived equations to predict the sentiments and emotions that arise from events. This theory relies on several large lexicons of words with affective ratings in a three-dimensional space of evaluation, potency, and activity (EPA). The equations and lexicons of ACT were evaluated on a newly collected news-headlines corpus. ACT lexicon was expanded using a label propagation algorithm, resulting in 86,604 new words. The predicted emotions for each news headline was then computed using the augmented lexicon and ACT equations. The results had a precision of 82%, 79%, and 68% towards the event, the subject, and object, respectively. These results are significantly higher than those of standard sentiment analysis techniques.
The present paper investigates linguistic norm-adherence in Belgian Dutch written and audiovisual translation. More particularly, it is measured to what extent language use in subtitles, in comparison to regular written translations and non-translations, conforms to explicit linguistic norms. Additionally, we analyze which effect different contextual parameters have on the extent of norm-adherence in Belgian Dutch subtitles. We use the Dutch Parallel Corpus and the SoNaR Corpus, and we analyze the data by means of profile-based correspondence analysis, yielding a visualization of norm-adherence distances between the different translation modes and non-translations. The results reveal that the parameters speaker type and source language significantly affect the degree of linguistic normadherence, whereas program genre has no influence. It is also shown that norm-adherence in subtitles holds a middle position between written translations and non-translations, which is explained in terms of target audience and communicative risk.
The purpose of this study was to reveal the effects of Westernized arrangements of traditional Korean folk music on music familiarity and preference. Two separate labs in one intact class were assigned to one of two treatment groups of either listening to traditional Korean folk songs ( n = 18) or listening to Western arrangements of the same Korean folk songs ( n = 22); a second intact class served as a control group with no listening ( n = 20). Before and after the listening treatment session, pre- and posttests were administered that included 12 music excerpts of current popular, Western classical, and traditional Korean music. Results showed that participants who listened to traditional folk songs demonstrated significant increases in both familiarity and preference ratings; however, those who listened to Westernized folk songs showed increases only in familiarity ratings but not preference ratings for the same Korean songs in traditional versions. An analysis of participants’ open-ended responses showed that affective–positive responses were used most frequently when explaining preference for traditional versions of Korean folk songs (28.1%) among the traditional Korean listening group; structural–negative reasons (47.8%) were the most frequent among the Westernized listening group.
The aim of this study is to give a modest presentation of sociolinguistic’state of the female speech in the South Parts of Albania. The difference of gender, as an important factor between people in global studies especially in the last 2 centuries, has risen a great interes. Although, the studies in Albanian language in this direction are fewer. For this study are exploited theoric materials in foreign languages and Albanian. The theoric datas are associated with concret examples taken from the Albanians speakers of this region. Through the extending of many social-linguistic’s issues that are related to the speaker with the communication’s situation, with social classes, with the semantic-lexical and morfo-syntactic selection have tried to describe the social-linguistic’s difference that happen during the organization of generally life (exp. the birth of a child, the wedding, the death, ect.) and the special elements of daily life (exp. The conversations in families, the games, the joking words, ect. Female sex adhere to certain norms, which in the course of cultural history, has become a discipline rather well protected by stringent sanctions. The repertoire of conduct norms in a feminine discourse, according to scales that exist in Albanian society which express respect, friendly or formal addressing enables the interlocutors. Apart from the theoretical research of the twentieth century, especially foreign and local people, this work is based on the elements of (South regions) and Albanian folk lyrics. DOI: 10.5901/mjss.2015.v6n6s2p168
The aim of this study was to analyse the psycholinguistic variables of the attributes and concepts involved in the recall of a concept. One hundred and twenty adults (18�40 years old) participated. A lexical recall task was administered by presenting a successive list of defining attributes. Forty concepts from different semantic categories were used. The attributes were obtained empirically from local Semantic Features Production Norms. The influence of the characteristics and attributes of the concepts on the number of participants who accessed the name of the concept and the correct guess trend was analysed. Significant values for Age of Acquisition, Presence of Distinctive Attributes and Presence of Taxonomic Attributes were observed. Results show that concepts which are acquired earliest are more easily recalled; presenting taxonomic categories narrows the search and the presence of distinctive attributes allow differentiating between such concepts within a category.
This work introduces SYMPAThy, a data representation model in which the combinatorial properties of a lexical item are described by merging surface and deeper linguistic information. The proposed approach is then evaluated by comparing, for a sample list of verbal idioms, a set of SYMPAThy-based fixedness indexes against the relevant speaker-elicited indexes available in the descriptive norms collected by Tabossi et al. (2011)
This paper predicts the mutual intelligibility of 15 Chinese dialects from multiple objective distance measures. Empirical mutual intelligibility measures were obtained from functional intelligibility tests at the sentence level from 15 listeners for each of 15 Chinese dialects. We computed various proximity measures on the basis of shared phonemes and tones in the sound inventories of the 15 dialects. Next, Levenshtein (string-edit) distance measures were computed on the 764 common syllabic units (zi in Pinyin, i.e., a meaningful character or morpheme with a complete transcription of segments and tone) shared by the same 15 Chinese dialects in the Dialect Sound Database of Modern Chinese (compiled by the Chinese Academy of Social Sciences). Unweighted and perceptually weighted Levenshtein distance measures were computed. We also included objective similarity measures of phonological correspondence, based on the Zihui character list and of lexical affinity, based on the Cihui cross-dialect lexical database with all cognate and non-cognate expressions of 905 core concepts) that have been published by Cheng (1997). The best single predictor of mutual intelligibility between a pair of dialects was the percentage of cognates shared between them (r² = .548). Including all predictors afforded a highly accurate prediction of mutual intelligibility (R² = .874). A very reasonable prediction is afforded if we just add the lexical frequency of finals (syllable rhymes) shared by a pair of dialects (R² = .611). (PsycINFO Database Record (c) 2016 APA, all rights reserved)
Differences in how writing systems represent language raise important questions about whether there could be a universal functional architecture for reading across languages. In order to study potential language differences in the neural networks that support reading skill, we collected fMRI data from readers of alphabetic (English) and morpho-syllabic (Chinese) writing systems during two reading tasks. In one, participants read short stories under conditions that approximate natural reading, and in the other, participants decided whether individual stimuli were real words or not. Prior work comparing these two writing systems has overwhelmingly used meta-linguistic tasks, generally supporting the conclusion that the reading system is organized differently for skilled readers of Chinese and English. We observed that language differences in the reading network were greatly dependent on task. In lexical decision, a pattern consistent with prior research was observed in which the Middle )
Phylogenetic models, originally developed to demonstrate evolutionary biology, have been applied to a wide range of cultural data including natural language lexicons, manuscripts, folktales, material cultures, and religions. A fundamental question regarding the application of phylogenetic inference is whether trees are an appropriate approximation of cultural evolutionary history. Their validity in cultural applications has been scrutinized, particularly with respect to the lexicons of dialects in contact. Phylogenetic models organize evolutionary data into a series of branching events through time. However, branching events are typically not included in dialectological studies to interpret the distributions of lexical terms. Instead, dialectologists have offered spatial interpretations to represent lexical data. For example, new lexical items that emerge in a politico-cultural center are likely to spread to peripheries, but not vice versa. To explore the question of the tree model’s )
We suggest an information-theoretic approach for measuring stylistic coordination in dialogues. The proposed measure has a simple predictive interpretation and can account for various confounding factors through proper conditioning. We revisit some of the previous studies that reported strong signatures of stylistic accommodation, and find that a significant part of the observed coordination can be attributed to a simple confounding effect—length coordination. Specifically, longer utterances tend to be followed by longer responses, which gives rise to spurious correlations in the other stylistic features. We propose a test to distinguish correlations in length due to contextual factors (topic of conversation, user verbosity, etc.) and turn-by-turn coordination. We also suggest a test to identify whether stylistic coordination persists even after accounting for length coordination and contextual factors. [ABSTRACT FROM AUTHOR], Copyright of PLoS ONE is the property of Public Library of)
У статті розглядається проблема визначення поняття «семи» як компонента значення слів, подано їхню типологію в сучасній лінгвістичній літературі. Досліджено особливості семного складу лексичних одиниць на позначення добра в англійській мові та з’ясовано його характер. (The article deals with the problem of definition of «seme» as a word meanings’ component and their typology in modern linguistic literature. The peculiarities of the seme stock of lexical units denoting good in English as well as its character are analyzed together with the qualitative and quantitative characteristics of semes. The study shows that the seme stock of the nouns denoting good in Modern English can be divided into 6 subsets, reflecting diverse shades of their meanings. The order and organization of semes in the seme stock under study is hierarchical. The obtained matrix gives an opportunity to investigate the seme stock as the unity of semantic features that possesses a definite structure. In the course of our analysis the semes which make up the meanings of the lexical units have been divided into polyfunctional and monofunctional according to frequency of their appearance in the lexical meanings of the nouns denoting good. Lexical units denoting good occupy an important place within the lexical system of any language because this notion belongs to the moral values, indicating moral norms, assessment, ideal, valuable orientation and moral qualities of the personality. Present understanding of good is inseparably connected with processes taking place in the social and political spheres of life, and leading to changes in the consciousness and mental perception of an individual. Thus, good is considered to be anthropocentric and socio-pragmatic notion referring to a person and serving to satisfy his/her social and everyday needs.)
Розглядаються питання, які стосуються перекладу газетних текстів з французької мови на російську, визначається статус мови преси, аналізуються різні класифікація відомих лінгвістів щодо типів текстів публіцистичної спрямованості. Вивчаються прийоми і способи перекладу газетно-журнальної публіцистики, проводиться порів- няльна характеристика мови французької і російської преси. Визначаються основні особливості стилю періодики: її інформативність, експресивно-емоційний характер, використання розмовної, ненормативної лексики, скорочених слів, алюзії, а також залучення реальних політичних і громадських подій. Детально розглядається питання вживання фразеологізмів, прислів’їв, приказок, інших стилістичних прийомів. Визначається поняття «фразеологізм» і досліджуються його види, а також вивчаються окрім фразеологізмів, які вказані в словниках, окказіональні фразеологізми. (The article considers the problems of newspaper texts translation from French into Russian that are characterized by definite lexical and grammatical structures, typical only for press language expressions and formulas understandable for native speakers, but are not always explainable from the viewpoint of language norms. The French press language status is determined, which is part of the journalistic style, the analysis of the approaches to the language study is carried on. Different classifications of journalistic text types offered by well-known linguists are analyzed. On the basis of the text classification in L.S. Barhudarov edition the approach in solution of tasks put in the article is determined: to define the peculiarities of journalistic style. The methods and ways of translation of newspaper and magazine texts are examined; the comparative analysis of French and Russian press is conducted. The basic features of the style of the periodic press are determined: its informativeness, expressively emotional character, the use of colloquial and substandard vocabulary, abbreviated words, and allusion. The typical feature of press texts is the coverage of real political and meaningful for public events. The problem of the use of phraseological units, proverbs, sayings and other stylistic devices is examined in detail. The notion «phraseological unit» is determined here and its types are investigated profoundly. Besides phraseological units indicated in dictionaries the author analyses nonce phraseological units.)