Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
This paper presents results of dependency parsing of Old French, a language which is poorly standardized at the lexical level, and which displays a relatively free word order. The work is carried out on five distinct sample texts extracted from the dependency treebank Syntactic Reference Corpus of Medieval French (SRCMF). Following Achim Stein's previous work, we have trained the Mate parser on each sub-corpus and cross-validated the results. We show that the parsing efficiency is diminished by the greater lexical variation of Old French compared to parse results on modern French. In order to improve the result of the POS tagging step in the parsing process, we applied a pre-treatment to the data, comparing two distinct strategies: one using a slightly post-treated version of the TreeTagger trained on Old French by Stein, and a CRF trained on the texts, enriched with external resources. The CRF version outperforms every other approach.
Recurrent neural network language models have solved the problems of data sparseness and dimensionality disaster which exist in traditional N-gram models. RNNLMs have recently demonstrated state-of-the-art performance in speech recognition, machine translation and other tasks. In this paper, we improve the model performance by providing contextual word vectors in association with RNNLMs. This method can reinforce the ability of learning long-distance information using vectors training from Skip-gram model. The experimental results show that the proposed method can improve the perplexity performance significantly on Penn Treebank data. And we further apply the models to speech recognition task on the Wall Street Journal corpora, where we achieve obvious improvements in word-error-rate.
This paper presents a set of Bilingual Dictionary Drafting (BDD) methods including manual extraction from existing lexical databases and corpus based NLP tools, as well as their evaluation on the example of German-Basque as language pair. Our aim is twofold: to give support to a German-Basque bilingual dictionary project by providing draft Bilingual Glossaries and to provide lexicographers with insight into how useful BDD methods are. Results show that the analysed methods can greatly assist on bilingual dictionary writing, in the context of medium-density language pairs.
The lexicon dynamics deals with the changes that occur in language from a historical stage to another. As part of a living organism, words emerge and fade away, bearing the mark of the linguistic norms. The linguists’ preoccupation with the accuracy of language goes back in time and is related to lexicon, semantics, grammar, spelling etc. The present study is designed as a concise presentation of certain mistakes frequently met in the current written press, with a view to correcting them. Some deviations from the norms are minor, others major, whereas their causes are numerous. Recent loans, neologisms, confusion of styles and excessive use of clichés are only a few aspects to be approached in the present work.
Recent work has sparked new interest in type-supervised part-of-speech tagging, a data setting in which no labeled sen-tences are available, but the set of allowed tags is known for each word type. This paper describes observational initializa-tion, a novel technique for initializing EM when training a type-supervised HMM tagger. Our initializer allocates probabil-ity mass to unambiguous transitions in an unlabeled corpus, generating token-level observations from type-level supervision. Experimentally, observational initializa-tion gives state-of-the-art type-supervised tagging accuracy, providing an error re-duction of 56 % over uniform initialization on the Penn English Treebank. 1
Abstract We investigated discrimination in the context of evaluating advertisements, based on the suppression model (justification‐suppression model [ JSM ]) of prejudice expression. Previous research has demonstrated that when people are given an opportunity to give a high rating to an ad featuring a Black model, a sense of nonprejudice is created, which, in turn, provides an opportunity to discriminate subsequently without feeling prejudiced. We extended the JSM by investigating whether the acquisition of legitimacy credits (a moral authority earned by demonstrating nonprejudice) is a sufficient condition to release the expression of prejudice. We found that subjects who first evaluated a high‐quality ad featuring a Black model felt eligible to use legitimacy credits in subsequent evaluations. But in a subsequent study, participants who acquired these credits evaluated Black model ads more negatively than White model ads only when these ads were of low quality. The implications for evaluating the subtle way that prejudice affects rating of models of color in advertisements are discussed.
In this paper, the development and evaluation of the Urdu parser is presented along with the comparison of existing resources for the language variants Urdu/Hindi. This parser was given a linguistically rich grammar extracted from a treebank. This context free grammar with sufficient encoded information is comparable with the state of the art parsing requirements for morphologically rich and closely related language variants Urdu/Hindi. The extended parsing model and the linguistically rich grammar together provide us promising parsing results for both the language variants. The parser gives 87% of f-score, which outperforms the multi-path shift-reduce parser for Urdu and a simple Hindi dependency parser with 4.8% and 22% increase in recall, respectively.
Abstract Statistical parsers often require careful parameter tuning and feature selection. This is a nontrivial task for application developers who are not interested in parsing for its own sake, and it can be time-consuming even for experienced researchers. In this paper we present MaltOptimizer, a tool developed to automatically explore parameters and features for MaltParser, a transition-based dependency parsing system that can be used to train parser's given treebank data. MaltParser provides a wide range of parameters for optimization, including nine different parsing algorithms, an expressive feature specification language that can be used to define arbitrarily rich feature models, and two machine learning libraries, each with their own parameters. MaltOptimizer is an interactive system that performs parser optimization in three stages. First, it performs an analysis of the training set in order to select a suitable starting point for optimization. Second, it selects the best parsing algorithm and tunes the parameters of this algorithm. Finally, it performs feature selection and tunes machine learning parameters. Experiments on a wide range of data sets show that MaltOptimizer quickly produces models that consistently outperform default settings and often approach the accuracy achieved through careful manual optimization.
This paper gives an overview of the latest developments in computational syntactic analysis of Estonian. We present Estonian Dependency Treebank, an ongoing corpus annotation project. Although the treebank construction is still under way, we have used it for training MaltParser and experimenting with combining MaltParser with a rule-based Constraint Grammar parser for Estonian. MaltParser achieves unlabeled attachment score (UAS; correct links to head node) of 83.4% and label accuracy (LA) of 88.6%. Labeled attachment score (LAS) was 80.3%.
This work improves a novel Service Selection Method for the development of Service-Oriented Applications in the context of the Service-Oriented Computing (SOC) paradigm. We have defined a Semantic-Structural Scheme to assess Web Services on Interface Compatibility exploring the available information from WSDL documents. The structural information involves data types from return, parameters and exceptions. The semantic information concerns identifiers from parameters and operation names. The lexical database WordNet is used as a semantic basis. Two appraisal values were defined: compatibility gap and adaptability gap. The former is centered on functional aspects. The latter explains the adaptation effort to a successful integration. We validated those appraisals values through different experiments with a data-set of 465 real-life Web Services and measured the results using three metrics from the Information Retrieval field.
BACKGROUND: The standard clinical acquisition for left ventricular functional parameter analysis with cardiovascular magnetic resonance (CMR) uses a multi-breathhold multi-slice segmented balanced SSFP sequence. Performing multiple long breathholds in quick succession for ventricular coverage in the short-axis orientation can lead to fatigue and is challenging in patients with severe cardiac or respiratory disorders. This study combines the encoding efficiency of a six-fold undersampled 3D stack of spirals balanced SSFP sequence with 3D through-time spiral GRAPPA parallel imaging reconstruction. This 3D spiral method requires only one breathhold to collect the dynamic data. METHODS: Ten healthy volunteers were recruited for imaging at 3 T. The 3D spiral technique was compared against 2D imaging in terms of systolic left ventricular functional parameter values (Bland-Altman plots), total scan time (Welch's t-test) and qualitative image rating scores (Wilcoxon signed-rank test). RESULTS: Systolic left ventricular functional values were not significantly different (i.e. 3D-2D) between the methods. The 95% confidence interval for ejection fraction was -0.1 ± 1.6% (mean ± 1.96*SD). The total scan time for the 3D spiral technique was 48 s, which included one breathhold with an average duration of 14 s for the dynamic scan, plus 34 s to collect the calibration data under free-breathing conditions. The 2D method required an average of 5 min 40s for the same coverage of the left ventricle. The difference between 3D and 2D image rating scores was significantly different from zero (Wilcoxon signed-rank test, p < 0.05); however, the scores were at least 3 (i.e. average) or higher for 3D spiral imaging. CONCLUSION: The 3D through-time spiral GRAPPA method demonstrated equivalent systolic left ventricular functional parameter values, required significantly less total scan time and yielded acceptable image quality with respect to the 2D segmented multi-breathhold standard in this study. Moreover, the 3D spiral technique used just one breathhold for dynamic imaging, which is anticipated to reduce patient fatigue as part of the complete cardiac examination in future studies that include patients.
This study examines the methodology of global foreign accent ratings in studies on L2 speech production. In three experiments, we test how variation in raters, range within speech samples, as well as instructions and procedures affects ratings of accent in predominantly monolingual speakers of German, non-native speakers of German, as well as long-term emigrants from Germany, that is, L1 attriters. The findings show that rater differences do not result in systematic changes in rating patterns. In contrast, range effects and effects of familiarity with accented speech lead to shifts in absolute and relative ratings. Including more strongly foreign-accented samples leads to lower judgements for the entire group of L2 speakers compared to natives. Similarly, lower familiarity with foreign accent results in more variable and more strongly foreign-accented judgements. We discuss the implications for research on L2 pronunciation as well as for the interpretation of nativeness in L2 studies and language testing more generally.
This paper proposes a simple yet effective framework of soft cross-lingual syntax projection to transfer syntactic structures from source language to target language using monolingual treebanks and large-scale bilingual parallel text. Here, soft means that we only project reliable dependencies to compose high-quality target structures. The projected instances are then used as additional training data to improve the performance of supervised parsers. The major issues for this idea are 1) errors from the source-language parser and unsupervised word aligner; 2) intrinsic syntactic non-isomorphism between languages; 3) incomplete parse trees after projection. To handle the first two issues, we propose to use a probabilistic dependency parser trained on the target-language treebank, and prune out unlikely projected dependencies that have low marginal probabilities. To make use of the incomplete projected syntactic structures, we adopt a new learning technique based on ambiguous labelings. For a word that has no head words after projection, we enrich the projected structure with all other words as its candidate heads as long as the newly-added dependency does not cross any projected dependencies. In this way, the syntactic structure of a sentence becomes a parse forest (ambiguous labels) instead of a single parse tree. During training, the objective is to maximize the mixed likelihood of manually labeled instances and projected instances with ambiguous labelings. Experimental results on benchmark data show that our method significantly outperforms a strong baseline supervised parser and previous syntax projection methods. 1
This is the first attempt at characterizing reading difficulty in Hindi using naturally occurring sentences. We created the Potsdam-Allahabad Hindi Eyetracking Corpus by recording eye-movement data from 30 participants at the University of Allahabad, India. The target stimuli were 153 sentences selected from the beta version of the Hindi-Urdu treebank. We find that word- or low-level predictors (syllable length, unigram and bigram frequency) affect first-pass reading times, regression path duration, total reading time, and outgoing saccade length. An increase in syllable length results in longer fixations, and an increase in word unigram and bigram frequency leads to shorter fixations. Longer syllable length and higher frequency lead to longer outgoing saccades. We also find that two predictors of sentence comprehension difficulty, integration and storage cost, have an effect on reading difficulty. Integration cost (Gibson, 2000) was approximated by calculating the distance (in words) between a dependent and head; and storage cost (Gibson, 2000), which measures difficulty of maintaining predictions, was estimated by counting the number of predicted heads at each point in the sentence. We find that integration cost mainly affects outgoing saccade length, and storage cost affects total reading times and outgoing saccade length. Thus, word-level predictors have an effect in both early and late measures of reading time, while predictors of sentence comprehension difficulty tend to affect later measures. This is, to our knowledge, the first demonstration using eye-tracking that both integration and storage cost influence reading difficulty.
Methylphenidate mainly enhances dopamine neurotransmission whereas 3,4-methylenedioxymethamphetamine (MDMA, "ecstasy") mainly enhances serotonin neurotransmission. However, both drugs also induce a weaker increase of cerebral noradrenaline exerting sympathomimetic properties. Dopaminergic psychostimulants are reported to increase sexual drive, while serotonergic drugs typically impair sexual arousal and functions. Additionally, serotonin has also been shown to modulate cognitive perception of romantic relationships. Whether methylphenidate or MDMA alter sexual arousal or cognitive appraisal of intimate relationships is not known. Thus, we evaluated effects of methylphenidate (40 mg) and MDMA (75 mg) on subjective sexual arousal by viewing erotic pictures and on perception of romantic relationships of unknown couples in a double-blind, randomized, placebo-controlled, crossover study in 30 healthy adults. Methylphenidate, but not MDMA, increased ratings of sexual arousal for explicit sexual stimuli. The participants also sought to increase the presentation time of implicit sexual stimuli by button press after methylphenidate treatment compared with placebo. Plasma levels of testosterone, estrogen, and progesterone were not associated with sexual arousal ratings. Neither MDMA nor methylphenidate altered appraisal of romantic relationships of others. The findings indicate that pharmacological stimulation of dopaminergic but not of serotonergic neurotransmission enhances sexual drive. Whether sexual perception is altered in subjects misusing methylphenidate e.g., for cognitive enhancement or as treatment for attention deficit hyperactivity disorder is of high interest and warrants further investigation.
We describe a new dependency parser for English tweets, TWEEBOPARSER. The parser builds on several contributions: new syntactic annotations for a corpus of tweets (TWEEBANK), with conventions informed by the domain; adaptations to a statistical parsing algorithm; and a new approach to exploiting out-of-domain Penn Treebank data. Our experiments show that the parser achieves over 80% unlabeled attachment accuracy on our new, high-quality test set and measure the benefit of our contributions. Our dataset and parser can be found at http://www.ark.cs.cmu.edu/TweetNLP.
According to Tsinghua Chinese Treebank annotation methods, the authors extracted relation words and marked their categories. Then syntax, lexical and position features of automatic syntax tree with and without functional marker were extracted to recognize and classify relation words. Experiment results show that relative recognition accuracy is 95.7%, and relation words classification F1 is 77.2%.
In this paper, we analyze the impact of various dependency representations for various constructions on the general parsing accuracy and on the parsing accuracy of these constructions. We focus on the analysis of coordination constructions, complex predicates, and punctuation mark attachment. We use Latvian Treebank as a dataset, thus, providing insight for an inflective language with a rather free word order. Experiments with MaltParser, a transition-based parser, show clear difference in learnability of various representations for the considered constructions. Future work would include carrying out comparable experiments with a graph-based dependency parser like MSTParser.
This paper mainly introduced the research on constructing Mongolian Treebank based on phrase structure grammar. Having Considered related Mongolian Treebank work and Mongolian words characteristics, we developed a Mongolian syntactic tagset. The tagset includes two kinds of tags. One is syntactic constituent tag and the other is grammatical relation tag. On the basis of the tagset, we developed the Mongolian Treebank auxiliary processing system. Finally, we built a Treebank that contains 3645 sentences and did an experiment on this Treebank.
English. Network theory provides a suitable framework to model the structure of language as a complex system. Based on a network built from a Latin dependency treebank, this paper applies methods for network analysis to show the key role of the verb sum (to be) in the overall structure of the network. Italiano. La teoria dei grafi fornisce un valido supporto alla modellizzazione strutturale del sistema linguistico. Basandosi su un network costruito a partire da una treebank a dipendenze del latino, l’articolo applica diversi metodi di analisi dei grafi, mostrando l’importanza del ruolo rivestito dal verbo sum (essere) nella struttura complessiva del network.
We present HamleDT - a HArmonized Multi-LanguagE Dependency Treebank. HamleDT is a compilation of existing dependency treebanks (or depen- dency conversions of other treebanks), transformed so that they all conform to the same annotation style. In the present article, we provide a thorough investigation and discussion of a number of phenomena that are comparable across languages, though their annotation in treebanks often differs. We claim that transformation procedures can be designed to automatically identify most such phenomena and convert them to a unified annotation style. This unification is beneficial both to comparative corpus linguistics and to machine learning of syntactic parsing.
Recent years have seen an increased interest in and availability of parallel corpora. Large corpora from international organizations (e.g. European Union, United Nations, European Patent Office), or from multilingual Internet sites (e.g. OpenSubtitles) are now easily available and are used for statistical machine translation but also for online search by different user groups. This paper gives an overview of different usages and different types of search systems. In the past, parallel corpus search systems were based on sentence-aligned corpora. We argue that automatic word alignment allows for major innovations in searching parallel corpora. Some online query systems already employ word alignment for sorting translation variants, but none supports the full query functionality that has been developed for parallel treebanks. We propose to develop such a system for efficiently searching large parallel corpora with a powerful query language.
Comunicació presentada al 9th International Conference on Language Resources and Evaluation (LREC'14), celebrat del 26 al 31 de maig de 2014 a Reykjavík, Islàndia.
Lexical Database the Japanese WordNet is a useful tool in natural language processing. However, it is officially announced that Japanese WordNet contains 5% errors. In this paper, we discuss error detection methods in the Japanese WordNet. キーワード Thesaurus, WordNet, Japanese WordNet, 1. はじめに 日本語 WordNet[1,2]は Princeton 大学が開発した WordNet[3]を用いた言語データベースである。日本語 WordNet は自然言語処理において有用であり、様々な 実験に使用されている [4]。フリーの Web シソーラス サービスにおいて、日本語 WordNet は一般的に使用さ れている。しかしながら、現行の日本語 WordNet は間 違いを 5%ほど含んでいると作成者らが認めており [2]、 それらの間違いが日本語 WordNet の使いやすさに影響 を及ぼしている可能性がある。 本論文では、われわれが検証した日本語 WordNet の 間違い探知手法において議論する。間違い探知は日本 語 WordNet の間違い修正の第一段階である。この手法 は、大規模言語データベースの作成に有用であると考 える。我々は特に日本語 WordNet の似たような間違い 1 Weblio, http://ejje.weblio.jp の発見を主眼にしており、この間違いのことを「類義 語の間違い」と呼んでいる。 英語でない WordNet や WordNet に似た言語データベ ースの作成という点において、複数のプロジェクトが 行 わ れ て い る 。 日 本 語 WordNet や Chinese Open WordNet[5]は、ブートストラップの段階で、Princeton WordNet のマッピング手法を用いて半自動生成されて いる。 また、 Universal WordNet[6]や Babel Net[7]、 Open Multilingual WordNet[8]といった、WordNet の拡張によ る統合、多言語概念字句データベースの生成の試みも なされている。概念と語句、または複数の概念間の関 係は、Wikipedia やタグ付けコーパスのような様々な資 源から自動的に抽出することが可能である。それらに よって得られた統合データベースの品質は、生成者自 身や、ネットワークコミュニティによって評価されて きた。 WordNet は、オントロジーのひとつとしてみなすこ とができる。多言語 WordNet を生成する場合には、言 語数に応じたオントロジー間のマッピングをする必要 がある。そのため、オントロジーの間違いの検出と修 正、オントロジー間のマッピングに関する研究がなさ れてきた。これらの研究において、オントロジー内で 分類が間違っているものや、冗長もしくは不適切であ る、または間違った関係性を生成されている箇所を修 正する試みがなされてきた。 間違い検出の手法として、日本語 WordNet のみを使 用した手法を動詞に適用した場合をベースラインとし て提示する [9]。また、コーパスを用いた単語をベクト ル化し、これらのコサイン類似度によって名詞の間違 い検出の手法として使用できないかを議論する。 本論文では、第 2 節で WordNet と日本語 WordNet の 説明、第 3 節で本論文のコンセプトと WordNet の構造 における「同義語の間違い」の一例を紹介する。第 4 節では間違いの抽出法に関するわれわれの手法の説明、 第 5 節では手法を用いた場合の結果の提示を行う。第 6 節では、本手法の Princeton WordNet における応用例 と、関心を持っている別手法に関しての説明、第 7 節 で word2vec を用いた単語のベクトル化とそれらを用 いた間違い検出の実験、第 8 節に今後の展望と課題を 述べる。 2. WordNet と日本語 WordNet 2.1. Princeton WordNet Princeton WordNet は英語の大規模言語データベース である。名詞、動詞、形容詞、副詞といった品詞ごと に、明確なコンセプトを持った「Synset」という認知同 義語のセットに纏められる。各 Synset は固有の ID に よって管理されており、Gloss と呼ばれる、Synset の簡 単な意味を説明するテキストがリンクされている。 Synset は概念 -意味関係もしくは字句トークン関係で 相互リンクを持っている。単語が持つ意味を Synset に よってグループ化することができるため、WordNet は シソーラスとして使用できる。多義である単語が存在 するため、単語は複数の Synset に属することがある。 2.2. 日本語 WordNet 日本語 WordNet は Princeton WordNet を基にした、日 本語の語彙データベースである。日本語 WordNet のプ 2 Wikipedia, http://ja.wikipedia.org ロジェクトの目的は、誰でも自由に使用可能な大規模 日本語データベースを提供することである。このデー タベースは 2006 年から開発されている。 日本語 WordNet の構造は、Princeton WordNet に準拠 している [1]。しかし、日本語と英語という言語の違い が存在するため、日本語 WordNet は Princeton WordNet に含まれていないオリジナルの Synset を含んでいる。 また、日本語 WordNet は、シソーラスとしての精度よ り多数の概念を包括することに主眼を置いている。 現行の日本語 WordNet の規模は以下のとおりである。 ・57,238 概念(Synset 数) ・93,834 語(日本語) ・158,058 語義(単語 -synset ペア数) 図 1 日本語 WordNet の Synset-同義語間リンク例 日本語とリンクを持つ Synset は日本語の gloss を持 っている。日本語 WordNet のカバー範囲の拡張のため に、SUMO や Wikipedia、GoiTaikei[10]といった他のリ ソースが使用されている。 2.3. 他言語の WordNet と WordNet の拡張 Princeton WordNet を基にした、様々な言語の言語デ ータベース作成プロジェクトが存在する。一部のプロ ジェクトでは WordNet、Wikipedia、Wiktionary及びそ の他の言語資源を用いて、多言語のごくデータベース を作成しようと試みている。 既存の言語資源と新しいデータベース間のマッピ ングの正確さは、それによって出力されるデータベー スの整合性の正しさを証明する指標になるので非常に 重要である。新しい言語の WordNet を作成することは、 他の言語からなる新しいオントロジーで表現されてい る、既存のオントロジーからマッピングで作成すると みなすことができる。 3 Wiktionary, http://ja.wiktionary.org 3. 日本語 WordNet の間違い 間違いの訂正は、新しく作成したオントロジーや、 オントロジー間のマッピングの整合性の確認において 重要である。日本語 WordNet の現行のバージョンでは、 約 5%の間違いが含まれている。また、Chinese Open WordNet も、それに匹敵するエラー率である。本節の 残りでは、同義語における間違いにおける、エラーの 種類に焦点を当てる。 3.1. 同義語の間違い WordNet の構造において、「同義語の間違い」を、語 wmiss が属している synset(S とする)の Gloss と合致 しない語であると定義する。 図 2 では、Synset 02651424-v について図示している。 この Synset は「泊める」、「収容」、「宿る」、「持ち込む」 という 4 つの同義語を持っている。
Machine assistance is vital to managing the cost of corpus annotation projects. Identifying effective forms of machine assistance through principled evaluation is particularly important and challenging in under-resourced domains and highly heterogeneous corpora, as the quality of machine assistance varies. We perform a fine-grained evaluation of two machine-assistance techniques in the context of an under-resourced corpus annotation project. This evaluation requires a carefully controlled user study crafted to test a number of specific hypotheses. We show that human annotators performing morphological analysis of text in a Semitic language perform their task significantly more accurately and quickly when even mediocre pre-annotations are provided. When pre-annotations are at least 70 % accurate, annotator speed and accuracy show statistically significant relative improvements of 25–35 and 5–7 %, respectively. However, controlled user studies are too costly to be suitable for under-resourced corpus annotation projects. Thus, we also present an alternative analysis methodology that models the data as a combination of latent variables in a Bayesian framework. We show that modeling the effects of interesting confounding factors can generate useful insights. In particular, correction propagation appears to be most effective for our task when implemented with minimal user involvement. More importantly, by explicitly accounting for confounding variables, this approach has the potential to yield fine-grained evaluations using data collected in a natural environment outside of costly controlled user studies.
We present IceMorph, a semi-supervised morphosyntactic analyzer of Old Icelandic. In addition to machine-read corpora and dictionaries, it applies a small set of declension prototypes to map corpus words to dictionary entries. A web-based GUI allows expert users to modify and augment data through an online process. A machine learning module incorporates prototype data, edit-distance metrics, and expert feedback to continuously update part-of-speech and morphosyntactic classification. An advantage of the analyzer is its ability to achieve competitive classification accuracy with minimum training data. [ABSTRACT FROM AUTHOR], Copyright of PLoS ONE is the property of Public Library of Science and its content may not be copied or emailed to multiple sites or posted to a listserv without the copyright holder's express written permission. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the co)
Purpose: Acoustic and perceptual studies show a number of differences between the voices of radio performers and controls. Despite this, the vocal fold kinematics underlying these differences are largely unknown. Using high-speed videoendoscopy, this study sought to determine whether the vocal vibration features of radio performers differed from those of non-performing controls. Method: Using high-speed videoendoscopy, recordings of a mid-phonatory/i/ in 16 male radio performers (aged 25–52 years) and 16 age-matched controls (aged 25–52 years) were collected. Videos were extracted and analysed semi-automatically using High-Speed Video Program, obtaining measures of fundamental frequency (f0), open quotient and speed quotient. Post-hoc analyses of sound pressure level (SPL) were also performed (n = 19). Pearson's correlations were calculated between SPL and both speed and open quotients. Results: Male radio performers had a significantly higher speed quotient than their matched control)
The present investigation examines the development of children's diagnostic reasoning abilities when such inferences involve belief revision about uncertain potential causes. Four- to 7-year-olds observed an event occur that was due to one of four potential causes. Some of those potential causes were revealed to be efficacious; others were revealed to be inefficacious, but there was always one potential cause presented with unknown efficacy. While all children could make appropriate predictive inferences about this situation, 4- and 5-year-olds were less capable of making correct diagnostic inferences about the cause of the event under these circumstances than older children. We discuss possible mechanisms for this development, as well as speculate on the relation between these findings and literature in children's scientific reasoning. [ABSTRACT FROM AUTHOR], Copyright of PLoS ONE is the property of Public Library of Science and its content may not be copied or emailed to multiple si)
We investigate the structure of spatial knowledge that spontaneously develops during free exploration of a novel environment. We present evidence that this structure is similar to a labeled graph: a network of topological connections between places, labeled with local metric information. In contrast to route knowledge, we find that the most frequent routes and detours to target locations had not been traveled during learning. Contrary to purely topological knowledge, participants typically traveled the shortest metric distance to a target, rather than topologically equivalent but longer paths. The results are consistent with the proposal that people learn a labeled graph of their environment. [ABSTRACT FROM AUTHOR], Copyright of PLoS ONE is the property of Public Library of Science and its content may not be copied or emailed to multiple sites or posted to a listserv without the copyright holder's express written permission. However, users may print, download, or email articles for in)
This book offers an exciting new perspective on the origins of language. Language is conceptualized as a collective invention, on the model of writing or the wheel, and the book places social and cultural dynamics at the centre of its evolution: language emerged and further developed in human communities already suffused with meaning and communication, mimesis, ritual, song and dance, coparenting, new divisions of labour, and revolutionary changes in social relations. The book thus challenges assumptions about the causal relations between genes, capacities, social communication, and innovation: the biological capacities are taken to evolve incrementally on the basis of cognitive plasticity, in a process that recruits previous adaptations and fine-tunes them to serve novel communicative ends. Topics include the ability brought about by language to tell lies, which must have confronted our ancestors with new problems of public trust; the dynamics of social-cognitive co-evolution; the role of gesture and mimesis in linguistic communication; studies of how monkeys and apes express their feelings or thoughts; play, laughter, dance, song, ritual, and other social displays among extant hunter-gatherers; the social nature of language acquisition and innovation; normativity and the emergence of linguistic norms; the interaction of language and emotions; and novel perspectives on the timeframe for language evolution. The contributors are leading international scholars from linguistics, anthropology, paleontology, primatology, psychology, evolutionary biology, artificial intelligence, archaeology, and cognitive science. (PsycINFO Database Record (c) 2016 APA, all rights reserved)
Reviewed by: Récits du corps au Maroc et au Japon ed. by Marc Kober and Khalid Zekri Gaëlle Corvaisier Kober, Marc, et Khalid Zekri, coords. Récits du corps au Maroc et au Japon. Paris: L’Harmattan, 2011. isbn 9782296557208. 200 p. Avec Récits du corps au Maroc et au Japon, Marc Kober et Khalid Zekri posent des questions essentielles pour la littérature francophone contemporaine dans un contexte postcolonial volontairement décentré d’une hégémonie culturelle occidentale, ici européenne. L’un des postulats de cet ouvrage, produit du Centre d’Étude des Nouveaux Espaces Littéraires de l’Université Paris 13 à Villetaneuse en France, est d’observer, de voir et de donner à voir (et à lire) un corps “oriental” afin de le distinguer des habitus nationaux voire régionaux. Par une volonté d’analyse polysémique prudente se déjouant, dans la mesure du possible, d’un européocentrisme prégnant, les auteurs de cet ouvrage questionnent la validité d’une démarche comparatiste entre aires proche-orientale et extrême-orientale aux précédents limités. La révélation d’un corps oriental comme dénominateur commun à un corpus littéraire et visuel (photographie, cinéma, bande dessinée) ne sera néanmoins pas de mise. Il est plutôt question d’analyser comment une réflexion historique, socio-culturelle, politique, religieuse et identitaire affecte, marque et montre des corps hybrides dans une aire culturelle plus globale. Les corps inscrits dans un corpus maroco-japonais ont-ils la possibilité de se parler et de se voir? Qu’ont-ils en commun? Y a-t-il un regard extraeuropéen sur les représentations du corps comme “objet social, historique ou psychanalytique” (7)? En quoi ce regard affecterait-il le travail introspectif et représentatif de l’artiste? Et s’il n’y avait pas de corps oriental à proprement parler, pourrait-on parler de corps (ou de corpus) national? L’existence de rituels similaires (les bains et le hammam; la honte d’être vu nu et la hchouma par exemple) permettrait-elle de concilier des visions du corps féminin intrinsèquement [End Page 217] différentes entre monde arabo-islamique, où son existence en changement est codifiée par la collectivité masculine et religieuse, et espace japonais mythique, religieux et fantastique dans lequel le corps féminin nu (parfois dénué d’érotisme) est omniprésent pour un lecteur occidental qui le quête? L’hétéronormativité fausserait-elle l’impact de la littérature féminine et de la littérature “queer” en les (re)présentant en tant qu’objets marginaux mettant à mal le principe d’appartenance identitaire unique? Comment aborder un corps militaire (principalement masculin) dont l’identité est à jamais marquée par une défaite brutale, et qui personnifie la souffrance de l’échec dans un monde postnucléaire? Et comment envisager le corps corporatif de l’ouvrier et de l’employé qui souffre d’un malaise identitaire dans le Japon des années 1960 et 1970 où modernisation rime avec nouvelle représentation et rejet des traditions? Le corps, cet “objet sémiologique” (15) est un lieu d’enjeux vitaux. C’est un élément perturbateur et perturbé, symptôme de son époque. Il personnifie l’implosion du corps social, il contredit les normes d’hier, il réécrit celles de demain. Il est vu à travers la lunette identitaire, historique et socio-culturelle de celui qui voit d’une manière qui n’est pas sans rappeler l’œuvre visuelle Étant donnés 1e la chute d’eau, 2e le gaz d’éclairage de Marcel Duchamp. Il emprunte à d’autres formats culturels afin d’assurer la survie de son message face à la censure. Il défie les définitions en offrant d’autres mots au champ lexical vernaculaire. Il explore et/ou déjoue les espaces physiologiques dans lesquels il est confiné pour poétiser sur une quête identitaire ambivalente dans laquelle “je est autre” selon la formule consacrée d’Arthur Rimbaud dans sa lettre à Paul Demeny datée du 15 mai 1871. En revisitant de nombreux textes dont des textes mythologiques et...
In the context of processing Bengali words through a computer, there may arise several issues that are directly linked with surface structure of words. These issues may create problems in manual and computer-based counting of number of words in a corpus. They can also create problems in morphological processing of words. These issues come up because there is hardly any consistency in orthographic representation of words in written Bengali texts. The high irregularities in writing of inflected words, proper names, adjectival forms, adverbial forms, compound words, reduplicated words, onomatopoeic words, hyphenated words, etc. present a daunting task before an investigator in normalizing the surface forms of words for generating a lexical database as well as developing a word processing system for the works of language technology.
The French Lexical Network (fr-LN) is a global model of the French lexicon presently under construction. The fr-LN accounts for lexical knowledge as a lexical network structured by paradigmatic and syntagmatic relations holding between lexical units. This paper describes how morphological knowledge is presently being introduced into the fr-LN through the implementation and lexicographic exploitation of a dynamic morphological model. Section 1 presents theoretical and practical justifications for the approach which we believe allows for a cognitively sound description of morphological data within semantically-oriented lexical databases. Section 2 gives an overview of the structure of the dynamic morphological model, which is constructed through two complementary processes: a Morphological Process--section 3--and a Lexicographic Process--section 4.