Papers reviewed and determined not to be word norm studies. Use the flag icon to report errors or suggest re-inclusion.
18265 papers
We present a system for cross-lingual parse disambiguation, exploiting the assumption that the meaning of a sentence remains unchanged during translation and the fact that different languages have different ambiguities. We simultaneously reduce ambiguity in multiple languages in a fully automatic way. Evaluation shows that the system reliably discards dispreferred parses from the raw parser output, which results in a pre-selection that can speed up manual treebanking. 1
Retrieval practice for some memory items from a given category can impair subsequent retrieval of unpracticed items from the same category (retrieval-induced forgetting, RIF). Inhibition of these items has been invoked as an explanation, and inhibition has also been proposed to cause stimulus devaluation. The present experiments investigated whether a similar devaluation effect can be observed in a RIF experiment for the unpracticed and presumably inhibited items. We report two experiments using the RIF paradigm, and both experiments yielded a RIF effect. At the same time, affective ratings of the very same items did not show signs of devaluation. These results run counter the idea that both RIF and devaluation effects are caused by a (similar) inhibitory mechanism, or at least they suggest differences between the mechanisms involved in external perceptual and internal memory selection.
A major computational burden, while performing document clustering, is the calculation of similarity measure between a pair of documents. Similarity measure is a function that assign a real number between 0 and 1 to a pair of documents, depending upon the degree of similarity between them. A value of zero means that the documents are completely dissimilar whereas a value of one indicates that the documents are practically identical. Traditionally, vector-based models have been used for computing the document similarity. The vector-based models represent several features present in documents. These approaches to similarity measures, in general, cannot account for the semantics of the document. Documents written in human languages contain contexts and the words used to describe these contexts are generally semantically related. Motivated by this fact, many researchers have proposed semantic-based similarity measures by utilizing text annotation through external thesauruses like WordNet (a lexical database). In this paper, we define a semantic similarity measure based on documents represented in topic maps. Topic maps are rapidly becoming an industrial standard for knowledge representation with a focus for later search and extraction. The documents are transformed into a topic map based coded knowledge and the similarity between a pair of documents is represented as a correlation between the common patterns. The experimental studies on the text mining datasets reveal that this new similarity measure is more effective as compared to commonly used similarity measures in text clustering.
State of the art parsers are currently trained on converted versions of Penn Treebank into dependency representations which however don’t include null elements. This is done to facilitate structural learning and prevent the probabilistic engine to postulate the existence of deprecated null elements everywhere (see [15]). However it is a fact that in this way, the semantics of the representation used and produced on runtime is inconsistent and will reduce dramatically its usefulness in real life applications like Information Extraction, Q/A and other semantically driven fields by hampering the mapping of a complete logical form. What systems have come up with are “Quasi”-logical forms or partial logical forms mapped directly from the surface representation in dependency structure. We show the most common problems derived from the conversion and then describe an algorithm that we have implemented to apply to our converted Italian Treebank, that can be used on any CONLL-style treebank or representation to produce an “almost complete” semantically consistent dependency treebank.
Unsupervised dependency parsing is one of the most challenging tasks in natural languages processing. The task involves finding the best possible dependency trees from raw sentences without getting any aid from annotated data. In this paper, we illustrate that by applying a supervised incremental parsing model to unsupervised parsing; parsing with a linear time complexity will be faster than the other methods. With only 15 training iterations with linear time complexity, we gain results comparable to those of other state of the art methods. By employing two simple universal linguistic rules inspired from the classical dependency grammar, we improve the results in some languages and get the state of the art results. We also test our model on a part of the ongoing Persian dependency treebank. This work is the first work done on the Persian language. 1
The paper presents and evaluates an efficient algorithm for measuring semantic similarity of texts. Calculating the level of semantic similarity of texts is a very difficult task and the proposed up to now methods suffer from computational complexity. This substantially limits their application area. The proposed algorithm tries to reduce the problem by merging a computationally efficient statistical approach to text analysis with a semantic component. The semantic properties of text words are extracted from the WordNet lexical database. The approach was tested using WordNets for two languages: English and Polish. The basic properties of this approach are also studied. The paper concludes with an analysis of the performance of the proposed method on a sample database and suggests some possible application areas.
Decoding pain in others is of high individual and social benefit in terms of harm avoidance and demands for accurate care and protection. The processing of facial expressions includes both specific neural activation and automatic congruent facial muscle reactions. While a considerable number of studies investigated the processing of emotional faces, few studies specifically focused on facial expressions of pain. Analyses of brain activity and facial responses elicited by the perception of facial pain expressions in contrast to other emotional expressions may unravel the processing specificities of pain-related information in healthy individuals and may contribute to explaining attentional biases in chronic pain patients. In the present study, 23 participants viewed short video clips of neutral, emotional (joy, fear), and painful facial expressions while affective ratings, event-related brain responses, and facial electromyography (Musculus corrugator supercilii, M. orbicularis oculi, M. zygomaticus major, M. levator labii) were recorded. An emotion recognition task indicated that participants accurately decoded all presented facial expressions. Electromyography analysis suggests a distinct pattern of facial response detected in response to happy faces only. However, emotion-modulated late positive potentials revealed a differential processing of pain expressions compared to the other facial expressions, including fear. Moreover, pain faces were rated as most negative and highly arousing. Results suggest a general processing bias in favor of pain expressions. Findings are discussed in light of attentional demands of pain-related information and communicative aspects of pain expressions.
We describe a transformation-based learning method for learning a sequence of mono-lingual tree transformations that improve the agreement between constituent trees and word alignments in bilingual corpora. Using the manually annotated English Chinese Transla-tion Treebank, we show how our method au-tomatically discovers transformations that ac-commodate differences in English and Chi-nese syntax. Furthermore, when transforma-tions are learned on automatically generated trees and alignments from the same domain as the training data for a syntactic MT system, the transformed trees achieve a 0.9 BLEU im-provement over baseline trees. 1
In this article, we investigate ambiguity in syntactic annotation. The ambiguity in question is inherent in a way that even human annotators interpret the meaning differently. In our experiment, we detect potential structurally ambiguous sentences with Constraint Grammar rules. In the linguistic phenomena we investigate, structural ambiguity is primarily caused by word order. The potentially ambiguous particle or adverbial is located between the main verb and the (participial) NP. After detecting the structures, we analyze how many of the potentially ambiguous cases are actually ambiguous using the double-blind method. We rank the sentences captured by the rules on a 1 to 5 scale to indicate which reading the annotator regards as the primary one. The results indicate that 67% of the sentences are ambiguous. Introducing ambiguity in the treebank/parsebank increases the informativeness of the representation since both correct analyses are presented.
After a period when the focus was essentially on mental architecture, the cognitive sciences are increasingly integrating the social dimension. The rise of a cognitive sociolinguistics is part of this trend. The article argues that this process requires a re-evaluation of some entrenched positions in linguistics: those that see linguistic norms as antithetical to a descriptive and variational linguistics. Once such a re-evaluation has taken place, however, the social recontextualization of cognition will enable linguistics (including sociolinguistics as an integral part), to eliminate the cracks in the foundations that were the result of suppressing the sociocultural underpinnings of linguistic facts. Structuralism, cognitivism and social constructionism introduced new and necessary distinctions, but in their strong forms they all turned into unnecessary divides. The article tries to show that an evolutionary account can reintegrate the opposed fragments into a whole picture that puts each of them in their ‘ecological position’ with respect to each other. Empirical usage facts should be seen in the context of operational norms in relation to which actual linguistic choices represent adaptations. Variational patterns should be seen in the context of structural categories without which there would be only ‘differences’ rather than variation. And emergence, individual choice, and flux should be seen in the context of the individual’s dependence on lineages of community practice sustained by collective norms.
Dependency parsing has attracted considerable interest from researchers and developers in natural language processing. However, to obtain a high‐accuracy dependency parser, supervised techniques require a large volume of hand‐annotated data, which are extremely expensive. This paper presents a simple and effective approach for improving dependency parsing with subtrees derived from unannotated data, which are easy to obtain. First, we use a baseline parser to parse large‐scale unannotated data. Then, we extract subtrees from dependency parse trees in the auto‐parsed data. Next, the extracted subtrees are classified into several sets according to their frequency. Finally, we design new features based on the subtree sets for parsing algorithms. To demonstrate the effectiveness of our proposed approach, we conduct experiments on the English Penn Treebank and Chinese Penn Treebank. The results show that our approach significantly outperforms baseline systems. It also achieves the best accuracy for the Chinese data and an accuracy competitive with the best known systems for the English data.
We present a detailed error analysis of a transition-based dependency parser trained on a Hindi dependency treebank. Parser error analysis has not been systematically examined from the point of view of treebanking before and this work intends to contribute in this area. We address two main questions in this paper: Can the parsing of certain structures be made easier by using alternative analyses for these structures? Are there certain linguistic cues implicit (or missing) in the current treebank that can be made explicit (or added) in order to make the parsing of complex constructions easier? These questions will guide us in examining the potential benefits of parser error analysis during treebanking. Through our experiments and analysis we were able to shed light on the causes of errors and subsequently have been able to improve the performance of the parser.
Among various neural network language models (NNLMs), recurrent neural network-based language models (RNNLMs) are very competitive in many cases. Most current RNNLMs only use one single feature stream, i.e., surface words. However, previous studies proved that language models with additional linguistic information achieve better performance. In this study, we extend RNNLM by explicitly integrating additional linguistic information, including morphological, syntactic, or semantic factors. Our proposed RNNLM is called a factored RNNLM that is expected to enhance RNNLMs. A number of experiments are carried out that show the factored RNNLM improves the performance for all considered tasks: consistent perplexity and word error rate (WER) reductions. In the Penn Treebank corpus, the relative improvements over n-gram LM and RNNLM are 29.0 % and 13.0%, respectively. In the IWSLT-2011 TED ASR test set, absolute WER reductions over RNNLM and n-gram LM reach 0.63 and 0.73 points. Title and Abstract in another language (Chinese) ddddddddddddddd ddddddddddddNNLMddddddddddddddRNNLMddd ddddddddddddddddddddRNNLMddddddddddddddd ddddddddddddddddddddddddddddddddddddddd ddddddddddddddddddRNNLMddddddddddddddddd dddddddddfRNNLMddddddddddddddddfRNNLMdddddd dddddddRNNLMdddddddddddddddddddddWERddddd ddddfRNNLMdddddddddnddddddRNNLMddddd29.0%d13.0%d dIWSLT-2011 TEDdddddddddfRNNLMddddddddddd0.63ddd dnddddddd0.73ddddRNNLMdd
Since the late 1800s, the Uruguayan Government has attempted to enforce cultural and linguistic norms along the border with Brazil through the prohibition of Portuguese, especially in schools, despite the fact that this is the heritage language of most border residents. This research focuses on the differential use of Spanish and Portuguese in Rivera, the largest city on the border. Using self-reported data and metalinguistic commentaries extracted from interviews with 63 Spanish–Portuguese bilinguals, the use of both languages in various domains (home, school, work spaces) and with diverse interlocutors (family, friends, co-workers, superiors) is analyzed. Quantitative and qualitative analysis reveals that Portuguese, which has been marginalized for decades, is more frequently used in the home with relatives and close friends. The use of Portuguese in more formal domains, including schools, is much less frequent. The results from this study corroborate a perception within the community that Portuguese lacks the prestige of Spanish and provide further evidence of its status as a primarily home language. The current research does not show a progressive shift toward Spanish in Rivera nor does it support claims by other researchers that this community is diglossic. (PsycINFO Database Record (c) 2016 APA, all rights reserved)
Context-free grammars are fundamental for the description of linguistic syntax. However, most artificial grammar learning experiments have explored learning of simpler finite-state grammars, while studies exploring context-free grammars have not assessed awareness and implicitness. This paper explores the implicit learning of context-free grammars employing features of hierarchical organization, recursive embedding and long-distance dependencies. The grammars also featured the distinction between left- and right-branching structures, as well as between centre- and tail-embedding, both distinctions found in natural languages. People acquired unconscious knowledge of relations between grammatical classes even for dependencies over long distances, in ways that went beyond learning simpler relations (e.g. n-grams) between individual words. The structural distinctions drawn from linguistics also proved important as performance was greater for tail-embedding than centre-embedding structures)
While humans are capable of mentally transcending the here and now, this faculty for mental time travel (MTT) is dependent upon an underlying cognitive representation of time. To this end, linguistic, cognitive and behavioral evidence has revealed that people understand temporal constructs by mapping them to concrete spatial domains (e.g. past = backward, future = forward). However, very little research has investigated factors that may determine the topographical characteristics of these spatiotemporal maps. Guided by the imperative role of episodic content for retrospective and prospective thought (i.e., MTT), here we explored the possibility that the spatialization of time is influenced by the amount of episodic detail a temporal unit contains. In two experiments, participants mapped temporal events along mediolateral (Experiment 1) and anterioposterior (Experiment 2) spatial planes. Importantly, the temporal units varied in self-relevance as they pertained to temporally proximal o)
The inhibition of unwanted behaviors is considered an effortful and controlled ability. However, inhibition also requires the detection of contexts indicating that old behaviors may be inappropriate - in other words, inhibition requires the ability to monitor context in the service of goals, which we refer to as context-monitoring. Using behavioral, neuroimaging, electrophysiological and computational approaches, we tested whether motoric stopping per se is the cognitivelycontrolled process supporting response inhibition, or whether context-monitoring may fill this role. Our results demonstrate that inhibition does not require control mechanisms beyond those involved in context-monitoring, and that such control mechanisms are the same regardless of stopping demands. These results challenge dominant accounts of inhibitory control, which posit that motoric stopping is the cognitively-controlled process of response inhibition, and clarify emerging debates on the frontal substrates of res)
Background: The Hospital Acquired Condition Strategy (HACS) denies payment for venous thromboembolism (VTE) after total knee arthroplasty (TKA). The intention is to reduce complications and associated costs, while improving the quality of care by mandating VTE prophylaxis. We applied a system dynamics model to estimate the impact of HACS on VTE rates, and potential unintended consequences such as increased rates of bleeding and infection and decreased access for patients who might benefit from TKA. Methods and Findings: The system dynamics model uses a series of patient stocks including the number needing TKA, deemed ineligible, receiving TKA, and harmed due to surgical complication. The flow of patients between stocks is determined by a series of causal elements such as rates of exclusion, surgery and complications. The number of patients harmed due to VTE, bleeding or exclusion were modeled by year by comparing patient stocks that results in scenarios with and without HACS. The perc)
Visual perceptual learning (VPL) is defined as visual performance improvement after visual experiences. VPL is often highly specific for a visual feature presented during training. Such specificity is observed in behavioral tuning function changes with the highest improvement centered on the trained feature and was originally thought to be evidence for changes in the early visual system associated with VPL. However, results of neurophysiological studies have been highly controversial concerning whether the plasticity underlying VPL occurs within the visual cortex. The controversy may be partially due to the lack of observation of neural tuning function changes in multiple visual areas in association with VPL. Here using human subjects we systematically compared behavioral tuning function changes after global motion detection training with decoded tuning function changes for 8 visual areas using pattern classification analysis on functional magnetic resonance imaging (fMRI))
How does language comprehension interact with motor activity? We investigated the conditions under which comprehending an action sentence affects people's balance. We performed two experiments to assess whether sentences describing forward or backward movement modulate the lateral movements made by subjects who made sensibility judgments about the sentences. In one experiment subjects were standing on a balance board and in the other they were seated on a balance board that was mounted on a chair. This allowed us to investigate whether the action compatibility effect (ACE) is robust and persists in the face of salient incompatibilities between sentence content and subject movement. Growth-curve analysis of the movement trajectories produced by the subjects in response to the sentences suggests that the ACE is indeed robust. Sentence content influenced movement trajectory despite salient inconsistencies between implied and actual movement. These results are interpreted in the context o)
Rapid vocabulary learning in children has been attributed to "fast mapping", with new words often claimed to be learned through a single presentation. As reported in 2004 in Science a border collie (Rico) not only learned to identify more than 200 words, but fast mapped the new words, remembering meanings after just one presentation. Our research tests the fast mapping interpretation of the Science paper based on Rico's results, while extending the demonstration of large vocabulary recognition to a lap dog. We tested a Yorkshire terrier (Bailey) with the same procedures as Rico, illustrating that Bailey accurately retrieved randomly selected toys from a set of 117 on voice command of the owner. Second we tested her retrieval based on two additional voices, one male, one female, with different accents that had never been involved in her training, again showing she was capable of recognition by voice command. Third, we did both exclusion-based training of new items (toys she had never s)
Orthographic neighborhood size (N size) effect in Chinese character naming has been studied in adults. In the present study, we aimed to explore the developmental characteristics of Chinese N size effect. One hundred and seventeen students (40 from the 3rd grade with mean age of 9 years; 40 from the 5th grade with mean age of 11 years; 37 from the 7th grade with mean age of 13 years) were recruited in the study. A naming task of Chinese characters was adopted to elucidate Nsize- effect development. Reaction times and error rates were recorded. Results showed that children in the 3rd grade named characters from large neighborhoods faster than named those from small neighborhoods, revealing a facilitatory N size effect; the 5th graders showed null N size effect; while the 7th graders showed an inhibitory N size effect, with longer reaction times for the characters from large neighborhoods than for those from small neighborhoods. The change from facilitation to inhibition of neighborhood)
La sophistication du discours patrimonial et la spécialisation croissante du domaine du savoir qu’il balise modifient le mécanisme symbolique du patrimoine: parallèlement aux transformations du champ lexical, on peut observer un matérialisme patrimonial qui, dicté par des normes règlementaires et une gouvernance de proximité, semble compromettre la « relique » au profit de nouvelles conceptions de ce que serait le patrimoine et de ses usages. Cette lecture permet d’entrevoir des enjeux peu discutés de la patrimonialisation contemporaine dès lors que l’on assume, bien sûr, que le patrimoine n’est autre qu’une représentation.
In the Diaries of Marin Sanudo, there are two strange reports sent to Venetian authorities from Hvar, in August and September 1512. The documents are strange because they are in Latin, and not, as is the norm for Cinquecento reports from Dalmatia by Venetian officials, in the Veneto dialect of Italian. The author of these reports is Sebastiano Giustinian, the provveditore generale of Dalmatia in 1512, on a mission to supress several local revolts. Sebastiano Giustinian 1459-1543 was a successful Venetian diplomat, serving in Hungary around 1500, on the Ferrarese court of Alfonso I d'Este in 1506, in Brescia at the time of Venetian defeat at Agnadello 1509. After his Istrian and Dalmatian engagement in 1510-1512, Giustinian will go on to England, to the court of Henry VIII 1514-1519; during this period Giustinian exchanged letters with Erasmus and Thomas More, to Crete 1520-1523 and to France 1526-1531, ending his career as the procurator of St Mark's in Venice. Giustinian was obviously well suited to courts and diplomacy; however, a peace-keeping mission in Dalmatia required abilities of a different type. At first, Giustinian successfully suppressed uprisals in Zadar, Sibenik, and Split, with a simple demonstration of Venetian military power. The rebellious citizens and peasants of Hvar, however, were by this time well organised guerillas. Giustinian employed paramilitaries from nearby Poljica, Brac and Trogir to attack Vrboska, a rebel village on Hvar, but this action ended in uncontrolled looting, which was not well received in Venice. Giustinian then tried something else, an almost theatrical public performance in Stari Grad, where he offered the inhabitants a choice between war and peace, celebrating the peace they have chosen in the cathedral of the City of Hvar. But soon afterwards Giustinian suffered a defeat by guerillas in Jelsa, with rebels later taking political action against him in the Venetian Senate. The Latin reports were written from Hvar, on August 3, 1512 this is a letter Giustinian sent to his son Marino, intending it for public circulation, and on September 2, 1512. The letter to Marino reports Giustinian's successes in Zadar and Sibenik; the report to the Senate is an apology, where Giustinian tries to balance the looting of Vrboska with good news from Split and Stari Grad. The performance in Stari Grad, obviously inspired by a scene from Livy Liv. 21, 18-19, when Q. Fabius Maximus in 218 B. C., holding two ends of his toga, theatrically offered the Carthaginians a choice between war and peace, is itself reported in a high humanist style, with lexical echoes from Curtius Rufus and Paulinus Petricordiae, with a Ciceronian antithesis between mansuetudo and severitas cf. Cic. off. 1,88, and Ambrosius, Epistles 9, 64, 10. In a similar way, the report from Zadar contrasts teachings from Scripture, from Aristotle and Cicero, presented by Giustinian in a public speech, with laughable cowardice of fearful rebels, who try to escape disguised as females, or hide in holes barely fit for mice. The rhetoric of Giustinian's reports is a characteristical Renaissance humanist strategy, based on words and wisdom of the Ancients, on a strong belief that the Antiquity can explain the present day and offer solutions for current problems. Seen in this light, Giustinian's reports from Hvar testify to a breakdown of humanist rhetoric; they are written in Latin and styled as humanist texts as long as the provveditore believed that the situation can conform to ancient models. When events get out of hand, Giustinian drops the Latin; neither the language nor its literary models are fit for reporting one's own defeats and unclean, tangled issues of impasse. Moreover, such defeats frustrate the very essence of Renaissance humanism, its idea that, if we can control words, we can control reality as well.
Language change takes place primarily via diffusion of linguistic variants in a population of individuals. Identifying selective pressures on this process is important not only to construe and predict changes, but also to inform theories of evolutionary dynamics of socio-cultural factors. In this paper, we advocate the Price equation from evolutionary biology and the Pó lyaurn dynamics from contagion studies as efficient ways to discover selective pressures. Using the Price equation to process the simulation results of a computer model that follows the Pó lya-urn dynamics, we analyze theoretically a variety of factors that could affect language change, including variant prestige, transmission error, individual influence and preference, and social structure. Among these factors, variant prestige is identified as the sole selective pressure, whereas others help modulate the degree of diffusion only if variant prestige is involved. This multidisciplinary study discerns the primary and co)
Background: Electronic health records are invaluable for medical research, but much of the information is recorded as unstructured free text which is time-consuming to review manually. Aim: To develop an algorithm to identify relevant free texts automatically based on labelled examples. Methods: We developed a novel machine learning algorithm, the 'Semi-supervised Set Covering Machine' (S3CM), and tested its ability to detect the presence of coronary angiogram results and ovarian cancer diagnoses in free text in the General Practice Research Database. For training the algorithm, we used texts classified as positive and negative according to their associated Read diagnostic codes, rather than by manual annotation. We evaluated the precision (positive predictive value) and recall (sensitivity) of S3CM in classifying unlabelled texts against the gold standard of manual review. We compared the performance of S3CM with the Transductive Vector Support Machine (TVSM), the original fully-supe)
This paper explores interoperability for data represented using the Graph Annotation Framework (GrAF) (Ide and Suderman, 2007) and the data formats utilized by two general-purpose annotation systems: the General Architecture for Text Engineering (GATE) (Cunningham et al., 2002) and the Unstructured Information Management Architecture (UIMA) (Ferrucci and Lally in Nat Lang Eng 10(3–4):327–348, 2004). GrAF is intended to serve as a “pivot” to enable interoperability among different formats, and both GATE and UIMA are at least implicitly designed with an eye toward interoperability with other formats and tools. We describe the steps required to perform a round-trip rendering from GrAF to GATE and GrAF to UIMA CAS and back again, and outline the commonalities as well as the differences and gaps that came to light in the process.
The advent of humanoid robots has enabled a new approach to investigating the acquisition of language, and we report on the development of robots able to acquire rudimentary linguistic skills. Our work focuses on early stages analogous to some characteristics of a human child of about 6 to 14 months, the transition from babbling to first word forms. We investigate one mechanism among many that may contribute to this process, a key factor being the sensitivity of learners to the statistical distribution of linguistic elements. As well as being necessary for learning word meanings, the acquisition of anchor word forms facilitates the segmentation of an acoustic stream through other mechanisms. In our experiments some salient one-syllable word forms are learnt by a humanoid robot in real-time interactions with naive participants. Words emerge from random syllabic babble through a learning process based on a dialogue between the robot and the human participant, whose speech is perceived b)
The small alpine district of East Tyrol (Austria) has an exceptional demographic history. It was contemporaneously inhabited by members of the Romance, the Slavic and the Germanic language groups for centuries. Since the Late Middle Ages, however, the population of the principally agrarian-oriented area is solely Germanic speaking. Historic facts about East Tyrol's colonization are rare, but spatial density-distribution analysis based on the etymology of place-names has facilitated accurate spatial mapping of the various language groups' former settlement regions. To test for present-day Y chromosome population substructure, molecular genetic data were compared to the information attained by the linguistic analysis of pasture names. The linguistic data were used for subdividing East Tyrol into two regions of former Romance (A) and Slavic (B) settlement. Samples from 270 East Tyrolean men were genotyped for 17 Y-chromosomal microsatellites (Y-STRs) and 27 single nucleotide polymorphism)
We examined whether language affects the strength of a visual representation in memory. Participants studied a picture, read a story about the depicted object, and then selected out of two pictures the one whose transparency level most resembled that of the previously presented picture. The stories contained two linguistic manipulations that have been demonstrated to affect concept availability in memory, i.e., object presence and goal-relevance. The results show that described absence of an object caused people to select the most transparent picture more often than described presence of the object. This effect was not moderated by goal-relevance, suggesting that our paradigm tapped into the perceptual quality of representations rather than, for example, their linguistic availability. We discuss the implications of these findings within a framework of grounded cognition. [ABSTRACT FROM AUTHOR], Copyright of PLoS ONE is the property of Public Library of Science and its content may not )
The strong association between music and speech has been supported by recent research focusing on musicians' superior abilities in second language learning and neural encoding of foreign speech sounds. However, evidence for a double association-the influence of linguistic background on music pitch processing and disorders-remains elusive. Because languages differ in their usage of elements (e.g., pitch) that are also essential for music, a unique opportunity for examining such language-to-music associations comes from a cross-cultural (linguistic) comparison of congenital amusia, a neurogenetic disorder affecting the music (pitch and rhythm) processing of about 5% of the Western population. In the present study, two populations (Hong Kong and Canada) were compared. One spoke a tone language in which differences in voice pitch correspond to differences in word meaning (in Hong Kong Cantonese, /si/ means 'teacher' and 'to try' when spoken in a high and mid pitch pattern, respectively). )
This study aims at investigating the HLA molecular variation across Switzerland in order to determine possible regional differences, which would be highly relevant to several purposes: optimizing donor recruitment strategies in hematopoietic stem cell transplantation (HSCT), providing reliable reference data in HLA and disease association studies, and understanding the population genetic background(s) of this culturally heterogeneous country. HLA molecular data of more than 20,000 HSCT donors from 9-13 recruitment centers of the whole country were analyzed. Allele and haplotype frequencies were estimated by using new computer tools adapted to the heterogeneity and ambiguity of the data. Nonparametric and resampling statistical tests were performed to assess Hardy-Weinberg equilibrium, selective neutrality and linkage disequilibrium among different loci, both in each recruitment center and in the whole national registry. Genetic variation was explored through genetic distance and hiera)
Many patterns displayed by the distribution of human linguistic groups are similar to the ecological organization described for biological species. It remains a challenge to identify simple and meaningful processes that describe these patterns. The population size distribution of human linguistic groups, for example, is well fitted by a log-normal distribution that may arise from stochastic demographic processes. As we show in this contribution, the distribution of the area size of home ranges of those groups also agrees with a log-normal function. Further, size and area are significantly correlated: the number of speakers p and the area a spanned by linguistic groups follow the allometric relation a ... pz, with an exponent z varying accross different world regions. The empirical evidence presented leads to the hypothesis that the distributions of p and a, and their mutual dependence, rely on demographic dynamics and on the result of conflicts over territory due to group growth. To s)
Background: Recent advances in automated assessment of basic vocabulary lists allow the construction of linguistic phylogenies useful for tracing dynamics of human population expansions, reconstructing ancestral cultures, and modeling transition rates of cultural traits over time. Methods: Here we investigate the Tupi expansion, a widely-dispersed language family in lowland South America, with a distance-based phylogeny based on 40-word vocabulary lists from 48 languages. We coded 11 cultural traits across the diverse Tupi family including traditional warfare patterns, post-marital residence, corporate structure, community size, paternity beliefs, sibling terminology, presence of canoes, tattooing, shamanism, men's houses, and lip plugs. Results/Discussion: The linguistic phylogeny supports a Tupi homeland in west-central Brazil with subsequent major expansions across much of lowland South America. Consistently, ancestral reconstructions of cultural traits over the linguistic phylogen)
Background: Continuity of care is widely acknowledged as a core value in family medicine. In this systematic review, we aimed to identify the instruments measuring continuity of care and to assess the quality of their measurement properties. Methods: We did a systematic review using the PubMed, Embase and PsycINFO databases, with an extensive search strategy including 'continuity of care', 'coordination of care', 'integration of care', 'patient centered care', 'case management' and its linguistic variations. We searched from 1995 to October 2011 and included articles describing the development and/ or evaluation of the measurement properties of instruments measuring one or more dimensions of continuity of care (1) care from the same provider who knows and follows the patient (personal continuity), (2) communication and cooperation between care providers in one care setting (team continuity), and (3) communication and cooperation between care providers in different care settings (cross)
The warp ikat method of making decorated textiles is one of the most geographically widespread in southeast Asia, being used by Austronesian peoples in Indonesia, Malaysia and the Philippines, and Daic peoples on the Asian mainland. In this study a dataset consisting of the decorative characters of 36 of these warp ikat weaving traditions is investigated using Bayesian and Neighbornet techniques, and the results are used to construct a phylogenetic tree and taxonomy for warp ikat weaving in southeast Asia. The results and analysis show that these diverse traditions have a common ancestor amongst neolithic cultures the Asian mainland, and parallels exist between the patterns of textile weaving descent and linguistic phylogeny for the Austronesian group. Ancestral state analysis is used to reconstruct some of the features of the ancestral weaving tradition. The widely held theory that weaving motifs originated in the late Bronze Age Dong-Son culture is shown to be inconsistent with the )
Annotating linguistic data has become a major field of interest, both for supplying the necessary data for machine learning approaches to NLP applications, and as a research issue in its own right. This comprises issues of technical formats, tools, and methodologies of annotation. We provide a brief overview of these notions and then introduce the papers assembled in this special issue.
Background: Medical research increasingly utilizes patient-reported outcome measures administered and scored in different languages. In order to pool or compare outcomes from different language versions, instruments should be measurement equivalent across linguistic groups. The objective of this study was to examine the cross-language measurement equivalence of the Patient Health Questionnaire-9 (PHQ-9) between English- and French-speaking Canadian patients with systemic sclerosis (SSc). Methods: The sample consisted of 739 English- and 221 French-speaking SSc patients. Multiple-Indicator Multiple-Cause (MIMIC) modeling was used to identify items displaying possible differential item functioning (DIF). Results: A one-factor model for the PHQ-9 fit the data well in both English- and French-speaking samples. Statistically significant DIF was found for 3 of 9 items on the PHQ-9. However, the overall estimate in depression latent scores between English- and French-speaking respondents was)
The human populations of the Iberian Peninsula are the varied result of a complex mixture of cultures throughout history, and are separated by clear social, cultural, linguistic or geographic barriers. The stronger genetic differences between closely related populations occur in the northern third of Spain, a phenomenon commonly known as "micro-differentiation". It has been argued and discussed how this form of genetic structuring can be related to both the rugged landscape and the ancient societies of Northern Iberia, but this is difficult to test in most regions due to the intense human mobility of previous centuries. Nevertheless, the Spanish autonomous community of Asturias shows a complex history which hints of a certain isolation of its population. This, joined together with a difficult terrain full of deep valleys and steep mountains, makes it suitable for performing a study of genetic structure, based on mitochondrial DNA and Y-Chromosome markers. Our analyses do not only show)
Computer use draws on linguistic abilities. Using this medium thus presents challenges for young people with Specific Language Impairment (SLI) and raises questions of whether computer-based tasks are appropriate for them. We consider theoretical arguments predicting impaired performance and negative outcomes relative to peers without SLI versus the possibility of positive gains. We examine the relationship between frequency of computer use (for leisure and educational purposes) and educational achievement; in particular examination performance at the end of compulsory education and level of educational progress two years later. Participants were 49 young people with SLI and 56 typically developing (TD) young people. At around age 17, the two groups did not differ in frequency of educational computer use or leisure computer use. There were no associations between computer use and educational outcomes in the TD group. In the SLI group, after PIQ was controlled for, educational computer)
Although several cognitive processes, including speech processing, have been studied during sleep, working memory (WM) has never been explored up to now. Our study assessed the capacity of WM by testing speech perception when the level of background noise and the sentential semantic length (SSL) (amount of semantic information required to perceive the incongruence of a sentence) were modulated. Speech perception was explored with the N400 component of the eventrelated potentials recorded to sentence final words (50% semantically congruent with the sentence, 50% semantically incongruent). During sleep stage 2 and paradoxical sleep: (1) without noise, a larger N400 was observed for (short and long SSL) sentences ending with a semantically incongruent word compared to a congruent word (i.e. an N400 effect); (2) with moderate noise, the N400 effect (observed at wake with short and long SSL sentences) was attenuated for long SSL sentences. Our results suggest that WM for linguistic informa)
The increasing number of experimental studies on second language (L2) processing, frequently with English as the L2, calls for a practical and valid measure of English vocabulary knowledge and proficiency. In a large-scale study with Dutch and Korean speakers of L2 English, we tested whether LexTALE, a 5-min vocabulary test, is a valid predictor of English vocabulary knowledge and, possibly, even of general English proficiency. Furthermore, the validity of LexTALE was compared with that of self-ratings of proficiency, a measure frequently used by L2 researchers. The results showed the following in both speaker groups: (1) LexTALE was a good predictor of English vocabulary knowledge; 2) it also correlated substantially with a measure of general English proficiency; and 3) LexTALE was generally superior to self-ratings in its predictions. LexTALE, but not self-ratings, also correlated highly with previous experimental data on two word recognition paradigms. The test can be carried out on or downloaded from www.lextale.com.
Student-constructed responses, such as essays, short-answer questions, and think-aloud protocols, provide a valuable opportunity to gauge student learning outcomes and comprehension strategies. However, given the challenges of grading student-constructed responses, instructors may be hesitant to use them. There have been major advances in the application of natural language processing of student-constructed responses. This literature review focuses on two dimensions that need to be considered when developing new systems. The first is type of response provided by the student—namely, meaning-making responses (e.g., think-aloud protocols, tutorial dialogue) and products of comprehension (e.g., essays, open-ended questions). The second corresponds to considerations of the type of natural language processing systems used and how they are applied to analyze the student responses. We argue that the appropriateness of the assessment protocols is, in part, constrained by the type of response and researchers should use hybrid systems that rely on multiple, convergent natural language algorithms.
Background: Normal reading requires eye guidance and activation of lexical representations so that words in text can be identified accurately. However, little is known about how the visual content of text supports eye guidance and lexical activation, and thereby enables normal reading to take place. Methods and Findings: To investigate this issue, we investigated eye movement performance when reading sentences displayed as normal and when the spatial frequency content of text was filtered to contain just one of 5 types of visual content: very coarse, coarse, medium, fine, and very fine. The effect of each type of visual content specifically on lexical activation was assessed using a target word of either high or low lexical frequency embedded in each sentence Results: No type of visual content produced normal eye movement performance but eye movement performance was closest to normal for medium and fine visual content. However, effects of lexical frequency emerged early in)
On average our eyes make 3-5 saccadic movements per second when we read, although their neural mechanism is still unclear. It is generally thought that saccades help redirect the retinal fovea to specific characters and words but that actual discrimination of information only occurs during periods of fixation. Indeed, it has been proposed that there is active and selective suppression of information processing during saccades to avoid experience of blurring due to the high-speed movement. Here, using a paradigm where a string of either lexical (Chinese) or non-lexical (alphabetic) characters are triggered by saccadic eye movements, we show that subjects can discriminate both while making saccadic eye movement. Moreover, discrimination accuracy is significantly better for characters scanned during the saccadic movement to a fixation point than those not scanned beyond it. Our results showed that character information can be processed during the saccade, therefore saccades during readin)
Visual lexical decision is a classical paradigm in psycholinguistics, and numerous studies have assessed the so-called ''lexicality effect'' (i.e., better performance with lexical than non-lexical stimuli). Far less is known about the dynamics of choice, because many studies measured overall reaction times, which are not informative about underlying processes. To unfold visual lexical decision in (over) time, we measured participants' hand movements toward one of two item alternatives by recording the streaming x,y coordinates of the computer mouse. Participants categorized four kinds of stimuli as ''lexical'' or ''non-lexical:'' high and low frequency words, pseudowords, and letter strings. Spatial attraction toward the opposite category was present for low frequency words and pseudowords. Increasing the ambiguity of the stimuli led to greater movement complexity and trajectory attraction to competitors, whereas no such effect was present for high frequency words and letter strings. )
Background: Extraction of linguistically relevant auditory features is critical for speech comprehension in complex auditory environments, in which the relationships between acoustic stimuli are often abstract and constant while the stimuli per se are varying. These relationships are referred to as the abstract auditory rule in speech and have been investigated for their underlying neural mechanisms at an attentive stage. However, the issue of whether or not there is a sensory intelligence that enables one to automatically encode abstract auditory rules in speech at a preattentive stage has not yet been thoroughly addressed. Methodology/Principal Findings: We chose Chinese lexical tones for the current study because they help to define word meaning and hence facilitate the fabrication of an abstract auditory rule in a speech sound stream. We continuously presented native Chinese speakers with Chinese vowels differing in formant, intensity, and level of pitch to construct a complex and)
The ASPM and MCPH1 genes have been implicated in the adaptive evolution of the human brain [Mekel-Bobrov N. et al., 2005. Ongoing adaptive evolution of ASPM, a brain size determinant in homo sapiens. Science 309; Evans P.D. et al., 2005. Microcephalin, a gene regulating brain size, continues to evolve adaptively in humans. Science 309]. Curiously, experimental attempts have failed to connect the implicated SNPs in these genes with higher-level brain functions. These results stand in contrast with a population-level study linking the population frequency of their alleles with the tendency to use lexical tones in a language [Dediu D., Ladd D.R., 2007. Linguistic tone is related to the population frequency of the adaptive haplogroups of two brain size genes, ASPM and microcephalin. Proc. Natl. Acad. Sci. U.S.A. 104]. In the present study, we found a significant correlation between the load of the derived alleles of ASPM and tone perception in a group of European Americans who did not spe)