Search results for "050101 languages & linguistics"
showing 10 items of 244 documents
Mediation effect of self-efficacy in the relationship between neuroticism and L2 attainment
2019
This article investigates the role self-efficacy plays in the relationship between Neuroticism and foreign language attainment (operationalised as final grades and self-perceived foreign language skills). To date, the role of personality in foreign language learning has not been clearly specified; moreover, self-efficacy related to this domain has not received sufficient attention. For the purpose of the paper it was proposed that the negative relationship between Neuroticism and attainment can be explained by self-efficacy. The study’s informants consisted of 495 secondary grammar school students at the intermediate to upperintermediate levels of English proficiency. The results revealed t…
The International Comparable Corpus: Challenges in building multilingual spoken and written comparable corpora
2021
This paper reports on the efforts of twelve national teams in building the International Comparable Corpus (ICC; https://korpus.cz/icc) that will contain highly comparable datasets of spoken, written and electronic registers. The languages currently covered are Czech, Finnish, French, German, Irish, Italian, Norwegian, Polish, Slovak, Swedish and, more recently, Chinese, as well as English, which is considered to be the pivot language. The goal of the project is to provide much-needed data for contrastive corpus-based linguistics. The ICC corpus is committed to the idea of re-using existing multilingual resources as much as possible and the design is modelled, with various adjustments, on t…
Co-occurrence of discourse markers in English : from juxtaposition to composition
2019
Abstract In this paper, we report on a qualitative analysis of co-occurring discourse markers, that is, sequences of adjacent discourse markers that belong to the same unit but may express different functions. We examine several formal and functional features of these co-occurring strings on the basis of corpus examples extracted from conversational data in English. In particular, we focus on scope, meaning-in-context (or functions), syntactic category and position. Our analysis reveals several degrees of integration: differences in scope allow us to differentiate juxtaposition and combination of markers. In the case of combination, difference in meaning integration allows us to distinguish…
Using discourse segmentation to account for the polyfunctionality of discourse markers:The case of well
2021
Abstract A large number of studies describe the many different functions of polyfunctional discourse markers like well in different contexts and from different theoretical perspectives. In the current paper, we propose to systematize the many different uses identified based on their position with respect to the discourse units they are associated with. Not only can previous findings on well be integrated into a single coherent representation of its uses and functions, but the positions with respect to the discourse units can also be associated with specific functions, thus shedding light on how the polyfunctionality of well is brought about.
Combinations of discourse markers with repairs and repetitions in English, French and Spanish
2020
Abstract Discourse markers have a central role in planning and repairing processes of speech production. They relate with fluency and disfluency phenomena such as pauses, repetitions and reformulations. Their polyfunctionality is challenging and few form-function mappings are stable cross-linguistically. This study combines a functional and a structural approach to discourse markers and their combination with and within repetitions and self-repairs in native English, French and Spanish, in order to establish the inter-relation between these three fluency-related devices and to find potentially universal patterns of use. Qualitative coding and quantitative analyses of categories of markers a…
Literacy and literacy practices: Plurilingual connected migrants and emerging literacy
2021
Abstract Recent migration towards Europe is characterized by the massive presence of adults whose educational paths have been interrupted and who are thus developing literacy for the first time in a new language. A literacy test elaborated at the University of Palermo, Italy, showed that, on a sample of 774 migrants, about 30 percent could not read and/or write short words. This test assessed the learners’ abilities to read and write, whether in the Roman alphabet or in other writing systems, and whether in Italian or in other languages of learners’ repertoires. These learners with emergent literacy mostly came from sub-Saharan Africa, an area characterized by diverse forms of multilinguali…
The Early Bird gets the Word
2019
Success in an increasingly globalized world sets requirements for versatile communication skills and understanding about other cultures. One of the keys to success is versatile language skills, on which the European Commission spoke out as early as in 1995, recommending that every European citizen should learn two foreign languages in addition to their mother tongue. Now, more than twenty years later, the launch of early A1 language teaching that is to begin in the first grade in Finland, in January 2020, is a significant step towards this goal. Studies show that early foreign language learning needs to be carefully carried out in order to achieve positive effects and the effects that have …
Implementing Ethics in AI: Initial Results of an Industrial Multiple Case Study
2019
Artificial intelligence (AI) is becoming increasingly widespread in system development endeavors. As AI systems affect various stakeholders due to their unique nature, the growing influence of these systems calls for ethical considerations. Academic discussion and practical examples of autonomous system failures have highlighted the need for implementing ethics in software development. However, research on methods and tools for implementing ethics into AI system design and development in practice is still lacking. This paper begins to address this focal problem by providing elements needed for producing a baseline for ethics in AI based software development. We do so by means of an industri…
Designing the Business Conversation Corpus
2020
While the progress of machine translation of written text has come far in the past several years thanks to the increasing availability of parallel corpora and corpora-based training technologies, automatic translation of spoken text and dialogues remains challenging even for modern systems. In this paper, we aim to boost the machine translation quality of conversational texts by introducing a newly constructed Japanese-English business conversation parallel corpus. A detailed analysis of the corpus is provided along with challenging examples for automatic translation. We also experiment with adding the corpus in a machine translation training scenario and show how the resulting system benef…
Inducing the Lyndon Array
2019
In this paper we propose a variant of the induced suffix sorting algorithm by Nong (TOIS, 2013) that computes simultaneously the Lyndon array and the suffix array of a text in $O(n)$ time using $\sigma + O(1)$ words of working space, where $n$ is the length of the text and $\sigma$ is the alphabet size. Our result improves the previous best space requirement for linear time computation of the Lyndon array. In fact, all the known linear algorithms for Lyndon array computation use suffix sorting as a preprocessing step and use $O(n)$ words of working space in addition to the Lyndon array and suffix array. Experimental results with real and synthetic datasets show that our algorithm is not onl…