221 to 230 of 263 Results
Jul 3, 2026
Abels, Klaus; Neeleman, Adriaan, 2026, "Supporting Data for: Linear Asymmetries and Incremental Parsing", https://doi.org/10.18710/HWWGL5, DataverseNO, V1
Linear asymmetries pose a challenge for syntactic theory but may yield to an explanation in terms of sentence parsing, given the inherent left-to-right nature of this process. We therefore explore the option of a parsing-based account of Greenberg’s (1963) Universal 20. Universal 20 describes a word order asymmetry in noun phrases which can be capt... |
Nov 22, 2023
Zhamaletdinova, Elmira, 2023, "Replication Data for: When modality and tense meet. The future marker budet ‘will’ in impersonal constructions with the modal adverb možno ‘be possible’", https://doi.org/10.18710/MOJBDK, DataverseNO, V1
Dataset description: This is a study of examples of Russian impersonal constructions with the modal word možno ‘can, be possible’ with and without the future copula budet ‘will be,’ i.e., možno + budet + INF and možno + INF. The data was collected in 2020-2021 from the old version of the Russian National Corpus (ruscorpora.ru). In the spreadsheet 0... |
May 9, 2025
Claassen, Simon; Enghels, Renata; Parafita Couto, M. Carmen, 2025, "Replication Data for: Intensification strategies in English-Spanish bilingual speech: Examining lexical and morphological markers in Miami bilinguals’ discourse", https://doi.org/10.18710/KZ5JKJ, DataverseNO, V1
Dataset abstract This dataset contains the three data files that the related publication is based on. In total, they contain 3000+ tokens of intensifying constructions. These constructions were extracted from the Miami Corpus, the Santa Barbara Corpus (specifically a subsample of the corpus containing all conversations involving non-Hispanic speake... |
Mar 20, 2025
Somers, Joren; Leuschner, Torsten; De Cuypere, Ludovic; Barðdal, Jóhanna, 2025, "Replication Data for: A corpus-based analysis of the Dat-Nom/Nom-Dat alternation in German", https://doi.org/10.18710/CRSJLY, DataverseNO, V1
Dataset abstract The dataset includes an annotated sample of N = 13292 German written sentences with a Nominative and a Dative argument. The sentences comprise 76 different verbs taking two alternating object orders: 5591 sentences occur with the Dat-Nom order, 8701 sentences occur with the Nom-Dat order. Each sentence is annotated for Object order... |
Mar 14, 2025
Knell, Georgia; Cipitria, Saioa; De Cuypere, Ludovic; Housen, Alex; Struys, Esli, 2025, "Stand-out: A systematic review of the role of salience in second language acquisition", https://doi.org/10.18710/EABCRW, DataverseNO, V1
Dataset description: This dataset contains the query results, and subsequent annotations, of a systematic review of empirical research on the role of salience in second language acquisition. Files includes a list of articles comprising the search results of our queries and their annotations after removing duplicates, a reference list of the include... |
Feb 20, 2025
Nesset, Tore, 2025, "Russian pora ‘time’ vs. vremja ‘time’", https://doi.org/10.18710/5NIX4N, DataverseNO, V1
In order to shed light on the distribution of the Russian nouns pora and vremja, both of which mean ‘time’, I created a database of examples of both nouns from the Russian National Corpus (RNC, syntactic subcorpus). The database comprises all examples where pora or vremja function as a grammatical subject, a grammatical object, or as an adverbial r... |
Feb 26, 2025
O'Neill, Paul, 2025, "Replication Data for: Defective verbs in Portuguese: a morphomic approach", https://doi.org/10.18710/TVYCZL, DataverseNO, V1
This data is used in an article which provides evidence via corpus data and statistical methods that defective verbs in Portuguese constitute a psychological reality for speakers. It looks at the different explanations for defective verbs in Portuguese and argues that the morphome-based explanation for Spanish defective verbs is the most appropriat... |
Nov 19, 2025
Vander Haegen, Flor, 2025, "German universal concessive conditionals with wh-clause-initial and wh-clause-medial marking", https://doi.org/10.18710/FVA2YV, DataverseNO, V1
These are the data analysed in Chapter 6 of Vander Haegen's dissertation entitled "Konstruktionsgrammatik und Variation. Eine Mikrotypologie universaler Irrelevanzgefüge im Gegenwartsdeutschen". The dataset includes an annotated sample of N = 3000 German written universal concessive conditionals (e.g. "Was immer auch passiert, ich bin für dich da!"... |
Jun 18, 2025
Vanhaverbeke, Margot; Enghels, Renata; Parafita Couto, M. Carmen; Ivanova, Iva, 2025, "Supporting Data for: Enhancing code-switching research through comparable corpora: Introducing the El Paso Bilingual Corpus", https://doi.org/10.18710/7LGSXY, DataverseNO, V1
Dataset description: This dataset contains two data files that the related publication is based on. In particular, the data file Dataset_Diminutives contains in total 1886 diminutive constructions extracted from the Bangor Miami Corpus and the El Paso Bilingual Corpus. These constructions are coded for intralinguistic variables relating to the ling... |
Jun 18, 2025
Makarova, Anastasia, 2025, "Replication Data for: Redundancy and rivalry in language. A case study of Russian diminutives", https://doi.org/10.18710/ND3SMQ, DataverseNO, V1
Dataset:The dataset includes examples of different diminutive constructions: morphological diminutives where the diminutive is expressed via suffixation (e.g. dom-ik 'house-dim'), analytical constructions with adjectives meaning 'small' (e.g. malen'kiy dom 'small house') and combinations of morphological diminutives with adjectives meaning 'small'... |
