Featured Dataverses

In order to use this feature you must have at least one published or linked dataverse.

Publish Dataverse

Are you sure you want to publish your dataverse? Once you do so it must remain published.

Publish Dataverse

This dataverse cannot be published because the dataverse it is in has not been published.

Delete Dataverse

Are you sure you want to delete your dataverse? You cannot undelete this dataverse.

Advanced Search

51 to 60 of 261 Results
Nov 9, 2021
Baten, Kristof; Van Hiel, Silke; De Cuypere, Ludovic, 2021, "Replication Data for: Vocabulary Development in a CLIL Context: A Comparison between French and English L2.", https://doi.org/10.18710/PXJX1F, DataverseNO, V1
Dataset abstract The dataset includes vocabulary test scores from 75 native Dutch speakers from Flanders (Belgium), learning both L2 French and L2 English in a CLIL (Content and Language Integrated Learning) context. CLIL refers to the teaching of subjects, such as history or economy, in a foreign language (Coyle et al. 2010). Participants were all...
Oct 24, 2023
Sönning, Lukas, 2023, "Background data (adapted from Jenset & McGillivray 2017) for: Down-sampling from hierarchically structured corpus data", https://doi.org/10.18710/5KCE4U, DataverseNO, V1
Dataset description This dataset, which is adapted from Jenset and McGillivray (2017), contains tabular files documenting the alternating usage of -(e)th and -(e)s to mark third-person verb inflection in Early Modern English. The data provided by Jenset and McGillivray (2017) are drawn from the PPCEME corpus (Kroch et al. 2004) and cover the period...
Jun 18, 2014
Gerstenberger, Ciprian-Virgil, 2014, "Romanian Weak Pronoun Choice Data", https://doi.org/10.18710/GSV27M, DataverseNO, V1
The following corpus study shows that soft linguistic constraints are hard to describe and operationalize. In specific contexts, some Romanian clitic pronouns allow a choice between phonological hosts such as in că-mi dai cartea vs. că îmi dai cartea both meaning [that you give me the book]. What determines the choice between subjunction că in că-m...
Mar 29, 2016
Endresen, Anna; Janda, Laura A.; Reynolds, Robert; Tyers, Francis M., 2016, "Replication data for: Who needs particles? A challenge to the classification of particles as a part of speech in Russian", https://doi.org/10.18710/700FNV, DataverseNO, V1
In 1985, Zwicky argued that “particle” is a pretheoretical notion that should be eliminated from linguistic analysis. We propose a reclassification of Russian particles that implements Zwicky’s directive. Russian particles lack a coherent conceptual basis as a category and many are ambiguous with respect to part of speech. Our corpus analysis of Ru...
Apr 5, 2016
Nesset, Tore, 2016, "Replication data for: Spøkelsesfiske, makrellfotball og traktoregg: norske sammensetninger og konseptuell integrasjon", https://doi.org/10.18710/R4E1EW, DataverseNO, V1
This database is part of a study of Norwegian compounds from the perspective of cognitive linguistics and conceptual integration published in the Norwegian journal Maal og Minne. The database contains a large number of compounds based on the word fiske ‘fishing’. Here is some information about the study from the article: "Denne artikkelen presenter...
Oct 30, 2018
Cvrček, Václav, 2018, "Multi-Dimensional Analysis of Czech", https://doi.org/10.18710/QAJKZW, DataverseNO, V1, UNF:6:5rqhrfGF8iJspOAQER3OCA== [fileUNF]
Original data for a general-purpose multi-dimensional analysis model of register variation in Czech. This post contains a CSV data set of 137 linguistic features measured on 3428 Czech text chunks, and an R script which performs a factor analysis on this data set. The results of this factor analysis were used as a basis for an 8-dimensional model o...
Mar 11, 2024
Lewandowski, Wojciech, 2018, "Replication data for: Constructions are not predictable but are motivated: evidence from the Spanish completive reflexive", https://doi.org/10.18710/4QHOBK, DataverseNO, V2, UNF:6:tokJXAhE3MEy0uanXSF5aQ== [fileUNF]
Many researchers seem to think that construction grammar posits the existence of just wholly idiosyncratic constructions or form-meaning pairings. However, this idea demonstrates a deep misunderstanding of the approach, since constructions rarely emerge sui generis. Rather, construction grammar aims to balance the fact that some linguistic uses can...
Oct 27, 2016
Hansen, Pernille, 2016, "Replication data for: What makes a word easy to acquire? The effects of word class, frequency, imageability and phonological neighbourhood density on lexical development", https://doi.org/10.18710/JEWIVW, DataverseNO, V1
The main dataset includes age of acquisition, vocabulary acquisition and two different sets of frequency data for words in the Norwegian adaptation of the MacArthur-Bates Communicative Development Inventories. In addition, a frequency list for child-directed speech based on two Norwegian CHILDES corpora is available in a separate file. Here, also w...
Feb 29, 2024
Knell, Georgia; De Cuypere, Ludovic; Manouilidou, Christina, 2024, "Method in the madnessless: Exploring factors that impact the processing of two-suffixed complex words", https://doi.org/10.18710/R9IZSV, DataverseNO, V1
Dataset abstract The data collected includes lexical decision data and reaction time data from 56 participants. Three sets of 30 two-suffixed pseudowords were created, each based on a type of grammatical constraint attested in the literature, and presented along with 120 existing two-suffixed English words and 30 nonwords. The data aims to shed lig...
May 11, 2022
Verroens, Filip; De Cuypere, Ludovic, 2022, "Replication Data for: French ingressives and (phasal) aspect. A frame-semantic corpus-based analysis", https://doi.org/10.18710/WVW9U4, DataverseNO, V1
Dataset abstract The dataset includes an annotated corpus sample of N = 2000 French sentences with se mettre à or commencer à (1000 observations of each verb). The sample was drawn from the literary corpus Frantext (FT) and the journalistic corpus Le Monde (1000 observations from both corpora). The sample is balanced for verb as well as corpus, so...
Add Data

Log in to create a dataverse or add a dataset.

Share Dataverse

Share this dataverse on your favorite social media networks.

Link Dataverse
Reset Modifications

Are you sure you want to reset the selected metadata fields? If you do this, any customizations (hidden, required, optional) you have done will no longer appear.