Datasets and R scripts for modelling Czech translation counterparts of Romance causative constructions

PID

This repository contains the datasets and code used in the study “Predicting translation counterparts in causative constructions.” The datasets consist of annotated examples of Italian and Spanish causative constructions and their Czech translation counterparts. The repository includes (i) full annotated datasets for Italian and Spanish, (ii) revised datasets used for statistical modelling, and (iii) the R script used to estimate Bayesian multinomial regression models using the brms package (Stan backend). The models estimate the probability of selecting a Czech translation counterpart (TYPE) as a function of verb valency (VALENCY) and complement class (COMP_CLASS), with random effects for VERB and TRANSLATOR. The repository also contains summaries of the fitted models.

Identifier
PID http://hdl.handle.net/11234/1-5841
Metadata Access http://lindat.mff.cuni.cz/repository/oai/request?verb=GetRecord&metadataPrefix=oai_dc&identifier=oai:lindat.mff.cuni.cz:11234/1-5841
Provenance
Creator Štichauer,Pavel; Čermák, Petr
Publisher Filozofická fakulta, Univerzita Karlova
Publication Year 2026
Rights Creative Commons - Attribution 4.0 International (CC BY 4.0); http://creativecommons.org/licenses/by/4.0/; PUB
OpenAccess true
Contact lindat-help(at)ufal.mff.cuni.cz
Representation
Language Italian; Spanish; Castilian
Resource Type lexicalConceptualResource
Format application/vnd.openxmlformats-officedocument.spreadsheetml.sheet; text/plain; application/octet-stream; text/csv; text/plain; charset=utf-8; downloadable_files_count: 7
Discipline Linguistics