Data · dataset · 2026
Adverb + COMP construction & adverbial clefts (Chilean and Peruvian Spanish)
Listed in REDU - Unicamp Institutional Research Data Repository
Description
The data reported herein, which form part of the FAPESP-funded project (23/16142-0), entitled “Diagnostics for Adverbs: (Micro)variation and Cartography”, test, in two varieties of Spanish (Chilean (CS) and Peruvian (PS), two well-documented properties of sentence adverbs: 1. their appearance before the complementiser in Adverb + COMP constructions and 2. their appearance in adverbial clefts. This dataset comprises two types of files: 1.
Files containing the prompts provided to the AI for the generation and linguistic and cultural adaptation of the sentences to be judged by native speakers. These files are identified by the word “Prompts” in the title and by the acronyms for each language/variety: CS (Chilean Spanish) and PS (Peruvian Spanish). Each file contains the prompt carefully prepared by the project executor, Prof.
Read the rest (3 more)
Aquiles Tescari Neto. The prompts specify the nature of the task to be undertaken by the AI, are theoretically oriented towards the target construction, and stipulate the type of sentences to be generated by the AI. In the FAPESP project within which these data were produced, the AI was used only to generate sentences and/or to adapt the generated sentences to the linguistic and cultural realities of Chile and Peru. 2.
Files containing the sentences already judged by native speakers of these varieties. Each file includes, in addition to the judged sentences, a page with the project metadata, with separate files containing the data for each language/variety. That is, there is one file per variety, each with its respective metadata.
These files are identified by the acronyms CS and PS preceded by “Corpus”. Once the data were generated and/or adapted by the AI, they were carefully reviewed by the project executor, Prof. Aquiles Tescari Neto, and submitted to acceptability/grammaticality judgements by native speakers of the respective varieties.
Links
Where it is published
- Dataverse dataset page redu.unicamp.br/dataset.xhtml?persistentId=doi%3A10.25824%2Fredu%2FI3QAVN ↗
landing page · from redu unicamp br
- DOI doi.org/10.25824/redu/i3qavn ↗
DOI / persistent id · from redu unicamp br
Catalogue records · 1
- Dataverse API redu.unicamp.br/api/datasets/:persistentId/?persistentId=doi%3A10.25824%2Fredu… ↗
metadata API · from redu unicamp br
Topics
- Stated by source
- Arts and Humanities
- From keywords
- Humanities
- Inferred from text
- Linguistics 74%
Provenance · 1 source records, 9 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| REDU - Unicamp Institutional Research Data Repository | doi:10.25824/redu/I3QAVN | 6 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| concepts[field].anzsrc:group:4704 | enrichment · redu unicamp br | taxonomy-embedding@1.1.0 | title+keywords+description (74%) |
| concepts[field].dataverse_subject:arts-and-humanities | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /subjects |
| concepts[field].local:field:humanities | mapping · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /subjects |
| created_date | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | |
| description | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /description |
| publication_date | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | |
| title | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /name |
| updated_date | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | |
| version_label | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 |