Data · dataset · 2026
Appearance of sentence adverbs in Wh-questions
Listed in REDU - Unicamp Institutional Research Data Repository
Description
The data reported herein form part of the FAPESP-funded project (23/16142-0) entitled "Diagnostics for Adverbs: (Micro)variation and Cartography" and test, in three varieties of Portuguese (Angolan (AO), Brazilian (BP) and Mozambican (MOZ)) and in two varieties of Spanish (Chilean (CS) and Peruvian (PS)), a well-documented property of sentence adverbs: their incompatibility with Wh-questions. The dataset is organised into four zipped folders, grouped by language family and by file type: - "2 Corpus Appearance of Sentence Adverbs in Yes-No Questions (AO, BP, MOZ).zip" — Corpus files containing the sentences tested in the Wh-question context in AO, BP and MOZ, already judged by native speakers of each variety, each with its respective metadata page. - "2 Corpus Appearance of Sentence Adverbs in Yes-No Questions (CS, PS).zip" — Corpus files containing the sentences tested in the Wh-question context in CS and PS, already judged by native speakers of each variety, each with its respective metadata page. - "2 Prompts Appearance of Sentence Adverbs in Yes-No Questions (AO, BP, MOZ).zip" — Prompt files provided to the AI for the generation and/or linguistic and cultural adaptation of the sentences tested in AO, BP and MOZ.
Each prompt was carefully prepared by the project executor, Prof. Aquiles Tescari Neto, and specifies the nature of the task to be undertaken by the AI, the target construction and the type of sentences to be generated. In this project, the AI was used solely to generate sentences for Brazilian Portuguese and to adapt them to the linguistic and cultural realities of Angola and Mozambique. - "2 Prompts Appearance of Sentence Adverbs in Yes-No Questions (CS, PS).zip"— Prompt files provided to the AI for the generation, translation and/or linguistic and cultural adaptation of the sentences tested in CS and PS, following the same procedure as described for the Portuguese varieties.
Read the rest (1 more)
Once the data were generated, translated and/or adapted by the AI, they were carefully reviewed by the project executor, Prof. Aquiles Tescari Neto, and submitted to acceptability/grammaticality judgements by native speakers of the respective varieties, following standard practice in Generative Grammar.
Links
Where it is published
- Dataverse dataset page redu.unicamp.br/dataset.xhtml?persistentId=doi%3A10.25824%2Fredu%2FNVYJUD ↗
landing page · from redu unicamp br
- DOI doi.org/10.25824/redu/nvyjud ↗
DOI / persistent id · from redu unicamp br
Catalogue records · 1
- Dataverse API redu.unicamp.br/api/datasets/:persistentId/?persistentId=doi%3A10.25824%2Fredu… ↗
metadata API · from redu unicamp br
Topics
- Stated by source
- Arts and Humanities
- From keywords
- Humanities
- Inferred from text
- Linguistics 73%
Provenance · 1 source records, 9 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| REDU - Unicamp Institutional Research Data Repository | doi:10.25824/redu/NVYJUD | 5 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| concepts[field].anzsrc:group:4704 | enrichment · redu unicamp br | taxonomy-embedding@1.1.0 | title+keywords+description (73%) |
| concepts[field].dataverse_subject:arts-and-humanities | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /subjects |
| concepts[field].local:field:humanities | mapping · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /subjects |
| created_date | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | |
| description | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /description |
| publication_date | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | |
| title | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /name |
| updated_date | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | |
| version_label | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 |