Constarium
← Search

Data · dataset · 2026

SLANGUAGE Tiv Parallel Corpus v1

Listed in Hugging Face Datasets

SLANGUAGE Tiv Parallel Corpus v1 Dataset Description The SLANGUAGE Tiv Parallel Corpus is the first commercially developed parallel English–Tiv text dataset, created by STECH Global Limited through the SAVA Languages (SLANGUAGE) project.

Description

Tiv is a tonal Bantoid language spoken by approximately 8 million people, primarily in Benue State, Nigeria. Despite its significant speaker population, Tiv remains severely underrepresented in AI and NLP research globally.

This… See the full description on the dataset page: huggingface.co/datasets/SLANGUAGE/SLANGUAGE_TIV_001.

Links

Get the data

Catalogue records · 1

Topics

Stated by source
translation
Inferred from text
Text 75%
Provenance · 1 source records, 10 field assertions
SourceKeyLast seenRaw
Hugging Face DatasetsSLANGUAGE/SLANGUAGE_TIV_0018 d agoJSON v1
FieldAssertionExtractorEvidence
access_levelsource · Hugging Faceconnector:huggingface@1.0.0/gated
concepts[field].local:field:computer-science-aimapping · Hugging Faceconnector:huggingface@1.0.0
concepts[modality].local:modality:textenrichment · Hugging Facekeyword-concept-rules@1.0.0title+description (75%)
concepts[task].hf_task:translationsource · Hugging Faceconnector:huggingface@1.0.0/tags[task_categories:*]
created_datesource · Hugging Faceconnector:huggingface@1.0.0
descriptionsource · Hugging Faceconnector:huggingface@1.0.0/description
licensesource · Hugging Faceconnector:huggingface@1.0.0/tags[license:*]
publication_datesource · Hugging Faceconnector:huggingface@1.0.0
titlesource · Hugging Faceconnector:huggingface@1.0.0/id
updated_datesource · Hugging Faceconnector:huggingface@1.0.0