Constarium
← Search

Data · dataset · 2026

FSAR-Cap:Large-scale fine-grained SAR image captioning dataset

Listed in ScienceDB

The FSAR Cap dataset aims to build an image text reference corpus with fine semantic description capabilities for SAR image semantic understanding and cross modal modeling, promoting the development of SAR image automatic interpretation, image captioning, and remote sensing multimodal models.

Description

This dataset is constructed based on the FAIR-CSAR object detection dataset, consisting of 14480 SAR images and 72400 accompanying descriptive texts.

FSAR Cap adopts a two-stage annotation method: first, it utilizes detection results and spatial location information to automatically generate basic descriptions through multiple templates; Then, combined with manual verification and language model polishing. In the end, each image will generate 5 descriptive sentences with different styles and complementary information, covering target types, quantities, positional relationships, external features, etc. As the first large-scale semantic annotation dataset with fine-grained description hierarchy for SAR images, FSAR Cap not only improves the semantic expression quality of SAR images, but also provides a unified and high-quality data benchmark for image captioning, remote sensing visual language model training, multimodal inference, and SAR natural language alignment research in the SAR field, laying the foundation for the further development of SAR automated interpretation and intelligent semantic understanding technology system. 

Links

Where it is published

Catalogue records · 1

Topics

Inferred from text
Image 75% · Satellite remote sensing 65% · Text 75%
Provenance · 1 source records, 13 field assertions
SourceKeyLast seenRaw
ScienceDB10.57760/sciencedb.radars.001018 d agoJSON v1
FieldAssertionExtractorEvidence
access_levelsource · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:earth-environmentalmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:engineeringmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:humanitiesmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:life-sciencesmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:social-sciencemapping · scidb cnconnector:scidb_cn@1.0.0
concepts[modality].local:modality:imageenrichment · scidb cnkeyword-concept-rules@1.0.0title+description (75%)
concepts[modality].local:modality:remote-sensingenrichment · scidb cnkeyword-concept-rules@1.0.0title+description (65%)
concepts[modality].local:modality:textenrichment · scidb cnkeyword-concept-rules@1.0.0title+description (75%)
descriptionsource · scidb cnconnector:scidb_cn@1.0.0/metadata/dc/description
licensesource · scidb cnconnector:scidb_cn@1.0.0/metadata/dc/rights
publication_datesource · scidb cnconnector:scidb_cn@1.0.0
titlesource · scidb cnconnector:scidb_cn@1.0.0/metadata/dc/title