Constarium
← Search

Data · dataset · 2026

TGCNet: A Fusion Double-Branch Network with Transformer Guiding CNN for Semantic Segmentation of Remote Sensing Images

Listed in ScienceDB

In recent years, various combined CNN and Transformer networks have achieved ideal results in the field of remote sensing image semantic segmentation.

Description

Such networks often use CNN and Transformer to obtain local details and global dependencies of the image, respectively. In order to further make the model pay more attention to the local features of remote sensing images, and solve the problems such as omission and wrong recognition caused by the difficulty of edge information extraction in the semantic segmentation task of remote sensing images, in this paper, a fusion double-branch network with Transformer giding CNN (TGCNet) is proposed.

The whole frame of the model adopts encoder-decoder structure. In the encoder, Transformer branch in addition to normal coding, the global information acquired by Transformer is stratifiedly fused with the local spatial details corresponding to each layer for feature fusion, and input to the next layer of CNN coding. In addition, there is a slight difference between the feature output of Transformer and that of CNN, and feature loss may occur when feature fusion is carried out between the two outputs.

Read the rest (2 more)

In order to reduce or eliminate such feature difference, we adopt the idea of making the structure of the two branches as consistent as possible. Taking NCB block as the CNN branch of the model solves this problem well. In order to prove the effectiveness of TGCNet in semantic segmentation of remote sensing images, we compared TGCNet with seven other classical semantic segmentation models on WHU building dataset and CHN6-CUG roads dataset.The experimental results show that the mIoU of TGCNet on the two datasets is 90.23% and 82.88%, respectively, which is better than the other seven models.

At The same time, the parametric quantities of TGCNet is 4.76M and FLOPs is 16.71G, better than any other network except SETR.

Links

Where it is published

Catalogue records · 1

Topics

Inferred from text
Image 75% · Machine learning 71% · Satellite remote sensing 65%
Provenance · 1 source records, 14 field assertions
SourceKeyLast seenRaw
ScienceDB10.57760/sciencedb.347469 d agoJSON v1
FieldAssertionExtractorEvidence
access_levelsource · scidb cnconnector:scidb_cn@1.0.0
concepts[field].anzsrc:group:4611enrichment · scidb cntaxonomy-embedding@1.0.0title+keywords+description (71%)
concepts[field].local:field:computer-science-aimapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:earth-environmentalmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:engineeringmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:humanitiesmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:life-sciencesmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:social-sciencemapping · scidb cnconnector:scidb_cn@1.0.0
concepts[modality].local:modality:imageenrichment · scidb cnkeyword-concept-rules@1.0.0title+description (75%)
concepts[modality].local:modality:remote-sensingenrichment · scidb cnkeyword-concept-rules@1.0.0title+description (65%)
descriptionsource · scidb cnconnector:scidb_cn@1.0.0/metadata/dc/description
licensesource · scidb cnconnector:scidb_cn@1.0.0/metadata/dc/rights
publication_datesource · scidb cnconnector:scidb_cn@1.0.0
titlesource · scidb cnconnector:scidb_cn@1.0.0/metadata/dc/title