CGAT-Net: Context-Aware Graph Attention Transformer Network for Scene Sketch Recognition
Description
Sketches often lack sufficient detail or quality necessary for standalone recognition, making their identification challenging without contextual information. While context understanding is commonly studied in computer vision applications like object detection or image classification, it remains under-explored in the sketch domain. Existing research primarily focuses on recognizing sketch objects in isolation, with little attention given to scene-level sketch understanding. To address this gap, we introduce a Context-Aware Graph Attention Transformer Network (CGAT-Net), which leverages visual and spatial relationships among objects to obtain a more accurate classification within a scene. This is the first study in scene sketch recognition that utilizes object relations in a Transformer-based network to incorporate context understanding. Extensive experiments show that CGAT-Net surpasses current state-of-the-art single-sketch classifiers, underscoring the value of contextual information in enhancing individual sketch recognition. Our code and trained model weights can be accessed from https://github.com/aleynakutuk6/CGAT-Net.
Files
bib-3331de60-7b91-4aa8-98eb-f746f7b827e5.txt
Files
(196 Bytes)
| Name | Size | Download all |
|---|---|---|
|
md5:79919722a1c6ef7867e1617f0307993f
|
196 Bytes | Preview Download |