Back to knowledge graph
science
machine-learning
Confidence 75%

Visual attention network

Visual attention network While originally designed for natural language processing tasks, the self-attention mechanism has recently taken various computer vision areas by storm. However, the 2D nature of images brings three challenges for applying self-attention in computer vision: (1) treating images as 1D sequences neglects their 2D structures; (2) the quadratic complexity is too expensive for high-resolution images; (3) it only captures spatial adaptability but ignores channel adaptability. ...

Anonymous preview shows an excerpt only. Sign in to read the full item.

Cited 1030 times
Visual attention network | Awareness Public Knowledge