عودة إلى خريطة المعرفة
science
machine-learning
الضمانة 75%

الخلاصة

Visual attention network While originally designed for natural language processing tasks, the self-attention mechanism has recently taken various computer vision areas by storm. However, the 2D nature of images brings three challenges for applying self-attention in computer vision: (1) treating images as 1D sequences neglects their 2D structures; (2) the quadratic complexity is too expensive for high-resolution images; (3) it only captures spatial adaptability but ignores channel adaptability. ...

المصدر:
تم الإشارة إليه 1030 مرة