Investigating transformers in the decomposition of polygonal shapes as point collections

Alfieri, Andrea; Lin, Yancong; van Gemert, Jan C.

Computer Science > Computer Vision and Pattern Recognition

arXiv:2108.07533 (cs)

[Submitted on 17 Aug 2021]

Title:Investigating transformers in the decomposition of polygonal shapes as point collections

Authors:Andrea Alfieri, Yancong Lin, Jan C. van Gemert

View PDF

Abstract:Transformers can generate predictions in two approaches: 1. auto-regressively by conditioning each sequence element on the previous ones, or 2. directly produce an output sequences in parallel. While research has mostly explored upon this difference on sequential tasks in NLP, we study the difference between auto-regressive and parallel prediction on visual set prediction tasks, and in particular on polygonal shapes in images because polygons are representative of numerous types of objects, such as buildings or obstacles for aerial vehicles. This is challenging for deep learning architectures as a polygon can consist of a varying carnality of points. We provide evidence on the importance of natural orders for Transformers, and show the benefit of decomposing complex polygons into collections of points in an auto-regressive manner.

Comments:	DLGC@ICCVW 2021
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2108.07533 [cs.CV]
	(or arXiv:2108.07533v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2108.07533

Submission history

From: Yancong Lin [view email]
[v1] Tue, 17 Aug 2021 09:36:24 UTC (966 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CV

< prev | next >

new | recent | 2021-08

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Jan C. van Gemert

export BibTeX citation

Computer Science > Computer Vision and Pattern Recognition

Title:Investigating transformers in the decomposition of polygonal shapes as point collections

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Investigating transformers in the decomposition of polygonal shapes as point collections

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators