Transcription of SeqFormer: a Frustratingly Simple Model for Video Instance ...
{{id}} {{{paragraph}}}
SeqFormer: a Frustratingly Simple Model for Video Instance Segmentation Junfeng Wu1 Yi Jiang2 Wenqing Zhang1 Xiang Bai1 Song Bai2. 1 2. Huazhong University of Science and Technology ByteDance [ ] 15 Dec 2021. 60. Abstract SeqFormer 55. In this work, we present SeqFormer, a Frustratingly sim- YouTube-VIS 2019 AP. ple Model for Video Instance segmentation. SeqFormer fol- 50. Propose-Reduce lows the principle of vision transformer that models in- stance relationships among Video frames. Nevertheless, we 45 IFC. observe that a stand-alone Instance query suffices for cap- 40 VisTR. turing a time sequence of instances in a Video , but atten- CrossVIS.
ing vision task that aims to simultaneously perform detec-tion, classification, segmentation, and tracking of object in-stances in videos. Compared to image instance segmenta-tion [6], video instance segmentation is much more chal-lenging since it requires accurate tracking of objects across an entire video.
Domain:
Source:
Link to this page:
Please notify us if you found a problem with this document:
{{id}} {{{paragraph}}}