Detailed Information

Cited 0 time in webofscience Cited 0 time in scopus
Metadata Downloads

Deep Transformer Based Video Inpainting Using Fast Fourier Tokenizationopen access

Authors
Kim, TaewanKim, JinwooOh, HeeseokKang, Jiwoo
Issue Date
Feb-2024
Publisher
IEEE-INST ELECTRICAL ELECTRONICS ENGINEERS INC
Keywords
video completion; free-form inpainting; object removal; adversarial learning
Citation
IEEE ACCESS, v.12, pp 21723 - 21736
Pages
14
Journal Title
IEEE ACCESS
Volume
12
Start Page
21723
End Page
21736
URI
https://scholarworks.sookmyung.ac.kr/handle/2020.sw.sookmyung/159847
DOI
10.1109/ACCESS.2024.3361283
ISSN
2169-3536
Abstract
Bridging distant space-time interactions is important for high-quality video inpainting with large moving masks. Most existing technologies exploit patch similarities within the frames, or leaverage large-scale training data to fill the hole along spatial and temporal dimensions. Recent works introduce promissing Transformer architecture into deep video inpainting to escape from the dominanace of nearby interactions and achieve superior performance than their baselines. However, such methods still struggle to complete larger holes containing complicated scenes. To alleviate this issue, we first employ a fast Fourier convolutions, which cover the frame-wide receptive field, for token representation. Then, the token passes through the seperated spatio-temporal transformer to explicitly moel the long-range context relations and simultaneously complete the missing regions in all input frames. By formulating video inpainting as a directionless sequence-to-sequence prediction task, our model fills visually consistent content, even under conditions such as large missing areas or complex geometries. Furthermore, our spatio-temporal transformer iteratively fills the hole from the boundary enabling it to exploit rich contextual information. We validate the superiority of the proposed model by using standard stationary masks and more realistic moving object masks. Both qualitative and quantitative results show that our model compares favorably against the state-of-the-art algorithms.
Files in This Item
Go to Link
Appears in
Collections
ICT융합공학부 > IT공학전공 > 1. Journal Articles

qrcode

Items in ScholarWorks are protected by copyright, with all rights reserved, unless otherwise indicated.

Related Researcher

Researcher Kang, Jiwoo photo

Kang, Jiwoo
공과대학 (인공지능공학부)
Read more

Altmetrics

Total Views & Downloads

BROWSE