跳到主要导航 跳到搜索 跳到主要内容

A New Deep Wavefront Based Model for Text Localization in 3D Video

  • Lokesh Nandanwar
  • , Palaiahnakote Shivakumara
  • , Raghavendra Ramachandra
  • , Tong Lu*
  • , Umapada Pal
  • , Apostolos Antonacopoulos
  • , Yue Lu
  • *此作品的通讯作者
  • University of Malaya
  • Norwegian University of Science and Technology
  • Nanjing University
  • Indian Statistical Institute
  • University of Salford

科研成果: 期刊稿件文章同行评审

摘要

With the evolution of electronic devices, such as 3D cameras, addressing the challenges of text localization in 3D video (e.g., for indexing) is increasingly drawing the attention of the multimedia and video processing community. Existing methods focus on 2D video and their performance in the presence of the challenges in 3D video, such as shadow areas associated with text and irregularly sized and shaped text, degrades. This paper proposes the first approach that successfully addresses the challenges of 3D video in addition to those of 2D. It employs a number of innovations, among which, the first is the Generalized Gradient Vector Flow (GGVF) for dominant points detection. The second is the Wavefront concept for text candidate point detection from those dominant points. In addition, an Adaptive B-Spline Polygon Curve Network (ABS-Net) is proposed for accurate text localization in 3D videos by constructing tight fitting bounding polygons using text candidate points. Extensive experiments on custom (3D video) and standard datasets (2D video and scene text) show that the proposed method is practical and useful, and overall outperforms existing state-of-the-art methods.

源语言英语
页(从-至)3375-3389
页数15
期刊IEEE Transactions on Circuits and Systems for Video Technology
32
6
DOI
出版状态已出版 - 1 6月 2022

指纹

探究 'A New Deep Wavefront Based Model for Text Localization in 3D Video' 的科研主题。它们共同构成独一无二的指纹。

引用此