跳到主要导航 跳到搜索 跳到主要内容

HSNet: hierarchical semantics network for scene parsing

  • Xin Tan
  • , Jiachen Xu
  • , Ying Cao*
  • , Ke Xu
  • , Lizhuang Ma
  • , Rynson W.H. Lau*
  • *此作品的通讯作者
  • Shanghai Jiao Tong University
  • City University of Hong Kong

科研成果: 期刊稿件文章同行评审

摘要

Scene parsing is one of the fundamental tasks in computer vision. Humans tend to perceive a scene in a hierarchical manner, i.e., first identifying the coarse category (e.g., vehicle) of a group of objects and then the fine category (e.g., bicycle, truck or car) of each of them. Despite recent tremendous progress on scene parsing, such a hierarchical semantics prior (HSP) has not been explicitly exploited. In this paper, we aim to introduce the HSP into scene parsing, by proposing a hierarchical semantics network (HSNet). Our key contribution is a bidirectional cross-level feature matching framework, which enables us to learn multi-level, hierarchy-aware features via forward feature transfer and backward feature regularization. In the forward stage, we train a coarse-to-fine module to learn fine-category features that explicitly encode hierarchical semantics information. In the backward stage, we introduce a fine-to-coarse module to collapse fine-category features to coarse-category features that are used to regularize the feature learning of our network. Experimental results on Cityscapes and Pascal Context show that our method achieves state-of-the-art performances. Our visualization also shows that our learned features capture semantic hierarchy favorably.

源语言英语
页(从-至)2543-2554
页数12
期刊Visual Computer
39
7
DOI
出版状态已出版 - 7月 2023
已对外发布

指纹

探究 'HSNet: hierarchical semantics network for scene parsing' 的科研主题。它们共同构成独一无二的指纹。

引用此