手语的视觉语义-结构化解析方法研究

Studies on Visual Semantic-Structural Parsing of Sign Languages

  • 摘要: 人类用视觉认知手语,是通过不变的结构来描述物体各部分及其关系,实现从整体到部分的拆解或各部分的重新组合.本文通过模拟听障人对手语的认知方式,先计算其视野中的视觉显著区域,再将注意力焦点种子集中于手部提高其有效性,将生成的语义超像素作为部分,再把部分映射为手指骨架或将手掌缩为一个关节点,而得到子手语集.按解剖结构拆解手语图像(状态),初步实现了对一组相似手语的视觉语义-结构化解析.

     

    Abstract: Sign languages are visually perceived, which describes the parts of an object and their relations based on the invariable structure. It is realized by disassembling the whole into parts and subsequently recombining them. In our research, the salient visual regions were firstly calculated in the views of the hearing-impaired person by simulating their cognitive mode about sign languages. Afterwards, the visual seeds of focus were concentrated on the hands to improve its effectiveness. Finally, we obtained the sub-sign language by making the generated semantic superpixels as the parts, which were mapped into the finger skeleton, or shrinking the palm to a node. By disassembling the images (states) of sign languages according to their anatomical structures, the visual semantic-structural parsing is preliminarily realized for a set of words of the similar sign language.

     

/

返回文章
返回