Multi-Modal Human-Aware Image Caption System for Intelligent Service Robotics Applications
Ren C. Luo, Yu‐Ting Hsu, Huan-Jun Ye
- 发表年份
- 2019
- 引用次数
- 11
摘要
Image captioning is a high-level task that generates the context scenario descriptions from an image. There are many ways to implement such kinds of ability that can generate a smooth sentence about the context scenario. Because of the open and extensive domain of the dataset, seldom of them can really help people to get the meaningful information of a specified scenario. In this paper, we propose a novel framework called Human-Aware Context Generator (HACG). We adopt the concept of dividing and conquer, which combines the ability of face recognition and facial expression recognition while retaining the capability of image caption model. This model aims to offer a significant context scenario sentence to the people, especially for those who are urgent to master the overall situation.
关键词
相关论文
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Fractional Differential Equations
Igor Podlubný
2025
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991