상세 보기
Empirical Comparison of Deep Learning Networks on Backbone Method of Human Pose Estimation
- 림빈보니카;
- 김준섭;
- 최유주;
- 홍민
초록
Accurate estimation of human pose relies on backbone method in which its role is to extract feature map. Up to dated, the method of backbone feature extraction is conducted by the plain convolutional neural networks named by CNN and the residual neural networks named by Resnet, both of which have various architectures and performances. The CNN family network such as VGG which is well-known as a multiple stacked hidden layers architecture of deep learning methods, is base and simple while Resnet which is a bottleneck layers architecture yields fewer parameters and outperform. They have achieved inspired results as a backbone network in human pose estimation. However, they were used then followed by different pose estimation networks named by pose parsing module. Therefore, in this paper, we present a comparison between the plain CNN family network (VGG) and bottleneck network (Resnet) as a backbone method in the same pose parsing module. We investigate their performances such as number of parameters, loss score, precision and recall. We experiment them in the bottom-up method of human pose estimation system by adapted the pose parsing module of openpose. Our experimental results show that the backbone method using VGG network outperforms the Resent network with fewer parameter, lower loss score and higher accuracy of precision and recall.
키워드
- 제목
- Empirical Comparison of Deep Learning Networks on Backbone Method of Human Pose Estimation
- 제목 (타언어)
- Empirical Comparison of Deep Learning Networks on Backbone Method of Human Pose Estimation
- 저자
- 림빈보니카; 김준섭; 최유주; 홍민
- 발행일
- 2020
- 저널명
- 인터넷정보학회논문지
- 권
- 21
- 호
- 5
- 페이지
- 21 ~ 29