Reality Capture Technologies (LiDAR, RGB-D, Vision)
Reality Capture Technologies (LiDAR, RGB-D, Vision)
复制标题
现实捕捉技术(LiDAR、RGB-D、视觉)
DOI:
10.1061/9780784482438.019
复制
发表时间:
2019
期刊:
影响因子:
--
通讯作者:
Kamat, Vineet R.
中科院分区:
文献类型:
--
作者:
Liang, Ci-Jyun;Lundeen, Kurt M.;McGee, Wes;Menassa, Carol C.;Lee, SangHyun;Kamat, Vineet R.
Struck-by accidents are potential safety concerns on construction sites and require a robust machine pose estimation. The development of deep learning methods has enhanced the human pose estimation that can be adapted for articulated machines. These methods require abundant dataset for training, which is challenging and time-consuming to obtain on-site. This paper proposes a fast data collection approach to build the dataset for excavator pose estimation. It uses two industrial robot arms as the excavator and the camera monopod to collect different excavator pose data. The 3D annotation can be obtained from the robot's embedded encoders. The 2D pose is annotated manually. For evaluation, 2,500 pose images were collected and trained with the stacked hourglass network. The results showed that the dataset is suitable for the excavator pose estimation network training in a controlled environment, which leads to the potential of the dataset augmenting with real construction site images.