详细信息
Generation of Synthetic Images of Randomly Stacked Object Scenes for Network Training Applications ( SCI-EXPANDED收录)
文献类型:期刊文献
英文题名:Generation of Synthetic Images of Randomly Stacked Object Scenes for Network Training Applications
作者:Zhang, Yajun[1];Yi, Jianjun[1];Zhang, Jiahao[1];Chen, Yuanhao[1];He, Liang[2]
机构:[1]East China Univ Sci & Technol, Shanghai 200237, Peoples R China;[2]Shanghai Aerosp Control Technol Inst, Shanghai 201109, Peoples R China
年份:2021
卷号:27
期号:2
起止页码:425
外文期刊名:INTELLIGENT AUTOMATION AND SOFT COMPUTING
收录:;WOS:【SCI-EXPANDED(收录号:WOS:000624954000009)】;
基金:Financial support for this work was provided by the National Natural Science Foundation of China [Grant No. 51575186, Jianjun Yi, www.nsfc.gov.cn] and the Shanghai Science and Technology Action Plan [Grant Nos. 18DZ1204000, 18510745500, and 18510730600, Jianjun Yi, www.sh.gov.cn].
语种:英文
外文关键词:Synthetic dataset; stacked object scenes; OpenGL; Bullet physics engine; image recognition; parts position
摘要:Image recognition algorithms based on deep learning have been widely developed in recent years owing to their capability of automatically capturing recognition features from image datasets and constantly improving the accuracy and efficiency of the image recognition process. However, the task of training deep learning networks is time-consuming and expensive because large training datasets are generally required, and extensive manpower is needed to annotate each of the images in the training dataset to support the supervised learning process. This task is particularly arduous when the image scenes involve randomly stacked objects. The present work addresses this issue by developing a synthetic training dataset generation method based on OpenGL and the Bullet physics engine which can automatically generate annotated synthetic datasets by simulating the freefall of a collection of objects under the force of gravity. Rigorous statistical comparison of a real image dataset of staked scenes with a synthetic image dataset generated by the proposed approach demonstrates that the two datasets exhibit no significant differences. Moreover, the object detection performances obtained by three popular network architectures trained using the synthetic dataset generated by the proposed approach are demonstrated to be much better than the results of training conducted using a synthetic dataset generated by a conventional cut and paste approach, and these performances are also competitive with the results of training conducted using a dataset composed of real images.
参考文献:
正在载入数据...
