详细信息

EMT-NAS: Transferring architectural knowledge between tasks from different datasets  ( EI收录)  

文献类型:期刊文献

英文题名:EMT-NAS: Transferring architectural knowledge between tasks from different datasets

作者:Liao, Peng[1]; Jin, Yaochu[1,2]; Du, Wenli[1]

机构:[1] ECUST, Key Laboratory of Smart Manufacturing in Energy Chemical Process, Ministry of Education, China; [2] Bielefeld University, Faculty of Technology, Germany

年份:2023

卷号:2023-June

起止页码:3643

外文期刊名:Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition

收录:EI(收录号:20250817900321)

语种:英文

外文关键词:Adversarial machine learning - Deep learning - Multi-task learning

摘要:The success of multitask learning (MTL) can largely be attributed to the shared representation of related tasks, allowing the models to better generalise. In deep learning, this is usually achieved by sharing a common neural network architecture and jointly training the weights. However, the joint training of weighting parameters on multiple related tasks may lead to performance degradation, known as negative transfer. To address this issue, this work proposes an evolutionary multitasking neural architecture search (EMT-NAS) algorithm to accelerate the search process by transferring architectural knowledge across multiple related tasks. In EMT-NAS, unlike the traditional MTL, the model for each task has a personalised network architecture and its own weights, thus offering the capability of effectively alleviating negative transfer. A fitness re-evaluation method is suggested to alleviate fluctuations in performance evaluations resulting from parameter sharing and the mini-batch gradient descent training method, thereby avoiding losing promising solutions during the search process. To rigorously verify the performance of EMT-NAS, the classification tasks used in the empirical assessments are derived from different datasets, including the CIFAR-10 and CIFAR-100, and four MedMNIST datasets. Extensive comparative experiments on different numbers of tasks demonstrate that EMT-NAS takes 8% and up to 40% on CIFAR and MedMNIST, respectively, less time to find competitive neural architectures than its single-task counterparts. ? 2023 IEEE.

参考文献:

正在载入数据...

版权所有©华东理工大学 重庆维普资讯有限公司 渝B2-20050021-7 
渝公网安备 50019002500408号 违法和不良信息举报中心