{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2024,7,30]],"date-time":"2024-07-30T02:50:23Z","timestamp":1722307823017},"reference-count":72,"publisher":"World Scientific Pub Co Pte Lt","issue":"06","funder":[{"name":"General Technology Fundamental Research United Fund","award":["U1736211"]},{"DOI":"10.13039\/501100003453","name":"Natural Science Foundation of Guangdong Province","doi-asserted-by":"crossref","award":["2019A1515011076"],"id":[{"id":"10.13039\/501100003453","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Innovation Group of Guangdong Education Department","award":["2020KCXTD014","2018KCXTD019"]},{"name":"National Natural Science Foundation of Hubei Province","award":["2018CFA024"]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["61702280","61933013","62041603","62076139"],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"National Natural Science Foundation of Jiangsu Province","award":["BK20170900"]},{"DOI":"10.13039\/501100012152","name":"National Postdoctoral Program for Innovative Talents","doi-asserted-by":"crossref","award":["BX20180146"],"id":[{"id":"10.13039\/501100012152","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100002858","name":"China Postdoctoral Science Foundation","doi-asserted-by":"crossref","award":["2019M661901"],"id":[{"id":"10.13039\/501100002858","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100010242","name":"Jiangsu Planned Projects for Postdoctoral Research Funds","doi-asserted-by":"crossref","award":["2019K024","2020M671678"],"id":[{"id":"10.13039\/501100010242","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Postdoctoral Science Foundation of Zhejiang Province","award":["ZJ2019072"]},{"DOI":"10.13039\/501100004479","name":"Natural Science Foundation of Jiangxi Province","doi-asserted-by":"crossref","award":["20202BABL202036"],"id":[{"id":"10.13039\/501100004479","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Zhejiang Lab","award":["20210AB03"]},{"name":"Postgraduate Research and Practice Innovation Program of Jiangsu Province","award":["KYCX17_0794"]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Int. J. Soft. Eng. Knowl. Eng."],"published-print":{"date-parts":[[2021,6]]},"abstract":" The heterogeneous defect prediction (HDP) technique can predict defects in a target company using heterogeneous metric data from external company, which has received substantial research attention. However, existing HDP methods assume that source data is labeled but labeling data is expensive. Semi-supervised defect prediction technique can perform defect prediction with few labeled data. In this paper, we investigate a new problem\u00a0\u2014 semi-supervised HDP (SHDP). To solve this problem, we propose a new approach named cost-sensitive kernel semi-supervised correlation analysis (CKSCA) as a solution of SHDP problem. It introduces unified metric representation and canonical correlation analysis to make the data distributions of different company projects more similar. CKSCA also designs a cost-sensitive kernel semi-supervised discriminant analysis mechanism to utilize the limited labeled data and sufficient real-life unlabeled data from different companies. Besides we collect lots of open-source projects from GitHub website to construct a new large-scale unlabeled dataset called GITHUB dataset. It contains 26,407 modules and is greater than each public project dataset. It has been public online and can be extended continuously. Experiments on the GITHUB dataset and other public datasets indicate that unlabeled GITHUB data can help prediction model improve prediction performance, and CKSCA is effective and efficient for solving SHDP problem. <\/jats:p>","DOI":"10.1142\/s0218194021500273","type":"journal-article","created":{"date-parts":[[2021,6,22]],"date-time":"2021-06-22T03:43:23Z","timestamp":1624333403000},"page":"889-916","source":"Crossref","is-referenced-by-count":5,"title":["Semi-supervised Heterogeneous Defect Prediction with Open-source Projects on GitHub"],"prefix":"10.1142","volume":"31","author":[{"given":"Ying","family":"Sun","sequence":"first","affiliation":[{"name":"School of Computer Science, Nanjing University of Posts and Telecommunications, Nanjing 210023, P.\u00a0R.\u00a0China"}]},{"given":"Xiao-Yuan","family":"Jing","sequence":"additional","affiliation":[{"name":"School of Computer Science, Nanjing University of Posts and Telecommunications, Nanjing 210023, P.\u00a0R.\u00a0China"},{"name":"School of Computer Science, Wuhan University, Wuhan 430072, P.\u00a0R.\u00a0China"}]},{"given":"Fei","family":"Wu","sequence":"additional","affiliation":[{"name":"School of Computer Science, Nanjing University of Posts and Telecommunications, Nanjing 210023, P.\u00a0R.\u00a0China"}]},{"given":"Xiwei","family":"Dong","sequence":"additional","affiliation":[{"name":"School of Computer Science, Nanjing University of Posts and Telecommunications, Nanjing 210023, P.\u00a0R.\u00a0China"},{"name":"School of Computer and Big Data Science, Jiujiang University, Jiujiang 332005, P.\u00a0R.\u00a0China"}]},{"given":"Yanfei","family":"Sun","sequence":"additional","affiliation":[{"name":"School of Internet of Things, Nanjing University of Posts and Telecommunications, Nanjing 210023, P.\u00a0R.\u00a0China"},{"name":"Jiangsu Engineering Research Center of HPC and Intelligent Processing, Nanjing 210003, P.\u00a0R.\u00a0China"}]},{"given":"Ruchuan","family":"Wang","sequence":"additional","affiliation":[{"name":"School of Computer Science, Nanjing University of Posts and Telecommunications, Nanjing 210023, P.\u00a0R.\u00a0China"}]}],"member":"219","published-online":{"date-parts":[[2021,6,21]]},"reference":[{"key":"S0218194021500273BIB001","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2011.103"},{"key":"S0218194021500273BIB002","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2014.2322358"},{"key":"S0218194021500273BIB003","doi-asserted-by":"publisher","DOI":"10.1109\/COMPSAC.2015.58"},{"key":"S0218194021500273BIB004","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2016.2599161"},{"key":"S0218194021500273BIB005","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2016.2584050"},{"key":"S0218194021500273BIB006","doi-asserted-by":"publisher","DOI":"10.1145\/3183339"},{"key":"S0218194021500273BIB007","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE.2019.00069"},{"key":"S0218194021500273BIB008","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2019.2892959"},{"key":"S0218194021500273BIB009","doi-asserted-by":"publisher","DOI":"10.1007\/s10515-017-0220-7"},{"key":"S0218194021500273BIB010","doi-asserted-by":"publisher","DOI":"10.1109\/ICSME.2017.19"},{"key":"S0218194021500273BIB011","doi-asserted-by":"publisher","DOI":"10.1016\/j.infsof.2018.10.004"},{"key":"S0218194021500273BIB012","doi-asserted-by":"publisher","DOI":"10.1007\/s10664-011-9173-9"},{"key":"S0218194021500273BIB013","doi-asserted-by":"publisher","DOI":"10.1111\/j.1468-0394.2009.00509.x"},{"key":"S0218194021500273BIB014","doi-asserted-by":"publisher","DOI":"10.1007\/s11390-011-9439-0"},{"key":"S0218194021500273BIB015","doi-asserted-by":"publisher","DOI":"10.1145\/2020390.2020405"},{"key":"S0218194021500273BIB016","doi-asserted-by":"publisher","DOI":"10.1016\/j.knosys.2020.105742"},{"key":"S0218194021500273BIB017","doi-asserted-by":"publisher","DOI":"10.1109\/TNN.2009.2015974"},{"key":"S0218194021500273BIB018","doi-asserted-by":"publisher","DOI":"10.1145\/1595696.1595713"},{"key":"S0218194021500273BIB019","doi-asserted-by":"publisher","DOI":"10.1007\/s10664-008-9103-7"},{"key":"S0218194021500273BIB020","doi-asserted-by":"publisher","DOI":"10.1007\/s10515-011-0090-3"},{"key":"S0218194021500273BIB021","doi-asserted-by":"publisher","DOI":"10.1145\/2786805.2786814"},{"key":"S0218194021500273BIB022","doi-asserted-by":"publisher","DOI":"10.1109\/MSR.2010.5463279"},{"key":"S0218194021500273BIB023","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2013.11"},{"key":"S0218194021500273BIB024","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE.2013.6606584"},{"key":"S0218194021500273BIB025","doi-asserted-by":"publisher","DOI":"10.1145\/2025113.2025120"},{"key":"S0218194021500273BIB027","doi-asserted-by":"publisher","DOI":"10.1145\/2786805.2786813"},{"key":"S0218194021500273BIB028","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2017.2720603"},{"key":"S0218194021500273BIB029","doi-asserted-by":"publisher","DOI":"10.3233\/IFS-141220"},{"key":"S0218194021500273BIB030","doi-asserted-by":"publisher","DOI":"10.1515\/jisys-2013-0030"},{"key":"S0218194021500273BIB031","doi-asserted-by":"publisher","DOI":"10.1007\/s10515-016-0194-x"},{"key":"S0218194021500273BIB032","doi-asserted-by":"publisher","DOI":"10.1007\/s00500-018-3093-1"},{"key":"S0218194021500273BIB033","doi-asserted-by":"publisher","DOI":"10.1007\/s10586-018-1696-z"},{"key":"S0218194021500273BIB034","doi-asserted-by":"publisher","DOI":"10.1016\/j.asoc.2020.106163"},{"key":"S0218194021500273BIB035","doi-asserted-by":"publisher","DOI":"10.1109\/ESEM.2013.20"},{"key":"S0218194021500273BIB036","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2017.2724538"},{"key":"S0218194021500273BIB037","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2017.2770124"},{"key":"S0218194021500273BIB038","doi-asserted-by":"publisher","DOI":"10.1109\/TR.2018.2804922"},{"issue":"8","key":"S0218194021500273BIB039","first-page":"5","volume":"74","author":"Singh P.","year":"2013","journal-title":"Int. J. Comput. Appl."},{"key":"S0218194021500273BIB040","doi-asserted-by":"publisher","DOI":"10.1016\/j.infsof.2015.01.014"},{"key":"S0218194021500273BIB041","doi-asserted-by":"publisher","DOI":"10.1109\/TSMCA.2006.889473"},{"key":"S0218194021500273BIB042","doi-asserted-by":"publisher","DOI":"10.1145\/2351676.2351734"},{"key":"S0218194021500273BIB043","doi-asserted-by":"publisher","DOI":"10.1109\/ICPC.2015.15"},{"key":"S0218194021500273BIB044","doi-asserted-by":"publisher","DOI":"10.1016\/j.infsof.2020.106364"},{"key":"S0218194021500273BIB045","doi-asserted-by":"publisher","DOI":"10.3115\/1119176.1119180"},{"key":"S0218194021500273BIB046","doi-asserted-by":"publisher","DOI":"10.1109\/ACVMOT.2005.107"},{"issue":"11","key":"S0218194021500273BIB047","first-page":"2399","volume":"7","author":"Belkin M.","year":"2006","journal-title":"J. Mach. Learn. Res."},{"issue":"1","key":"S0218194021500273BIB048","first-page":"175","volume":"37","author":"Li Y.-F.","year":"2014","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"S0218194021500273BIB049","doi-asserted-by":"publisher","DOI":"10.1145\/1706591.1706599"},{"key":"S0218194021500273BIB050","doi-asserted-by":"publisher","DOI":"10.1049\/ic.2011.0012"},{"key":"S0218194021500273BIB051","doi-asserted-by":"publisher","DOI":"10.1145\/2915970.2916007"},{"key":"S0218194021500273BIB052","doi-asserted-by":"publisher","DOI":"10.1145\/3180155.3180197"},{"issue":"2","key":"S0218194021500273BIB053","first-page":"111","volume":"1","author":"Kotsiantis S.","year":"2006","journal-title":"Int. J. Comput. Sci."},{"key":"S0218194021500273BIB054","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2007.383054"},{"key":"S0218194021500273BIB055","first-page":"153","volume-title":"Advances in Neural Information Processing Systems","author":"He X.","year":"2004"},{"key":"S0218194021500273BIB056","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.1976.233837"},{"key":"S0218194021500273BIB057","volume-title":"Elements of Software Science","author":"Halstead M. H.","year":"1977"},{"key":"S0218194021500273BIB058","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE.2013.6606589"},{"key":"S0218194021500273BIB059","doi-asserted-by":"publisher","DOI":"10.1007\/s10664-014-9346-4"},{"key":"S0218194021500273BIB060","doi-asserted-by":"publisher","DOI":"10.1109\/COMPSAC.2014.65"},{"key":"S0218194021500273BIB061","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2019.113156"},{"key":"S0218194021500273BIB062","doi-asserted-by":"publisher","DOI":"10.1109\/TR.2013.2259203"},{"issue":"1","key":"S0218194021500273BIB063","first-page":"1","volume":"7","author":"Dem\u0161ar J.","year":"2006","journal-title":"J. Mach. Learn. Res."},{"key":"S0218194021500273BIB064","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2020.2978819"},{"key":"S0218194021500273BIB065","first-page":"103","volume":"3","author":"Abdi H.","year":"2007","journal-title":"Encycl. Meas. Stat."},{"key":"S0218194021500273BIB066","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2016.2597849"},{"key":"S0218194021500273BIB067","doi-asserted-by":"publisher","DOI":"10.1007\/978-1-4612-4380-9_16"},{"key":"S0218194021500273BIB068","doi-asserted-by":"publisher","DOI":"10.1016\/j.jss.2019.03.012"},{"key":"S0218194021500273BIB069","doi-asserted-by":"publisher","DOI":"10.1037\/0033-2909.114.3.494"},{"key":"S0218194021500273BIB070","first-page":"1","volume-title":"Annual Meeting of the Florida Association of Institutional Research","author":"Romano J.","year":"2006"},{"key":"S0218194021500273BIB071","doi-asserted-by":"publisher","DOI":"10.1016\/j.infsof.2019.07.003"},{"key":"S0218194021500273BIB072","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2008.35"},{"key":"S0218194021500273BIB073","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE.2015.91"}],"container-title":["International Journal of Software Engineering and Knowledge Engineering"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.worldscientific.com\/doi\/pdf\/10.1142\/S0218194021500273","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2021,6,22]],"date-time":"2021-06-22T03:44:37Z","timestamp":1624333477000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.worldscientific.com\/doi\/abs\/10.1142\/S0218194021500273"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,6]]},"references-count":72,"journal-issue":{"issue":"06","published-print":{"date-parts":[[2021,6]]}},"alternative-id":["10.1142\/S0218194021500273"],"URL":"https:\/\/doi.org\/10.1142\/s0218194021500273","relation":{},"ISSN":["0218-1940","1793-6403"],"issn-type":[{"value":"0218-1940","type":"print"},{"value":"1793-6403","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,6]]}}}