{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2024,3,2]],"date-time":"2024-03-02T22:00:34Z","timestamp":1709416834997},"reference-count":39,"publisher":"Hindawi Limited","license":[{"start":{"date-parts":[[2020,5,20]],"date-time":"2020-05-20T00:00:00Z","timestamp":1589932800000},"content-version":"unspecified","delay-in-days":0,"URL":"http:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Complexity"],"published-print":{"date-parts":[[2020,5,20]]},"abstract":"Constrained clustering is intended to improve accuracy and personalization based on the constraints expressed by an Oracle. In this paper, a new constrained clustering algorithm is proposed and some of the informative data pairs are selected during an iterative process. Then, they are presented to the Oracle and their relation is answered with \u201cMust-link (ML) or Cannot-link (CL).\u201d In each iteration, first, the support vector machine (SVM) is utilized based on the label produced by the current clustering. According to the distance of each document from the hyperplane, the distance matrix is created. Also, based on cosine similarity of word2vector of each document, the similarity matrix is created. Two types of probability (similarity and degree of similarity) are calculated and they are smoothed for belonging to neighborhoods. Neighborhoods form the samples that are labeled by Oracle, to be in the same cluster. Finally, at the end of each iteration, the data with a greater level of uncertainty (in term of probability) is selected for questioning the oracle. In order to evaluate, the proposed method is compared with famous state-of-the-art methods based on two criteria and over a standard dataset. The result demonstrates an increased accuracy and stability of the obtained result with fewer questions.<\/jats:p>","DOI":"10.1155\/2020\/3207306","type":"journal-article","created":{"date-parts":[[2020,5,20]],"date-time":"2020-05-20T23:31:55Z","timestamp":1590017515000},"page":"1-16","source":"Crossref","is-referenced-by-count":4,"title":["Active Learning for Constrained Document Clustering with Uncertainty Region"],"prefix":"10.1155","volume":"2020","author":[{"ORCID":"http:\/\/orcid.org\/0000-0001-5898-0871","authenticated-orcid":true,"given":"M. A.","family":"Balafar","sequence":"first","affiliation":[{"name":"Department of IT, Faculty of Engineering, University of Tabriz, Tabriz, Iran"}]},{"given":"R.","family":"Hazratgholizadeh","sequence":"additional","affiliation":[{"name":"Department of IT, Faculty of Engineering, University of Tabriz, Tabriz, Iran"}]},{"given":"M. R. F.","family":"Derakhshi","sequence":"additional","affiliation":[{"name":"Department of Computer, Faculty of Engineering, University of Tabriz, Tabriz, Iran"}]}],"member":"98","reference":[{"key":"1","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2016.11.003"},{"key":"2","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2018.10.026"},{"key":"3","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2012.12.004"},{"key":"4","doi-asserted-by":"publisher","DOI":"10.1016\/j.procs.2017.05.206"},{"key":"5","doi-asserted-by":"publisher","DOI":"10.1155\/2019\/4275720"},{"key":"8","doi-asserted-by":"publisher","DOI":"10.1155\/2019\/6876173"},{"key":"9","doi-asserted-by":"publisher","DOI":"10.1016\/j.patrec.2009.09.011"},{"key":"10","doi-asserted-by":"publisher","DOI":"10.1007\/s10586-018-2199-7"},{"key":"11","year":"2009"},{"key":"12","doi-asserted-by":"publisher","DOI":"10.3233\/ida-173781"},{"key":"14","first-page":"319","volume-title":"Active learning method for constraint-based clustering algorithms","volume":"9659","year":"2016"},{"key":"15","doi-asserted-by":"publisher","DOI":"10.1007\/s10994-017-5643-7"},{"key":"16","doi-asserted-by":"publisher","DOI":"10.1016\/j.cam.2018.04.035"},{"key":"18","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-013-0680-6"},{"key":"19","first-page":"207","volume-title":"A survey of constrained clustering","year":"2016"},{"key":"20","doi-asserted-by":"publisher","DOI":"10.3233\/ida-2010-0461"},{"key":"22","doi-asserted-by":"publisher","DOI":"10.1016\/j.knosys.2013.01.032"},{"key":"24","first-page":"140","volume-title":"Constraint selection by committee: an ensemble approach to identifying informative constraints for semi-supervised clustering","volume":"4701","year":"2007"},{"key":"25","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2015.05.007"},{"key":"26","doi-asserted-by":"publisher","DOI":"10.1109\/tkde.2013.22"},{"key":"27","doi-asserted-by":"publisher","DOI":"10.5120\/16899-6972"},{"key":"28","doi-asserted-by":"publisher","DOI":"10.31645\/jisrc\/(2015).13.1.0009"},{"key":"29","doi-asserted-by":"publisher","DOI":"10.3390\/s19173728"},{"key":"30","doi-asserted-by":"publisher","DOI":"10.1109\/tmi.2017.2746879"},{"key":"31","doi-asserted-by":"publisher","DOI":"10.1007\/s00779-018-1183-9"},{"key":"33","doi-asserted-by":"publisher","DOI":"10.3390\/info9040100"},{"key":"34","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2018.08.001"},{"key":"35","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2014.12.009"},{"key":"38","doi-asserted-by":"publisher","DOI":"10.1109\/tpami.2016.2539965"},{"key":"39","doi-asserted-by":"publisher","DOI":"10.1016\/j.asoc.2017.01.023"},{"key":"40","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2017.01.001"},{"key":"41","doi-asserted-by":"publisher","DOI":"10.1007\/s13042-016-0628-6"},{"key":"42","doi-asserted-by":"publisher","DOI":"10.1109\/tkde.2018.2818729"},{"key":"43","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2013.09.034"},{"key":"44","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2014.09.106"},{"key":"45","doi-asserted-by":"publisher","DOI":"10.1016\/j.patrec.2014.02.014"},{"key":"47","doi-asserted-by":"publisher","DOI":"10.1177\/0165551518816302"},{"key":"48","doi-asserted-by":"publisher","DOI":"10.1007\/s10115-011-0389-1"},{"key":"49","doi-asserted-by":"publisher","DOI":"10.1016\/j.ins.2015.03.038"}],"container-title":["Complexity"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/downloads.hindawi.com\/journals\/complexity\/2020\/3207306.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/downloads.hindawi.com\/journals\/complexity\/2020\/3207306.xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/downloads.hindawi.com\/journals\/complexity\/2020\/3207306.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2020,5,20]],"date-time":"2020-05-20T23:31:58Z","timestamp":1590017518000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.hindawi.com\/journals\/complexity\/2020\/3207306\/"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,5,20]]},"references-count":39,"alternative-id":["3207306","3207306"],"URL":"https:\/\/doi.org\/10.1155\/2020\/3207306","relation":{},"ISSN":["1076-2787","1099-0526"],"issn-type":[{"value":"1076-2787","type":"print"},{"value":"1099-0526","type":"electronic"}],"subject":[],"published":{"date-parts":[[2020,5,20]]}}}