[Purpose/significance] Semantic annotation of digital images is an effective way to solve this problem. The foundation of semantic image annotation is the construction of semantic models. This paper intends to review the existing mainstream semantic models for image annotation, and explore their advantages and disadvantages.[Method/process] Firstly, four representative semantic models for image annotation were reviewed, including Eakins model, Jaimes & Chang model, Kong model and Panofsky model, using literature survey, and then the first three models from three aspects (i.e. semantic level, extensibility and application range) were compared and analyzed using comparative analysis.[Result/conclusion] Through the above analysis, it can be concluded that Eakins model has the most comprehensive semantic level, the strongest semantic expression ability and the widest application range, whereas Kong model is the most scalable and adaptable one.
[1] SMEULDERS A W M, WORRING M, SANTINI S, et al. Content-based image retrieval at the end of the early years[J]. IEEE transactions on pattern analysis and machine intelligence, 2000, 22(12):1349-1379.
[2] EAKINS J P. Retrieval of still images by content[M]. Lectures on information retrieval. Springer, Berlin, Heidelberg, 2000:111-138.
[3] JAIMES A, CHANG S F. A conceptual framework for indexing visual information at multiple levels[J]. Proceeding of SPIE-The International Society for Optical Engineering. San Jose:IS&T/SPIE Internet imaging, 2000, 3964:2-16.
[4] 蔡昌许. 基于语义的图像标注与检索系统研究[D].武汉:武汉大学,2005.
[5] CHUNG E K, YOON J W. Image needs in the context of image use:an exploratory study[J]. Journal of information science, 2011, 37(2):163-177.
[6] HARE J S, LEWIS P H, ENSER P G B, et al. Mind the gap:another look at the problem of the semantic gap in image retrieval[J]. Multimedia Content Analysis Management & Retrieval, 2006, spie v.
[7] KRAUSE M G. Intellectual problems of indexing picture collections[J]. Audiovisual librarian, 1988, 14(2):73-81.
[8] BURFORD B, BRIGGS P, EAKINS J P. A taxonomy of the image:on the classification of content for image retrieval[J]. Visual communication, 2003, 2(2):123-161.
[9] BADR Y, CHBEIR R. Automatic image description based on textual data[M]//Journal on data semantics VⅡ. Berlin, Heidelberg:Springer, 2006.
[10] EAKINS J P. Automatic image content retrieval-are we getting anywhere?[C]//Proceeding of Third International Conference on Electronic Library and Visual Information Research. De Mont fort University. Milton Keynes:Aslib,1996:123-125.
[11] EAKINS J P. Design criteria for a shape retrieval system[J]. Computers in industry, 1993, 21(2):167-184.
[12] EAKINS J P, GRAHAM M E, BOARDMAN J M, et al. Retrieval of trade mark images by shape feature[C]//Proceeding of first International conference on electronic library and visual information system research. Milton Keynes:De Montfort University, 1996:101-109.
[13] PETKOVIC D. Query by image content[C]//Oral presentation to storage and retrieval for image and video databases. California:San Jose,1996.
[14] GUDIVADA V N, RAHAVAN V V. Content-based image retrieval systems[J]. IEEE computer, 1995, 28(9):18-22.
[15] 武人杰. 图像层次语义描述的初步研究[J]. 电脑开发与应用,2011(5):12-14.
[16] 张捷. 图像语义标注[J]. 电脑开发与应用,2012(1):10-12.
[17] 陆泉,丁恒. 基于情感的图像检索研究综述[J]. 情报理论与实践,2013(2):119-124.
[18] 黄质纯. 基于语义的图像检索及相关技术的研究[D].广州:华南理工大学,2012.
[19] HONG D, WU J, SINGH S S. Refining image retrieval based on context-driven methods[C]//Storage and retrieval for image and video databases VⅡ. 1998:581-592.
[20] 于永新. 基于本体的图像语义识别和检索研究[D].天津:天津大学,2009.
[21] 彭杨. 基于本体的动画素材图像语义标注研究[D].长沙:湖南师范大学,2009.
[22] JAIMES A, CHANG S F. Model-based classification of visual information for content-based retrieval[C]//Proceedings of SPIE-The International Society for Optical Engineering. 1998:402-414.
[23] TOUSCH A M, HERBIN S, AUDIBERT J Y. Semantic hierarchies for image annotation:a survey[J]. Pattern recognition, 2012, 45(1):333-345.
[24] HOLLINK L, SCHREIBER A T, WIELINGA B J, et al. Classification of user image descriptions[J]. International journal of human-computer studies, 2004, 61(5):601-626.
[25] YOON J W, CHUNG E K. Understanding image needs in daily life by analyzing questions in a social Q&A site[J]. Journal of the Association for Information Science & Technology, 2011, 62(11):2201-2213.
[26] KONG H, HWANG M, KIM P. The study on the semantic image retrieval based on the personalized ontology[J]. International journal of information technology, 2006, 12(2):35-46.
[27] 邓涛,郭雷,杨卫莉. 基于本体的图像语义标注与检索模型[J]. 计算机工程,2008(17):188-190.
[28] 史婷婷,闫大顺,沈玉利. 基于个性化本体的图像语义标注和检索[J]. 计算机应用,2010(1):90-93.
[29] PANOFSKY E. Meaning in the visual art:papers in and on art history[M]. New York:Doubleday Anchor Books, 1955:39-40.
[30] SHATFORD S. Analyzing the subject of a picture:a theoretical approach[J]. Cataloging & classification quarterly, 1986, 6(3):39-62.
[31] 黄崑,王珊珊,耿骞. 国外图像特征研究进展与启示[J]. 图书情报工作,2015,59(8):138-146.
[32] CHOI Y, RASMUSSEN E M. Searching for images:the analysis of users' queries for image retrieval in American history[J]. Journal of the Association for Information Science and Technology, 2003, 54(6):498-511.
[33] CONDUIT N, RAFFERTY P. Constructing an image indexing template for the children's society:users' queries and archivists' practice[J]. Journal of documentation, 2007, 63(6):898-919.
[34] RAFFERTY P, HIDDERLEY R. Flickr and democratic indexing:dialogic approaches to indexing[J].Aslib Proceedings, 2007, 59(4/5):397-410.
[35] FAUZI F, BELKHATIR M. Multifaceted conceptual image indexing on the World Wide Web[J]. Information processing & management, 2013, 49(2):420-440.
[36] 张梅,郝佳,阎艳,等. 基于本体的知识建模技术[J]. 北京理工大学学报,2010(12):1405-1408,1431.
[37] 张杨,房斌,徐传运. 基于本体和描述逻辑的图像语义识别[C]//南宁:全国安全关键技术与应用学术会议. 2009.
[38] BRACHMAN R J, SCHMOLZE J G. An overview of the KL-ONE knowledge representation system[J]//Cognitive science, 1985, 9(2):171-216.
[39] SIMOU N, TZOUVARAS V, AVRITHIS Y, et al. A visual descriptor ontology for multimedia reasoning[C]//Proceedings of workshop on image analysis for multimedia interactive services. Montreux, 2005:13-15.
[40] 王晓光,徐雷,李纲. 敦煌壁画数字图像语义描述方法研究[J]. 中国图书馆学报,2014,40(1):50-59.
[41] 徐雷,王晓光. 叙事型图像语义标注模型研究[J]. 中国图书馆学报,2017,43(5):70-83.
[42] JORGENSEN C, JAIMES A, BENITEZ A B, et al. A conceptual framework and empirical research for classifying visual descriptors[J]. Journal of the Association for Information Science and Technology, 2001, 52(11):938-947.