基于潜语义主题加强的跨媒体检索算法 Cross-media retrieval based on latent semantic topic reinforce期刊界 All Journals 搜尽天下杂志传播学术成果专业期刊搜索期刊信息化学术搜索

基于潜语义主题加强的跨媒体检索算法

引用本文：	黄育,张鸿.基于潜语义主题加强的跨媒体检索算法[J].计算机应用,2017,37(4):1061-1064.

作者姓名：	黄育张鸿

作者单位：	1. 武汉科技大学计算机科学与技术学院, 武汉 430065;2. 智能信息处理与实时工业系统湖北省重点实验室(武汉科技大学), 武汉 430065

基金项目：	国家自然科学基金资助项目（61003127，61373109）。

摘要：	针对不同模态数据对相同语义主题表达存在差异性，以及传统跨媒体检索算法忽略了不同模态数据能以合作的方式探索数据的内在语义信息等问题，提出了一种新的基于潜语义主题加强的跨媒体检索（LSTR）算法。首先，利用隐狄利克雷分布（LDA）模型构造文本语义空间，然后以词袋（BoW）模型来表达文本对应的图像；其次，使用多分类逻辑回归对图像和文本分类，用得到的基于多分类的后验概率表示文本和图像的潜语义主题；最后，利用文本潜语义主题去正则化图像的潜语义主题，使图像的潜语义主题得到加强，同时使它们之间的语义关联最大化。在Wikipedia数据集上，文本检索图像和图像检索文本的平均查准率为57.0%，比典型相关性分析（CCA）、SM（Semantic Matching）、SCM（Semantic Correlation Matching）算法的平均查准率分别提高了35.1%、34.8%、32.1%。实验结果表明LSTR算法能有效地提高跨媒体检索的平均查准率。
关键词：	跨媒体检索潜语义主题多分类逻辑回归后验概率正则化
收稿时间：	2016-09-23
修稿时间：	2016-12-22
Cross-media retrieval based on latent semantic topic reinforce

HUANG Yu,ZHANG Hong.Cross-media retrieval based on latent semantic topic reinforce[J].journal of Computer Applications,2017,37(4):1061-1064.

Authors:	HUANG Yu ZHANG Hong

Affiliation:	1. School of Computer Science and Technology, Wuhan University of Science and Technology, Wuhan Hubei 430065, China;2. Hubei Province Key Laboratory of Intelligent Information Processing and Real-time Industrial System(Wuhan University of Science and Technology), Wuhan Hubei 430065, China

Abstract:	As an important and challenging problem in the multimedia area, common semantic topic has different expression across different modalities, and exploring the intrinsic semantic information from different modalities in a collaborative manner was usually neglected by traditional cross-media retrieval methods. To address this problem, a Latent Semantic Topic Reinforce cross-media retrieval (LSTR) method was proposed. Firstly, the text semantic was represented based on Latent Dirichlet Allocation (LDA) and the corresponding images were represented with Bag of Words (BoW) model. Secondly, multiclass logistic regression was used to classify both texts and images, and the posterior probability under the learned classifiers was exploited to indicate the latent semantic topic of images and texts. Finally, the learned posterior probability was used to regularize their image counterparts to reinforce the image semantic topics, which greatly improved the semantic similarity between them. In the Wikipedia data set, the mean Average Precision (mAP) of retrieving text with image and retrieving image with text is 57.0%, which is 35.1%, 34.8% and 32.1% higher than that of the Canonical Correlation Analysis (CCA), Semantic Matching (SM) and Semantic Correlation Matching (SCM) method respectively. Experimental results show that the proposed method can effectively improve the average precision of cross-media retrieval.

Keywords:	cross-media retrieval latent semantic topic multiclass logistic regression posterior probability regularization

	点击此处可从《计算机应用》浏览原始摘要信息
	点击此处可从《计算机应用》下载全文

设为首页 | 免责声明 | 关于勤云 | 加入收藏