###
DOI:
计算机系统应用英文版:2012,21(12):203-205,185
本文二维码信息
码上扫一扫!
中文深层网络的模式匹配和接口集成
(武汉大学 计算机学院, 武汉 430072)
Schema Matching and Interface Integration for Chinese Deep Web
(School of Computer, Wuhan University, Wuhan 430072, China)
摘要
图/表
参考文献
相似文献
本文已被:浏览 1553次   下载 2884
Received:April 23, 2012    Revised:May 29, 2012
中文摘要: 目前国内外在深层网络方面的研究几乎都围绕英文环境进行, 还没有针对中文深层网络的研究. 提出了对中文深层网络进行模式匹配和接口集成的方法. 该方法首先创建一个用来存储同义词、超义词和子义词的字典, 然后使用基于规则的分词算法将从接口中抽取的属性分成词. 对于每一个属性, 从定义的字典中找到其对应的所有同义词、超义词和子义词, 生成一条相应的记录并存储到列表中, 再从每条记录中选取出现次数最多的属性作为联合接口的属性.
Abstract:Many researches about deep web focus on the deep web with English language, ignoring that with Chinese. In this paper, we present our work in schema matching and interface integration for Chinese deep web. We create a dictionary, which stores synonyms, hypernyms and hyponyms, at the very beginning. After interface extracting, we use Principle-based Segmentation algorithm to segment each attribute into words. Then, for each attribute, we look up the pre-created dictionary to find all its synonyms, hypernyms and hyponyms, form a record and store them in a list. Furthermore, we keep a counter for each attribute in the list to record times it appearing in the local interfaces. At last, we choose from each record a synonym with the largest count number as the attribute of union interface.
文章编号:     中图分类号:    文献标志码:
基金项目:国家自然科学基金(60970018)
引用文本:
张晶星.中文深层网络的模式匹配和接口集成.计算机系统应用,2012,21(12):203-205,185
ZHANG Jing-Xing.Schema Matching and Interface Integration for Chinese Deep Web.COMPUTER SYSTEMS APPLICATIONS,2012,21(12):203-205,185