Search CORE

3,301 research outputs found

No-But-Semantic-Match: Computing Semantically Matched XML Keyword Search Results

Author: Islam Md. Saiful
Liu Chengfei
Moser Irene
Naseriparsa Mehdi
Publication venue
Publication date: 06/03/2017
Field of study

Users are rarely familiar with the content of a data source they are querying, and therefore cannot avoid using keywords that do not exist in the data source. Traditional systems may respond with an empty result, causing dissatisfaction, while the data source in effect holds semantically related content. In this paper we study this no-but-semantic-match problem on XML keyword search and propose a solution which enables us to present the top-k semantically related results to the user. Our solution involves two steps: (a) extracting semantically related candidate queries from the original query and (b) processing candidate queries and retrieving the top-k semantically related results. Candidate queries are generated by replacement of non-mapped keywords with candidate keywords obtained from an ontological knowledge base. Candidate results are scored using their cohesiveness and their similarity to the original query. Since the number of queries to process can be large, with each result having to be analyzed, we propose pruning techniques to retrieve the top-

k

results efficiently. We develop two query processing algorithms based on our pruning techniques. Further, we exploit a property of the candidate queries to propose a technique for processing multiple queries in batch, which improves the performance substantially. Extensive experiments on two real datasets verify the effectiveness and efficiency of the proposed approaches.Comment: 24 pages, 21 figures, 6 tables, submitted to The VLDB Journal for possible publicatio

arXiv.org e-Print Archive

Crossref

Federation ResearchOnline

Reasoning & Querying – State of the Art

Author: Bry François
Furche Tim
Weiand Klara
Publication venue
Publication date: 31/08/2008
Field of study

Various query languages for Web and Semantic Web data, both for practical use and as an area of research in the scientific community, have emerged in recent years. At the same time, the broad adoption of the internet where keyword search is used in many applications, e.g. search engines, has familiarized casual users with using keyword queries to retrieve information on the internet. Unlike this easy-to-use querying, traditional query languages require knowledge of the language itself as well as of the data to be queried. Keyword-based query languages for XML and RDF bridge the gap between the two, aiming at enabling simple querying of semi-structured data, which is relevant e.g. in the context of the emerging Semantic Web. This article presents an overview of the field of keyword querying for XML and RDF

Open Access LMU

Web Queries: From a Web of Data to a Semantic Web?

Author: Bry François
Furche Tim
Vossen Gottfried
Weiand Klara
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2009
Field of study

Open Access LMU

N-GRAM BASED QUERY STRUCTURING SYSTEM FOR EFFECTIVE XML RETRIEVAL

Author: Abubakar Roko
Muhammad Bui Aminu
Saidu Ibrahim
Shehu Asma’u
Publication venue: 'Revista Mexicana de Biodiversidad'
Publication date: 25/08/2019
Field of study

Query structuring systems are keyword search systems recently used for the effective retrieval of XML documents. Existing systems fail to put keyword query ambiguity prob-lems into consideration during query pre-processing and return irrelevant predicate nodes. As a result, these sys-tems return irrelevant results. In this research, an XML keyword search system, called N-gram based XML query structuring system (NBXQSS) is developed to improve the performance of keyword searches. The NBXQSS uses an N-gram Based Query Segmentation (NBQS) method which interprets a user query as a list of semantic units to help resolve ambiguity. The system also introduces an improved predicate identification algorithm (IPIA) to return rele-vant predicates. The IPIA uses a proposed function to com-pute the query term proximity and ordering. The effective-ness of the NBXQS is demonstrated through experimental performance study on some real-world XML documents. The results show that the developed system performs bet-ter compared to the existing system in terms of precision

International Journal of Advanced Computer Technology

Finding Patterns in a Knowledge Base using Keywords to Compose Table Answers

Author: Chakrabarti Kaushik
Chaudhuri Surajit
Ding Bolin
Yang Mohan
Publication venue
Publication date: 03/09/2014
Field of study

We aim to provide table answers to keyword queries against knowledge bases. For queries referring to multiple entities, like "Washington cities population" and "Mel Gibson movies", it is better to represent each relevant answer as a table which aggregates a set of entities or entity-joins within the same table scheme or pattern. In this paper, we study how to find highly relevant patterns in a knowledge base for user-given keyword queries to compose table answers. A knowledge base can be modeled as a directed graph called knowledge graph, where nodes represent entities in the knowledge base and edges represent the relationships among them. Each node/edge is labeled with type and text. A pattern is an aggregation of subtrees which contain all keywords in the texts and have the same structure and types on node/edges. We propose efficient algorithms to find patterns that are relevant to the query for a class of scoring functions. We show the hardness of the problem in theory, and propose path-based indexes that are affordable in memory. Two query-processing algorithms are proposed: one is fast in practice for small queries (with small patterns as answers) by utilizing the indexes; and the other one is better in theory, with running time linear in the sizes of indexes and answers, which can handle large queries better. We also conduct extensive experimental study to compare our approaches with a naive adaption of known techniques.Comment: VLDB 201

arXiv.org e-Print Archive

CiteSeerX

Data Model and Query Constructs for Versatile Web Query Languages

Author: Bry François
Furche Tim
Linse Benedikt
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2006
Field of study

As the Semantic Web is gaining momentum, the need for truly versatile query languages becomes increasingly apparent. A Web query language is called versatile if it can access in the same query program data in different formats (e.g. XML and RDF). Most query languages are not versatile: they have not been specifically designed to cope with both worlds, providing a uniform language and common constructs to query and transform data in various formats. Moreover, most of them do not provide a flexible data model that is powerful enough to naturally convey both Semantic Web data formats (especially RDF and Topic Maps) and XML. This article highlights challenges related to the data model and language constructs for querying both standard Web and Semantic Web data with an emphasis on facilitating sophisticated reasoning. It is shown that Xcerpt’s data model and querying constructs are particularly well-suited for the Semantic Web, but that some adjustments of the Xcerpt syntax allow for even more effective and natural querying of RDF and Topic Maps

CiteSeerX

Crossref

Open Access LMU

No-but-semantic-match : computing semantically matched xml keyword search results

Author: Islam Md Saiful
Liu Chengfei
Moser Irene
Naseriparsa Mehdi
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 01/01/2018
Field of study

Federation ResearchOnline