Context-based literature digital collection search

被引:8
|
作者
Ratprasartporn, Nattakarn [1 ]
Po, Jonathan [1 ]
Cakmak, Ali [1 ]
Bani-Ahmad, Sulieman [1 ]
Ozsoyoglu, Gultekin [1 ]
机构
[1] Case Western Reserve Univ, Dept Elect Engn & Comp Sci, Cleveland, OH 44106 USA
来源
VLDB JOURNAL | 2009年 / 18卷 / 01期
关键词
Context-based search; Digital collections; Ontology; Context score; Ranking; ALGORITHM; DOCUMENTS;
D O I
10.1007/s00778-008-0099-9
中图分类号
TP3 [计算技术、计算机技术];
学科分类号
0812 ;
摘要
We identify two issues with searching literature digital collections within digital libraries: (a) there are no effective paper-scoring and ranking mechanisms. Without a scoring and ranking system, users are often forced to scan a large and diverse set of publications listed as search results and potentially miss the important ones. (b) Topic diffusion is a common problem: publications returned by a keyword-based search query often fall into multiple topic areas, not all of which are of interest to users. This paper proposes a new literature digital collection search paradigm that effectively ranks search outputs, while controlling the diversity of keyword-based search query output topics. Our approach is as follows. First, during pre-querying, publications are assigned into pre-specified ontology-based contexts, and query-independent context scores are attached to papers with respect to the assigned contexts. When a query is posed, relevant contexts are selected, search is performed within the selected contexts, context scores of publications are revised into relevancy scores with respect to the query at hand and the context that they are in, and query outputs are ranked within each relevant context. This way, we (1) minimize query output topic diversity, (2) reduce query output size, (3) decrease user time spent scanning query results, and (4) increase query output ranking accuracy. Using genomics-oriented PubMed publications as the testbed and Gene Ontology terms as contexts, our experiments indicate that the proposed context-based search approach produces search results with up to 50% higher precision, and reduces the query output size by up to 70%.
引用
收藏
页码:277 / 301
页数:25
相关论文
共 50 条
  • [1] Context-based literature digital collection search
    Nattakarn Ratprasartporn
    Jonathan Po
    Ali Cakmak
    Sulieman Bani-Ahmad
    Gultekin Ozsoyoglu
    The VLDB Journal, 2009, 18 : 277 - 301
  • [2] Evaluating different ranking functions for context-based literature search
    Ratprasartpom, Nattakam
    Bani-Ahmad, Sulieman
    Cakmak, Ali
    Po, Jonathan
    Ozsoyoglu, Gultekin
    2007 IEEE 23RD INTERNATIONAL CONFERENCE ON DATA ENGINEERING WORKSHOP, VOLS 1-2, 2007, : 261 - 268
  • [3] Context-Based Search in Software Development
    Antunes, Bruno
    Cordeiro, Joel
    Gomes, Paulo
    20TH EUROPEAN CONFERENCE ON ARTIFICIAL INTELLIGENCE (ECAI 2012), 2012, 242 : 937 - 942
  • [4] An Efficient Search for Context-Based Chatbots
    Atiyah, Ayah
    Jusoh, Shaidah
    Almajali, Sufyan
    2018 8TH INTERNATIONAL CONFERENCE ON COMPUTER SCIENCE AND INFORMATION TECHNOLOGY (CSIT), 2018, : 125 - 130
  • [5] DataGopher: Context-based Search for Research Datasets
    Singhal, Ayush
    Kasturi, Ravindra
    Srivastava, Jaideep
    2014 IEEE 15TH INTERNATIONAL CONFERENCE ON INFORMATION REUSE AND INTEGRATION (IRI), 2014, : 749 - 756
  • [6] Context-Based Clustering of Image Search Results
    Wang, Hongqi
    Missura, Olana
    Gaertner, Thomas
    Wrobel, Stefan
    KI 2009: ADVANCES IN ARTIFICIAL INTELLIGENCE, PROCEEDINGS, 2009, 5803 : 153 - 160
  • [7] Multiobjective Evolutionary Algorithms for Context-Based Search
    Cecchini, Rocio L.
    Lorenzetti, Carlos M.
    Maguitman, Ana G.
    Brignole, Nelida B.
    JOURNAL OF THE AMERICAN SOCIETY FOR INFORMATION SCIENCE AND TECHNOLOGY, 2010, 61 (06): : 1258 - 1274
  • [8] Context-based Matching Refinement for Person Search
    Han, Byeong-Ju
    Yang, Jae-Won
    Lee, Oggyu
    Sim, Jae-Young
    2021 ASIA-PACIFIC SIGNAL AND INFORMATION PROCESSING ASSOCIATION ANNUAL SUMMIT AND CONFERENCE (APSIPA ASC), 2021, : 1607 - 1610
  • [9] EMCODIST: A Context-based Search Tool for Email Archives
    Venkata, Santhilata Kuppili
    Decker, Stephanie
    Kirsch, David A.
    Nix, Adam
    2021 IEEE INTERNATIONAL CONFERENCE ON BIG DATA (BIG DATA), 2021, : 2281 - 2290
  • [10] Exploiting rich context: An incremental approach to context-based Web search
    Leake, D
    Maguitman, A
    Reichherzer, T
    MODELING AND USING CONTEXT, PROCEEDINGS, 2005, 3554 : 254 - 267