Visual Bayesian Fusion to Navigate a Data Lake

被引:0
|
作者
Singh, Karamjit [1 ]
Paneri, Kaushal [1 ]
Pandey, Aditeya [1 ]
Gupta, Garima [1 ]
Sharma, Geetika [1 ]
Agarwal, Puneet [1 ]
Shroff, Gautam [1 ]
机构
[1] Tata Consultancy Serv Ltd, TCS Res, Gurgaon, India
关键词
D O I
暂无
中图分类号
TP301 [理论、方法];
学科分类号
081202 ;
摘要
The evolution from traditional business intelligence to big data analytics has witnessed the emergence of 'Data Lakes' in which data is ingested in raw form rather than into traditional data warehouses. With the increasing availability of many more pieces of information about each entity of interest, e.g., a customer, often from diverse sources (socialmedia, mobility, internet-of-things), fusing, visualizing and deriving insights from such data pose a number of challenges: First, disparate datasets often lack a natural join key. Next, datasets may describe measures at different levels of granularity, e.g., individual vs. aggregate data, and finally, different datasets may be derived from physically distinct populations. Moreover, once data has been fused, queries are often an inefficient and inaccurate mechanism to derive insight from high-dimensional data. In this paper we describe iFuse, a data-fusion based visual analytics platform for navigating a data lake to derive insights. We rely on Bayesian graphical models to provide useful rudder with which to fuse and analyze disparate islands of data in a systematic manner. Our platform allows for rich interactive visualizations, querying and keyword-based search within and across datasets or models, as well as intuitive visual interfaces for value-imputation or model-based predictions. We illustrate the use of our platform in multiple scenarios, including two public data challenges as well as a real-life industry use-case involving the probabilistic fusion of datasets that lack a natural join-key.
引用
收藏
页码:987 / 994
页数:8
相关论文
共 50 条
  • [31] Exact and Approximate Heterogeneous Bayesian Decentralized Data Fusion
    Dagan, Ofer
    Ahmed, Nisar R.
    IEEE TRANSACTIONS ON ROBOTICS, 2023, 39 (02) : 1136 - 1150
  • [32] Rotor aerodynamic data fusion based on Bayesian framework
    Yang H.
    Chen S.
    Gao Z.
    Jiang Q.
    Zhang W.
    Hangkong Xuebao/Acta Aeronautica et Astronautica Sinica, 2024, 45 (08):
  • [33] Homogeneous functionals and Bayesian data fusion with unknown correlation
    Taylor, Clark N.
    Bishop, Adrian N.
    INFORMATION FUSION, 2019, 45 : 179 - 189
  • [34] Decision level data fusion using Bayesian inference
    Parish, J
    SENSOR FUSION: ARCHITECTURES, ALGORITHMS, AND APPLICATIONS, 1997, 3067 : 50 - 61
  • [35] Bayesian data fusion and credit assignment in vision and fMRI data analysis
    Schrater, PR
    COMPUTATIONAL IMAGING, 2003, 5016 : 24 - 35
  • [36] Data Fusion Approach for Learning Transcriptional Bayesian Networks
    Sauta, Elisabetta
    Demartini, Andrea
    Vitali, Francesca
    Riva, Alberto
    Bellazzi, Riccardo
    ARTIFICIAL INTELLIGENCE IN MEDICINE, AIME 2017, 2017, 10259 : 76 - 80
  • [37] A Distributed Bayesian Data Fusion Algorithm With Uniform Consistency
    Li, Yingke
    Zhou, Enlu
    Zhang, Fumin
    IEEE TRANSACTIONS ON AUTOMATIC CONTROL, 2024, 69 (09) : 6176 - 6182
  • [38] Multispectral image data fusion under a Bayesian approach
    Int J Remote Sens, 8 (1457-1471):
  • [39] Multisource data fusion for bandlimited signals:: a Bayesian perspective
    Jalobeanu, A.
    Gutierrez, J. A.
    BAYESIAN INFERENCE AND MAXIMUM ENTROPY METHODS IN SCIENCE AND ENGINEERING, 2006, 872 : 391 - +
  • [40] On Bayesian Tracking and Data Fusion: A Tutorial Introduction with Examples
    Koch, Wolfgang
    IEEE AEROSPACE AND ELECTRONIC SYSTEMS MAGAZINE, 2010, 25 (07) : 29 - 51