The information bottleneck problem and its applications in machine learning

被引:82
|
作者
Goldfeld Z. [1 ]
Polyanskiy Y. [2 ]
机构
[1] The Electrical and Computer Engineering Department, Cornell University, Ithaca, 14850, NY
[2] The Department of Electrical Engineering and Computer Science, Massachusetts Institute of Technology, Cambridge, 02139, MA
关键词
Deep learning; Information bottleneck; Machine learning; Mutual information; Neural networks;
D O I
10.1109/JSAIT.2020.2991561
中图分类号
学科分类号
摘要
Inference capabilities of machine learning (ML) systems skyrocketed in recent years, now playing a pivotal role in various aspect of society. The goal in statistical learning is to use data to obtain simple algorithms for predicting a random variable Y from a correlated observation X. Since the dimension of X is typically huge, computationally feasible solutions should summarize it into a lower-dimensional feature vector T, from which Y is predicted. The algorithm will successfully make the prediction if T is a good proxy of Y, despite the said dimensionality-reduction. A myriad of ML algorithms (mostly employing deep learning (DL)) for finding such representations T based on real-world data are now available. While these methods are effective in practice, their success is hindered by the lack of a comprehensive theory to explain it. The information bottleneck (IB) theory recently emerged as a bold information-theoretic paradigm for analyzing DL systems. Adopting mutual information as the figure of merit, it suggests that the best representation T should be maximally informative about Y while minimizing the mutual information with X. In this tutorial we survey the information-theoretic origins of this abstract principle, and its recent impact on DL. For the latter, we cover implications of the IB problem on DL theory, as well as practical algorithms inspired by it. Our goal is to provide a unified and cohesive description. A clear view of current knowledge is important for further leveraging IB and other information-theoretic ideas to study DL models. © 2020 IEEE.
引用
收藏
页码:19 / 38
页数:19
相关论文
共 50 条
  • [21] An Information Bottleneck Problem with Renyi's Entropy
    Weng, Jian-Jia
    Alajaji, Fady
    Linder, Tamas
    2021 IEEE INTERNATIONAL SYMPOSIUM ON INFORMATION THEORY (ISIT), 2021, : 2489 - 2494
  • [22] Machine Learning and Its Applications in Wireless Communications
    Lv, Jiaqi
    Na, Zhenyu
    Liu, Xin
    Deng, Zhian
    COMMUNICATIONS, SIGNAL PROCESSING, AND SYSTEMS, 2019, 463 : 2429 - 2436
  • [23] Machine Learning: A Review of the Algorithms and Its Applications
    Dhall, Devanshi
    Kaur, Ravinder
    Juneja, Mamta
    PROCEEDINGS OF RECENT INNOVATIONS IN COMPUTING, ICRIC 2019, 2020, 597 : 47 - 63
  • [24] Machine learning and its applications for plasmonics in biology
    Moon, Gwiyeong
    Lee, Jongha
    Lee, Hyunwoong
    Yoo, Hajun
    Ko, Kwanhwi
    Im, Seongmin
    Kim, Donghyun
    CELL REPORTS PHYSICAL SCIENCE, 2022, 3 (09):
  • [25] Game Theory and Its Applications in Machine Learning
    Rekha, J. Ujwala
    Chatrapati, K. Shahu
    Babu, A. Vinaya
    INFORMATION SYSTEMS DESIGN AND INTELLIGENT APPLICATIONS, VOL 3, INDIA 2016, 2016, 435 : 195 - 207
  • [26] A Unified Definition of Mutual Information with Applications in Machine Learning
    Zeng, Guoping
    MATHEMATICAL PROBLEMS IN ENGINEERING, 2015, 2015
  • [27] Special issue on Information Visualization in Machine Learning and Applications
    Zhang, Du
    Orgun, Mehmet A.
    Zhang, Kang
    JOURNAL OF VISUAL LANGUAGES AND COMPUTING, 2016, 33 : 1 - 2
  • [28] Submodular Combinatorial Information Measures with Applications in Machine Learning
    Iyer, Rishabh
    Khargonkar, Ninad
    Bilmes, Jeff
    Asnani, Himanshu
    ALGORITHMIC LEARNING THEORY, VOL 132, 2021, 132
  • [29] Machine Learning Approaches to Information Retrieval and Its Applications to the Web, Medical Informatics and Health Care
    Huang, Xiangji
    2008 IEEE INTERNATIONAL CONFERENCE ON GRANULAR COMPUTING, VOLS 1 AND 2, 2008, : 39 - 40
  • [30] Machine Learning and its applications in e-Learning systems
    Krendzelak, M.
    12TH IEEE INTERNATIONAL CONFERENCE ON EMERGING ELEARNING TECHNOLOGIES AND APPLICATIONS (ICETA 2014), 2014, : 267 - 269