查看文章

arxiv.org 中的 [PDF]

The information bottleneck method

作者

Naftali Tishby, Fernando C Pereira, William Bialek

发表日期

2000/4/24

期刊

arXiv preprint physics/0004057

简介

We define the relevant information in a signal as being the information that this signal provides about another signal $y\in \Y$. Examples include the information that face images provide about the names of the people portrayed, or the information that speech sounds provide about the words spoken. Understanding the signal requires more than just predicting , it also requires specifying which features of $\X$ play a role in the prediction. We formalize this problem as that of finding a short code for $\X$ that preserves the maximum information about $\Y$. That is, we squeeze the information that $\X$ provides about $\Y$ through a `bottleneck' formed by a limited set of codewords $\tX$. This constrained optimization problem can be seen as a generalization of rate distortion theory in which the distortion measure $d(x,\x)$ emerges from the joint statistics of $\X$ and $\Y$. This approach yields an exact set of self consistent equations for the coding rules $X \to \tX$ and $\tX \to \Y$. Solutions to these equations can be found by a convergent re-estimation method that generalizes the Blahut-Arimoto algorithm. Our variational principle provides a surprisingly rich framework for discussing a variety of problems in signal processing and learning, as will be described in detail elsewhere.

引用总数

被引用次数：3851

20012002200320042005200620072008200920102011201220132014201520162017201820192020202120222023202425 45 65 72 77 74 90 93 96 117 79 85 125 91 90 95 134 201 250 310 368 439 499 303

学术搜索中的文章

The information bottleneck method

N Tishby, FC Pereira, W Bialek - arXiv preprint physics/0004057, 2000

被引用次数：3851 相关文章所有 34 个版本