A decision-theoretic extension of stochastic complexity and its applications to learning

被引：52

作者：

Yamanishi, K ^{[1
]}

机构：

[1] NEC Res Inst, Princeton, NJ 08540 USA

来源：

IEEE TRANSACTIONS ON INFORMATION THEORY | 1998年 / 44卷 / 04期

关键词：

aggregating strategy; batch-learning; complexity regularization; extended stochastic complexity; MDL principle; on-line prediction; stochastic complexity;

D O I：

10.1109/18.681319

中图分类号：

TP [自动化技术、计算机技术];

学科分类号：

0812 ;

摘要：

Rissanen has introduced stochastic complexity to define the amount of information in a given data sequence relative to a given hypothesis class of probability densities, where the information is measured in terms of the logarithmic loss associated with universal data compression. This paper introduces the notion of extended stochastic complexity (ESC) and demonstrates its effectiveness in design and analysis of learning algorithms in on-line prediction and batch-learning scenarios. ESC can be thought of as an extension of Rissanen's stochastic complexity to the decision-theoretic setting where a general real-valued function is used as a hypothesis and a general loss function is used as a distortion measure. As an application of ESC to online prediction, this paper shows that a sequential realization of ESC produces an on-line prediction algorithm called Vovk's aggregating strategy, which can be thought of as an extension of the Bayes algorithm. We derive upper bounds on the cumulative loss for the aggregating strategy both of an expected form and a worst case form in the case where the hypothesis class is continuous. As an application of ESC to batch-learning, this paper shows that a batch-approximation of ESC induces a batch-learning algorithm called the minimum L-complexity algorithm (MLC), which is an extension of the minimum description length (MDL) principle. We derive upper bounds on the statistical risk for MLC, which are least to date. Through ESC we give a unifying view of the most effective learning algorithms that have recently been explored in computational learning theory.

引用

页码：1424 / 1439

页数：16

共 50 条

[21] Decision-theoretic refinement planning
Hillner, BE
MEDICAL DECISION MAKING, 1996, 16 (04) : 419 - 420
[22] A Decision-Theoretic Model of Assistance
Fern, Alan
Natarajan, Sriraam
Judah, Kshitij
Tadepalli, Prasad
20TH INTERNATIONAL JOINT CONFERENCE ON ARTIFICIAL INTELLIGENCE, 2007, : 1879 - 1884
[23] Decision-theoretic file carving
Gladyshev, Pavel
James, Joshua I.
DIGITAL INVESTIGATION, 2017, 22 : 46 - 61
[24] Depression: A Decision-Theoretic Analysis
Huys, Quentin J. M.
Daw, Nathaniel D.
Dayan, Peter
ANNUAL REVIEW OF NEUROSCIENCE, VOL 38, 2015, 38 : 1 - 23
[25] The Decision-Theoretic Lockean Thesis
Locke, Dustin Troy
INQUIRY-AN INTERDISCIPLINARY JOURNAL OF PHILOSOPHY, 2014, 57 (01): : 28 - 54
[26] Decision-theoretic reliability sensitivity
Straub, Daniel
Ehre, Max
Papaioannou, Iason
RELIABILITY ENGINEERING & SYSTEM SAFETY, 2022, 221
[27] Decision-theoretic approach to risk
Young, S.C.
IMA Journal of Mathematics Applied in Business and Industry, 1993, 5 (03):
[28] A decision-theoretic generalization of on-line learning and an application to boosting
Freund, Y
Schapire, RE
JOURNAL OF COMPUTER AND SYSTEM SCIENCES, 1997, 55 (01) : 119 - 139
[29] Causal decision theory and decision-theoretic causation
Hitchcock, CR
NOUS, 1996, 30 (04): : 508 - 526
[30] Decision-theoretic cooperative sensor planning
Cook, DJ
Gmytrasiewicz, P
Holder, LB
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 1996, 18 (10) : 1013 - 1023

← 1 2 3 4 5 →