What size test set gives good error rate estimates?

被引：89

作者：

Guyon, I

Makhoul, J

Schwartz, R

Vapnik, V

机构：

[1] AT&T Bell Labs, Red Bank, NJ 07701 USA

[2] BBN Syst & Technol Corp, Cambridge, MA 02138 USA

来源：

IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE | 1998年 / 20卷 / 01期

关键词：

pattern recognition; test set; test set size; benchmark; hypothesis testing; designed experiment; statistical significance; estimation; guaranteed estimators; recognition error;

D O I：

10.1109/34.655649

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

We address the problem of determining what size test set guarantees statistically significant results in a character recognition task, as a function of the expected error rate. We provide a statistical analysis showing that if, for example, the expected character error rate is around 1 percent, then, with a test set of at least 10,000 statistically independent handwritten characters (which could be obtained by taking 100 characters from each of 100 different writers), we guarantee, with 95 percent confidence, that: (1) The expected value of the character error rate is not worse than 1.25 E, where Eis the empirical character error rate of the best recognizer, calculated on the test set; and (2) a difference of 0.3 E between the error rates of two recognizers is significant. We developed this framework with character recognition applications in mind, but it applies as well to speech recognition and to other pattern recognition problems.

引用

页码：52 / 64

页数：13

共 10 条

[1]

[Anonymous], 1974, Introduction to the Theory of Statistics

[2]

BOTTOU L, 1992, TM1135992012405 AT T

[3] A MEASURE OF ASYMPTOTIC EFFICIENCY FOR TESTS OF A HYPOTHESIS BASED ON THE SUM OF OBSERVATIONS [J].

CHERNOFF, H .

ANNALS OF MATHEMATICAL STATISTICS, 1952, 23 (04) :493-507

[4]

GEIST J, 1994, NISTIR5452 NIST US D

[5]

Gillick L., 1989, P ICASSP

[6]

GUYON I, 1994, P 12 INT C PATT REC

[7]

GUYON I, 1992, PIXELS FEATURES, V3, P493

[8]

GUYON I, IN PRESS OVERVIEW SY

[9] PROBABILITY-INEQUALITIES FOR SUMS OF BOUNDED RANDOM-VARIABLES [J].

HOEFFDING, W .

JOURNAL OF THE AMERICAN STATISTICAL ASSOCIATION, 1963, 58 (301) :13-+

[10]

WILKINSON RA, 1992, NISTIR4912 NIST US D

← 1 →