Introduction
The 2000 HUB5 English Evaluation, Linguistic Data Consortium (LDC)
catalog number LDC2002S09 and ISBN 1-58563-225-2, is part of an
ongoing series of periodic evaluations conducted by NIST. These
evaluations provide an important contribution to the direction
of research efforts and the calibration of technical
capabilities. They are intended to be of interest to all
researchers working on the general problem of conversational
speech recognition. To this end the evaluation was designed to
be simple, to focus on core speech technology issues, to be
fully supported, and to be accessible.
Additional documentation is available at the 2000 NIST Evaluation
Plan for Recognition of Conversational Speech Over the
Telephone website.
Data
This publications contains 40 sphere files encoded in two channel
interleaved mulaw for a total of 644,996,352 bytes (615 Mbytes) of
sphere data. The sphere headers have been modified from the original
evaluation data by the addition of sample checksums to the 20
CALLHOME data files.
An included documentation table contains information on the speech
segments.
The transcripts for this speech may be found in 2000 HUB5 English Evaluation Transcripts(LDC2002T43).
Updates
There are no updates at this time.
Content Copyright
Portions © 1997, 2000, 2002 Trustees of the University of Pennsylvania |