Introduction
This publication contains the speech data and transcripts used
as the EvalSet data in NIST's 1997 Hub5 English Evaluation. The
DevSet data for this evaluation consists of the EvalSet data from
the CALLHOME
American English Speech corpus.
The 1997 Hub5 English evaluation was part of an ongoing series of
periodic evaluations conducted by NIST. These evaluations provide an
important contribution to the direction of research efforts and the
calibration of technical capabilities. They are intended to be of interest
to all researchers working on the general problem of conversational speech
recognition. To this end the evaluation was designed to be simple, to focus
on core speech technology issues, to be fully supported, and to be
accessible.
More information is available at NIST's web site for The 1997 Hub-5E Spring Evaluation.
Data
This publication contains 40 audio sphere files, each with a corresponding transcript. The sphere headers have been modified from the original
evaluation data by the addition of sample checksums to the CALLHOME
data files.
An included documentation table contains information on the speech
segments to be processed as follows:
...
Updates
There are no updates at this time.
Content Copyright
Portions © 1996-2002 Trustees of the University of Pennsylvania. |