2000 HUB5 English Evaluation Speech

Item Name: 2000 HUB5 English Evaluation Speech
Authors: LDC
LDC Catalog No.: LDC2002S09
ISBN: 1-58563-225-2
Data Type: speech
Data Source(s): telephone conversations
Language(s): English
Language ID(s): eng
Distribution: 1 CD
Member fee: $0 for 2002 members
Non-member Fee: N/A (Members Only)
Reduced-License Fee: N/A
Extra-Copy Fee: US $150.00
Online documentation: yes
Licensing Instructions: Subscription & Standard Members, and Non-Members
Citation: LDC
2000 HUB5 English Evaluation Speech
Linguistic Data Consortium, Philadelphia


The 2000 HUB5 English Evaluation, Linguistic Data Consortium (LDC) catalog number LDC2002S09 and ISBN 1-58563-225-2, is part of an ongoing series of periodic evaluations conducted by NIST. These evaluations provide an important contribution to the direction of research efforts and the calibration of technical capabilities. They are intended to be of interest to all researchers working on the general problem of conversational speech recognition. To this end the evaluation was designed to be simple, to focus on core speech technology issues, to be fully supported, and to be accessible.

Additional documentation is available at the 2000 NIST Evaluation Plan for Recognition of Conversational Speech Over the Telephone website.


This publications contains 40 sphere files encoded in two channel interleaved mulaw for a total of 644,996,352 bytes (615 Mbytes) of sphere data. The sphere headers have been modified from the original evaluation data by the addition of sample checksums to the 20 CALLHOME data files.

An included documentation table contains information on the speech segments.

The transcripts for this speech may be found in 2000 HUB5 English Evaluation Transcripts(LDC2002T43).


There are no updates at this time.

Content Copyright

Portions 1997, 2000, 2002 Trustees of the University of Pennsylvania