WS06AFSR Manual Transcriptions, v1
----------------------------------

As part of the Johns Hopkins 2006 Summer Workshop project on articulatory feature-based ASR, a small set of manual transcriptions at the articulatory feature level were collected, to be used as "ground truth" for testing feature classifiers and forced alignments.  The data collection and analysis are described in 

K. Livescu, A. Bezman, N. Borges, L. Yung, O. Cetin, J. Frankel, S. King, M. Magimai-Doss, X. Chi, and L. Lavoie, "Manual transcription of conversational speech at the articulatory feature level", in Proc. ICASSP, Honolulu, April 2007.

If you use the data, please cite this paper in any resulting publications.  Further details, such as feature set definitions and complete transcription conventions, can be found on the project wiki, http://people.csail.mit.edu/klivescu/twiki/bin/view.cgi/WS06.

Some notes on the data:

   * IMPORTANT:  The .wd and .phn files are not part of the transcriptions; they are being provided for completeness and convenience only.  The .phn files were generated during the 1st pass "hybrid" labeling.  They were not modified after the 1st pass and may not match the feature tiers.  We make no claims about their accuracy or usefulness!  The .wd files were generated from the Mississippi State University Switchboard alignments.  However, some of the alignments may have been modified by the WS06 transcribers.

   * We are from time to time still finding errors in the transcription files.  Please let us know if you find errors.  We may release an updated version if there are further corrections.

   * The download does not include the waveform files.  The STP waveforms are from the Switchboard database; the SVB ones are segments of Switchboard utterances as defined in the SVitchboard distribution.  If you have a license for Switchboard 1, we will be glad to provide the waveforms as well for convenience--please contact Karen Livescu at klivescu@csail.mit.edu

   * The transcriptions are in a simple ASCII format.  For easier viewing, two WaveSurfer configuration files are included:  one for viewing a single transcriber's labels and one for viewing both transcribers' side by side.  These have been tested with WaveSurfer 1.8.5 on Windows XP.  Place the config files in your WaveSurfer config directory (in version 1.8 the default is "Documents and Settings\<username>\.wavesurfer\1.8\configurations"), open the desired wav file, and choose one of these two configs.  WaveSurfer can be downloaded from KTH (as of this writing, http://www.speech.kth.se/wavesurfer/).

   * When opening wav files in WaveSurfer, set Sample Rate to 8000 and Read Offset to 1024 bytes when prompted.

   * The data are divided into subdirectories as follows:

      SVB/   The SVitchboard utterances
         ll/  Transcriptions done by Lisa Lavoie
         xc/  Transcriptions done by Xuemin Chi

      STP/   The STP utterances
         allfeature/  Utterances transcribed using an all-feature format
         hybrid/  Utterances transcribed with the hybrid format in the 1st pass.

   * The transcription files are named <wavfile-tag>.<feature-name> and each line is "<start-time> <end-time> <feature-value>".  Times are in seconds.

   * The .wd files were generated from the Mississippi State University word alignments.  We are grateful to them for these alignments.  However, note that some of the alignments may have been modified by the WS06 transcribers.  We are responsible for any problems with the .wd files.
