Publications

September 1, 2005

Comparing HMM, Maximum Entropy, and Conditional Random Fields for Disfluency Detection

We compare a generative hidden Markov model (HMM)-based approach and two conditional models — a maximum entropy (Maxent) model and a conditional random field (CRF) — for detecting disfluencies in…

Publications, Speech & natural language publications
September 1, 2005

Distinguishing Deceptive from Non-Deceptive Speech

ByAndreas Kathol, Martin Graciarena

We present results from a study seeking to distinguish deceptive from non-deceptive speech using machine learning techniques on features extracted from a large corpus of deceptive and non-deceptive speech. We…

Publications, Speech & natural language publications
September 1, 2005

Improved Discriminative Training Using Phone Lattices

We present an efficient discriminative training procedure utilizing phone lattices. Different approaches to expediting lattice generation, statistics collection, and convergence were studied.

Publications, Speech & natural language publications
September 1, 2005

Speech Translation for Low-Resource Languages: The Case of Pashto

ByKristin Precoda, Dimitra Vergyri, Andreas Kathol

We present a number of challenges and solutions that have arisen in the development of a speech translation system for American English and Pashto, highlighting those specific to a very…

Publications, Speech & natural language publications
September 1, 2005

Generation of fast interpreters for Huffman compressed bytecode

Our approach uses canonical Huffman codes to generate compact opcodes with custom-sized operand fields and with a virtual machine that directly executes this compact code. In effect, this automatically creates…

Cyber & formal methods publications, Publications
September 1, 2005

Leveraging Speaker-dependent Variation of Adaptation

This work introduces an automatic procedure for determining the size of regression class trees for individual speakers using an ensemble of speaker-level features to control the number of transformations, if…

Publications, Speech & natural language publications
September 1, 2005

Class-dependent Score Combination for Speaker Recognition

In this work, we are presenting a class-based score combination technique that relies on clustering of both the target models and the test utterances in a vector space defined by…

Publications, Speech & natural language publications
September 1, 2005

Using MLP Features in SRI’s Conversational Speech Recognition System

We describe the development of a speech recognition system for conversational telephone speech (CTS) that incorporates acoustic features estimated by multilayer perceptrons (MLP). The acoustic features are based on frame-level…

Publications, Speech & natural language publications
September 1, 2005

MLLR Transforms as Features in Speaker Recognition

We explore the use of adaptation transforms employed in speech recognition systems as features for speaker recognition. This approach is attractive because, unlike standard frame-based cepstral speaker recognition models, it…

Publications, Speech & natural language publications
September 1, 2005

Robust Feature Compensation in Nonstationary and Multiple Noise Environments

ByMartin Graciarena, Horacio Franco, Victor Abrash

We extend the POF algorithm to allow a more accurate way to select noisy-to-clean feature mappings, by allowing different combinations of speech and noise to have combination-specific mappings selected depending…

Publications, Speech & natural language publications
September 1, 2005

Two Experiments Comparing Reading with Listening for Human Processing of Conversational Telephone Speech

We report on results of two experiments designed to compare subjects’ ability to extract information from audio recordings of conversational telephone speech (CTS) with their ability to extract information from…

Publications, Speech & natural language publications
September 1, 2005

Spoken Language Understanding

SLU systems contain an automatic speech recognition (ASR) component and must be robust to noise due to the spontaneous nature of spoken language and the errors introduced by ASR. SLU…

Publications, Speech & natural language publications