Speech & natural language publications
-
Using MLP Features in SRI’s Conversational Speech Recognition System
We describe the development of a speech recognition system for conversational telephone speech (CTS) that incorporates acoustic features estimated by multilayer perceptrons (MLP). The acoustic features are based on frame-level…
-
Class-dependent Score Combination for Speaker Recognition
In this work, we are presenting a class-based score combination technique that relies on clustering of both the target models and the test utterances in a vector space defined by…
-
Development of a Conversational Telephone Speech Recognizer for Levantine Arabic
In this paper, we describe the development of a large-vocabulary speech recognition system for Levantine Arabic, which was a new dialectal recognition task for our existing system. We discuss the…
-
Leveraging Speaker-dependent Variation of Adaptation
This work introduces an automatic procedure for determining the size of regression class trees for individual speakers using an ensemble of speaker-level features to control the number of transformations, if…
-
Spoken Language Understanding
SLU systems contain an automatic speech recognition (ASR) component and must be robust to noise due to the spontaneous nature of spoken language and the errors introduced by ASR. SLU…
-
Two Experiments Comparing Reading with Listening for Human Processing of Conversational Telephone Speech
We report on results of two experiments designed to compare subjects’ ability to extract information from audio recordings of conversational telephone speech (CTS) with their ability to extract information from…
-
Improved Discriminative Training Using Phone Lattices
We present an efficient discriminative training procedure utilizing phone lattices. Different approaches to expediting lattice generation, statistics collection, and convergence were studied.
-
Meeting Structure Annotation: Data and Tools
We present a set of annotations of hierarchical topic segmentations and action item sub-dialogues collected over 65 meetings from the ICSI and ISL meeting corpora, designed to support automatic meeting…
-
Collaborative and argumentative models of natural discussions
We report in this paper experiences and insights resulting from the first two years of work in two similar projects on meeting tracking and understanding. The projects are the DARPA-funded…
-
Using Conditional Random Fields for Sentence Boundary Detection in Speech
In this paper, we evaluate the use of a conditional random field (CRF) for this task and relate results with this model to our prior work. We evaluate across two…
-
Ontology-based multi-party meeting understanding
This paper describes current and planned research efforts towards developing multimodal discourse understanding for an automated personal office assistant.
-
Structural Metadata Research in the EARS Program
In this paper we provide a brief overview of research on structural metadata extraction in the DARPA EARS rich transcription program. Tasks include detection of sentence boundaries, filler words, and…