Skip to content

Research
Commercialization
About
- People
- News & Stories
- Events
- Our History
- Contact
Innovating In
Careers
日本支社

Research
Commercialization
About
- People
- News & Stories
- Events
- Our History
- Contact
Innovating In
Careers
日本支社

Search sri.com

August 1, 2011

Analysis and comparison of recent MLP features for LVCSR systems

Citation

F. Valente, M. M. Doss, and W. Wang, “Analysis and comparison of recent MLP features for LVCSR systems.” in Proc. Interspeech, 2011, pp. 1245–1248.

Abstract

MLP based front-ends have evolved in different ways in recent years beyond the seminal TANDEM-PLP features. This paper aims at providing a fair comparison of these recent progresses including the use of different long/short temporal inputs (PLP,MRASTA,wLP-TRAPS,DCT-TRAPS) and the use of complex architectures (bottleneck, hierarchy, multistream) that go beyond the conventional three layer MLP. Furthermore, the paper identifies which of these actually provide advantages over the conventional TANDEM-PLP. The investigation is carried on an LVCSR task for recognition of Mandarin Broadcast speech and results are analyzed in terms of Character Error Rate and phonetic confusions. Results reveal that as stand alone features, multistream front-ends can outperform by 10% conventional MFCC while TANDEM-PLP only improve by 1%. On the other hand, when used in concatenation with MFCC features, hierarchical/bottleneck front-ends reduce the character error rate by +18% relative compared to +14% relative from TANDEM-PLP. The various input long-term representations recently developed provide comparable performances.

Index Terms: TANDEM features, Multilayer Perceptron, Acoustic features, GALE project, LVCSR.

↓ Download

Read more from SRI

April 15, 2025

SRI’s PARC Forum is now a podcast

Episodes featuring guests such as Julie Packard, Eric Schmidt, and Astro Teller can now be streamed on podcasting apps.
April 11, 2025

With DomiNite®, extreme low-light imaging goes digital

A new camera from SRI is poised to disrupt the decades-long dominance of analog low-light imaging.
April 10, 2025

New research shows kids learn STEM skills from PBS KIDS

SRI research demonstrates that screen-based educational content can measurably improve how young children understand the world around them.

Join Our Team

Build your own legacy

Explore careers

Hire Us

Solutions to your most complex challenges

Send an inquiry

Contact Us

General inquiries

Get the latest news from SRI

Commercialization

Media Inquiries

333 Ravenswood Ave
Menlo Park, CA 94025 USA

+1 (650) 859-2000

© 2025 SRI INTERNATIONAL

Manage Cookie Consent

To provide the best experiences, we use technologies like cookies to store and/or access device information. Consenting to these technologies will allow us to process data such as browsing behavior or unique IDs on this site. Not consenting or withdrawing consent, may adversely affect certain features and functions.

Functional Functional Always active

The technical storage or access is strictly necessary for the legitimate purpose of enabling the use of a specific service explicitly requested by the subscriber or user, or for the sole purpose of carrying out the transmission of a communication over an electronic communications network.

Preferences Preferences

The technical storage or access is necessary for the legitimate purpose of storing preferences that are not requested by the subscriber or user.

Statistics Statistics

The technical storage or access that is used exclusively for statistical purposes. The technical storage or access that is used exclusively for anonymous statistical purposes. Without a subpoena, voluntary compliance on the part of your Internet Service Provider, or additional records from a third party, information stored or retrieved for this purpose alone cannot usually be used to identify you.

Marketing Marketing

The technical storage or access is required to create user profiles to send advertising, or to track the user on a website or across several websites for similar marketing purposes.

Manage options Manage services Manage {vendor_count} vendors Read more about these purposes

View preferences

{title} {title} {title}