Blick ins Buch

Speech Separation by Humans and Machines (eBook)

eBook Download: PDF

2006 | 2004
XXIV, 319 Seiten
Springer US (Verlag)
978-0-387-22794-8 (ISBN)

Lese- und Medienproben

Ebook-Leseprobe (PDF)

This book is appropriate for those specializing in speech science, hearing science, neuroscience, or computer science and engineers working on applications such as automatic speech recognition, cochlear implants, hands-free telephones, sound recording, multimedia indexing and retrieval.

There is a serious problem in the recognition of sounds. It derives from the fact that they do not usually occur in isolation but in an environment in which a number of sound sources (voices, traffic, footsteps, music on the radio, and so on) are active at the same time. When these sounds arrive at the ear of the listener, the complex pressure waves coming from the separate sources add together to produce a single, more complex pressure wave that is the sum of the individual waves. The problem is how to form separate mental descriptions of the component sounds, despite the fact that the "e;mixture wave"e; does not directly reveal the waves that have been summed to form it. The name auditory scene analysis (ASA) refers to the process whereby the auditory systems of humans and other animals are able to solve this mixture problem. The process is believed to be quite general, not specific to speech sounds or any other type of sounds, and to exist in many species other than humans. It seems to involve assigning spectral energy to distinct "e;auditory objects"e; and "e;streams"e; that serve as the mental representations of distinct sound sources in the environment and the patterns that they make as they change over time. How this energy is assigned will affect the perceived n- ber of auditory sources, their perceived timbres, loudnesses, positions in space, and pitches.

Speech Segregation: Problems and Perspectives.- Auditory Scene Analysis.- Speech separation.- Recurrent Timing Nets for F0-based Speaker Separation.- Blind Source Separation Using Graphical Models.- Speech Recognizer Based Maximum Likelihood Beamforming.- Exploiting Redundancy to Construct Listening Systems.- Automatic Speech Processing by Inference in Generative Models.- Signal Separation Motivated by Human Auditory Perception: Applications to Automatic Speech Recognition.- Speech Segregation Using an Event-synchronous Auditory Image and STRAIGHT.- Underlying Principles of a High-quality Speech Manipulation System STRAIGHT and Its Application to Speech Segregation.- On Ideal Binary Mask As the Computational Goal of Auditory Scene Analysis.- The History and Future of CASA.- Techniques for Robust Speech Recognition in Noisy and Reverberant Conditions.- Source Separation, Localization, and Comprehension in Humans, Machines, and Human-machine Systems.- The Cancellation Principle in Acoustic Scene Analysis.- Informational and Energetic Masking Effects in Multitalker Speech Perception.- Masking the Feature Information In Multi-stream Speech-analogue Displays.- Interplay Between Visual and Audio Scene Analysis.- Evaluating Speech Separation Systems.- Making Sense of Everyday Speech: a Glimpsing Account.

Erscheint lt. Verlag	16.1.2006
Zusatzinfo	XXIV, 319 p.
Verlagsort	New York
Sprache	englisch
Themenwelt	Geisteswissenschaften ► Psychologie ► Biopsychologie / Neurowissenschaften
	Informatik ► Software Entwicklung ► User Interfaces (HCI)
	Technik ► Elektrotechnik / Energietechnik
Schlagworte	Cognition • Computer Science • Development • Information • Multimedia • Neuroscience • quality • Science • Speech processing • Speech Recognition
ISBN-10	0-387-22794-6 / 0387227946
ISBN-13	978-0-387-22794-8 / 9780387227948

Haben Sie eine Frage zum Produkt?

PDF (Wasserzeichen)
Größe: 16,1 MB

DRM: Digitales Wasserzeichen
Dieses eBook enthält ein digitales Wasserzeichen und ist damit für Sie personalisiert. Bei einer missbräuchlichen Weitergabe des eBooks an Dritte ist eine Rückverfolgung an die Quelle möglich.

Dateiformat: PDF (Portable Document Format)
Mit einem festen Seitenlayout eignet sich die PDF besonders für Fachbücher mit Spalten, Tabellen und Abbildungen. Eine PDF kann auf fast allen Geräten angezeigt werden, ist aber für kleine Displays (Smartphone, eReader) nur eingeschränkt geeignet.

Systemvoraussetzungen:
PC/Mac: Mit einem PC oder Mac können Sie dieses eBook lesen. Sie benötigen dafür einen PDF-Viewer - z.B. den Adobe Reader oder Adobe Digital Editions.
eReader: Dieses eBook kann mit (fast) allen eBook-Readern gelesen werden. Mit dem amazon-Kindle ist es aber nicht kompatibel.
Smartphone/Tablet: Egal ob Apple oder Android, dieses eBook können Sie lesen. Sie benötigen dafür einen PDF-Viewer - z.B. die kostenlose Adobe Digital Editions-App.

Buying eBooks from abroad
For tax law reasons we can sell eBooks just within Germany and Switzerland. Regrettably we cannot fulfill eBook-orders from other countries.

Print-Ausgabe

Buch | Softcover

CHF 217,15