Cargando…

Speech Separation by Humans and Machines

The "cocktail-party effect" - the ability to focus on one voice in a sea of noises - is a highly sophisticated skill that is usually effortless to listeners but largely impossible for machines. Investigating and unraveling this capacity spans numerous fields including psychology, physiolog...

Descripción completa

Detalles Bibliográficos
Clasificación:	Libro Electrónico
Autor Corporativo:	SpringerLink (Online service)
Otros Autores:	Divenyi, Pierre (Editor )
Formato:	Electrónico eBook
Idioma:	Inglés
Publicado:	New York, NY : Springer US : Imprint: Springer, 2005.
Edición:	1st ed. 2005.
Temas:	Signal processing. User interfaces (Computer systems). Human-computer interaction. Engineering. Signal, Speech and Image Processing . User Interfaces and Human Computer Interaction. Technology and Engineering.
Acceso en línea:	Texto Completo

MARC


LEADER	00000nam a22000005i 4500
001	978-0-387-22794-8
003	DE-He213
005	20220118100349.0
007	cr nn 008mamaa
008	100301s2005 xxu\| s \|\|\|\| 0\|eng d
020			\|a 9780387227948 \|9 978-0-387-22794-8
024	7		\|a 10.1007/b99695 \|2 doi
050		4	\|a TK5102.9
072		7	\|a TJF \|2 bicssc
072		7	\|a UYS \|2 bicssc
072		7	\|a TEC008000 \|2 bisacsh
072		7	\|a TJF \|2 thema
072		7	\|a UYS \|2 thema
082	0	4	\|a 621.382 \|2 23
245	1	0	\|a Speech Separation by Humans and Machines \|h [electronic resource] / \|c edited by Pierre Divenyi.
250			\|a 1st ed. 2005.
264		1	\|a New York, NY : \|b Springer US : \|b Imprint: Springer, \|c 2005.
300			\|a XXIV, 319 p. \|b online resource.
336			\|a text \|b txt \|2 rdacontent
337			\|a computer \|b c \|2 rdamedia
338			\|a online resource \|b cr \|2 rdacarrier
347			\|a text file \|b PDF \|2 rda
505	0		\|a Speech Segregation: Problems and Perspectives -- Auditory Scene Analysis -- Speech separation -- Recurrent Timing Nets for F0-based Speaker Separation -- Blind Source Separation Using Graphical Models -- Speech Recognizer Based Maximum Likelihood Beamforming -- Exploiting Redundancy to Construct Listening Systems -- Automatic Speech Processing by Inference in Generative Models -- Signal Separation Motivated by Human Auditory Perception: Applications to Automatic Speech Recognition -- Speech Segregation Using an Event-synchronous Auditory Image and STRAIGHT -- Underlying Principles of a High-quality Speech Manipulation System STRAIGHT and Its Application to Speech Segregation -- On Ideal Binary Mask As the Computational Goal of Auditory Scene Analysis -- The History and Future of CASA -- Techniques for Robust Speech Recognition in Noisy and Reverberant Conditions -- Source Separation, Localization, and Comprehension in Humans, Machines, and Human-machine Systems -- The Cancellation Principle in Acoustic Scene Analysis -- Informational and Energetic Masking Effects in Multitalker Speech Perception -- Masking the Feature Information In Multi-stream Speech-analogue Displays -- Interplay Between Visual and Audio Scene Analysis -- Evaluating Speech Separation Systems -- Making Sense of Everyday Speech: a Glimpsing Account.
520			\|a The "cocktail-party effect" - the ability to focus on one voice in a sea of noises - is a highly sophisticated skill that is usually effortless to listeners but largely impossible for machines. Investigating and unraveling this capacity spans numerous fields including psychology, physiology, engineering, and computer science. All these perspectives are brought together in this volume which, for the first time, provides a comprehensive and authoritative discussion of our understanding of how humans separate speech, and the state of the art in approaching these abilities with machines. This material is drawn from an October 2003 workshop, sponsored by the National Science Foundation, on speech separation. Leading authorities from around the world were invited to present their perspectives and discuss the points of contact to other perspectives. The result is a clear and uniform overview of this problem, and a primer in what is emerging as an important, active and successful area for the development of new techniques and applications. Chapters include historical and current summaries of relevant research in behavioral science, neuroscience and engineering, along with more in-depth descriptions of several of the most exciting current research projects and techniques, including the latest experimental results illuminating how listeners organize the mixtures of sound they hear, and the most powerful and successful signal processing and machine learning techniques for the separation of real-world recordings of sound mixtures by one or more microphones. There is no comparable collection that seeks to bring together the underlying experimental science and the wide variety of technical approaches to give an integrated picture of the problem and solutions to speech separation. Those specializing in speech science, hearing science, neuroscience, or computer science and engineers working on applications such as automatic speech recognition, cochlear implants, hands-free telephones, sound recording, multimedia indexing and retrieval will find Speech Separation by Humans and Machines a useful and inspiring read.
650		0	\|a Signal processing.
650		0	\|a User interfaces (Computer systems).
650		0	\|a Human-computer interaction.
650		0	\|a Engineering.
650	1	4	\|a Signal, Speech and Image Processing .
650	2	4	\|a User Interfaces and Human Computer Interaction.
650	2	4	\|a Technology and Engineering.
700	1		\|a Divenyi, Pierre. \|e editor. \|4 edt \|4 http://id.loc.gov/vocabulary/relators/edt
710	2		\|a SpringerLink (Online service)
773	0		\|t Springer Nature eBook
776	0	8	\|i Printed edition: \|z 9780387522456
776	0	8	\|i Printed edition: \|z 9781441954602
776	0	8	\|i Printed edition: \|z 9781402080012
856	4	0	\|u https://doi.uam.elogim.com/10.1007/b99695 \|z Texto Completo
912			\|a ZDB-2-ENG
912			\|a ZDB-2-SXE
950			\|a Engineering (SpringerNature-11647)
950			\|a Engineering (R0) (SpringerNature-43712)

Speech Separation by Humans and Machines

MARC

Ejemplares similares