Back to Search Start Over

Listening with generative models.

Authors :
Cusimano M
Hewitt LB
McDermott JH
Source :
Cognition [Cognition] 2024 Dec; Vol. 253, pp. 105874. Date of Electronic Publication: 2024 Aug 30.
Publication Year :
2024

Abstract

Perception has long been envisioned to use an internal model of the world to explain the causes of sensory signals. However, such accounts have historically not been testable, typically requiring intractable search through the space of possible explanations. Using auditory scenes as a case study, we leveraged contemporary computational tools to infer explanations of sounds in a candidate internal generative model of the auditory world (ecologically inspired audio synthesizers). Model inferences accounted for many classic illusions. Unlike traditional accounts of auditory illusions, the model is applicable to any sound, and exhibited human-like perceptual organization for real-world sound mixtures. The combination of stimulus-computability and interpretable model structure enabled 'rich falsification', revealing additional assumptions about sound generation needed to account for perception. The results show how generative models can account for the perception of both classic illusions and everyday sensory signals, and illustrate the opportunities and challenges involved in incorporating them into theories of perception.<br /> (Copyright © 2024 The Authors. Published by Elsevier B.V. All rights reserved.)

Details

Language :
English
ISSN :
1873-7838
Volume :
253
Database :
MEDLINE
Journal :
Cognition
Publication Type :
Academic Journal
Accession number :
39216190
Full Text :
https://doi.org/10.1016/j.cognition.2024.105874