Finding, visualizing, and quantifying latent structure across diverse animal vocal repertoires.
Finding, visualizing, and quantifying latent structure across diverse animal vocal repertoires.
复制标题
发现,可视化和量化各种动物声曲目的潜在结构。
DOI:
10.1371/journal.pcbi.1008228
复制
发表时间:
2020-10
影响因子:
4.3
通讯作者:
Gentner TQ
中科院分区:
文献类型:
--
作者:
Sainburg T;Thielk M;Gentner TQ
Animals produce vocalizations that range in complexity from a single repeated call to hundreds of unique vocal elements patterned in sequences unfolding over hours. Characterizing complex vocalizations can require considerable effort and a deep intuition about each species’ vocal behavior. Even with a great deal of experience, human characterizations of animal communication can be affected by human perceptual biases. We present a set of computational methods for projecting animal vocalizations into low dimensional latent representational spaces that are directly learned from the spectrograms of vocal signals. We apply these methods to diverse datasets from over 20 species, including humans, bats, songbirds, mice, cetaceans, and nonhuman primates. Latent projections uncover complex features of data in visually intuitive and quantifiable ways, enabling high-powered comparative analyses of vocal acoustics. We introduce methods for analyzing vocalizations as both discrete sequences and as continuous latent variables. Each method can be used to disentangle complex spectro-temporal structure and observe long-timescale organization in communication. Of the thousands of species that communicate vocally, the repertoires of only a tiny minority have been characterized or studied in detail. This is due, in large part, to traditional analysis methods that require a high level of expertise that is hard to develop and often species-specific. Here, we present a set of unsupervised methods to project animal vocalizations into latent feature spaces to quantitatively compare and develop visual intuitions about animal vocalizations. We demonstrate these methods across a series of analyses over 19 datasets of animal vocalizations from 29 different species, including songbirds, mice, monkeys, humans, and whales. We show how learned latent feature spaces untangle complex spectro-temporal structure, enable cross-species comparisons, and uncover high-level attributes of vocalizations such as stereotypy in vocal element clusters, population regiolects, coarticulation, and individual identity.
登录
查看更多内容
影响因子:
46.9
作者:
Becht, Etienne;McInnes, Leland;Newell, Evan W.
通讯作者:
Newell, Evan W.
影响因子:
1.3
作者:
ADRETHAUSBERGER, M;JENKINS, PF
通讯作者:
JENKINS, PF
DOI:
10.1073/pnas.1607601113
发表时间:
2016-10-18
影响因子:
11.1
作者:
Berman, Gordon J.;Bialek, William;Shaevitz, Joshua W.
通讯作者:
Shaevitz, Joshua W.
影响因子:
5.1
作者:
Arriaga, Julio G.;Cody, Martin L.;Taylor, Charles E.
通讯作者:
Taylor, Charles E.
DOI:
10.1073/pnas.1515380113
发表时间:
2016-02-09
影响因子:
11.1
作者:
Bregman, Micah R.;Patel, Aniruddh D.;Gentner, Timothy Q.
通讯作者:
Gentner, Timothy Q.