Search bioRxivSearch

Biology subjects

Okanoya, K.

Publications and source records attributed to Okanoya, K..

2 recordsLinked to original sources

USVSEG: A robust segmentation of rodents’ ultrasonic vocalization

Rodents ultrasonic vocalization (USV) provides useful information to assess their social behaviors. Despite of previous efforts for classifying subcategories of time-frequency patterns of USV syllables to associate with their functional relevances, detection of vocal elements from continuously recorded data have remained to be not well-optimized. We here propose a novel procedure for detecting USV segments in continuous sound data with background noises which were inevitably contaminated during observation of the social behavior. The proposed procedure utilizes a stable version of spectrogram and additional signal processing for better separation of vocal signals by reducing variation of the background noise. Our procedure also provides a precise time tracking of spectral peaks within each syllable. We showed that this procedure can be applied to a variety of USVs obtained from several rodent species. A performance test with an appropriate parameter set showed performance for detecting USV syllables than conventional methods.

animal behavior and cognition

A language-familiarity effect on the recognition of computer-transformed vocal emotional cues

People are more accurate in voice identification and emotion recognition in their native language than in other languages, a phenomenon known as the language familiarity effect (LFE). Previous work on cross-cultural inferences of emotional prosody has left it difficult to determine whether these native-language advantages arise from a true enhancement of the auditory capacity to extract socially relevant cues in familiar speech signals or, more simply, from cultural differences in how these emotions are expressed. In order to rule out such production differences, this work employed algorithmic voice transformations to create pairs of stimuli in the French and Japanese language which differed by exactly the same amount of prosodic expression. Even though the cues were strictly identical in both languages, they were better recognized when participants processed them in their native language. This advantage persisted in three types of stimulus degradation (jabberwocky, shuffled and reversed sentences). These results provide univocal evidence that production differences are not the sole drivers of LFEs in cross-cultural emotion perception, and suggest that it is the listeners lack of familiarity with the individual speech sounds of the other language, and not e.g. with their syntax or semantics, which impairs their processing of higher-level emotional cues.

neuroscience