Figures and data

Timecourse and results of the species categorization task and the control task testing for potential attentional biases.
(A) Detail of the timecourse of four trials of the species categorization task in non-representative order, including waveform and spectrogram graphs for one example stimulus of each species. (B) Behavioral results of the task (N=23) showing the probability (Probab.) of correctly categorizing (categ.) each Species’ vocalization, using the 6 acoustic covariates from Model 3 and reaction times as variables. (C) Histograms of the acoustic Mahalanobis distance data of each species including mean (numbers represent exact mean value). (D) Control task and its design in an independent sample of 28 participants using each species’ vocalization as an exogenous cue preceding the target to be detected as fast as possible, namely a short sine wave tone (600Hz). (E) Results of the control task (N=28), showing no biasing of attention by any of the species stimuli: no species triggered attentional capture, yielding to no attentional advantage for target detection. For the results plots, violin plots illustrate distribution fit, error bars the standard error of the mean (SEM). Points represent individual values. ITI: inter trial interval; Resp.: response; Hum: human; Chimp: chimpanzee; Bon: bonobo; Mac: macaque. **p<.01, ***p<.001, n.s.: non-significant.


Posthoc contrasts for behavioral effects of the Species by Context interaction

Activations, cluster size and coordinates for each contrast of interest of model 3 (vocalization loudness, intensity, chagoe in spectrum, F2 bandwidth contour, F0 power and intensity contour difference as trial-level covariate of no-interest) in the sample-specific temporal voice areas, wholebrain voxelwise p<.05 FDR corrected, k>10.

Wholebrain results with TVA outlines when contrasting the processing of chimpanzee to other species’ vocalizations with vocalization loudness, intensity, change in spectrum, F2 bandwidth contour, F0 power and intensity contour difference as trial-level covariates of no-interest (model 3).
Enhanced brain activity on a sagittal view with activity for [chimpanzee > human,bonobo,macaque] (dark blue to green), [chimpanzee > bonobo,macaque] (brown to red with light yellow outline) and [macaque > human,bonobo,chimpanzee] (red to yellow) vocalizations, with outlines of TVA for voice > animal sounds (A,B), voice > nature sounds (C,D), voice > music (E,F) and voice > noise (G,H). Brain activations are independent of the most discriminant low-level acoustic parameters of the stimuli set [33]. Data corrected for multiple comparisons using wholebrain voxelwise false discovery rate (FDR) at a threshold of p<.05. Hum: human; Chimp: chimpanzee; Bon: bonobo; Mac: macaque. White outline: sample-specific temporal voice areas (TVA; N=23); Dotted black outline: sample-specific TVA, per sound category; Blue outline: areas selective to chimpanzee calls. ‘a’ prefix: anterior; ‘m’ prefix: mid; ‘p’ prefix: posterior; STG: superior temporal gyrus; STS: superior temporal sulcus; L: left hemisphere; R: right hemisphere.

Synthesis of mid-to-anterior TVA clusters of activity recruited specifically by the processing of chimpanzee and macaque vocalizations (Models 1,2,3).
Anterior superior temporal gyrus (aSTG) and sulcus (aSTS) clusters recruited for the processing of chimpanzee calls as opposed to human voices, bonobo, macaque calls (pink: model 1; purple: model 2; blue: model 3) in the general TVA (A,B, N=98) as well as in the sample-specific TVA (C,D, N=23). Macaque results are only significant for Model 3 (teal: Macaque vs all other species). Model 1: mean of fundamental frequency and energy (covariates of no-interest, N=2); Model 2: acoustic distance (covariate of no-interest, N=1); Model 3: acoustic parameters that characterize low-level acoustics of our stimuli following a discriminant analysis (covariates of no-interest, N=6). Data are all corrected for multiple comparison using wholebrain voxelwise false discovery rate (FDR) at a threshold of p<.05 with t-values ranging from 5 to 10. Hum: human; Chimp: chimpanzee; Bon: bonobo; Mac: macaque. TVA: temporal voice areas. Prefix ‘a’: anterior; ‘m’: mid. L / R: left / right hemisphere.

Model-based correlates of the probability of correct species categorization, within sample-specific TVA (Model 4).
Correlates of the probability of correct species categorization computed using model-based analysis technique for all species, as illustrated on sagittal renders for all species, including human (A,B), then specifically for chimpanzee calls (C,D), bonobo calls (E,F) and macaque calls (G,H). These correlates were constrained to the bounds of the sample-specific TVA (N=23, black outline) using an inclusive masking procedure with correction for multiple comparison using voxelwise false discovery rate (FDR) at a threshold of p<.05. The colorbars represents T-value statistics. TVA: temporal voice areas. Prefix ‘a’: anterior; ‘m’: mid; ‘p’: posterior. STG: superior temporal gyrus; STS: superior temporal sulcus; MTG: middle temporal gyrus; PT: planum temporale. L / R: left / right hemisphere.