Listeners Are Biased towards Voices of Young Speakers and Female Speakers When Discriminating Voices

Saved in:
Bibliographic Details
Title: Listeners Are Biased towards Voices of Young Speakers and Female Speakers When Discriminating Voices
Language: English
Authors: Valeriia Vyshnevetska (ORCID 0009-0003-3355-9580), Nathalie Giroud, Meike Ramon, Volker Dellwo
Source: Cognitive Research: Principles and Implications. 2025 10.
Availability: Springer. Available from: Springer Nature. One New York Plaza, Suite 4600, New York, NY 10004. Tel: 800-777-4643; Tel: 212-460-1500; Fax: 212-460-1700; e-mail: customerservice@springernature.com; Web site: https://link.springer.com/
Peer Reviewed: Y
Page Count: 14
Publication Date: 2025
Document Type: Journal Articles
Reports - Research
Descriptors: Listening Skills, Bias, Auditory Discrimination, Females, Young Adults, Older Adults, Age Differences, Recognition (Psychology), Responses
DOI: 10.1186/s41235-025-00636-3
ISSN: 2365-7464
Abstract: In face processing, an own-age recognition advantage has frequently been reported whereby observers are better at recognizing faces of their own compared to other age groups. We wanted to know whether own-age effects exist in voice recognition. Two listener groups, younger adults (n = 42, 19-35 years, 21 males) and older adults (n = 32, 65-83 years, 14 males), completed a speaker discrimination task (same/different speakers), which included younger and older adult speakers of both sexes. Results revealed no interaction of the factors speaker and listener age and speaker and listener sex on listeners' sensitivity (d'). Main effects were significant for listener age (young adult listeners exhibited higher sensitivity than the older adult listeners) and speaker sex (listeners' sensitivity was higher for male compared to female voices). Crucially, response bias (c) revealed that listeners had a significantly higher 'same' bias when hearing younger speakers and female speakers. Our findings have implications for theories of voice identity processing and forensic contexts requiring discrimination of speakers' identity, e.g. earwitnesses telling apart younger and female speakers.
Abstractor: As Provided
Entry Date: 2025
Accession Number: EJ1473391
Database: ERIC
Full text is not displayed to guests.
FullText Links:
  – Type: pdflink
    Url: https://content.ebscohost.com/cds/retrieve?content=AQICAHj0k_4E0hTGH8RJwT4gCJyBsGNe_WN95AvKlDbXJGqwxwGAjKSD5xgt8MPH7T0Bt5LKAAAA4zCB4AYJKoZIhvcNAQcGoIHSMIHPAgEAMIHJBgkqhkiG9w0BBwEwHgYJYIZIAWUDBAEuMBEEDPnIiHilyZkIt8LKlwIBEICBm2TdSyPSAlwFI06yb4Kb_qUeqr4P3Pp-_zOE8sZRGFjmlnNJvz1XTtdL5KcZmoke-DG2AelJd30mPdmeJr8nuvtMuWLZy5cgJv7xHnwqK686UbVEYr_buUXCOR8haN_qR5q2jmUPxFZmt_G6ecnJfye6jIjO0HRe4weI3s3KdF10D1LmQZQor1c6gWA7f7-c1RDXHkARZbI567L2
Text:
  Availability: 1
  Value: <anid>AN0185782071;[k1e6]07jun.25;2025Jun10.03:08;v2.2.500</anid> <title id="AN0185782071-1">Listeners are biased towards voices of young speakers and female speakers when discriminating voices </title> <p>In face processing, an own-age recognition advantage has frequently been reported whereby observers are better at recognizing faces of their own compared to other age groups. We wanted to know whether own-age effects exist in voice recognition. Two listener groups, younger adults (n = 42, 19–35 years, 21 males) and older adults (n = 32, 65–83 years, 14 males), completed a speaker discrimination task (same/different speakers), which included younger and older adult speakers of both sexes. Results revealed no interaction of the factors speaker and listener age and speaker and listener sex on listeners' sensitivity (d′). Main effects were significant for listener age (young adult listeners exhibited higher sensitivity than the older adult listeners) and speaker sex (listeners' sensitivity was higher for male compared to female voices). Crucially, response bias (c) revealed that listeners had a significantly higher 'same' bias when hearing younger speakers and female speakers. Our findings have implications for theories of voice identity processing and forensic contexts requiring discrimination of speakers' identity, e.g. earwitnesses telling apart younger and female speakers.</p> <p>Keywords: Speaker recognition; Own-age effect; Response bias; Psychology and Cognitive Sciences Psychology</p> <hd id="AN0185782071-2">Significance statement</hd> <p>Recently, 'fake police officer' crimes were reported in many countries. In these scenarios, older individuals are contacted by younger persons pretending to be police officials and attempting to deceive older adults of their valuables. Crucially, the younger individuals do not represent official state security but are part of the arranged scam. What is more, the so-called deepfake voices are gaining popularity, by which novel speech utterances can be generated with an individual's voice using artificial intelligence techniques. Audio deepfakes have led to crimes whereby victims are tricked to believe that one of their family members or close friends is in urgent need of financial help, and in such scenarios, fraudsters' and victims' age may differ drastically. Thus, an earwitness at court might have to judge whether the voice of a suspect from a strongly different age group belongs to a speaker they spoke to on the telephone during the voice crime. However, it is unclear whether listeners are better at recognizing voices of their own age compared to other ages. This study showed that when hearing younger speaker pairs and female speaker pairs, listeners are significantly biased to saying that both excerpts stem from the same speaker. In voice crime cases discussed earlier, this could imply that earwitnesses might find it more challenging to discriminate younger speakers and female speakers, especially if the audio quality is poor.</p> <hd id="AN0185782071-3">Introduction</hd> <p>Voice recognition is a seemingly effortless everyday task which nevertheless can be affected by many speaker-, listener-, and channel-related factors such as listener's degree of familiarity with the voice (Hollien et al., [<reflink idref="bib36" id="ref1">36</reflink>]; Schmidt‐Nielsen & Stern, [<reflink idref="bib88" id="ref2">88</reflink>]), voice distinctiveness (Papcun et al., [<reflink idref="bib71" id="ref3">71</reflink>]), duration of exposure to a voice (Clifford, [<reflink idref="bib18" id="ref4">18</reflink>]; Foulkes & Barron, [<reflink idref="bib26" id="ref5">26</reflink>]; McGehee, [<reflink idref="bib61" id="ref6">61</reflink>]), the amount of time elapsed between exposure and test (Clifford, [<reflink idref="bib18" id="ref7">18</reflink>]), as well as communication channel quality (McDougall et al., [<reflink idref="bib60" id="ref8">60</reflink>]; Nolan et al., [<reflink idref="bib68" id="ref9">68</reflink>]; Rathborn et al., [<reflink idref="bib80" id="ref10">80</reflink>]), to name just a few (for a review of different factors, see Jessen ([<reflink idref="bib38" id="ref11">38</reflink>]); McDougall et al. ([<reflink idref="bib60" id="ref12">60</reflink>])). Listener's age has also been shown to affect various voice perception tasks, with younger adult listeners outperforming older adult listeners (Best et al., [<reflink idref="bib9" id="ref13">9</reflink>]; Goy et al., [<reflink idref="bib31" id="ref14">31</reflink>]; Kausler & Puckett, [<reflink idref="bib40" id="ref15">40</reflink>]; Moyse et al., [<reflink idref="bib64" id="ref16">64</reflink>]; Schvartz & Chatterjee, [<reflink idref="bib91" id="ref17">91</reflink>]; Zaltz & Kishon-Rabin, [<reflink idref="bib114" id="ref18">114</reflink>]). However, it remains unclear whether younger adults outperform older adults in recognizing voices of all ages or only of voices of their own-age group, a phenomenon referred to as own-group advantage in other contexts, predominantly face recognition (see below). This study explores own-age advantages in voice discrimination.</p> <p>Recently, voice crime scenarios became widespread whereby older adults are contacted on the phone by young individuals pretending to be police officers and informed about criminal activity in the nearby area (Action Fraud, [<reflink idref="bib27" id="ref19">27</reflink>]; Hadfield, [<reflink idref="bib33" id="ref20">33</reflink>]). As a security measure, older adult persons are instructed to hand over valuables they possess at home to the fictitious police officers or provide access to their bank account. Crucially, the young callers do not represent any official state security but are part of an arranged scam by which older adults get deceived of their possessions. Such scenario is known as the 'fake police officer' crime, and numerous variants of this type of voice crimes are frequently reported nowadays from many countries such as the UK (Ball, [<reflink idref="bib5" id="ref21">5</reflink>]; Loffreda, [<reflink idref="bib53" id="ref22">53</reflink>]), Germany (Schumacher, [<reflink idref="bib90" id="ref23">90</reflink>]) or Switzerland (Der Landbote, [<reflink idref="bib46" id="ref24">46</reflink>]; Swiss Banking Ombudsman, [<reflink idref="bib98" id="ref25">98</reflink>]) amongst others. What is more, the recent introduction of so-called deepfake voices, by which novel speech utterances can be generated with an individual's voice using deep neural network learning techniques, has led to crimes whereby victims are tricked to believe that one of their family members or close friends is in urgent need of financial help (Brewster, [<reflink idref="bib14" id="ref26">14</reflink>]; Flitter & Cowley, [<reflink idref="bib24" id="ref27">24</reflink>]; Khatsenkova, [<reflink idref="bib41" id="ref28">41</reflink>]). In such scenarios, a fraudsters' age appearing in the context of the call and the victims' age may differ drastically when an old person is frauded by a young caller, for example. Given these circumstances, we expect a strong increase in court cases, in which listeners appear as earwitnesses and are asked to give evidence whether the voice of a suspect is the voice of the person they spoke to on the phone during the frauded phone calls or otherwise give evidence in a formal voice identification procedure commonly known as a voice parade (McDougall, [<reflink idref="bib59" id="ref29">59</reflink>]; Robson, [<reflink idref="bib83" id="ref30">83</reflink>]). Such scenarios increase the diversity of person characteristics in interaction which may all have an impact on auditory speaker recognition and discrimination ability. Thus, a witness at court might have to judge whether the voice of a suspect from a strongly different age group belongs to a speaker they spoke to on the telephone during the voice crime. Listeners might have more experience with the voices of their own-age group due to relatively increased exposure compared to voices of other ages. Therefore, we asked whether listeners' age group positively impacts discrimination of voices from the same age group, henceforth the own-age advantage in voice discrimination.</p> <p>Own-group advantages (sometimes also known as own-group biases[<reflink idref="bib1" id="ref31">1</reflink>]) refer to a wide spectrum of phenomena by which an individual has an advantage in making perceptual judgements about a person based on some shared group attribution (in-group) compared to missing group attributions (out-group). These effects have dominantly been researched in the field of face identity processing (Denkinger & Kinn, [<reflink idref="bib22" id="ref32">22</reflink>]; Herlitz & Lovén, [<reflink idref="bib35" id="ref33">35</reflink>]; Mason, [<reflink idref="bib58" id="ref34">58</reflink>]; Meissner & Brigham, [<reflink idref="bib62" id="ref35">62</reflink>]; Rhodes & Anastasi, [<reflink idref="bib81" id="ref36">81</reflink>]; Sporer, [<reflink idref="bib94" id="ref37">94</reflink>]; Wright & Sladden, [<reflink idref="bib109" id="ref38">109</reflink>]). The predominant explanation of the effect suggests that observers have more experience with individual fine details of the own-group stimuli, which increases their recognition ability (Rhodes & Anastasi, [<reflink idref="bib81" id="ref39">81</reflink>]). Perhaps the best known example is the so-called own-race bias (Meissner & Brigham, [<reflink idref="bib62" id="ref40">62</reflink>]) by which faces of the perceivers own-ethnic group are better recognized compared to faces of a different ethnic group (often wrongly referred to as 'race'). However, studies on own-age recognition advantages have generated conflicting results: while a number of studies reported superior sensitivity for faces of observers' own-age group (Anastasi & Rhodes, [<reflink idref="bib1" id="ref41">1</reflink>]; Denkinger & Kinn, [<reflink idref="bib22" id="ref42">22</reflink>]; He et al., [<reflink idref="bib34" id="ref43">34</reflink>]; Wright & Stroud, [<reflink idref="bib110" id="ref44">110</reflink>]), other studies failed to observe it (Memon et al., [<reflink idref="bib63" id="ref45">63</reflink>]; Proietti et al., [<reflink idref="bib75" id="ref46">75</reflink>]; Rose et al., [<reflink idref="bib85" id="ref47">85</reflink>]). Further, some studies report the own-age recognition advantage in all of the investigated age groups (Wright & Stroud, [<reflink idref="bib110" id="ref48">110</reflink>]), while others observe it only for particular age groups (Anastasi & Rhodes, [<reflink idref="bib1" id="ref49">1</reflink>]; Denkinger & Kinn, [<reflink idref="bib22" id="ref50">22</reflink>]). It is worth noting that many tests used to assess face identity processing involve skewed sex and age compositions — in terms of either stimuli or participants (e.g. cf. Fysh et al. ([<reflink idref="bib28" id="ref51">28</reflink>]); Stacchi et al. ([<reflink idref="bib95" id="ref52">95</reflink>])) — and are characterized by low to modest reliability in neurotypical observers (Bobak et al., [<reflink idref="bib10" id="ref53">10</reflink>]).</p> <p>Own-group effects in <emph>voice</emph> identity processing received much less attention compared to factors influencing overall voice recognition performance discussed earlier. For example, it has been shown that listeners are better at describing accents which are geographically closer to their own (Braber et al., [<reflink idref="bib12" id="ref54">12</reflink>]; Tompkinson & Watt, [<reflink idref="bib100" id="ref55">100</reflink>]) and recognizing own-accent voices better than other-accent voices (Stevenage et al., [<reflink idref="bib97" id="ref56">97</reflink>]). Some studies demonstrated own-gender effects, whereby listeners were better at identifying voices of their own compared to other sexes (Roebuck & Wilding, [<reflink idref="bib84" id="ref57">84</reflink>]; Skuk & Schweinberger, [<reflink idref="bib93" id="ref58">93</reflink>]). Results on own-group effects in relation to speakers' and listeners' age are very limited. (Moyse et al., [<reflink idref="bib64" id="ref59">64</reflink>]) studied age estimation from voices and reported an own-group age estimation advantage for older adult listeners but not for younger adult listeners. Importantly, own-age advantages for voices (i.e., pertaining to voice recognition and discrimination) remain obscure.</p> <p>Voice crime in the past mostly involved young male speakers aged approximately 18 to 40 years (Michael Jessen, German Federal Police Office, personal communication). Female voices appear less often as evidence in crime, even though their numbers have increased over the past 15 years and currently constitute between 5 and 10% of cases in the UK and Germany (Kirsty McDougall, University of Cambridge, personal communication; Richard Rhodes, The Forensic Voice Centre, personal communication; Michael Jessen, German Federal Police Office, personal communication). Cases like the 'fake police officer' possibly introduce a new dimension of female voice crime as female voices may intuitively be more trustworthy (Schirmer et al., [<reflink idref="bib87" id="ref60">87</reflink>]) when convincing older adult targets of crime in opening the door to an alleged police officer. As such, possible own-age voice discrimination advantages cannot be studied without considering own-sex effects.</p> <p>In this study, we tested own-age effects in voice discrimination for male and female voices in male and female listeners using speaker discrimination task (i.e., same-/different-speaker judgements). We used speaker discrimination to avoid speaker- and listener-related familiarization effects. Listeners' ability to learn and recognize voices can be affected by many factors, including set size (Legge et al., [<reflink idref="bib51" id="ref61">51</reflink>]), distinctiveness (Papcun et al., [<reflink idref="bib71" id="ref62">71</reflink>]), listeners cognitive abilities (Best et al., [<reflink idref="bib9" id="ref63">9</reflink>]) and familiarity (Case et al., [<reflink idref="bib17" id="ref64">17</reflink>]; Hollien et al., [<reflink idref="bib36" id="ref65">36</reflink>]). To limit these complex and compounding effects, we opted for a same/different judgement (i.e., voice discrimination) task, which does not require listeners to form and consolidate abstract voice representations in memory. We applied signal detection theory (SDT) to quantify listeners' sensitivity (<emph>d</emph>′) as a measure of discrimination performance and response bias (<emph>c</emph>) as a measure of listeners' tendency to respond 'same' or 'different', when the stimulus is ambiguous to them, for example. Response bias has been overlooked in the past and is particularly crucial for forensic applications because it indicates listeners' tendencies to accept or reject a stimulus as familiar when in doubt, when stimulus quality is poor or in extreme cases without the presence of the stimulus itself.</p> <hd id="AN0185782071-4">Material and methods</hd> <p></p> <hd id="AN0185782071-5">Database and speakers</hd> <p>The materials for the current experiment were drawn from the TEVOID corpus (Dellwo et al., [<reflink idref="bib20" id="ref66">20</reflink>]; Pellegrino et al., [<reflink idref="bib73" id="ref67">73</reflink>]), which contains read and spontaneous sentence recordings from younger adults (henceforth YA speakers, range<subs>Age</subs> = 18–32 years, <emph>M</emph><subs>Age</subs> = 30.3 years, standard deviation (SD) = 6.6 years) and older adults (henceforth OA speakers, range<subs>Age</subs> = 66–81 years, <emph>M</emph><subs>Age</subs> = 71.7 years; SD = 4.9 years). All speakers were fluent native speakers of Zurich German, the Alemannic dialect spoken in the city and in most parts of the Canton of Zurich, and all sentence recordings were produced in Zurich German. The recordings were made in a sound treated booth using professional equipment and digitized at 44.1 kHz, 16 kbit/s bitrate. In this study, speech from 20 TEVOID speakers was included: 10 YA speakers (5 males and 5 females) and 10 OA speakers (5 males and 5 females).</p> <hd id="AN0185782071-6">Listeners</hd> <p>In total, 74 listeners completed the experiment, including 42 YA listeners (21 males, range<subs>Age</subs> = 19–35 years, <emph>M</emph><subs>Age</subs> = 26.7 years; <emph>SD</emph><subs>Age</subs> = 3.6 years) and 32 OA listeners (14 males, range<subs>Age</subs> = 65–83 years, <emph>M</emph><subs>Age</subs> = 73 years; <emph>SD</emph><subs>Age</subs> = 5.7 years). All listeners were native speakers of Swiss German and had lived in Zurich for a substantial number of years. They did not learn a second language before the age of seven. None of the listeners reported any history of speech or language deficits.</p> <p>To ensure that OA listeners' performance was not influenced by age-related cognitive impairment, we performed Montreal Cognitive Assessment (MoCA) (Nasreddine et al., [<reflink idref="bib67" id="ref68">67</reflink>]). 30 OA listeners had MoCA score ≥ 26 suggesting they had no cognitive impairment, while the MoCA data from two OA listeners were not recorded due to technical reasons. Hearing loss in OA listeners did not exceed moderate hearing loss: mean pure-tone audiometry (PTA) threshold was 19.5 dB, standard deviation 11.3 dB for the octave frequencies from 0.5 to 4 kHz. Listeners with > 50 dB hearing threshold in the better hearing ear were excluded, since this is the upper threshold for moderate hearing loss defined by WHO (World Health Organization, [<reflink idref="bib108" id="ref69">108</reflink>]). All OA listeners had symmetrical hearing (< 15 dB interaural threshold difference), and none of them were usering hearing aids. No PTA thresholds were measured for YA listeners.</p> <hd id="AN0185782071-7">Materials</hd> <p>We used a speaker discrimination task (same/different judgement) to investigate speaker discrimination performance in YA and OA listeners. To create same- and different-speaker pairs, we used read sentence recordings from the TEVOID corpus (see Sect. "Database and speakers'). In total, the pool of 1820 read sentence recordings was used to construct stimuli pairs (20 speakers × 91 sentences). To create speaker pairs, sentence stimuli were resampled to 10 kHz, and 800 ms snippets were extracted from each sentence midpoint using Hanning window over the frequency range of 80–5000 Hz with 40 Hz slope. Each speaker pair thus consisted of two 800 ms speech snippets separated by a 500 ms silent interval. While extracting snippets from a sentence midpoint may lead to a decrease in grammaticality and intelligibility, multiple studies show that voice recognition and discrimination is possible with unintelligible stimuli, for example, in time-reversed or noise-vocoded speech (Fleming et al., [<reflink idref="bib23" id="ref70">23</reflink>]; Garrido et al., [<reflink idref="bib29" id="ref71">29</reflink>]). Furthermore, we chose 800 ms snippet length for the current task since previous studies show that voice discrimination performance is optimal with stimuli length between 500 and 1000 ms and that further increase in duration does not increase the performance (Bricker & Pruzansky, [<reflink idref="bib15" id="ref72">15</reflink>]; Pollack et al., [<reflink idref="bib74" id="ref73">74</reflink>]). We limited the bandwidth of our stimuli to exclude any possible high-frequency artefacts that might be audible, and to present listeners exclusively with relevant speech and speaker information below 5 kHz. The remaining information contains sufficient speaker-specific voice detail for successful discrimination and recognition, as demonstrated exhaustively by studies using landline telephone speech as stimuli which has bandwidth of approximately 300–3400 kHz (Köster & Schiller, [<reflink idref="bib42" id="ref74">42</reflink>]; McDougall, [<reflink idref="bib59" id="ref75">59</reflink>]; Nolan et al., [<reflink idref="bib68" id="ref76">68</reflink>]; Rathborn et al., [<reflink idref="bib80" id="ref77">80</reflink>]; Schiller & Koster, [<reflink idref="bib86" id="ref78">86</reflink>]).</p> <p>Each listener received a unique subset of 80 stimuli pairs, in which equal number of same- and different-speaker pairs, younger and older speaker pairs, as well as female and male speaker pairs were included. In different-speaker pairs, stimuli were always matched for age and sex, so no speaker pairs contained mismatched stimuli by age and/or sex of the speakers. Sentence numbers within each speaker pair were mismatched. This way, in same-speaker pairs listeners never compared two identical stimuli tokens. In different-speaker pairs, listeners never compared linguistically identical sentences produced by two different speakers. Speaker pairs were created using Praat scripts, Praat version 6.1.51 (Boersma & Weenink, [<reflink idref="bib11" id="ref79">11</reflink>]).</p> <hd id="AN0185782071-8">Procedure</hd> <p>Testing took place in person at the Linguistic Research Infrastructure (LiRI) laboratory at the University of Zurich. Listeners performed the task in a soundproof booth, where they were seated at a desktop PC. Sound was played back through a loudspeaker, since not all OAs were comfortable with wearing headphones and since such listening conditions may be considered more realistic compared to listening to voices via closed-up headphones. The loudspeaker was situated to the left side of the PC monitor, at approximately 70 cm distance from the listeners, and they could adjust the loudness level to their comfort. The experiment was created and administered via the Gorilla experiment builder (Anwyl-Irvine et al., [<reflink idref="bib2" id="ref80">2</reflink>]). On every trial (<emph>N</emph> = 80), listeners heard a pair of audio snippets and were instructed to indicate whether both snippets stemmed from one speaker or from two different speakers using buttons on the screen. No other answer options were available. The audio was played automatically, and listeners heard stimuli in each trial only once before giving an answer to ensure that all listeners receive the same duration of voice input. Listeners were instructed to complete the task at their own pace, and no time limits were introduced for completing the task (it took on average 10 min to complete). Eight attention checks were also included in the task to ensure listeners stayed attentive during the experiment. During attention checks, listeners were shown animal cartoon pictures on the screen and instructed to type animals' names in the text field below.</p> <hd id="AN0185782071-9">Measures</hd> <p>Using signal detection theory (SDT), listeners' performance was quantified with measures of sensitivity (<emph>d</emph>′) and response bias (<emph>c</emph>) (Macmillan & Creelman, [<reflink idref="bib55" id="ref81">55</reflink>]; Stanislaw & Todorov, [<reflink idref="bib96" id="ref82">96</reflink>]). Sensitivity is broadly conceived as the ability to perceive a signal (Macmillan & Creelman, [<reflink idref="bib54" id="ref83">54</reflink>]), whereas response bias is defined as subjects' tendency to prefer one type of response over the other (Stanislaw & Todorov, [<reflink idref="bib96" id="ref84">96</reflink>]). <emph>d</emph>′ and <emph>c</emph> values were calculated per listener and condition, such that from each listener we obtained four <emph>d</emph>′ and four <emph>c</emph> values corresponding to each of the four experimental conditions (i.e., YA and OA speakers, as well as female and male speakers). Data were inspected for quality prior to performing statistical analyses. Data would have been excluded if 20% or more of attention checks were solved incorrectly and/or if performance in the discrimination task was at chance level or below (i.e., 50% or less correct responses), since this might indicate that listeners did not stay attentive or were unable to solve the task, for example, because of an inability to recognize voices. All listeners completed all attention checks correctly and performed significantly above chance level, so no data were discarded based on these exclusion criteria. Thus, the final dataset for analyses comprised 296 <emph>d</emph>′ and 296 <emph>c</emph> values (74 listeners × 4 conditions).</p> <hd id="AN0185782071-10">Acoustic analyses</hd> <p>To inspect overall acoustic characteristics of our voice sample, we investigated acoustic differences between male and female, younger and older voices. We extracted <emph>f</emph><subs>0</subs> contours from all 800 ms speech snippets which were presented to the listeners and calculated mean <emph>f</emph><subs>0</subs>, <emph>f</emph><subs>0</subs> range (<emph>f</emph><subs>0<emph>max</emph></subs>–<emph>f</emph><subs>0<emph>min</emph></subs>) and <emph>f</emph><subs>0</subs> coefficient of variation, a standardized measure of <emph>f</emph><subs>0</subs> variation computed as (<emph>f</emph><subs>0<emph>SD</emph></subs><emph>/f</emph><subs>0<emph>Mean</emph></subs><emph>)</emph>*100. All measures were calculated on a linear scale in Hz in Praat version 6.1.51 (Boersma & Weenink, [<reflink idref="bib11" id="ref85">11</reflink>]) using gender-specific pitch ranges (75–400 Hz for male, 120–600 Hz for female voices). Figure 1B presents distributions of individual speaker's <emph>f</emph><subs>0</subs> mean values in Hz.</p> <p>Graph: Fig. 1 A Density plots showing distributions of mean f0 values for younger and older, as well as male and female speakers. Dashed lines represent group means. B Boxplots showing mean, range and interquartile range of mean f0 values per speaker. 'f_old'—older female speakers, 'f_yng'—younger female speakers, 'm_old'—older male speakers, 'm_yng'—younger male speakers. 1, 2, 3, 4, 5—individual speaker IDs</p> <p> <emph>f</emph> <subs>0</subs> mean, <emph>f</emph><subs>0</subs> range and <emph>f</emph><subs>0</subs> coefficient of variation were modelled in three separate mixed-effects models (Baayen et al., [<reflink idref="bib3" id="ref86">3</reflink>]) using <emph>lmerTest</emph> package in R (Kuznetsova et al., [<reflink idref="bib45" id="ref87">45</reflink>]). All models had identical structure: fixed effects for speaker sex (categorical with two levels: male and female) and age (categorical with two levels: OA and YA), as well as by-speaker and by-sentence random slopes for age and sex. This model had the maximal random effect structure justified by our experimental design, which should optimize generalization of the findings (Barr et al., [<reflink idref="bib6" id="ref88">6</reflink>]). Significance was assessed using <emph>p</emph>-values from the Satterthwaite approximation for degrees of freedom in the <emph>lmerTest</emph> package. Two-way interaction between main effects of speaker age and sex on <emph>f</emph><subs>0</subs> mean, <emph>f</emph><subs>0</subs> range and <emph>f</emph><subs>0</subs> coefficient of variation was not significant in neither of the three models (all <emph>p</emph> > 0.05). We therefore fitted all three models with fixed effects of speaker sex and age without interaction on <emph>f</emph><subs>0</subs> mean, <emph>f</emph><subs>0</subs> range and <emph>f</emph><subs>0</subs> coefficient of variation, respectively.</p> <p>For the <emph>f</emph><subs>0</subs> mean model, only the effect of speaker sex was significant, whereby male speakers had significantly lower mean <emph>f</emph><subs>0</subs> compared to female speakers. (<emph>β</emph> = − 103.3, <emph>SE</emph> = 11.4, <emph>z</emph> = − 9.06, <emph>p</emph> < 0.0001), whereas speaker age effect was not significant (<emph>β</emph> = 3.1, <emph>SE</emph> = 5.2, <emph>z</emph> = 0.6, <emph>p</emph> > 0.05) (Fig. 1A). For <emph>f</emph><subs>0</subs> range model, both effects of speaker sex (<emph>β</emph> = 72.8, <emph>SE</emph> = 7.7, <emph>z</emph> = 9<emph>.</emph>4<emph>, p</emph> < 0.0001) and age (<emph>β</emph> = 24.6, <emph>SE</emph> = 5.3, <emph>z</emph> = 4<emph>.</emph>7<emph>, p</emph> = 0.004) were significant, whereby female speakers and older speakers had significantly larger <emph>f</emph><subs>0</subs> range compared to male speakers and younger speakers, respectively (Fig. 1A). Similarly, for <emph>f</emph><subs>0</subs> coefficient of variation model, both effects of speaker sex and age were significant, whereby female speakers (<emph>β</emph> = 2.8, <emph>SE</emph> = 0.7, <emph>z</emph> = 3<emph>.</emph>6<emph>, p</emph> = 0.009) and older speakers (<emph>β</emph> = 3.6, <emph>SE</emph> = 0.7, <emph>z</emph> = 5<emph>.</emph>2<emph>, p</emph> = 0.003) had significantly larger <emph>f</emph><subs>0</subs> range compared to male speakers and younger speakers, respectively (Fig. 1A).</p> <p>As additional analysis, we converted <emph>f</emph><subs>0</subs> mean, range and coefficient of variation values to logHz and rerun all regression analyses with logHz values as dependent variables. The conclusion from regression analyses with logHz were in the same direction as when using values on a linear scale: (<reflink idref="bib1" id="ref89">1</reflink>) no significant interactions between speaker age and sex variables in neither of the models; (<reflink idref="bib2" id="ref90">2</reflink>) for the <emph>f</emph><subs>0</subs> mean model, only the effect of speaker sex was significant; and (<reflink idref="bib3" id="ref91">3</reflink>) for the <emph>f</emph><subs>0</subs> range and <emph>f</emph><subs>0</subs> coefficient of variation models, both the effects of speaker age and sex were significant.</p> <hd id="AN0185782071-11">Results</hd> <p>All statistical analyses were conducted in R version 4.0.3 (R Core Team, [<reflink idref="bib77" id="ref92">77</reflink>]). To assess the effects of listener age, listener sex, speaker age and speaker sex on <emph>d</emph>′ and <emph>c</emph>, we used two four-way mixed ANOVAs (one for <emph>d</emph>′ and one for <emph>c</emph>, respectively) which test for all main effects and interactions between our four factors of interest: listener age (YA and OA listeners), listener sex (male and female), speaker age (YA and OA speakers) and speaker sex (male and female) on <emph>d</emph>′ and <emph>c</emph>.[<reflink idref="bib2" id="ref93">2</reflink>] Four-way interactions were not significant for neither <emph>d</emph>′ nor <emph>c</emph>. Also, no interactions which included listener sex factor were significant, and listener sex was also not significant as a main effect neither for <emph>d</emph>′ nor for <emph>c</emph> (Figs. 2B and 3B, respectively). Therefore, we collapse further results across male and female listeners and examine the interactions between the remaining three factors (i.e., listener age, speaker age and speaker sex) and their main effects on <emph>d</emph>′ and <emph>c</emph>. We used the three-way robust mixed ANOVAs which tests for all main effects and interactions using trimmed means (Sect. 7.1 in (Wilcox, [<reflink idref="bib105" id="ref94">105</reflink>])). Robust approaches offer higher statistical power and robustness to deviations from the assumed optimal distribution parameters, as suggested by previous studies with the experimental design similar to ours (Ramon, [<reflink idref="bib78" id="ref95">78</reflink>]; Ramon et al., [<reflink idref="bib79" id="ref96">79</reflink>]). Trimmed means is a robust approach typically used to minimize the standard error of the data containing outliers and small deviations from normality (Wilcox & Keselman, [<reflink idref="bib106" id="ref97">106</reflink>]). It is especially suitable for designs with unequal sample sizes, since one of its advantages is to deal with the unequal variances of the involved samples (Mair & Wilcox, [<reflink idref="bib57" id="ref98">57</reflink>]). The three-way robust ANOVAs for <emph>d</emph>′ and <emph>c</emph> were fitted using the function <emph>t</emph>3<emph>way</emph> from the <emph>WRS</emph>2 package (Mair & Wilcox, [<reflink idref="bib57" id="ref99">57</reflink>]).</p> <p>Graph: Fig. 2 Raincloud plots showing distributions and boxplots with median, range and interquartile range of sensitivity (d′) values for: A listener age (younger and older listeners), B listener sex (male and female listeners), C speaker age (younger and older speakers) and D speaker sex (male and female speakers). The shaded region in each plot shows the distribution of the data. Data from individual participants are indicated by transparent dots. YA = younger adults; OA = older adults. Significance codes: 0 '***' 0.001 '**' 0.01 '*' 0.05 '.' 0.1 '' 1, 'n.s.' not significant</p> <p>Graph: Fig. 3 Raincloud plots showing distributions and boxplots with median, range and interquartile range of response bias (c) values for: A listener age (younger and older listeners), B listener sex (male and female listeners), C speaker age (younger and older speakers) and D speaker sex (male and female speakers). The shaded region in each plot shows the distribution of the data. Data from individual participants are indicated by transparent dots. YA = younger adults; OA = older adults. Significance codes: 0 '***' 0.001 '**' 0.01 '*' 0.05 '.' 0.1 '' 1, 'n.s.' not significant</p> <p>Following Mair and Wilcox ([<reflink idref="bib57" id="ref100">57</reflink>]), effect sizes for the main effects for independent samples (i.e., listener age) were calculated using AKP-type effect sizes (<emph>d</emph><subs><emph>t</emph></subs>), a robust alternative to Cohen's <emph>d</emph> using the function <emph>akp.effect</emph> from the <emph>WRS</emph>2 package (Mair & Wilcox, [<reflink idref="bib57" id="ref101">57</reflink>]). AKP-type effect sizes of 0.2, 0.5 and 0.8 correspond to small, medium and large effect sizes, respectively (Wilcox, [<reflink idref="bib105" id="ref102">105</reflink>]). As suggested by Mair and Wilcox ([<reflink idref="bib57" id="ref103">57</reflink>]), effect sizes for main effects for dependent samples (i.e., speaker age and speaker sex) were produced using Yuen's trimmed mean <emph>t</emph> test for dependent samples calculated with function <emph>yuend</emph> from the <emph>MASS</emph> package (Venables & Ripley, [<reflink idref="bib104" id="ref104">104</reflink>]). It reports the explanatory measure of effect size (<emph>ξ</emph>), which is interpreted as follows: values of 0.15, 0.35 and 0.50 correspond to small, medium and large effect sizes, respectively (Sect. 5.3.4 in Wilcox ([<reflink idref="bib105" id="ref105">105</reflink>])). Below we report the results for <emph>d</emph>′ and<emph> c</emph> in detail.</p> <hd id="AN0185782071-12">Sensitivity (d′)</hd> <p>The results of the robust three-way ANOVA on <emph>d</emph>′ showed no significant three-way or two-way interactions between listener age, speaker age and speaker sex (all <emph>p</emph> > 0.05). The main effect of listener age was significant (<emph>F</emph>(<reflink idref="bib1" id="ref106">1</reflink>, 72) = 11.4, <emph>p</emph> < 0.001, <emph>d</emph><subs><emph>t</emph></subs> = 0.57), whereby YA listeners performed significantly better compared to OA listeners (Fig. 2A). However, the main effect of speaker age was not significant (<emph>F</emph>(<reflink idref="bib1" id="ref107">1</reflink>, 72) = 0.2, <emph>p</emph> = 0.7, <emph>ξ</emph> = 0.01) indicating that there was no significant difference between discrimination scores for YA and OA speakers (Fig. 2C). Lastly, the main effect of speaker sex was also significant (<emph>F</emph>(<reflink idref="bib1" id="ref108">1</reflink>, 72) = 76.6, <emph>p</emph> < 0.001, <emph>ξ</emph> = 0.79), whereby male speakers were discriminated significantly better than female speakers (Fig. 2D).</p> <hd id="AN0185782071-13">Response bias (c)</hd> <p>Similar to <emph>d</emph>′, a three-way robust ANOVA on <emph>c</emph> showed no statistically significant three-way or two-way interactions between listener age, speaker age and speaker sex on <emph>c</emph> (all <emph>p</emph> < 0.05). Likewise, the main effect of listener age was not significant (<emph>F</emph>(<reflink idref="bib1" id="ref109">1</reflink>, 72) = 1.2, <emph>p</emph> = 0.3, <emph>d</emph><subs><emph>t</emph></subs> = 0.22) suggesting the two listener groups did not differ significantly in terms of response bias (Fig. 3A). However, the main effect of speaker age was significant (<emph>F</emph>(<reflink idref="bib1" id="ref110">1</reflink>, 72) = 18.4, <emph>p</emph> < 0.0001, <emph>ξ</emph> = 0.47), whereby listeners were significantly more biased towards responding 'same' when hearing YA speaker pairs compared to OA speaker pairs (Fig. 3C). Lastly, the main effect of speaker sex was also significant (<emph>F</emph>(<reflink idref="bib1" id="ref111">1</reflink>, 72) = 24.3, <emph>p</emph> < 0.001, <emph>ξ</emph> = 0.53) suggesting that listeners were significantly more biased towards responding 'same' when hearing female speaker pairs compared to male speaker pairs (Fig. 3D). It should be noted that response bias values in all experimental conditions were significantly shifted above zero (assessed with one-sample t-tests, all <emph>p</emph> < 0.05), which could be a result of the experimental design, namely, short stimulus duration and the overall nature of a discrimination task (see Discussion). Crucially, however, the differences between YA and OA speakers, as well as between male and female speakers were significant.</p> <hd id="AN0185782071-14">Discussion</hd> <p>This study investigated own-age effects in voice discrimination. While the study did not find a performance difference in terms of a higher discrimination accuracy for the own-age group stimuli, we crucially discovered a listener bias revealing a preference for 'same' response in case of younger speakers and female speakers. It is possible that such bias can be accounted for by acoustic differences between male and female speakers. Males typically exhibit comparatively lower fundamental frequencies than females—a reflection of both physiological and cultural factors (Munson & Babel, [<reflink idref="bib65" id="ref112">65</reflink>]). Since the harmonic components of the glottis signal are simple multiples of <emph>f</emph><subs>0</subs>, the harmonic signal in male speakers is much tighter compared to female speakers (Simpson, [<reflink idref="bib92" id="ref113">92</reflink>]). It is unclear, however, whether this sparser sampling of harmonics in female voices (Munson & Babel, [<reflink idref="bib65" id="ref114">65</reflink>]) leads to less vocal tract individualities being revealed in spectral envelopes of female voices. It furthermore remains to be investigated if narrower spacing between harmonics in males is advantageous for voice identity recognition. Superior recognition performance for male compared to female voices has been previously reported for x-vector and i-vector based automatic speaker recognition systems in balanced male and female datasets (Kathiresan, [<reflink idref="bib39" id="ref115">39</reflink>]). However, human listeners and machines use different features to process voices (Park et al., [<reflink idref="bib72" id="ref116">72</reflink>]), and their performance is tested with different tasks. Future studies could systematically test the relationship between vocal tract sampling and recognition performance. One possibility would be to extract speakers' spectral envelopes and fill them with harmonic signal of different densities while leaving the spectral envelope unchanged. Different listener groups can then be trained to remember voice identities either by listening to spectrally dense stimuli or to spectrally undersampled stimuli. Afterwards, listeners will perform a voice recognition test, which can elucidate whether a spectrally dense signal is beneficial for learning and recognizing voice identities. In addition, future studies could address whether female voices are perceived as more similar by the listeners compared to male voices. If confirmed, such effect could imply that listeners develop a bias by experience in the sense that whenever the stimulus is ambiguous (or when the stimulus quality is poor), the choice is on 'same' rather than on 'different' response.</p> <p>In addition, studies show that female speakers have wider <emph>f</emph><subs>0</subs> range and variability than male speakers (Haan & van Heuven, [<reflink idref="bib32" id="ref117">32</reflink>]; Traunmüller & Eriksson, [<reflink idref="bib101" id="ref118">101</reflink>]), and acoustic analyses of our stimuli confirm that. It is possible that increased <emph>f</emph><subs>0</subs> range and variability make it more challenging for the listeners to 'tell speakers together' (Lavan et al., [<reflink idref="bib48" id="ref119">48</reflink>], [<reflink idref="bib49" id="ref120">49</reflink>]; Lavan et al., [<reflink idref="bib48" id="ref121">48</reflink>], [<reflink idref="bib49" id="ref122">49</reflink>]), i.e., generalize over highly variable samples and correctly attribute them to the same-speaker identity. This, in turn, might lead to a decreased discrimination accuracy of female compared to male voices. Another source for higher similarities in female voices might be the fact that they show a higher phonetic convergence compared to males, i.e., they change their vocal parameters to sound more similar to their interlocutors (Namy et al., [<reflink idref="bib66" id="ref123">66</reflink>]). This may be an additional factor that contributes to female voices being on the whole more similar than male voices and thus fostering a bias by experience. In other words, because listeners are already unconsciously aware of the tendency of female speakers to sound more similar to the interlocutors compared to male speakers, this may contribute to forming a perceptual bias that female speakers are per se more similar compared to male speakers. For a comprehensive review of various acoustic, linguistic and social factors contributing to variation in male and female speech, see (Babel & Munson, [<reflink idref="bib4" id="ref124">4</reflink>]; Munson & Babel, [<reflink idref="bib65" id="ref125">65</reflink>]).</p> <p>A factor that may contribute to the same-speaker bias in younger voices may possibly lie in older speakers being experienced as more different in everyday situations due to their development of a variety of source signal individualities in terms of laryngealizations over the years that are not yet present in younger speakers. Age-related non-pathological changes in voice include changes to fundamental frequency (<emph>f</emph><subs>0</subs>), increased <emph>f</emph><subs>0</subs> variability patterns within speakers (as confirmed by acoustic analyses of our stimuli), decreased harmonic to noise ratio, increased jitter and shimmer, as well as decreased acoustic intensity (for a review, see Goy et al. ([<reflink idref="bib31" id="ref126">31</reflink>]) and Schultz et al. ([<reflink idref="bib89" id="ref127">89</reflink>])). Such age-related voice changes may contribute to a bias that younger voices — not showing these distinctions — are per se more similar. In other words, younger voices might be perceived as being more similar in the presence of more variable voices of the older speakers.</p> <p>Voice perception relies on many different features beyond <emph>f</emph><subs>0</subs>. While detailed acoustic analysis of the stimuli is beyond the scope of this paper, future studies could investigate in detail the relationship between, for example, differences in mean <emph>f</emph><subs>0</subs> or <emph>f</emph><subs>0</subs> variation between speakers in a pair and listeners' responses. Such item-wise analysis could clarify whether pairs of speakers with large differences in <emph>f</emph><subs>0</subs> or <emph>f</emph><subs>0</subs> variation would be easier to discriminate than pairs in which speakers' <emph>f</emph><subs>0</subs> is more similar. Such findings might suggest that it is not speaker sex per se that is driving discrimination differences in sensitivity and bias, but rather speakers' <emph>f</emph><subs>0</subs> properties. If speakers have distinct <emph>f</emph><subs>0</subs> mean and smaller <emph>f</emph><subs>0</subs> variance, as is the case for male speakers, it is possible that speaker pairs will be well discriminated by pitch alone. On the other hand, if female speakers in the sample have higher <emph>f</emph><subs>0</subs> variance, then <emph>f</emph><subs>0</subs> mean would be a less reliable cue for voice discrimination: the same female speaker might have very different pitch values across different speech samples, while two different female speakers might have very similar mean <emph>f</emph><subs>0</subs> values (i.e., if there is high overlap in their <emph>f</emph><subs>0</subs> ranges). These hypotheses remain to be addressed by future research.</p> <p>Note that our participants generally adopted a loose response criterion, as evidenced in the more frequent 'same-speaker' response bias across all conditions (Fig. 3). Crucially, however, the differences between experimental condition in response bias were significant. Shorter stimuli offer the listener less detail to be compared and thus to arrive at a 'different-speaker' response, thus, it is plausible that the number of false 'same-speaker' responses should increase. Previous research shows that voice discrimination as compared to voice recognition tasks are associated with higher 'same' response biases. Kreiman and Papcun ([<reflink idref="bib43" id="ref128">43</reflink>]) directly compared voice recognition and discrimination performance and found that listeners were overall more biased towards 'same' response in discrimination task, but not in recognition task. The authors hypothesized that both stimulus duration (they used shorter stimuli in discrimination compared to recognition task) and task demands causing a shift in response criteria could account for these findings. Both across Kreiman and Papcun ([<reflink idref="bib43" id="ref129">43</reflink>]) and our study, listeners never compared two identical stimuli is same-speaker pairs or two sentences with the same content in different-speaker pairs. Thus, linguistic content between stimulus 1 and stimulus 2 in each speaker pair was different. This means listeners had to accept a certain amount of difference between stimuli as possibly belonging to a same-speaker identity, because speakers sound different when producing different linguistic structures and even when producing the same linguistic content multiple times (Kreiman & Papcun, [<reflink idref="bib43" id="ref130">43</reflink>]).</p> <p>Other findings of this study also brought important phenomena to light. In terms of sensitivity (<emph>d</emph>′), we found no interactions between listener and speaker age, which would have suggested the presence of an own-age discrimination advantage. This contrasts with the often reported own-age advantage for face identity processing (Denkinger & Kinn, [<reflink idref="bib22" id="ref131">22</reflink>]; Rhodes & Anastasi, [<reflink idref="bib81" id="ref132">81</reflink>]; Wright & Stroud, [<reflink idref="bib110" id="ref133">110</reflink>]), which might point towards differential processing of faces and voices, as suggested by previous research: for example, faces may provide more reliable identity information than voices (Brédart et al., [<reflink idref="bib13" id="ref134">13</reflink>]). Also, identity and sex information might be processed separately for faces, but not for voices (Burton & Bonner, [<reflink idref="bib16" id="ref135">16</reflink>]). Therefore, it is possible that an own-age recognition advantage for facial identity could follow different principles than those involved in voice processing. However, future studies should disambiguate an own-age processing advantage with different listener populations (e.g. children and middle-aged adults), languages and listening conditions (e.g. speech in noise).</p> <p>Our results also showed that YA listeners outperformed OA listeners in terms of sensitivity (<emph>d</emph>′), which is in line with previous literature about the listener age effect on various voice perception tasks (Best et al., [<reflink idref="bib9" id="ref136">9</reflink>]; Clifford, [<reflink idref="bib18" id="ref137">18</reflink>]; Kausler & Puckett, [<reflink idref="bib40" id="ref138">40</reflink>]; Yonan & Sommers, [<reflink idref="bib113" id="ref139">113</reflink>]; Zaltz & Kishon-Rabin, [<reflink idref="bib114" id="ref140">114</reflink>]). This might be expected given a general age-related decline in hearing and cognitive functions in older adults (Deary et al., [<reflink idref="bib19" id="ref141">19</reflink>]). A novel finding was observed regarding a speaker age effect, whereby listeners could discriminate YA and OA speakers equally well. Previous research shows that accurate speaker discrimination and recognition helps listeners to structure and process linguistic content of speech (Kreiman & Sidtis, [<reflink idref="bib44" id="ref142">44</reflink>]; Nygaard & Pisoni, [<reflink idref="bib69" id="ref143">69</reflink>]), therefore, it is equally important for listeners to accurately discriminate speakers of different ages to accurately process and understand speech.</p> <p>As for the effects of listener and speaker sex on sensitivity, we found no significant interaction between these factors on <emph>d</emph>′, which contrast with studies reporting an own-sex voice recognition advantage (Roebuck & Wilding, [<reflink idref="bib84" id="ref144">84</reflink>]; Skuk & Schweinberger, [<reflink idref="bib93" id="ref145">93</reflink>]; Wilding & Cook, [<reflink idref="bib107" id="ref146">107</reflink>]). However, previous findings of these interactions appear inconsistent, with the own-sex advantage either present in both male and female listeners (Roebuck & Wilding, [<reflink idref="bib84" id="ref147">84</reflink>]), or confined to only female (Wilding & Cook, [<reflink idref="bib107" id="ref148">107</reflink>]) or male (Skuk & Schweinberger, [<reflink idref="bib93" id="ref149">93</reflink>]) listeners. Instead, our results showed that male speakers were discriminated better than female speakers, which aligns with the results reported by Best et al. ([<reflink idref="bib9" id="ref150">9</reflink>]) and Thompson ([<reflink idref="bib99" id="ref151">99</reflink>]), as well as corroborates results from the automatic speaker recognition domain showing a consistent performance advantage for male compared to female speakers (Kathiresan, [<reflink idref="bib39" id="ref152">39</reflink>]). Male voices tend to have lower <emph>f</emph><subs>0</subs> compared to female voices as a result of longer vocal folds in men compared to women (Puts et al., [<reflink idref="bib76" id="ref153">76</reflink>]). Therefore, it is possible that male voices are easier to discriminate due to a denser spectrum of harmonics that better samples the individual vocal tract characteristics (Dellwo et al., [<reflink idref="bib21" id="ref154">21</reflink>]). On the other hand, male and female <emph>listeners</emph> did not differ in terms of sensitivity (<emph>d</emph>′), which is in line with previous literature (Clifford, [<reflink idref="bib18" id="ref155">18</reflink>]; Thompson, [<reflink idref="bib99" id="ref156">99</reflink>]; Yarmey & Matthys, [<reflink idref="bib111" id="ref157">111</reflink>]; Yarmey et al., [<reflink idref="bib112" id="ref158">112</reflink>]). Only the early work by McGehee (McGehee, [<reflink idref="bib61" id="ref159">61</reflink>]) showed that male listeners outperformed female listeners using a large sample of graduate students. However, our study included both younger and older adult listeners, which could explain differences in our results compared to those of McGehee ([<reflink idref="bib61" id="ref160">61</reflink>]).</p> <p>Evidence from the field of face identity processing suggests that own-age advantage can be explained by the more extensive exposure to the faces of observers' own-age group, which aids successful recognition (Rhodes & Anastasi, [<reflink idref="bib81" id="ref161">81</reflink>]). On the other hand, social–cognitive theories suggest that superior recognition for in-group faces is driven by an initial categorization of a face as belonging to an in-group (Rhodes & Anastasi, [<reflink idref="bib81" id="ref162">81</reflink>]). Categorizing a face as an in-group one aids an observer in successful encoding of face individual properties, which then facilitate recognition (Hugenberg et al., [<reflink idref="bib37" id="ref163">37</reflink>]). By contrast, if a face is categorized as an out-group one, the subsequent encoding focuses on the category-level, rather than individual-level features (Levin, [<reflink idref="bib52" id="ref164">52</reflink>]).</p> <p>In voice identity processing, it could be that own-age effects will be amplified by the <emph>explicit</emph> training and experience with the voices of the own-age group. In everyday situations, we typically acquire voice information implicitly, without consciously attending to it, therefore, some indexical cues might remain unattended to by the listeners. On the other hand, if listeners are instructed to actively memorize individual properties of voices of their own-age group, the own-age effects might be detectable. Our procedure did not involve any training or prior exposure to voices, but this can be addressed by future studies using voice recognition/identification tasks. Voice discrimination and recognition are distinct but related abilities supported by partially dissociated cognitive processes and response strategies: while identifying familiar voices involves holistic pattern recognition and matching it to a name or a person, discriminating unfamiliar voices relies on feature analysis and comparison of basic acoustic parameters between the compared voices (Maguinness et al., [<reflink idref="bib56" id="ref165">56</reflink>]; Van Lancker & Kreiman, [<reflink idref="bib102" id="ref166">102</reflink>]). Such a partial dissociation between voice discrimination and recognition is further supported by evidence from brain-lesioned patients with impaired ability to discriminate voices but intact ability to recognize familiar voices and vice versa (Maguinness et al., [<reflink idref="bib56" id="ref167">56</reflink>]; Van Lancker et al., [<reflink idref="bib103" id="ref168">103</reflink>]). Therefore, it is possible that a different pattern of results would emerge in terms of own-age advantages when using a recognition instead of discrimination task. Also, as discussed above, differences can be expected in response bias, especially if voice recognition task involves open set design (i.e., listeners are informed that a target speaker may or may not be in the recognition set), in which case listeners may adopt an overall stricter response criterion (see discussion in Kreiman and Papcun ([<reflink idref="bib43" id="ref169">43</reflink>])).</p> <p>To summarize, this study examined own-age effects in voice discrimination using measures of sensitivity (<emph>d</emph>′) and response bias (<emph>c</emph>). Our results indicated that YA listeners discriminate speakers better than OA listeners and that male speakers were discriminated better than female speakers. We also showed that listeners are significantly biased towards responding 'same' when hearing young speaker pairs and female speaker pairs. These results might also be relevant for a variety of forensic procedures involving voice evidence. First, our stimuli had limited bandwidth (80–5000 Hz) and thus resembled realistic casework audio material better than studio-quality stimuli often used in speaker recognition tasks. Furthermore, it is possible that in voice crime cases involving young speakers and female speakers it might be more challenging for the earwitnesses to tell apart younger and female voices, especially when the stimulus quality is poor and exposure to the voice is brief, which is often the case for forensic voice evidence. Previous work shows that the risk of bias increases with decreasing stimulus quality (Forensic Science Regulator, [<reflink idref="bib25" id="ref170">25</reflink>]), and it is essential that forensic voice experts take into account all possible sources of bias when assessing an earwitness' testimony in court. This is especially relevant given that most of voice crime cases involve young speakers and that the number of voice crime cases with female speakers has increased in recent years, especially in the financial crime domain and in 'fake police officer' crimes discussed in Introduction.</p> <p>Even though our experiment design differs from the typical voice parade procedure, the observed findings are of interest for forensic experts when constructing voice parades for earwitness identification testimonies in court. In a voice parade, an earwitness is asked to identify the voice of the speaker they heard at the crime scene from a collection of recorded speech samples by a suspect and a number of foils (McDougall, [<reflink idref="bib59" id="ref171">59</reflink>]). Voice parades are prepared using thorough experimental procedures: typically, a sample of a suspect's speech is compiled from excerpts extracted from a recorded police interview, and speech samples from foil speakers are constructed from recordings of similar quality, speaking style and duration (McDougall, [<reflink idref="bib59" id="ref172">59</reflink>]). The foil speech samples must be screened to ensure that foil speakers do not stand out in terms of accent, pitch and speaking rate compared to the suspect's voice (Home Office, [<reflink idref="bib70" id="ref173">70</reflink>]). In other words, voice parade procedures require a rigorous phonetic screening to select suitable foil speakers. Thus, research evidence obtained under controlled laboratory conditions is important for daily forensic investigations related to formal voice identification procedures.</p> <p>In addition, future research might investigate whether forensic voice experts are also liable to such biases for younger voices and/or female voices, since this would have an impact on forensic phonetic casework on a much broader level. However, (Bartle & Dellwo, [<reflink idref="bib7" id="ref174">7</reflink>]) showed that voice experts tend to have a conservative bias when unsure, meaning that they tended to respond that samples come from different speakers. Thus, expert listeners should be less liable to a positive response bias as compared to lay listeners. Also, expert listeners should also be aware of the range of cognitive biases which may affect recognition performance and use some strategies to mitigate the influence of biases (Gold & French, [<reflink idref="bib30" id="ref175">30</reflink>]; Rhodes, [<reflink idref="bib82" id="ref176">82</reflink>]). Further, the analysis carried out by the expert listeners in forensic speaker comparison is categorically different from ad hoc same/different judgments in our experiment. Forensic voice experts systematically use a wide range of analytic methods and take decisions after applying typically complex acoustic, auditory and automatic procedures.</p> <p>At least two components contribute to the perception of person similarity: the signal itself (i.e., the acoustic distinguishability of voices) and the perceived distance between voices (i.e., perceptual factors impacting whether voices are perceived as more or less different). Recently, researchers became increasingly interested in equating acoustic differences between voices with their perceived differences for human listeners. State-of-the-art deep neural networks, for example, produce vectors from acoustic voice samples that allow maximum classification and recognition accuracy. However, to what degree human perceptual judgements align with distances produced by neural networks is unclear. Despite the increased awareness in acoustic and perceptual factors, they are often viewed as somehow mechanically contributing to voice recognition. Current models of voice perception suggest that voices are represented in terms of their acoustic deviation from a voice prototype conceptualized as average of all voices heard by a listener (Belin et al., [<reflink idref="bib8" id="ref177">8</reflink>]; Latinus et al., [<reflink idref="bib47" id="ref178">47</reflink>]; Lavner et al., [<reflink idref="bib50" id="ref179">50</reflink>]; Maguinness et al., [<reflink idref="bib56" id="ref180">56</reflink>]). Our results highlight that voices may be judged as more or less similar based on characteristics other than stimulus acoustics and demonstrate that cognitive biases can be an important component of voice perception. However, cognitive biases have not been paid much attention to in the past and are not accounted for by current voice perception models. It would thus be interesting to explore further their origin and the impact they have on recognition accuracy and perception of other person characteristics.</p> <p>It also seems plausible that there are evolutionary mechanisms employed that lead to listeners being biased towards perceiving some groups of listeners with higher similarity than others. In the case of younger speakers this might be rooted in mechanisms by which younger individuals are possibly more viewed as part of a group and not as individuals to the same degree as older adults. This can be addressed by future studies.</p> <hd id="AN0185782071-15">Acknowledgements</hd> <p>The data were collected at the Phonetics Laboratory and Linguistics Research Infrastructure (LiRI) laboratory at the University of Zurich. We thank Sandra Schwab for the assistance with the statistical analyses for this study.</p> <hd id="AN0185782071-16">Author contributions</hd> <p>Valeriia Vyshnevetska was involved in conceptualization, data curation, formal analysis, investigation, methodology, visualization, writing—original draft and writing—reviewing and editing. Nathalie Giroud was responsible for methodology and writing—reviewing and editing. Meike Ramon took part in formal analysis, resources, validation and writing—reviewing and editing. Volker Dellwo participated in conceptualization, funding acquisition, methodology, resources, supervision, validation, writing—original draft and writing—reviewing and editing.</p> <hd id="AN0185782071-17">Funding</hd> <p>VV and VD were supported by grants # 185399 (Indexical Dynamics) and PCEFP1_186841 (EVOPHON) from Swiss National Science Foundation (SNSF). NG was supported by a Promoting Women in Academia (PRIMA) grant from SNSF # 185715. MR was supported by a PRIMA grant from SNSF PR00P1 179872. The funders had no role in study design, data collection, analysis and interpretation, or preparation of the manuscript.</p> <hd id="AN0185782071-18">Data availability</hd> <p>Due to the data protection guidelines of the University of Zurich, the raw audio recordings used during this study are not publicly available. However, mel-frequency cepstral coefficients (MFCCs) of the audio recordings are available for scientific purposes from the first author upon request. Behavioural data and analyses code are publicly available in the study's Open Science Framework repository: https://osf.io/sg96w/</p> <hd id="AN0185782071-19">Declarations</hd> <p></p> <hd id="AN0185782071-20">Ethics statement</hd> <p>All listeners gave an informed consent before participating and received monetary compensation for their participation. The study was approved by Ethics Committee of the Faculty of Arts and Social Sciences at the University of Zurich. The research was performed in accordance with the Declaration of Helsinki.</p> <hd id="AN0185782071-21">Consent for publication</hd> <p>Consent for publishing the behavioural data was obtained from all participants.</p> <hd id="AN0185782071-22">Competing interests</hd> <p>None.</p> <hd id="AN0185782071-23">Publisher's Note</hd> <p>Springer Nature remains neutral with regard to jurisdictional claims in published maps and institutional affiliations.</p> <ref id="AN0185782071-24"> <title> References </title> <blist> <bibl id="bib1" idref="ref31" type="bt">1</bibl> <bibtext> Anastasi JS, Rhodes MG. An own-age bias in face recognition for children and older adults. Psychonomic Bulletin & Review. 2005; 12; 6: 1043-1047. 10.3758/BF03206441</bibtext> </blist> <blist> <bibl id="bib2" idref="ref80" type="bt">2</bibl> <bibtext> Anwyl-Irvine AL, Massonnié J, Flitton A, Kirkham N, Evershed JK. Gorilla in our midst: An online behavioral experiment builder. Behavior Research Methods. 2020; 52; 1: 388-407. 10.3758/s13428-019-01237-x. 31016684</bibtext> </blist> <blist> <bibl id="bib3" idref="ref86" type="bt">3</bibl> <bibtext> Baayen RH, Davidson DJ, Bates DM. Mixed-effects modeling with crossed random effects for subjects and items. Journal of Memory and Language. 2008; 59; 4: 390-412. 10.1016/j.jml.2007.12.005</bibtext> </blist> <blist> <bibl id="bib4" idref="ref124" type="bt">4</bibl> <bibtext> Babel M, Munson B. Producing socially meaningful linguistic variation. Oxford University Press. 2014. 10.1093/oxfordhb/9780199735471.013.022</bibtext> </blist> <blist> <bibl id="bib5" idref="ref21" type="bt">5</bibl> <bibtext> Ball, E. (2023). Fake police officer scammers swipe hundreds from Tewkesbury pensioners. https://<ulink href="http://www.gloucestershirelive.co.uk/news/gloucester-news/fake-police-officer-scammers-swipe-8637967">www.gloucestershirelive.co.uk/news/gloucester-news/fake-police-officer-scammers-swipe-8637967</ulink></bibtext> </blist> <blist> <bibl id="bib6" idref="ref88" type="bt">6</bibl> <bibtext> Barr DJ, Levy R, Scheepers C, Tily HJ. Random effects structure for confirmatory hypothesis testing: Keep it maximal. Journal of Memory and Language. 2013; 68; 3: 255-278. 10.1016/j.jml.2012.11.001</bibtext> </blist> <blist> <bibl id="bib7" idref="ref174" type="bt">7</bibl> <bibtext> Bartle A, Dellwo V. Auditory speaker discrimination by forensic phoneticians and naive listeners in voiced and whispered speech. International Journal of Speech Language and the Law. 2015; 22; 2: 229-248. 10.1558/ijsll.v22i2.23101</bibtext> </blist> <blist> <bibl id="bib8" idref="ref177" type="bt">8</bibl> <bibtext> Belin P, Bestelmeyer PEG, Latinus M, Watson R. Understanding voice perception. British Journal of Psychology. 2011; 102; 4: 711-725. 10.1111/j.2044-8295.2011.02041.x. 21988380</bibtext> </blist> <blist> <bibl id="bib9" idref="ref13" type="bt">9</bibl> <bibtext> Best V, Ahlstrom JB, Mason CR, Roverud E, Perrachione TK, Kidd G, Dubno JR. Talker identification: Effects of masking, hearing loss, and age. The Journal of the Acoustical Society of America. 2018; 143; 2: 1085-1092. 10.1121/1.5024333. 29495693. 5820061</bibtext> </blist> <blist> <bibtext> Bobak AK, Jones AL, Hilker Z, Mestry N, Bate S, Hancock PJB. Data-driven studies in face identity processing rely on the quality of the tests and data sets. Cortex. 2023; 166: 348-364. 10.1016/j.cortex.2023.05.018. 37481857</bibtext> </blist> <blist> <bibtext> Boersma, P, & Weenink, D. (2025). Praat: Doing phonetics by computer [Computer software]. <ulink href="http://www.praat.org/">http://www.praat.org/</ulink></bibtext> </blist> <blist> <bibtext> Braber N, Smith H, Wright D, Hardy A, Robson J. Assessing the specificity and accuracy of accent judgments by lay listeners. Language and Speech. 2023; 66; 2: 267-290. 10.1177/00238309221101560. 35723130</bibtext> </blist> <blist> <bibtext> Brédart S, Barsics C, Hanley R. Recalling semantic information about personally known faces and voices. European Journal of Cognitive Psychology. 2009; 21; 7: 1013-1021. 10.1080/09541440802591821</bibtext> </blist> <blist> <bibtext> Brewster, T. (2021). Fraudsters cloned company director's voice in $35 Million heist, police find. Forbes. https://<ulink href="http://www.forbes.com/sites/thomasbrewster/2021/10/14/huge-bank-fraud-uses-deep-fake-voice-tech-to-steal-millions/?sh=352ab2e07559">www.forbes.com/sites/thomasbrewster/2021/10/14/huge-bank-fraud-uses-deep-fake-voice-tech-to-steal-millions/?sh=352ab2e07559</ulink></bibtext> </blist> <blist> <bibtext> Bricker PD, Pruzansky S. Effects of stimulus content and duration on talker identification. The Journal of the Acoustical Society of America. 1966; 40; 6: 1441-1449. 10.1121/1.1910246. 5975580</bibtext> </blist> <blist> <bibtext> Burton AM, Bonner L. Familiarity influences judgments of sex: The case of voice recognition. Perception. 2004; 33; 6: 747-752. 10.1068/p3458. 15330367</bibtext> </blist> <blist> <bibtext> Case J, Seyfarth S, Levi SV. Does implicit voice learning improve spoken language processing? Implications for clinical practice. Journal of Speech, Language, and Hearing Research. 2018; 61; 5: 1251-1260. 10.1044/2018_JSLHR-L-17-0298. 29800358. 6195079</bibtext> </blist> <blist> <bibtext> Clifford BR. Voice identification by human listeners: On earwitness reliability. Law and Human Behavior. 1980; 4; 4: 373-394. 10.1007/BF01040628</bibtext> </blist> <blist> <bibtext> Deary IJ, Corley J, Gow AJ, Harris SE, Houlihan LM, Marioni RE, Penke L, Rafnsson SB, Starr JM. Age-associated cognitive decline. British Medical Bulletin. 2009; 92; 1: 135-152. 10.1093/bmb/ldp033. 19776035</bibtext> </blist> <blist> <bibtext> Dellwo, V, Leemann, A, & Kolly, M.-J. (2012). Speaker idiosyncratic rhythmic features in the speech signal. In Interspeech conference proceedings (pp. 1–4). https://doi.org/10.5167/UZH-68554</bibtext> </blist> <blist> <bibtext> Dellwo V, Kathiresan T, Pellegrino E, He L, Schwab S, Maurer D. Influences of fundamental oscillation on speaker identification in vocalic utterances by humans and computers. Interspeech. 2018; 20: 3795-3799. 10.21437/Interspeech.2018-2331</bibtext> </blist> <blist> <bibtext> Denkinger B, Kinn M. Own-age bias and positivity effects in facial recognition. Experimental Aging Research. 2018; 44; 5: 411-426. 10.1080/0361073X.2018.1521493. 30285572</bibtext> </blist> <blist> <bibtext> Fleming D, Giordano BL, Caldara R, Belin P. A language-familiarity effect for speaker discrimination without comprehension. Proceedings of the National Academy of Sciences. 2014; 111; 38: 13795-13798. 10.1073/pnas.1401383111</bibtext> </blist> <blist> <bibtext> Flitter, E, & Cowley, S. (2023). Voice deepfakes are coming for your bank balance. New York Times. https://<ulink href="http://www.nytimes.com/2023/08/30/business/voice-deepfakes-bank-scams.html">www.nytimes.com/2023/08/30/business/voice-deepfakes-bank-scams.html</ulink></bibtext> </blist> <blist> <bibtext> Forensic Science Regulator. (2020). Forensic science regulator guidance. Cognitive bias effects relevant to forensic science examinations. https://<ulink href="http://www.gov.uk/government/publications/cognitive-bias-effects-relevant-to-forensic-science-examinations">www.gov.uk/government/publications/cognitive-bias-effects-relevant-to-forensic-science-examinations</ulink></bibtext> </blist> <blist> <bibtext> Foulkes P, Barron A. Telephone speaker recognition amongst members of a close social network. International Journal of Speech, Language and the Law. 2000; 7; 2: 2. 10.1558/sll.2000.7.2.180</bibtext> </blist> <blist> <bibtext> Action Fraud. (2015). Fake police officers targeting elderly. https://thecrimepreventionwebsite.com/action-fraud-notified-scams/816/fake-police-officers-targeting-elderly/</bibtext> </blist> <blist> <bibtext> Fysh MC, Stacchi L, Ramon M. Differences between and within individuals, and subprocesses of face cognition: Implications for theory, research and personnel selection. Royal Society Open Science. 2020; 7; 9. 10.1098/rsos.200233. 33047013. 7540753200233</bibtext> </blist> <blist> <bibtext> Garrido L, Eisner F, McGettigan C, Stewart L, Sauter D, Hanley JR, Schweinberger SR, Warren JD, Duchaine B. Developmental phonagnosia: A selective deficit of vocal identity recognition. Neuropsychologia. 2009; 47; 1: 123-131. 10.1016/j.neuropsychologia.2008.08.003. 18765243</bibtext> </blist> <blist> <bibtext> Gold E, French P. International practices in forensic speaker comparisons: Second survey. International Journal of Speech Language and the Law. 2019; 26; 1: 1-20. 10.1558/ijsll.38028</bibtext> </blist> <blist> <bibtext> Goy H, Kathleen Pichora-Fuller M, Van Lieshout P. Effects of age on speech and voice quality ratings. The Journal of the Acoustical Society of America. 2016; 139; 4: 1648-1659. 10.1121/1.4945094. 27106312</bibtext> </blist> <blist> <bibtext> Haan, J, & van Heuven, V. J. (1999). Male versus female pitch range in Dutch questions. ICPhS99 (pp. 1581–1584).</bibtext> </blist> <blist> <bibtext> Hadfield, C. (2024). Eight elderly men and woman told to hand over thousands of pounds to 'police officers'. Liverpool ECHO. https://<ulink href="http://www.liverpoolecho.co.uk/news/liverpool-news/fake-police-officer-told-woman-28440428">www.liverpoolecho.co.uk/news/liverpool-news/fake-police-officer-told-woman-28440428</ulink></bibtext> </blist> <blist> <bibtext> He Y, Ebner NC, Johnson MK. What predicts the own-age bias in face recognition memory?. Social Cognition. 2011; 29; 1: 97-109. 10.1521/soco.2011.29.1.97. 21415928. 3057073</bibtext> </blist> <blist> <bibtext> Herlitz A, Lovén J. Sex differences and the own-gender bias in face recognition: A meta-analytic review. Visual Cognition. 2013; 21; 9–10: 1306-1336. 10.1080/13506285.2013.823140</bibtext> </blist> <blist> <bibtext> Hollien H, Majewski W, Doherty ET. Perceptual identification of voices under normal, stress and disguise speaking conditions. Journal of Phonetics. 1982; 10; 2: 139-148. 10.1016/S0095-4470(19)30953-2</bibtext> </blist> <blist> <bibtext> Hugenberg K, Young SG, Bernstein MJ, Sacco DF. The categorization-individuation model: An integrative account of the other-race recognition deficit. Psychological Review. 2010; 117; 4: 1168-1187. 10.1037/a0020463. 20822290</bibtext> </blist> <blist> <bibtext> Jessen M. Forensic phonetics. Language and Linguistics Compass. 2008; 2; 4: 671-711. 10.1111/j.1749-818X.2008.00066.x</bibtext> </blist> <blist> <bibtext> Kathiresan TBernardasci C, Dipino D, Garassino D, Negrinelli S, Pellegrino E, Schmid S. Gender bias in voice recognition: An i- and x-vector-based gender-specific automatic speaker recognition study. L'individualità del parlante nelle scienze fonetiche: Applicazioni tecnologiche e forensi. 2021; Officinaventuno: 113-122. 10.17469/O2108AISV000006; 8</bibtext> </blist> <blist> <bibtext> Kausler DH, Puckett JM. Adult age differences in memory for sex of voice. Journal of Gerontology. 1981; 36; 1: 44-50. 10.1093/geronj/36.1.44. 7451836</bibtext> </blist> <blist> <bibtext> Khatsenkova, S. (2023). Audio deepfake scams: Criminals are using AI to sound like family and people are falling for it. Euronews. https://<ulink href="http://www.euronews.com/next/2023/03/25/audio-deepfake-scams-criminals-are-using-ai-to-sound-like-family-and-people-are-falling-fo">www.euronews.com/next/2023/03/25/audio-deepfake-scams-criminals-are-using-ai-to-sound-like-family-and-people-are-falling-fo</ulink></bibtext> </blist> <blist> <bibtext> Köster O, Schiller NO. Different influences of the native language of a listener on speaker recognition. International Journal of Speech Language and the Law. 1997; 4; 1: 18-28. 10.1558/ijsll.v4i1.18</bibtext> </blist> <blist> <bibtext> Kreiman J, Papcun G. Comparing discrimination and recognition of unfamiliar voices. Speech Communication. 1991; 10; 3: 265-275. 10.1016/0167-6393(91)90016-M</bibtext> </blist> <blist> <bibtext> Kreiman J, Sidtis D. Foundations of voice studies: An interdisciplinary approach to voice production and perception. 20111; Wiley. 10.1002/9781444395068</bibtext> </blist> <blist> <bibtext> Kuznetsova A, Brockhoff PB, Christensen RHB. lmerTest package: Tests in linear mixed effects models. Journal of Statistical Software. 2017. 10.18637/jss.v082.i13</bibtext> </blist> <blist> <bibtext> Der Landbote. (2023). Telefonbetrügerin auf frischer Tat ertappt. https://<ulink href="http://www.landbote.ch/kriminalitaet-in-winterthur-telefonbetruegerin-auf-frischer-tat-ertappt-253219737455">www.landbote.ch/kriminalitaet-in-winterthur-telefonbetruegerin-auf-frischer-tat-ertappt-253219737455</ulink></bibtext> </blist> <blist> <bibtext> Latinus M, McAleer P, Bestelmeyer PEG, Belin P. Norm-based coding of voice identity in human auditory cortex. Current Biology. 2013; 23; 12: 1075-1080. 10.1016/j.cub.2013.04.055. 23707425. 3690478</bibtext> </blist> <blist> <bibtext> Lavan N, Burston LFK, Garrido L. How many voices did you hear? Natural variability disrupts identity perception from unfamiliar voices. British Journal of Psychology. 2019; 110; 3: 576-593. 10.1111/bjop.12348. 30221374</bibtext> </blist> <blist> <bibtext> Lavan N, Burton AM, Scott SK, McGettigan C. Flexible voices: Identity perception from variable vocal signals. Psychonomic Bulletin & Review. 2019; 26; 1: 90-102. 10.3758/s13423-018-1497-7</bibtext> </blist> <blist> <bibtext> Lavner Y, Rosenhouse J, Gath I. The prototype model in speaker identification by human listeners. International Journal of Speech Technology. 2001; 4; 1: 63-74. 10.1023/A:1009656816383</bibtext> </blist> <blist> <bibtext> Legge GE, Grosmann C, Pieper CM. Learning unfamiliar voices. Journal of Experimental Psychology: Learning, Memory, and Cognition. 1984; 10; 2: 298-303. 10.1037/0278-7393.10.2.298</bibtext> </blist> <blist> <bibtext> Levin DT. Race as a visual feature: Using visual search and perceptual discrimination tasks to understand face categories and the cross-race recognition deficit. Journal of Experimental Psychology: General. 2000; 129; 4: 559-574. 10.1037/0096-3445.129.4.559. 11142869</bibtext> </blist> <blist> <bibtext> Loffreda, D. (2023). The names fake police officer scammers use to con thousands out of elderly. https://<ulink href="http://www.derbytelegraph.co.uk/news/local-news/names-fake-police-officer-scammers-8046574">www.derbytelegraph.co.uk/news/local-news/names-fake-police-officer-scammers-8046574</ulink></bibtext> </blist> <blist> <bibtext> Macmillan NA, Creelman CD. Detection theory: A user's guide. 1991; Cambridge University Press</bibtext> </blist> <blist> <bibtext> Macmillan NA, Creelman CD. Detection theory. 2004; Psychology Press. 10.4324/9781410611147</bibtext> </blist> <blist> <bibtext> Maguinness C, Roswandowitz C, Von Kriegstein K. Understanding the mechanisms of familiar voice-identity recognition in the human brain. Neuropsychologia. 2018; 116: 179-193. 10.1016/j.neuropsychologia.2018.03.039. 29614253</bibtext> </blist> <blist> <bibtext> Mair P, Wilcox R. Robust statistical methods in R using the WRS2 package. Behavior Research Methods. 2020; 52; 2: 464-488. 10.3758/s13428-019-01246-w. 31152384</bibtext> </blist> <blist> <bibtext> Mason SE. Age and gender as factors in facial recognition and identification. Experimental Aging Research. 1986; 12; 3: 151-154. 10.1080/03610738608259453. 3830234</bibtext> </blist> <blist> <bibtext> McDougall KBernardasci C, Dipino D, Garassino D, Negrinelli S, Pellegrino E, Schmid S. Ear-catching versus eye-catching? Some developments and current challenges in earwitness identification evidence. L'individualità del parlante nelle scienze fonetiche: Applicazioni tecnologiche e forensi. 2021; Officinaventuno: 33-56. 10.17469/O2108AISV000002; 8</bibtext> </blist> <blist> <bibtext> McDougall K, Nolan F, Hudson T. Telephone transmission and earwitnesses: Performance on voice parades controlled for voice similarity. Phonetica. 2015; 72; 4: 257-272. 10.1159/000439385. 26633169</bibtext> </blist> <blist> <bibtext> McGehee F. The reliability of the identification of the human voice. The Journal of General Psychology. 1937; 17; 2: 249-271. 10.1080/00221309.1937.9917999</bibtext> </blist> <blist> <bibtext> Meissner CA, Brigham JC. Thirty years of investigating the own-race bias in memory for faces: A meta-analytic review. Psychology, Public Policy, and Law. 2001; 7; 1: 3-35. 10.1037/1076-8971.7.1.3</bibtext> </blist> <blist> <bibtext> Memon A, Bartlett J, Rose R, Gray C. The aging eyewitness: Effects of age on face, delay, and source-memory ability. The Journals of Gerontology Series b: Psychological Sciences and Social Sciences. 2003; 58; 6: P338-P345. 10.1093/geronb/58.6.P338. 14614118</bibtext> </blist> <blist> <bibtext> Moyse E, Beaufort A, Brédart S. Evidence for an own-age bias in age estimation from voices in older persons. European Journal of Ageing. 2014; 11; 3: 241-247. 10.1007/s10433-014-0305-0. 28804330. 5549201</bibtext> </blist> <blist> <bibtext> Munson B, Babel MKatz WF, Assmann PF. The phonetics of sex and gender. The Routledge handbook of phonetics. 20191; Routledge: 499-525. 10.4324/9780429056253-19</bibtext> </blist> <blist> <bibtext> Namy LL, Nygaard LC, Sauerteig D. Gender differences in vocal accommodation: The role of perception. Journal of Language and Social Psychology. 2002; 21; 4: 422-432. 10.1177/026192702237958</bibtext> </blist> <blist> <bibtext> Nasreddine ZS, Phillips NA, Bedirian V, Charbonneau S, Whitehead V, Collin I, Cummings JL, Chertkow H. The Montreal cognitive assessment, MoCA: A brief screening tool for mild cognitive impairment. Journal of the American Geriatrics Society. 2005; 53; 4: 695-699. 10.1111/j.1532-5415.2005.53221.x. 15817019</bibtext> </blist> <blist> <bibtext> Nolan F, McDougall K, Hudson T. Effects of the telephone on perceived voice similarity: Implications for voice line-ups. International Journal of Speech Language and the Law. 2013; 20; 2: 229-246. 10.1558/ijsll.v20i2.229</bibtext> </blist> <blist> <bibtext> Nygaard LC, Pisoni DB. Talker-specific learning in speech perception. Perception & Psychophysics. 1998; 60; 3: 355-376. 10.3758/BF03206860</bibtext> </blist> <blist> <bibtext> Home Office. (2003). Advice on the use of voice identification parades. Home Office. https://webarchive.nationalarchives.gov.uk/ukgwa/20130308000037/<ulink href="http://www.homeoffice.gov.uk/about-us/corporate-publications-strategy/home-office-circulars/circulars-2003/057-2003/">http://www.homeoffice.gov.uk/about-us/corporate-publications-strategy/home-office-circulars/circulars-2003/057-2003/</ulink></bibtext> </blist> <blist> <bibtext> Papcun G, Kreiman J, Davis A. Long-term memory for unfamiliar voices. The Journal of the Acoustical Society of America. 1989; 85; 2: 913-925. 10.1121/1.397564. 2926007</bibtext> </blist> <blist> <bibtext> Park SJ, Yeung G, Vesselinova N, Kreiman J, Keating PA, Alwan A. Towards understanding speaker discrimination abilities in humans and machines for text-independent short utterances of different speech styles. The Journal of the Acoustical Society of America. 2018; 144; 1: 375-386. 10.1121/1.5045323. 30075658. 6062421</bibtext> </blist> <blist> <bibtext> Pellegrino E, He L, Dellwo V. Age-related rhythmic variations: The role of syllable intensity variability. Travaux Neuchâtelois De Linguistique. 2021; 74: 167-185. 10.26034/tranel.2021.2924</bibtext> </blist> <blist> <bibtext> Pollack I, Pickett JM, Sumby WH. On the identification of speakers by voice. The Journal of the Acoustical Society of America. 1954; 26; 3: 403-406. 10.1121/1.1907349</bibtext> </blist> <blist> <bibtext> Proietti V, Laurence S, Matthews CM, Zhou X, Mondloch CJ. Attending to identity cues reduces the own-age but not the own-race recognition advantage. Vision Research. 2019; 157: 184-191. 10.1016/j.visres.2017.11.010. 29454885</bibtext> </blist> <blist> <bibtext> Puts DA, Hill AK, Bailey DH, Walker RS, Rendall D, Wheatley JR, Welling LLM, Dawood K, Cárdenas R, Burriss RP, Jablonski NG, Shriver MD, Weiss D, Lameira AR, Apicella CL, Owren MJ, Barelli C, Glenn ME, Ramos-Fernandez G. Sexual selection on male vocal fundamental frequency in humans and other anthropoids. Proceedings of the Royal Society B: Biological Sciences. 2016; 283; 1829: 20152830. 10.1098/rspb.2015.2830. 4855375</bibtext> </blist> <blist> <bibtext> R Core Team. (2024). R: A language and environment for statistical computing [Computer software]. R Foundation for Statistical Computing. https://<ulink href="http://www.R-project.org/">www.R-project.org/</ulink></bibtext> </blist> <blist> <bibtext> Ramon M. Differential processing of vertical interfeature relations due to real-life experience with personally familiar daces. Perception. 2015; 44; 4: 368-382. 10.1068/p7909. 26492723</bibtext> </blist> <blist> <bibtext> Ramon M, Caharel S, Rossion B. The speed of recognition of personally familiar faces. Perception. 2011; 40; 4: 437-449. 10.1068/p6794. 21805919</bibtext> </blist> <blist> <bibtext> Rathborn HA, Bull RH, Clifford BR. Voice recognition over the telephone. Journal of Police Science and Administration. 1981; 9; 3: 280-284</bibtext> </blist> <blist> <bibtext> Rhodes MG, Anastasi JS. The own-age bias in face recognition: A meta-analytic and theoretical review. Psychological Bulletin. 2012; 138; 1: 146-174. 10.1037/a0025750. 22061689</bibtext> </blist> <blist> <bibtext> Rhodes R. Cognitive bias in forensic speech science. 2014; IAFPA</bibtext> </blist> <blist> <bibtext> Robson R. A fair hearing? The use of voice identification parades in criminal investigations in England and Wales. Criminal Law Review. 2017; 1: 36-50</bibtext> </blist> <blist> <bibtext> Roebuck R, Wilding J. Effects of vowel variety and sample lenght on identification of a speaker in a line-up. Applied Cognitive Psychology. 1993; 7; 6: 475-481. 10.1002/acp.2350070603</bibtext> </blist> <blist> <bibtext> Rose RA, Bull R, Vrij A. Non-biased lineup instructions do matter—A problem for older witnesses. Psychology, Crime & Law. 2005; 11; 2: 147-159. 10.1080/10683160512331316307</bibtext> </blist> <blist> <bibtext> Schiller NO, Koster O. Evaluation of a foreign speaker in forensic phonetics: A report. International Journal of Speech Language and the Law. 1996; 3; 1: 176-185. 10.1558/ijsll.v3i1.176</bibtext> </blist> <blist> <bibtext> Schirmer A, Chiu MH, Lo C, Feng Y-J, Penney TB. Angry, old, male—and trustworthy? How expressive and person voice characteristics shape listener trust. PLoS ONE. 2020; 15; 5. 10.1371/journal.pone.0232431. 32365066. 7197804e0232431</bibtext> </blist> <blist> <bibtext> Schmidt-Nielsen A, Stern KR. Identification of known voices as a function of familiarity and narrow-band coding. The Journal of the Acoustical Society of America. 1985; 77; 2: 658-663. 10.1121/1.391884</bibtext> </blist> <blist> <bibtext> Schultz BG, Rojas S, St John M, Kefalianos E, Vogel AP. A cross-sectional study of perceptual and acoustic voice characteristics in healthy aging. Journal of Voice. 2023; 37; 6: 969.e23-969.e41. 10.1016/j.jvoice.2021.06.007. 34272139</bibtext> </blist> <blist> <bibtext> Schumacher, E. (2019). 'Fake police' stealing from Germany's elderly. https://<ulink href="http://www.dw.com/en/fake-police-steal-hundreds-of-thousands-from-germanys-elderly/a-47523341#:~:text=The%20criminal%20syndicate%2C%20which%20worked,been%20logged%20by%20the%20authorities">www.dw.com/en/fake-police-steal-hundreds-of-thousands-from-germanys-elderly/a-47523341#:~:text=The%20criminal%20syndicate%2C%20which%20worked,been%20logged%20by%20the%20authorities</ulink></bibtext> </blist> <blist> <bibtext> Schvartz KC, Chatterjee M. Gender identification in younger and older adults: Use of spectral and temporal cues in noise-vocoded speech. Ear & Hearing. 2012; 33; 3: 411-420. 10.1097/AUD.0b013e31823d78dc</bibtext> </blist> <blist> <bibtext> Simpson AP. Phonetic differences between male and female speech. Language and Linguistics Compass. 2009; 3; 2: 621-640. 10.1111/j.1749-818X.2009.00125.x</bibtext> </blist> <blist> <bibtext> Skuk VG, Schweinberger SR. Gender differences in familiar voice identification. Hearing Research. 2013; 296: 131-140. 10.1016/j.heares.2012.11.004. 23168357</bibtext> </blist> <blist> <bibtext> Sporer SL. Recognizing faces of other ethnic groups: An integration of theories. Psychology, Public Policy, and Law. 2001; 7; 1: 36-97. 10.1037/1076-8971.7.1.36</bibtext> </blist> <blist> <bibtext> Stacchi L, Huguenin-Elie E, Caldara R, Ramon M. Normative data for two challenging tests of face matching under ecological conditions. Cognitive Research: Principles and Implications. 2020; 5; 1: 8. 10.1186/s41235-019-0205-0. 32076893</bibtext> </blist> <blist> <bibtext> Stanislaw H, Todorov N. Calculation of signal detection theory measures. Behavior Research Methods, Instruments, & Computers. 1999; 31; 1: 137-149. 10.3758/BF03207704</bibtext> </blist> <blist> <bibtext> Stevenage SV, Clarke G, McNeill A. The "other-accent" effect in voice recognition. Journal of Cognitive Psychology. 2012; 24; 6: 647-653. 10.1080/20445911.2012.675321</bibtext> </blist> <blist> <bibtext> Swiss Banking Ombudsman. (2023). Claim for damages after a fraud by false police officers. https://bankingombudsman.ch/en/claim-for-damages-after-a-fraud-by-false-police-officers/</bibtext> </blist> <blist> <bibtext> Thompson C. Voice identification: Speaker identifiability and a correction of the record regarding sex effects. Human Learning: Journal of Practical Research & Applications. 1985; 4; 1: 19-27</bibtext> </blist> <blist> <bibtext> Tompkinson J, Watt D. Assessing the abilities of phonetically untrained listeners to determine pitch and speaker accent in unfamiliar voices. Language and Law/linguagem e Direito. 2018; 5; 1: 19-37</bibtext> </blist> <blist> <bibtext> Traunmüller H, Eriksson A. The frequency range of the voice fundamental in the speech of male and female adults. 1995; Stockholm University</bibtext> </blist> <blist> <bibtext> Van Lancker D, Kreiman J. Voice discrimination and recognition are separate abilities. Neuropsychologia. 1987; 25; 5: 829-834. 10.1016/0028-3932(87)90120-5. 3431677</bibtext> </blist> <blist> <bibtext> Van Lancker DR, Cummings JL, Kreiman J, Dobkin BH. Phonagnosia: A dissociation between familiar and unfamiliar voices. Cortex. 1988; 24; 2: 195-209. 10.1016/S0010-9452(88)80029-7. 3416603</bibtext> </blist> <blist> <bibtext> Venables WN, Ripley BD. Modern applied statistics with S. 20024; Springer. 10.1007/978-0-387-21706-2</bibtext> </blist> <blist> <bibtext> Wilcox RR. Introduction to robust estimation and hypothesis testing. 20215; Elsevier</bibtext> </blist> <blist> <bibtext> Wilcox RR, Keselman HJ. Modern robust data analysis methods: Measures of central tendency. Psychological Methods. 2003; 8; 3: 254-274. 10.1037/1082-989X.8.3.254. 14596490</bibtext> </blist> <blist> <bibtext> Wilding J, Cook S. Sex differences and individual consistency in voice identification. Perceptual and Motor Skills. 2000; 91; 2: 535-538. 10.2466/pms.2000.91.2.535. 11065315</bibtext> </blist> <blist> <bibtext> World Health Organization. World report on hearing. 2021; World Health Organization</bibtext> </blist> <blist> <bibtext> Wright DB, Sladden B. An own gender bias and the importance of hair in face recognition. Acta Psychologica. 2003; 114; 1: 101-114. 10.1016/S0001-6918(03)00052-0. 12927345</bibtext> </blist> <blist> <bibtext> Wright DB, Stroud JN. Age differences in lineup identification accuracy: people are better with their own age. Law and Human Behavior. 2002; 26; 6: 641-654. 10.1023/A:1020981501383. 12508699</bibtext> </blist> <blist> <bibtext> Yarmey AD, Matthys E. Voice identification of an abductor. Applied Cognitive Psychology. 1992; 6; 5: 367-377. 10.1002/acp.2350060502</bibtext> </blist> <blist> <bibtext> Yarmey AD, Yarmey AL, Yarmey MJ, Parliament L. Commonsense beliefs and the identification of familiar voices. Applied Cognitive Psychology. 2001; 15; 3: 283-299. 10.1002/acp.702</bibtext> </blist> <blist> <bibtext> Yonan CA, Sommers MS. The effects of talker familiarity on spoken word identification in younger and older listeners. Psychology and Aging. 2000; 15; 1: 88-99. 10.1037/0882-7974.15.1.88. 10755292</bibtext> </blist> <blist> <bibtext> Zaltz Y, Kishon-Rabin L. Difficulties experienced by older listeners in utilizing voice cues for speaker discrimination. Frontiers in Psychology. 2022; 13. 10.3389/fpsyg.2022.797422. 35310278. 8928022797422</bibtext> </blist> </ref> <ref id="AN0185782071-25"> <title> Footnotes </title> <blist> <bibtext> We refrain from using the term 'bias' for this effect since it is ambiguous: it can either mean a recognition advantage for own-group stimuli (i.e., so-called performance bias) or a response bias, whereby individuals tend to choose one of the response options in the experiment significantly more.</bibtext> </blist> <blist> <bibtext> Next to the chosen statistical procedure, we considered numerous other plausible models. One of them was linear regression. However, we found that this procedure was not suitable for our data since visual inspection revealed outliers. Another procedure, linear mixed effect regression, was also not applicable since single-trial data were collapsed during calculations of d′ and c measures. Also, nonparametric alternatives to ANOVA such as Kruskal–Wallis or Friedman's tests were not suitable for our experimental design since they do not allow to test for interactions. Lastly, there appears to be no robust alternative for the four-way mixed ANOVAs in R, which would mitigate the presence of outliers. This is why we tested the interactions and main effects of our four factors with normal four-way mixed ANOVAs.</bibtext> </blist> </ref> <aug> <p>By Valeriia Vyshnevetska; Nathalie Giroud; Meike Ramon and Volker Dellwo</p> <p>Reported by Author; Author; Author; Author</p> </aug> <nolink nlid="nl1" bibid="bib36" firstref="ref1"></nolink> <nolink nlid="nl2" bibid="bib88" firstref="ref2"></nolink> <nolink nlid="nl3" bibid="bib71" firstref="ref3"></nolink> <nolink nlid="nl4" bibid="bib18" firstref="ref4"></nolink> <nolink nlid="nl5" bibid="bib26" firstref="ref5"></nolink> <nolink nlid="nl6" bibid="bib61" firstref="ref6"></nolink> <nolink nlid="nl7" bibid="bib60" firstref="ref8"></nolink> <nolink nlid="nl8" bibid="bib68" firstref="ref9"></nolink> <nolink nlid="nl9" bibid="bib80" firstref="ref10"></nolink> <nolink nlid="nl10" bibid="bib38" firstref="ref11"></nolink> <nolink nlid="nl11" bibid="bib31" firstref="ref14"></nolink> <nolink nlid="nl12" bibid="bib40" firstref="ref15"></nolink> <nolink nlid="nl13" bibid="bib64" firstref="ref16"></nolink> <nolink nlid="nl14" bibid="bib91" firstref="ref17"></nolink> <nolink nlid="nl15" bibid="bib114" firstref="ref18"></nolink> <nolink nlid="nl16" bibid="bib27" firstref="ref19"></nolink> <nolink nlid="nl17" bibid="bib33" firstref="ref20"></nolink> <nolink nlid="nl18" bibid="bib53" firstref="ref22"></nolink> <nolink nlid="nl19" bibid="bib90" firstref="ref23"></nolink> <nolink nlid="nl20" bibid="bib46" firstref="ref24"></nolink> <nolink nlid="nl21" bibid="bib98" firstref="ref25"></nolink> <nolink nlid="nl22" bibid="bib14" firstref="ref26"></nolink> <nolink nlid="nl23" bibid="bib24" firstref="ref27"></nolink> <nolink nlid="nl24" bibid="bib41" firstref="ref28"></nolink> <nolink nlid="nl25" bibid="bib59" firstref="ref29"></nolink> <nolink nlid="nl26" bibid="bib83" firstref="ref30"></nolink> <nolink nlid="nl27" bibid="bib22" firstref="ref32"></nolink> <nolink nlid="nl28" bibid="bib35" firstref="ref33"></nolink> <nolink nlid="nl29" bibid="bib58" firstref="ref34"></nolink> <nolink nlid="nl30" bibid="bib62" firstref="ref35"></nolink> <nolink nlid="nl31" bibid="bib81" firstref="ref36"></nolink> <nolink nlid="nl32" bibid="bib94" firstref="ref37"></nolink> <nolink nlid="nl33" bibid="bib109" firstref="ref38"></nolink> <nolink nlid="nl34" bibid="bib34" firstref="ref43"></nolink> <nolink nlid="nl35" bibid="bib110" firstref="ref44"></nolink> <nolink nlid="nl36" bibid="bib63" firstref="ref45"></nolink> <nolink nlid="nl37" bibid="bib75" firstref="ref46"></nolink> <nolink nlid="nl38" bibid="bib85" firstref="ref47"></nolink> <nolink nlid="nl39" bibid="bib28" firstref="ref51"></nolink> <nolink nlid="nl40" bibid="bib95" firstref="ref52"></nolink> <nolink nlid="nl41" bibid="bib10" firstref="ref53"></nolink> <nolink nlid="nl42" bibid="bib12" firstref="ref54"></nolink> <nolink nlid="nl43" bibid="bib100" firstref="ref55"></nolink> <nolink nlid="nl44" bibid="bib97" firstref="ref56"></nolink> <nolink nlid="nl45" bibid="bib84" firstref="ref57"></nolink> <nolink nlid="nl46" bibid="bib93" firstref="ref58"></nolink> <nolink nlid="nl47" bibid="bib87" firstref="ref60"></nolink> <nolink nlid="nl48" bibid="bib51" firstref="ref61"></nolink> <nolink nlid="nl49" bibid="bib17" firstref="ref64"></nolink> <nolink nlid="nl50" bibid="bib20" firstref="ref66"></nolink> <nolink nlid="nl51" bibid="bib73" firstref="ref67"></nolink> <nolink nlid="nl52" bibid="bib67" firstref="ref68"></nolink> <nolink nlid="nl53" bibid="bib108" firstref="ref69"></nolink> <nolink nlid="nl54" bibid="bib23" firstref="ref70"></nolink> <nolink nlid="nl55" bibid="bib29" firstref="ref71"></nolink> <nolink nlid="nl56" bibid="bib15" firstref="ref72"></nolink> <nolink nlid="nl57" bibid="bib74" firstref="ref73"></nolink> <nolink nlid="nl58" bibid="bib42" firstref="ref74"></nolink> <nolink nlid="nl59" bibid="bib86" firstref="ref78"></nolink> <nolink nlid="nl60" bibid="bib11" firstref="ref79"></nolink> <nolink nlid="nl61" bibid="bib55" firstref="ref81"></nolink> <nolink nlid="nl62" bibid="bib96" firstref="ref82"></nolink> <nolink nlid="nl63" bibid="bib54" firstref="ref83"></nolink> <nolink nlid="nl64" bibid="bib45" firstref="ref87"></nolink> <nolink nlid="nl65" bibid="bib77" firstref="ref92"></nolink> <nolink nlid="nl66" bibid="bib105" firstref="ref94"></nolink> <nolink nlid="nl67" bibid="bib78" firstref="ref95"></nolink> <nolink nlid="nl68" bibid="bib79" firstref="ref96"></nolink> <nolink nlid="nl69" bibid="bib106" firstref="ref97"></nolink> <nolink nlid="nl70" bibid="bib57" firstref="ref98"></nolink> <nolink nlid="nl71" bibid="bib104" firstref="ref104"></nolink> <nolink nlid="nl72" bibid="bib65" firstref="ref112"></nolink> <nolink nlid="nl73" bibid="bib92" firstref="ref113"></nolink> <nolink nlid="nl74" bibid="bib39" firstref="ref115"></nolink> <nolink nlid="nl75" bibid="bib72" firstref="ref116"></nolink> <nolink nlid="nl76" bibid="bib32" firstref="ref117"></nolink> <nolink nlid="nl77" bibid="bib101" firstref="ref118"></nolink> <nolink nlid="nl78" bibid="bib48" firstref="ref119"></nolink> <nolink nlid="nl79" bibid="bib49" firstref="ref120"></nolink> <nolink nlid="nl80" bibid="bib66" firstref="ref123"></nolink> <nolink nlid="nl81" bibid="bib89" firstref="ref127"></nolink> <nolink nlid="nl82" bibid="bib43" firstref="ref128"></nolink> <nolink nlid="nl83" bibid="bib13" firstref="ref134"></nolink> <nolink nlid="nl84" bibid="bib16" firstref="ref135"></nolink> <nolink nlid="nl85" bibid="bib113" firstref="ref139"></nolink> <nolink nlid="nl86" bibid="bib19" firstref="ref141"></nolink> <nolink nlid="nl87" bibid="bib44" firstref="ref142"></nolink> <nolink nlid="nl88" bibid="bib69" firstref="ref143"></nolink> <nolink nlid="nl89" bibid="bib107" firstref="ref146"></nolink> <nolink nlid="nl90" bibid="bib99" firstref="ref151"></nolink> <nolink nlid="nl91" bibid="bib76" firstref="ref153"></nolink> <nolink nlid="nl92" bibid="bib21" firstref="ref154"></nolink> <nolink nlid="nl93" bibid="bib111" firstref="ref157"></nolink> <nolink nlid="nl94" bibid="bib112" firstref="ref158"></nolink> <nolink nlid="nl95" bibid="bib37" firstref="ref163"></nolink> <nolink nlid="nl96" bibid="bib52" firstref="ref164"></nolink> <nolink nlid="nl97" bibid="bib56" firstref="ref165"></nolink> <nolink nlid="nl98" bibid="bib102" firstref="ref166"></nolink> <nolink nlid="nl99" bibid="bib103" firstref="ref168"></nolink> <nolink nlid="nl100" bibid="bib25" firstref="ref170"></nolink> <nolink nlid="nl101" bibid="bib70" firstref="ref173"></nolink> <nolink nlid="nl102" bibid="bib30" firstref="ref175"></nolink> <nolink nlid="nl103" bibid="bib82" firstref="ref176"></nolink> <nolink nlid="nl104" bibid="bib47" firstref="ref178"></nolink> <nolink nlid="nl105" bibid="bib50" firstref="ref179"></nolink>
Header DbId: eric
DbLabel: ERIC
An: EJ1473391
AccessLevel: 3
PubType: Academic Journal
PubTypeId: academicJournal
PreciseRelevancyScore: 0
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: Listeners Are Biased towards Voices of Young Speakers and Female Speakers When Discriminating Voices
– Name: Language
  Label: Language
  Group: Lang
  Data: English
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22Valeriia+Vyshnevetska%22">Valeriia Vyshnevetska</searchLink> (ORCID <externalLink term="http://orcid.org/0009-0003-3355-9580">0009-0003-3355-9580</externalLink>)<br /><searchLink fieldCode="AR" term="%22Nathalie+Giroud%22">Nathalie Giroud</searchLink><br /><searchLink fieldCode="AR" term="%22Meike+Ramon%22">Meike Ramon</searchLink><br /><searchLink fieldCode="AR" term="%22Volker+Dellwo%22">Volker Dellwo</searchLink>
– Name: TitleSource
  Label: Source
  Group: Src
  Data: <searchLink fieldCode="SO" term="%22Cognitive+Research%3A+Principles+and+Implications%22"><i>Cognitive Research: Principles and Implications</i></searchLink>. 2025 10.
– Name: Avail
  Label: Availability
  Group: Avail
  Data: Springer. Available from: Springer Nature. One New York Plaza, Suite 4600, New York, NY 10004. Tel: 800-777-4643; Tel: 212-460-1500; Fax: 212-460-1700; e-mail: customerservice@springernature.com; Web site: https://link.springer.com/
– Name: PeerReviewed
  Label: Peer Reviewed
  Group: SrcInfo
  Data: Y
– Name: Pages
  Label: Page Count
  Group: Src
  Data: 14
– Name: DatePubCY
  Label: Publication Date
  Group: Date
  Data: 2025
– Name: TypeDocument
  Label: Document Type
  Group: TypDoc
  Data: Journal Articles<br />Reports - Research
– Name: Subject
  Label: Descriptors
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Listening+Skills%22">Listening Skills</searchLink><br /><searchLink fieldCode="DE" term="%22Bias%22">Bias</searchLink><br /><searchLink fieldCode="DE" term="%22Auditory+Discrimination%22">Auditory Discrimination</searchLink><br /><searchLink fieldCode="DE" term="%22Females%22">Females</searchLink><br /><searchLink fieldCode="DE" term="%22Young+Adults%22">Young Adults</searchLink><br /><searchLink fieldCode="DE" term="%22Older+Adults%22">Older Adults</searchLink><br /><searchLink fieldCode="DE" term="%22Age+Differences%22">Age Differences</searchLink><br /><searchLink fieldCode="DE" term="%22Recognition+%28Psychology%29%22">Recognition (Psychology)</searchLink><br /><searchLink fieldCode="DE" term="%22Responses%22">Responses</searchLink>
– Name: DOI
  Label: DOI
  Group: ID
  Data: 10.1186/s41235-025-00636-3
– Name: ISSN
  Label: ISSN
  Group: ISSN
  Data: 2365-7464
– Name: Abstract
  Label: Abstract
  Group: Ab
  Data: In face processing, an own-age recognition advantage has frequently been reported whereby observers are better at recognizing faces of their own compared to other age groups. We wanted to know whether own-age effects exist in voice recognition. Two listener groups, younger adults (n = 42, 19-35 years, 21 males) and older adults (n = 32, 65-83 years, 14 males), completed a speaker discrimination task (same/different speakers), which included younger and older adult speakers of both sexes. Results revealed no interaction of the factors speaker and listener age and speaker and listener sex on listeners' sensitivity (d'). Main effects were significant for listener age (young adult listeners exhibited higher sensitivity than the older adult listeners) and speaker sex (listeners' sensitivity was higher for male compared to female voices). Crucially, response bias (c) revealed that listeners had a significantly higher 'same' bias when hearing younger speakers and female speakers. Our findings have implications for theories of voice identity processing and forensic contexts requiring discrimination of speakers' identity, e.g. earwitnesses telling apart younger and female speakers.
– Name: AbstractInfo
  Label: Abstractor
  Group: Ab
  Data: As Provided
– Name: DateEntry
  Label: Entry Date
  Group: Date
  Data: 2025
– Name: AN
  Label: Accession Number
  Group: ID
  Data: EJ1473391
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=eric&AN=EJ1473391
RecordInfo BibRecord:
  BibEntity:
    Identifiers:
      – Type: doi
        Value: 10.1186/s41235-025-00636-3
    Languages:
      – Text: English
    PhysicalDescription:
      Pagination:
        PageCount: 14
    Subjects:
      – SubjectFull: Listening Skills
        Type: general
      – SubjectFull: Bias
        Type: general
      – SubjectFull: Auditory Discrimination
        Type: general
      – SubjectFull: Females
        Type: general
      – SubjectFull: Young Adults
        Type: general
      – SubjectFull: Older Adults
        Type: general
      – SubjectFull: Age Differences
        Type: general
      – SubjectFull: Recognition (Psychology)
        Type: general
      – SubjectFull: Responses
        Type: general
    Titles:
      – TitleFull: Listeners Are Biased towards Voices of Young Speakers and Female Speakers When Discriminating Voices
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: Valeriia Vyshnevetska
      – PersonEntity:
          Name:
            NameFull: Nathalie Giroud
      – PersonEntity:
          Name:
            NameFull: Meike Ramon
      – PersonEntity:
          Name:
            NameFull: Volker Dellwo
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 01
              M: 12
              Type: published
              Y: 2025
          Identifiers:
            – Type: issn-electronic
              Value: 2365-7464
          Numbering:
            – Type: volume
              Value: 10
          Titles:
            – TitleFull: Cognitive Research: Principles and Implications
              Type: main
ResultId 1