Face Masks and Fake Masks: The Effect of Real and Superimposed Masks on Face Matching with Super-Recognisers, Typical Observers, and Algorithms

Saved in:
Bibliographic Details
Title: Face Masks and Fake Masks: The Effect of Real and Superimposed Masks on Face Matching with Super-Recognisers, Typical Observers, and Algorithms
Language: English
Authors: Kay L. Ritchie (ORCID 0000-0002-1348-760X), Daniel J. Carragher, Josh P. Davis, Katie Read, Ryan E. Jenkins, Eilidh Noyes, Katie L. H. Gray, Peter J. B. Hancock
Source: Cognitive Research: Principles and Implications. 2024 9.
Availability: Springer. Available from: Springer Nature. One New York Plaza, Suite 4600, New York, NY 10004. Tel: 800-777-4643; Tel: 212-460-1500; Fax: 212-460-1700; e-mail: customerservice@springernature.com; Web site: https://link.springer.com/
Peer Reviewed: Y
Page Count: 13
Publication Date: 2024
Document Type: Journal Articles
Reports - Research
Descriptors: Artificial Intelligence, Recognition (Psychology), Clothing, Health Behavior, Observation, Human Body, Visual Acuity, Visual Stimuli, COVID-19, Pandemics
DOI: 10.1186/s41235-024-00532-2
ISSN: 2365-7464
Abstract: Mask wearing has been required in various settings since the outbreak of COVID-19, and research has shown that identity judgements are difficult for faces wearing masks. To date, however, the majority of experiments on face identification with masked faces tested humans and computer algorithms using images with superimposed masks rather than images of people wearing real face coverings. In three experiments we test humans (control participants and super-recognisers) and algorithms with images showing different types of face coverings. In all experiments we tested matching concealed or unconcealed faces to an unconcealed reference image, and we found a consistent decrease in face matching accuracy with masked compared to unconcealed faces. In Experiment 1, typical human observers were most accurate at face matching with unconcealed images, and poorer for three different types of superimposed mask conditions. In Experiment 2, we tested both typical observers and super-recognisers with superimposed and real face masks, and found that performance was poorer for real compared to superimposed masks. The same pattern was observed in Experiment 3 with algorithms. Our results highlight the importance of testing both humans and algorithms with real face masks, as using only superimposed masks may underestimate their detrimental effect on face identification.
Abstractor: As Provided
Notes: https://osf.io/qgxhs/?view_only=6c6e8368c49d4d4fb634ada0671a7972
Entry Date: 2024
Accession Number: EJ1410458
Database: ERIC
Full text is not displayed to guests.
FullText Links:
  – Type: pdflink
    Url: https://content.ebscohost.com/cds/retrieve?content=AQICAHj0k_4E0hTGH8RJwT4gCJyBsGNe_WN95AvKlDbXJGqwxwHICRLyqtNrKMzteangSFOmAAAA4zCB4AYJKoZIhvcNAQcGoIHSMIHPAgEAMIHJBgkqhkiG9w0BBwEwHgYJYIZIAWUDBAEuMBEEDOA5_qZ8gB1lHS4sRQIBEICBm6uN8yMZw-nrPFRT0DDlK9yomzM_T9h6NiG9BvOiAfeM7b-Z-e7eq1ERtA73oxgsqatavQMWJeFzpzpvPsuzwxtzth0P4xKUn77MitWUTgTKBxnyeYJl6oAqva8kku7GDwsIGC1NN5T3LYBhilgAAsQLjRqmKqDetPHbbuQjpjDNl775FXPSF2nLPK29-XJsUbJ660trokDnk4v_
Text:
  Availability: 1
  Value: <anid>AN0175199277;[k1e6]02feb.24;2024Feb05.05:55;v2.2.500</anid> <title id="AN0175199277-1">Face masks and fake masks: the effect of real and superimposed masks on face matching with super-recognisers, typical observers, and algorithms </title> <p>Mask wearing has been required in various settings since the outbreak of COVID-19, and research has shown that identity judgements are difficult for faces wearing masks. To date, however, the majority of experiments on face identification with masked faces tested humans and computer algorithms using images with superimposed masks rather than images of people wearing real face coverings. In three experiments we test humans (control participants and super-recognisers) and algorithms with images showing different types of face coverings. In all experiments we tested matching concealed or unconcealed faces to an unconcealed reference image, and we found a consistent decrease in face matching accuracy with masked compared to unconcealed faces. In Experiment 1, typical human observers were most accurate at face matching with unconcealed images, and poorer for three different types of superimposed mask conditions. In Experiment 2, we tested both typical observers and super-recognisers with superimposed and real face masks, and found that performance was poorer for real compared to superimposed masks. The same pattern was observed in Experiment 3 with algorithms. Our results highlight the importance of testing both humans and algorithms with real face masks, as using only superimposed masks may underestimate their detrimental effect on face identification.</p> <p>Keywords: Face masks; Face matching; Super-recognisers; Automatic face recognition</p> <p>Supplementary Information The online version contains supplementary material available at https://doi.org/10.1186/s41235-024-00532-2.</p> <hd id="AN0175199277-2">Introduction</hd> <p></p> <hd id="AN0175199277-3">Unfamiliar face matching</hd> <p>While humans are very good at recognising the faces of familiar people (e.g. Bruce, [<reflink idref="bib6" id="ref1">6</reflink>]; Bruce et al., [<reflink idref="bib7" id="ref2">7</reflink>]; Burton et al., [<reflink idref="bib10" id="ref3">10</reflink>]), we are far poorer at recognising unfamiliar people. In a typical face matching task, participants are shown two images and are asked to judge whether they depict the same person or two different people. Unfamiliar face matching performance has been shown to be poor both in the laboratory (Clutterbuck & Johnston, [<reflink idref="bib14" id="ref4">14</reflink>], [<reflink idref="bib15" id="ref5">15</reflink>]; Megreya & Burton, [<reflink idref="bib40" id="ref6">40</reflink>]; Ritchie et al., [<reflink idref="bib52" id="ref7">52</reflink>], [<reflink idref="bib50" id="ref8">50</reflink>], [<reflink idref="bib49" id="ref9">49</reflink>]; Sandford & Ritchie, [<reflink idref="bib55" id="ref10">55</reflink>]), and in live tasks matching a physically present unfamiliar person to a photograph (Davis & Valentine, [<reflink idref="bib18" id="ref11">18</reflink>]; Kemp et al., [<reflink idref="bib34" id="ref12">34</reflink>]; Megreya & Burton, [<reflink idref="bib40" id="ref13">40</reflink>]; Ritchie et al., [<reflink idref="bib51" id="ref14">51</reflink>]). Unfamiliar face matching performance is poor even in people who are employed to make identity decisions from images, such as checkout assistants (Kemp et al., [<reflink idref="bib34" id="ref15">34</reflink>]), passport officers (White et al., [<reflink idref="bib64" id="ref16">64</reflink>]), and police officers (Burton et al., [<reflink idref="bib10" id="ref17">10</reflink>]).</p> <p>The addition of everyday paraphernalia such as glasses and sunglasses to one image in the pair has been shown to reduce face matching accuracy (Graham & Ritchie, [<reflink idref="bib29" id="ref18">29</reflink>]; Kramer & Ritchie, [<reflink idref="bib35" id="ref19">35</reflink>]; Noyes et al., [<reflink idref="bib45" id="ref20">45</reflink>]). Face masks have also been shown to impair face identification (Fitousi et al., [<reflink idref="bib23" id="ref21">23</reflink>]; Freud et al., [<reflink idref="bib25" id="ref22">25</reflink>], [<reflink idref="bib24" id="ref23">24</reflink>]) and face matching (Carragher & Hancock, [<reflink idref="bib12" id="ref24">12</reflink>]; Dhamecha et al., [<reflink idref="bib20" id="ref25">20</reflink>]; Estudillo et al., [<reflink idref="bib21" id="ref26">21</reflink>]; Noyes et al., [<reflink idref="bib45" id="ref27">45</reflink>]), with masks causing more of a reduction in accuracy than sunglasses (Noyes et al., [<reflink idref="bib45" id="ref28">45</reflink>]). It is not clear, however, precisely why face masks cause an impairment to face matching performance. The current study seeks to shed light on the mechanisms underlying this effect by testing face matching using different types of lower face occlusions.</p> <hd id="AN0175199277-4">Super-recognisers</hd> <p>Although unfamiliar face matching is generally poor, some people are able to perform with far higher accuracy than the general population. First described as having exceptional face memory (Russell et al., [<reflink idref="bib54" id="ref29">54</reflink>]), these people are referred to as super-recognisers (see Noyes et al., [<reflink idref="bib47" id="ref30">47</reflink>] for a review). Although there are individual differences between super-recognisers, at the group level they perform with consistently higher accuracy than control participants (Bobak et al., [<reflink idref="bib3" id="ref31">3</reflink>], [<reflink idref="bib4" id="ref32">4</reflink>]; Bobak et al., [<reflink idref="bib3" id="ref33">3</reflink>], [<reflink idref="bib4" id="ref34">4</reflink>]; Davis et al, [<reflink idref="bib17" id="ref35">17</reflink>]; Noyes et al., [<reflink idref="bib46" id="ref36">46</reflink>]; Phillips et al., [<reflink idref="bib48" id="ref37">48</reflink>]). A recent study showed that super-recognisers are also more accurate than control participants at face matching with images wearing face masks (Noyes et al., [<reflink idref="bib45" id="ref38">45</reflink>]). The current study extends this work by testing both control participants and super-recognisers with different types of face coverings.</p> <hd id="AN0175199277-5">Algorithms</hd> <p>In recent years, there has been a rapid improvement in the performance of facial recognition algorithms through the use of 'Deep Convolutional Neural Networks' (DCNNs; e.g. Cao et al., [<reflink idref="bib11" id="ref39">11</reflink>]; Kemelmacher-Shlizerman et al., [<reflink idref="bib33" id="ref40">33</reflink>]; Taigman et al., [<reflink idref="bib60" id="ref41">60</reflink>]). One study tested algorithms made in 2015, 2016 and 2017 and showed a monotonic increase in performance from the oldest (68% accurate) to the newest (96% accurate; Phillips et al., [<reflink idref="bib48" id="ref42">48</reflink>]). Face masks present a new challenge for algorithm face identification. A recent competition receiving 18 submissions found that eight did not meet the baseline criterion for verification errors (Boutros et al., [<reflink idref="bib5" id="ref43">5</reflink>]). The National Institute of Standards and Technology (NIST) in the USA runs a regular Face Recognition Vendor Test (FRVT) which is a standard test of facial recognition algorithms. The FRVT has consistently reported improvements in algorithm face identification with algorithms achieving higher accuracy than humans (NIST, [<reflink idref="bib42" id="ref44">42</reflink>]). NIST now also runs an 'FRVT Face Mask Effects' looking specifically at algorithm identification from masked faces. Algorithms are presented faces with superimposed masks and are tasked with identifying the person from a database of unmasked images (NIST, [<reflink idref="bib41" id="ref45">41</reflink>]). Updates to the test show that some developers have adapted their algorithms to better cope with face masks, although the shape, colour, and coverage of the different masks used in the test affects some algorithms' ability both to detect the face in the first place, and then to correctly identify the person pictured (Ngan et al., [<reflink idref="bib43" id="ref46">43</reflink>]).</p> <hd id="AN0175199277-6">Types of face coverings</hd> <p>While some previous studies of human face identification ability with face masks have used images of people wearing real masks (Dhamecha et al., [<reflink idref="bib20" id="ref47">20</reflink>]; Fitousi et al., [<reflink idref="bib23" id="ref48">23</reflink>]; Noyes et al., [<reflink idref="bib45" id="ref49">45</reflink>]), the majority have used pre-existing images with masks superimposed on to them (Carragher & Hancock, [<reflink idref="bib12" id="ref50">12</reflink>]; Estudillo et al., [<reflink idref="bib21" id="ref51">21</reflink>]; Freud et al., [<reflink idref="bib25" id="ref52">25</reflink>], [<reflink idref="bib24" id="ref53">24</reflink>]). Some recent computer vision research has used real face masks (e.g. Jeevan et al., [<reflink idref="bib31" id="ref54">31</reflink>]; Lionnie et al. [<reflink idref="bib36" id="ref55">36</reflink>]), but the NIST FRVT Face Mask Effects test uses superimposed masks as the test images (Ngan et al., [<reflink idref="bib43" id="ref56">43</reflink>]).</p> <p>It is not clear whether superimposed and real face masks produce different deficits in either human or computer face matching performance, and this difference is important for both theoretical understanding of face perception, and for understanding the impact of masks in applied face recognition practice. We have previously argued that one study using real face masks (Noyes et al., [<reflink idref="bib45" id="ref57">45</reflink>]) found a smaller reduction in face matching accuracy than a study using superimposed face masks (Carragher & Hancock, [<reflink idref="bib12" id="ref58">12</reflink>]) because it is possible that some elements of the person's real face shape are still available to the viewer in real mask images but are covered in superimposed mask images. Although we predominantly use face texture to recognise other people (e.g. Burton et al., [<reflink idref="bib8" id="ref59">8</reflink>]), some element of face shape information may be useful (Rogers et al., [<reflink idref="bib53" id="ref60">53</reflink>]). Alternatively, it is possible that real face masks introduce extra texture information which may be more disruptive for face processing than superimposed masks, and the previously observed differences in findings (Carragher & Hancock, [<reflink idref="bib12" id="ref61">12</reflink>]; Noyes et al., [<reflink idref="bib45" id="ref62">45</reflink>]) were simply due to different task demands and methodologies.</p> <hd id="AN0175199277-7">The current studies</hd> <p>It is not clear exactly why face masks cause such a marked impairment in human face matching performance. One possibility is that masks cover facial features that are useful for identification (Towler et al., [<reflink idref="bib63" id="ref63">63</reflink>]). But previous research suggests that the upper half of the face, which remains visible when wearing a face covering, tends to be more useful for identification than the lower half (Fisher & Cox, [<reflink idref="bib22" id="ref64">22</reflink>]; McKelvie, [<reflink idref="bib39" id="ref65">39</reflink>]). Alternatively, covering the features of the lower face might interfere with the holistic processes that are used in face recognition (Maurer et al., [<reflink idref="bib38" id="ref66">38</reflink>]; Tanaka & Farah, [<reflink idref="bib61" id="ref67">61</reflink>]). In support of this possibility, Freud et al. ([<reflink idref="bib25" id="ref68">25</reflink>]) report that holistic processing is impaired for faces wearing a face mask (see also Stajduhar et al., [<reflink idref="bib58" id="ref69">58</reflink>]). However, face matching can be aided by featural comparisons (Towler et al., [<reflink idref="bib63" id="ref70">63</reflink>]; White et al., [<reflink idref="bib65" id="ref71">65</reflink>]), which can occur without holistic processing (Towler et al., [<reflink idref="bib62" id="ref72">62</reflink>]). Recent research has shown that featural comparisons can lead to modest improvements in masked face matching performance (Carragher et al., [<reflink idref="bib13" id="ref73">13</reflink>]). The final possibility considered here is that the face mask serves as a source of distraction by attracting attention to the mask and away from the visible facial features.</p> <p>In Experiment 1, we compare human unfamiliar face matching with different types of superimposed lower face occlusions. In Experiment 2, we compare unfamiliar face matching by control participants and super-recognisers with superimposed and real face masks, and in Experiment 3, we test algorithm performance with the real and superimposed masks.</p> <hd id="AN0175199277-8">Experiment 1: face matching with different types of superimposed lower face occlusions</hd> <p>This experiment was designed to investigate whether different types of superimposed face masks modulate the degree of impairment caused to unfamiliar face matching performance. In a within-participants design, observers completed a matching task in which one face in each pair was always presented unmasked, while the other face was selected from the following mask conditions: control (unmasked), fitted mask (the mask closely followed the shape of the face), loose mask (the mask occluded a large square shape, including the neck) and the top half only (the entire lower half of the image was removed). First, we expect that performance will be higher for the control condition than all others, replicating the basic finding that face masks impair matching performance (Carragher & Hancock, [<reflink idref="bib12" id="ref74">12</reflink>]; Noyes et al., [<reflink idref="bib45" id="ref75">45</reflink>]). Comparisons between the mask conditions could potentially reveal the mechanism by which masks impair face matching performance. Higher accuracy in the fitted mask condition compared to the loose mask condition would suggest that observers can extract information about facial shape from the mask. Alternatively, significantly better performance in the top half only condition compared to the two mask conditions (fitted, loose), would suggest that masks are a source of attentional distraction. Finally, no difference between the three manipulated conditions (fitted mask, loose mask, top only) would be consistent with two different explanations; either that face masks impair matching performance because they cover facial features that are important for identification, or because they impede holistic processing. These final possibilities are inextricably linked because covering facial features will, by definition, also interfere with holistic processing.</p> <hd id="AN0175199277-9">Method</hd> <p></p> <hd id="AN0175199277-10">Participants</hd> <p>From a convenience sample of volunteers recruited via email and social media, we received complete data from 79 participants (22 male, 57 female; mean age: 34 years; SD: 16 years; range: 18–67 years). All participants were naïve to the aims of the study. This research was approved by the General University Ethics Panel at the University of Stirling, and all participants gave informed consent.</p> <hd id="AN0175199277-11">Stimuli</hd> <p>The face masks in the current study were plain colour patches that were fitted to the faces automatically using custom written code (see Fig. 1). Automatically located landmark points were fine-tuned manually. The same landmark points below the eyes and over the bridge of the nose were used to establish the top of the mask in each mask condition (fitted, loose, top only). The fitted mask was created by filling the landmark points that follow the shape of the jaw with a plain pale blue patch (RGB 143, 205, 205), which is most similar to the FRVT Face Mask Effects' 'wide, medium coverage' mask (Ngan et al., [<reflink idref="bib43" id="ref76">43</reflink>]). The loose masks were created by extending the occlusion 10 pixels down below the bottom of the jaw, square below the widest point at the ears. The top only condition was created by cropping the image below the top of the mask.</p> <p>Graph: Fig. 1Examples of the a Control b Fitted Mask c Loose Mask and d Top Only stimuli used in Experiment 1. The images depict an identity who was not included in the experiment, but has given permission for their images to be used</p> <p>The faces for the current experiment came from two separate face matching tests. Half of the trials were the unfamiliar face pairs from the Stirling Famous Face Matching Task created by Carragher and Hancock ([<reflink idref="bib12" id="ref77">12</reflink>]), making this the Stirling Unfamiliar Face Matching Task (SUFMT). These face pairs are images of amateur models that were downloaded from various online sources. The SUFMT consists of 40 image pairs, of which 20 are identity matches. The match and mismatch trials are evenly split for face sex. Each image only appears once within the SUFMT. The remaining trials came from the short version of the Kent Face Matching Test (KFMT; Fysh & Bindemann, [<reflink idref="bib27" id="ref78">27</reflink>]). The KFMT also consists of 40 trials, of which 20 are matches and 20 are mismatches. Each image pair consists of one smaller image that is typical of a student ID card, and one larger high-quality portrait image. The KFMT also consists of male and female face pairs. Thus, the experiment consisted of 80 trials in total, of which 40 were identity matches.</p> <p>Trials from the SUFMT and KFMT were intermixed and randomised. Because all participants completed the same two tasks, we did not compare performance between the two tests. Allocation of trial pairs to mask conditions (control, fitted mask, loose mask, top only) was randomised between participants, such that all pairs were presented in each mask condition across participants. All participants completed 20 trials of each mask condition, of which 10 were match trials and 10 were mismatch trials. Face pairs in the fitted mask, loose mask and top only conditions consisted of one full-view face and one altered face. This image arrangement is consistent with the scenario in which a masked individual presents an official photo-ID document for inspection. In the KFMT, the smaller ID image was always unmasked, while the larger image was shown in each mask condition. All images were presented in colour. Images from the SUFMT were 420 × 595 px in size. Images from the KFMT were presented in their original sizes (Fysh & Bindemann, [<reflink idref="bib27" id="ref79">27</reflink>]); small (142 × 192 px), large (283 × 332 px).</p> <hd id="AN0175199277-12">Procedure</hd> <p>Participants completed the experiment on their personal computers via a web link. The experiment was run using Qualtrics survey software. Participants were informed that their task was to decide whether the two simultaneously presented images showed the same person or two different people. Responses were made using a 6-choice scale, which conveyed the identification decision ("Same", "Different") and confidence ("Certain", "Think", "Guess"). There was no time limit to give a response. All trial types were intermixed and presented in a random order in a single experimental block that consisted of all 80 trials. The experiment took approximately 15 min (<emph>M</emph> = 899 s, SD = 363 s) to complete.</p> <hd id="AN0175199277-13">Analysis</hd> <p>We analysed the data using signal detection measures of sensitivity (d′) and response bias (criterion). Sensitivity measures how well participants can discriminate match pairs from mismatches, with higher values indicating better performance (Macmillan & Creelman, [<reflink idref="bib37" id="ref80">37</reflink>]). Criterion is a measure of response bias, which shows whether participants had an overall tendency to report that pairs were a match ("same") or mismatch ("different"). Positive criterion values indicate a bias to respond "different" across all trials (i.e. a conservative criterion), whereas negative values signal a "match" response bias (i.e. a liberal criterion). To calculate both measures, we collapsed across the confidence component of our scale, leaving only "same" and "different" responses (e.g. "Certainly Same", "Think Same" and "Guess Same" were counted as "same"). These simplified responses correspond to hits (correctly responding "same" on a match trial) and false alarms (incorrectly responding "same" on a mismatch trial) which are used to calculate both d′ and criterion (Macmillan & Creelman, [<reflink idref="bib37" id="ref81">37</reflink>]; Stanislaw & Todorov, [<reflink idref="bib59" id="ref82">59</reflink>]). In both Experiment 1 and 2, we corrected for hits of 1 using the formula 1–1/(2N) and false alarms of 0 using the formula 1/(2N) where N is the number of trials in each condition. The number of trials was the same in each condition in each experiment, giving a maximum d′ value of 3.29. In addition to traditional frequentist hypothesis testing, we included Bayes factors calculated in JASP (JASP Team, [<reflink idref="bib30" id="ref83">30</reflink>]) with default prior width, which allowed us to quantify the extent to which the data support the alternative hypothesis (BF<subs>10</subs>). We interpret BFs of less than 3.0 as anecdotal evidence of the alternative hypothesis (e.g. Jeffreys, [<reflink idref="bib32" id="ref84">32</reflink>]).</p> <hd id="AN0175199277-14">Results and discussion</hd> <p>All data for all experiments is available at https://osf.io/qgxhs/?view_only=6c6e8368c49d4d4fb634ada0671a7972</p> <p>We present descriptive statistics here for ease of reading—full analysis of accuracy as defined by per cent correct can be found in the Additional file 1. In Experiment 1, face matching accuracy in each condition varied as follows: control (no concealment), 40% to 95% out of 20 (<emph>M</emph> = 69%, SD = 11%); fitted mask, 35% to 85% (<emph>M</emph> = 62%, SD = 10%); loose mask, 30% to 85% (<emph>M</emph> = 61%, SD = 12%); and top only, 30% to 90% (<emph>M</emph> = 60%, SD = 12%).</p> <hd id="AN0175199277-15">Sensitivity</hd> <p>Our main analysis uses signal detection theory as is common in the literature. A repeated measures ANOVA revealed a significant effect of mask condition on d′, <emph>F</emph>(<reflink idref="bib3" id="ref85">3</reflink>, 234) = 13.55, <emph>p</emph> < 0.001, <emph>η</emph><subs><emph>p</emph></subs><sups>2</sups> = 0.15, BF<subs>10</subs> > 1000 (see Fig. 2). Bonferroni corrected post-hoc comparisons showed that sensitivity was significantly higher in the control condition compared to all other conditions (all <emph>p</emph>s < 0.001, all BF<subs>10</subs> > 400), which did not differ from each other (all <emph>p</emph>s > 0.999, all BF<subs>10</subs> < 1). The pattern of results is the same when the results are analysed using per cent correct, for both overall accuracy (collapsing across match and mismatch trials), and for match trials. However, there was no effect of mask condition on mismatch trials accuracy (see Additional file 1: Sect. 1).</p> <p>Graph: Fig. 2Sensitivity (d ′) and criterion scores for Experiment 1</p> <hd id="AN0175199277-16">Criterion</hd> <p>There was a non-significant effect of mask condition on response bias, <emph>F</emph>(<reflink idref="bib3" id="ref86">3</reflink>, 234) = 2.12, <emph>p</emph> = 0.098, <emph>η</emph><subs><emph>p</emph></subs><sups>2</sups> = 0.03, BF<subs>10</subs> = 0.22.</p> <p>Sensitivity was highest in the control condition and fell significantly for the three mask conditions, which did not differ from each other. These results suggest that the shape of the superimposed mask does not influence the degree of impairment to matching performance. Our findings suggest that masks impair performance either because they occlude facial features that carry identity information, or because they disrupt holistic processing. However, this experiment only examined the effect of superimposed masks. It is possible that real masks introduce extra information, either attracting attention to the mask, or adding additional spurious texture information to the face. Therefore, it is possible that images of faces wearing real face masks may lead to reduced face matching ability compared to superimposed masks. Alternatively, as we have previously suggested (Noyes et al., [<reflink idref="bib45" id="ref87">45</reflink>]), it is possible that real face masks might preserve some information about face shape, which could be useful for identification (see Rogers et al., [<reflink idref="bib53" id="ref88">53</reflink>]). Therefore, in the following experiment we tested unfamiliar face matching with real and superimposed face masks.</p> <hd id="AN0175199277-17">Experiment 2: face matching with real and superimposed masks</hd> <p>This experiment tested both typical participants and super-recognisers. Both sets of participants were recruited from a large database of participants used in previous research (e.g. Belanova et al., [<reflink idref="bib2" id="ref89">2</reflink>]; Noyes et al., [<reflink idref="bib45" id="ref90">45</reflink>]; Satchell et al., [<reflink idref="bib56" id="ref91">56</reflink>]). Importantly, none of the participants who took part in this study had taken part in our previous test of masked face matching (Noyes et al., [<reflink idref="bib45" id="ref92">45</reflink>]). Here we aimed to examine the effect of real and superimposed masks on typical participants' and super-recognisers' unfamiliar face matching performance.</p> <p>Previous research using super-recognisers has tended to assess their ability using two tests: the Glasgow Face Matching Test: short version (GFMT, Burton et al., [<reflink idref="bib9" id="ref93">9</reflink>]) and the Cambridge Face Memory Test: Extended (CFMT + , Russell et al., [<reflink idref="bib54" id="ref94">54</reflink>]). The GFMT has recently been criticised for being a relatively easy test (e.g. Ramon, 2021), therefore here, we add a third test to the initial recruitment battery, the Kent Face Matching Test (KFMT, Fysh & Bindemann, [<reflink idref="bib27" id="ref95">27</reflink>]), which is a more difficult test of face matching than the GFMT.</p> <p>Our super-recognisers are defined as individuals scoring 100% (40 out of 40) on the GFMT (Burton et al., [<reflink idref="bib9" id="ref96">9</reflink>]), 93% (95 or more out of 102) on the CFMT + (Russell et al., [<reflink idref="bib54" id="ref97">54</reflink>]) and 82.5% (33 or more out of 40) on the KFMT (Fysh & Bindemann, [<reflink idref="bib27" id="ref98">27</reflink>]). Less than 5% of people achieve perfect performance on the GFMT (Burton et al., [<reflink idref="bib9" id="ref99">9</reflink>]), while an estimated 2% score 95 or above on the CFMT + (Bobak et al., [<reflink idref="bib3" id="ref100">3</reflink>], [<reflink idref="bib4" id="ref101">4</reflink>]; Russell et al., [<reflink idref="bib54" id="ref102">54</reflink>]), and average performance on the KFMT is 66.22%, taking the mean of performance reported in three studies (Fysh, [<reflink idref="bib26" id="ref103">26</reflink>]; Fysh & Bindemann, [<reflink idref="bib27" id="ref104">27</reflink>]; Gentry & Bindemann, [<reflink idref="bib28" id="ref105">28</reflink>]).</p> <p>During the original database recruitment process, many participants did not meet the criteria to be classed as super-recognisers. Typical-ability participants were invited from this second group who had previously scored within approximately 1 standard deviation of the normal population mean on the GFMT (i.e. 28–36: Burton et al., [<reflink idref="bib9" id="ref106">9</reflink>]), CFMT + (i.e. 58–83: Bobak et al., [<reflink idref="bib3" id="ref107">3</reflink>], [<reflink idref="bib4" id="ref108">4</reflink>]) and the KFMT (i.e. 24–29: Fysh, [<reflink idref="bib26" id="ref109">26</reflink>]; Fysh & Bindemann, [<reflink idref="bib27" id="ref110">27</reflink>]; Gentry & Bindemann, [<reflink idref="bib28" id="ref111">28</reflink>]).</p> <hd id="AN0175199277-18">Method</hd> <p></p> <hd id="AN0175199277-19">Participants</hd> <p>The control group were recruited from a large database of interested participants from the UK used in previous research (Belanova et al., [<reflink idref="bib2" id="ref112">2</reflink>]; Noyes et al., [<reflink idref="bib45" id="ref113">45</reflink>]; Satchell et al., [<reflink idref="bib56" id="ref114">56</reflink>]). We received complete data from 175 control participants (55 male, 118 female, 2 other; mean age 45 years; SD: 14 years; age range 18–75 years). The control participants had a mean GFMT score of 33.89/40 (SD = 2.06), a mean CFMT + score of 73.17 (SD = 6.81), and a mean KFMT score of 27.05 (SD = 1.60) as assessed in a previous battery of unpublished tests.</p> <p>The super-recognisers were recruited from the same large database as the control participants. We received complete data from 136 super-recognisers (43 male, 91 female, 2 other; mean age 39 years; SD: 9 years; age range 24–60 years). The super-recognisers all scored 40/40 on the GFMT, had a mean CFMT + score of 97.32 (SD = 1.97), and a mean KFMT score of 34.90 (SD = 1.53) as assessed in a previous battery of unpublished tests. No participants were given monetary compensation for taking part. The experiment received ethical approval from the University of Reading (ref: 2021–093-KG).</p> <hd id="AN0175199277-20">Stimuli</hd> <p>The stimuli were images of people who had volunteered photographs of themselves for this research project. Models were recruited from the same large database as the participants, and none of the models also acted as participants. Models were asked to provide multiple images of themselves both with and without face masks. The images supplied by 60 models (21 male, 39 female) were used to create the stimuli pairs in four concealment conditions: a) reference image (unconcealed), b) unconcealed image, c) superimposed mask image (this was the unconcealed image (b) with a face mask superimposed on to the face), and d) real mask image (see Fig. 3). Reference images always depicted the identity with a different background to the unconcealed and real mask images. We did not remove the backgrounds from the images, therefore the same background in the reference and test images may have provided a cue that the images showed the same person. As in our previous research on face matching with masked faces (Noyes et al., [<reflink idref="bib45" id="ref115">45</reflink>]), the unconcealed reference image chosen for each model was front-facing and showed a neutral expression (where possible). A different identity 'foil' image was selected from the same model database for each identity to serve as the reference image in mismatch trials. The foil identities were chosen to match the same verbal description as the target identity e.g. "young woman, dark hair". A subset of 20 of the identities was used in a recent study of forensic facial examiners (Noyes et al., [<reflink idref="bib44" id="ref116">44</reflink>]).</p> <p>Graph: Fig. 3Examples of the a Reference b Unconcealed c Superimposed Mask and d Real Mask stimuli used in Experiment 2. The images depict an identity who was not included in the experiment, but has given permission for their images to be used</p> <p>Superimposed masks were added to the unconcealed images by open source software (Anwar & Raychowdhury, [<reflink idref="bib1" id="ref117">1</reflink>]https://github.com/aqeelanwar/MaskTheFace) that uses standard face landmarking code to locate the relevant part of the face and superimpose a mask image. A variety of mask types are available; we used the standard surgical mask, as illustrated in Fig. 3c. This mask is most like the NIST FRVT Face Mask Effects 'wide, medium coverage' mask which is particularly important for Experiment 3 which uses these same stimuli.</p> <hd id="AN0175199277-21">Procedure</hd> <p>The stimuli were presented side by side in pairs. In all trials, the image on the left was the reference image for match trials, and the foil image for mismatch trials. The image on the right was either the unconcealed, superimposed mask, or real mask image. The assignment of identities to conditions was counterbalanced between participants, and each participant saw each identity only once. Participants saw ten trials in each concealment condition (unconcealed, superimposed mask, real mask) for each trial type (match, mismatch), making a total of 60 trials. On each trial, participants were asked to indicate whether the two images showed the same person or two different people.</p> <hd id="AN0175199277-22">Results and discussion</hd> <p>Again, we present descriptive statistics here for ease of reading—full analysis of accuracy as defined by per cent correct can be found in the Additional file 1. In Experiment 2, as a group, the face matching scores (out of 60) for super-recognisers (range = 44–60, <emph>M</emph> = 54, SD = 3) were higher than controls (range = 37–57, <emph>M</emph> = 48, SD = 4). Accuracy across both groups of participants in each condition was as follows: unconcealed, (range = 55–100%, <emph>M</emph> = 89%, SD = 10%); superimposed mask, (range = 50–100%, <emph>M</emph> = 83%, SD = 10%); and real mask, (range = 45–100%, <emph>M</emph> = 81%, SD = 11%).</p> <hd id="AN0175199277-23">Sensitivity</hd> <p>As in Experiment 1 our main analysis uses signal detection theory. Again, we corrected for hits of 1 and false alarms of 0, giving a maximum d′ value of 3.29. A mixed ANOVA with the within subjects factor of mask condition (unconcealed, superimposed mask, real mask) and the between subjects factor of participant group (control, super-recogniser) revealed a significant effect of mask condition on d′, <emph>F</emph>(<reflink idref="bib3" id="ref118">3</reflink>, 618) = 66.06, <emph>p</emph> < 0.001, <emph>η</emph><subs><emph>p</emph></subs><sups>2</sups> = 0.18, BF<subs>10</subs> > 1000, see Fig. 4. Bonferroni corrected post-hoc comparisons showed that sensitivity was significantly higher in the unconcealed condition (<emph>M</emph> = 2.45, SD = 0.70) compared to both the superimposed mask condition (<emph>M</emph> = 2.05, SD = 0.71, <emph>t</emph>(<reflink idref="bib310" id="ref119">310</reflink>) = 8.77, <emph>p</emph> < 0.001, BF<subs>10</subs> > 1000), and the real mask condition (<emph>M</emph> = 1.90, SD = 0.78, <emph>t</emph>(<reflink idref="bib310" id="ref120">310</reflink>) = 10.68, <emph>p</emph> < 0.001, BF<subs>10</subs> > 1000). The comparison between superimposed and real masks was also significant, whereby sensitivity was higher with superimposed compared to real masks, <emph>t</emph>(<reflink idref="bib310" id="ref121">310</reflink>) = 3.08, <emph>p</emph> = 0.006, BF<subs>10</subs> = 6.51. There was a significant main effect of participant group whereby the super-recognisers as a group performed more accurately (<emph>M</emph> = 2.53) than the control participants (<emph>M</emph> = 1.83), <emph>F</emph>(<reflink idref="bib3" id="ref122">3</reflink>, 309) = 224.36 <emph>p</emph> < 0.001, <emph>η</emph><subs><emph>p</emph></subs><sups>2</sups> = 0.42, BF<subs>10</subs> > 1000. The interaction was non-significant <emph>F</emph>(<reflink idref="bib3" id="ref123">3</reflink>, 618) = 1.98, <emph>p</emph> = 0.139, <emph>η</emph><subs><emph>p</emph></subs><sups>2</sups> < 0.01, BF<subs>10</subs> = 0.59).</p> <p>Graph: Fig. 4Sensitivity (d ′) and criterion scores for Experiment 2</p> <p>The pattern of results is the same when the results are analysed using per cent correct, for overall accuracy (collapsing across match and mismatch trials), match, and mismatch trials (see Additional file 1: Sect. 2).</p> <hd id="AN0175199277-24">Criterion</hd> <p>A mixed ANOVA with the within subjects factor of mask condition (unconcealed, superimposed mask, real mask) and the between subjects factor of participant group (control, super-recogniser) showed a non-significant main effect of mask condition on response bias, <emph>F</emph>(<reflink idref="bib3" id="ref124">3</reflink>, 618) = 2.12, <emph>p</emph> = 0.225, <emph>η</emph><subs><emph>p</emph></subs><sups>2</sups> < 0.01, BF<subs>10</subs> = 0.04. There was a non-significant main effect of participant group, <emph>F</emph>(<reflink idref="bib3" id="ref125">3</reflink>, 309) = 1.96, <emph>p</emph> = 0.163, <emph>η</emph><subs><emph>p</emph></subs><sups>2</sups> < 0.01, BF<subs>10</subs> = 0.23 and a non-significant interaction <emph>F</emph>(<reflink idref="bib3" id="ref126">3</reflink>, 618) = 0.01, <emph>p</emph> = 0.994, <emph>η</emph><subs><emph>p</emph></subs><sups>2</sups> < 0.01, BF<subs>10</subs> < 0.01.</p> <p>In this experiment we found that human observers performed most accurately with unconcealed faces, then with superimposed masks, and were least accurate with real face masks. These results demonstrate the importance of the face covering used when testing face matching ability. Both superimposed and real face masks impaired performance, possibly because they attract attention to the mask, or because they disrupt holistic processing. It is unclear why real face masks impaired performance more than superimposed masks, but this could be a result of spurious texture information being introduced by the mask, disrupting face matching ability to a greater extent than superimposed masks. Experiment 3 tested the effects of both types of face masks on algorithm performance.</p> <hd id="AN0175199277-25">Experiment 3: algorithm performance</hd> <p>In this experiment, we tested four face recognition algorithms with face images covered by both superimposed and real face masks. We wished to repeat the pairings (reference image compared to unconcealed, real mask, and superimposed mask) shown to the human participants, for a direct comparison. However, as computers do not grow tired with time on task, we are able to test other pairings. In particular, we were interested in further exploring cases involving an unmasked reference image and a test image wearing a real mask. Carragher and Hancock ([<reflink idref="bib12" id="ref127">12</reflink>]) reported that although Deep Convolutional Neural Networks (DCNNs) were able to accurately match faces with superimposed masks, their performance for pairs in which one face was unobstructed and the other was wearing a mask was far below that for pairs where both faces were masked. Therefore, might it help the algorithms to superimpose a fake mask on the reference image?</p> <p>All four algorithms that we tested are DCNNs that make image computations in a broadly similar way. An input image of a face (here the images from our matching task) is processed to generate a vector of 512 real-value numbers that make up the system's representation of that face image, sometimes termed an embedding. To decide whether two faces show the same identity, the two vectors are compared. If they are similar enough, the faces are declared a match. There is a variety of ways to compute the similarity of the two vectors: all the algorithms here use the cosine of the angle between the vectors. This gives a value of 1 for a perfect match, when the angle is zero, and zero when the vectors are orthogonal (90 degrees apart). The score can go negative, if the angle between the vectors is greater than 90 degrees.</p> <p>The threshold for deciding that two faces match is a critical aspect of the system. A high threshold reduces the likelihood of declaring an incorrect match (i.e. a false positive). However, it also increases the likelihood of incorrectly rejecting a true match (i.e. a miss). An ideal algorithm would give complete separation between the similarity scores of match and mismatch pairs, with a threshold being set in the gap between the two distributions. In practice, when a DCNN is used with a large database there will likely be some overlap of these distributions, so a threshold is typically set to give an acceptable false positive rate, for example 1 in 10,000 comparisons. What is deemed acceptable will depend on the application and desired level of security. Here, we use the default recommended thresholds for each algorithm (as provided by the developers). Note that none of these algorithms had been designed specifically for use with masked faces. A mask only on one face seems likely to decrease the similarity score for a given pair (Carragher & Hancock, [<reflink idref="bib12" id="ref128">12</reflink>]), so there may be a case for using a lower threshold to declare a match in this circumstance, but we do not explore that possibility here.</p> <hd id="AN0175199277-26">Method</hd> <p></p> <hd id="AN0175199277-27">Stimuli</hd> <p>The stimuli were those used in Experiment 2. We also added an additional condition, in which we superimposed a fake mask onto the reference image and paired those with the real mask images. We therefore had four conditions: Unconcealed, Superimposed, Real, and Masked-Reference.</p> <hd id="AN0175199277-28">Software</hd> <p>We tested four different automatic face recognition (AFR) algorithms, all based on deep convolutional networks. Two are (as yet) unpublished research algorithms, made available to us through the FACER2VM project ('Face Matching for Automatic Identity Retrieval, Recognition, Verification and Management' EPSRC grant no. EP/N007743/1); one from Imperial College London (ICL), the other from the University of Surrey (SU). The other two are FaceNet (Schroff et al., [<reflink idref="bib57" id="ref129">57</reflink>]) and ARCFace (Deng et al., [<reflink idref="bib19" id="ref130">19</reflink>]), as implemented in Deepface (https://github.com/serengil/deepface). These final algorithms were state of the art in their day and are included as an indication of how AFR performance is improving.</p> <hd id="AN0175199277-29">Procedure</hd> <p>Each image was submitted to each AFR separately and the resultant vector stored. The four similarity scores (Reference–Unconcealed; Reference–Superimposed; Reference–Real; and Masked Reference–Real) were then computed locally in Matlab.</p> <hd id="AN0175199277-30">Results</hd> <p>D-prime and Criterion scores are shown in Fig. 5. The two research algorithms achieve 100% accuracy in some conditions. This requires an adjustment to d′ that assumes half an error across the 60 trials, resulting in a maximum d′ of 4.79 (Stanislaw & Todorov, [<reflink idref="bib59" id="ref131">59</reflink>]).</p> <p>Graph: Fig. 5Sensitivity (d ′) and criterion scores for the four AFR algorithms in the four test conditions. There are no error bars as the algorithms are deterministic</p> <hd id="AN0175199277-31">Sensitivity</hd> <p>There is a consistent order evident across the four algorithms, with the ICL system better than SU, which is better than the two older algorithms (and ARCFace is somewhat better than FaceNet). Note that inferential statistics cannot be conducted: the algorithms are deterministic so there is no variability to test. Adding a mask to the reference image improved sensitivity for the two research algorithms but not the older algorithms. Importantly for our research question, sensitivity for all four of the algorithms was lower for real face masks compared to superimposed masks.</p> <hd id="AN0175199277-32">Criterion</hd> <p>The most obvious effect among the criterion values shown in Fig. 5 is that in the two masked conditions, the criterion for all the algorithms is increasingly conservative, meaning a shift towards reporting mismatch. Perhaps this result is not surprising, as none of these algorithms were designed to work with masks. With a mask across one face, the two faces appear more different to the AFR algorithms. Conversely, when a mask is added to the reference image, the criterion drops for all algorithms, going strongly negative for the two older algorithms. They see the mask on each face and interpret it as greater similarity between the two, resulting in a shift towards reporting a match, with little change in sensitivity. In the unconcealed condition, both research algorithms performed perfectly, resulting in zero bias. The two older algorithms have a negative criterion, indicating a bias towards reporting a match.</p> <hd id="AN0175199277-33">Analysis of mask size</hd> <p>In Experiments 2 and 3 we have shown that both humans and algorithms are poorer at face matching with real masks compared to superimposed masks. We sought to determine whether, in our stimulus set, real masks covered a greater area of the face than the superimposed masks. We used sketchandcalc.com to determine the area of the face covered by the masks. A paired samples <emph>t</emph>-test showed that real masks (mean percentage of face covered = 48.17%) did cover more of the face than our superimposed masks (mean percentage of face covered = 39.38%) <emph>t</emph>(<reflink idref="bib59" id="ref132">59</reflink>) = 13.12, <emph>p</emph> < 0.001, Cohen's d = 1.69, BF<subs>10</subs> > 1000. To determine whether mask size explains performance, we correlated mask size with item accuracy (per cent of participants responding correctly to each item). For this analysis we used only control participant data as we did not find group differences in the main task. Mask area was not correlated with item accuracy <emph>r</emph>(<reflink idref="bib118" id="ref133">118</reflink>) = 0.02, <emph>p</emph> = 0.803, BF<subs>10</subs> = 0.12. In addition we correlated change in mask size (real mask minus superimposed mask percentage of face covered) with change in accuracy per item (superimposed mask minus real mask accuracy) and found a non-significant correlation <emph>r</emph>(<reflink idref="bib60" id="ref134">60</reflink>) = − 0.01, <emph>p</emph> = 0.951, BF<subs>10</subs> = 0.16. Mask size therefore does not explain accuracy on our task. Below we discuss possible explanations for our effects.</p> <hd id="AN0175199277-34">Discussion</hd> <p>In three experiments, we have shown that face masks impair face matching performance for both typical human observers and super-recognisers, as well as four AFR algorithms, replicating previous work (Boutros et al., [<reflink idref="bib5" id="ref135">5</reflink>]; Carragher & Hancock, [<reflink idref="bib12" id="ref136">12</reflink>]; Dhamecha et al., [<reflink idref="bib20" id="ref137">20</reflink>]; Estudillo et al., [<reflink idref="bib21" id="ref138">21</reflink>]; Noyes et al., [<reflink idref="bib45" id="ref139">45</reflink>]). It is worth noting that we do not suggest that human observers and algorithms are equivalent or are performing the task in the same way. It is possible that humans approach this task in a way akin to a deep neural network, but this is a topic which requires further research. Importantly, irrespective of the mechanisms driving performance, Experiments 2 and 3 showed that both humans and algorithms are poorer at matching faces when one image in the pair wears a real face mask compared to a superimposed face mask. This highlights the importance of the type of face coverings used when testing both humans and computer algorithms. Our data suggest that the current tendency to rely on superimposed face coverings in research could be underestimating the degree of impairment real face masks cause in real-world settings.</p> <p>In Experiment 1, sensitivity was highest in the control condition and fell significantly for the three mask conditions—which did not differ from each other. This finding suggests that the shape of the superimposed face covering does not influence the degree of impairment to matching performance. In Experiments 2 and 3, both humans and algorithms were more impaired in the real face mask condition than the superimposed mask condition. The explanation as to why face coverings impair face matching performance, and why real masks impair performance more than superimposed masks remains unclear. Real face masks are not standardised, and in this study each model identity wore their own face mask (we did not provide a standard mask). Superimposed masks, in contrast, are applied in a uniform way across faces. Real face masks therefore add more variability in a number of dimensions than superimposed masks. Each real face mask is fitted differently to each face, whereas the technique used here and elsewhere (Ngan et al., [<reflink idref="bib43" id="ref140">43</reflink>]) to fit superimposed masks to faces ensures a tight fit. In Experiment 1, a loose-fitting mask and even the complete removal of the bottom half of the image did not result in additional impairment beyond the fitted mask, and in Experiment 2 although we found that our real masks covered more of the face than the superimposed masks, mask size was not correlated with item accuracy. Therefore mask size alone does not explain our findings in Experiment 2 where real masks resulted in a larger impairment than superimposed masks. The variability of the fit of real masks is not captured with superimposed masks, which may introduce more noise, resulting in a greater impairment for real masks. Importantly, real masks introduce extra variability in terms of texture information to the face which may disrupt processing. It is also possible that in wearing a real face mask, other aspects of the face are slightly changed such as the ears are pulled forward, which may also produce greater variability in the images, resulting in the impairment in face matching which we see here. We would suggest that future research may wish to standardise real masks, for example by having every model wear a surgical mask of the same type. This would not overcome the issue of standard masks covering more of one person's face than another, or more of the face than a superimposed mask, but would remove the variability in mask texture. These issues all highlight that real masks fit the face differently to superimposed masks, and emphasise the importance of using real face masks rather than superimposed masks for research and in applied settings.</p> <p>In this paper, we sought to explore the different effects of real and superimposed masks on face matching performance. However, it is important to understand why either type of obstruction affects face matching performance. It is possible that both types of masks cover features that are useful for identification, interfere with holistic processing, and attract attention. The evidence for face masks attracting attention is mixed. One study found evidence from EEG that more attentional mechanisms are involved when viewing faces wearing masks compared to unconcealed faces (Żochowska et al., [<reflink idref="bib66" id="ref141">66</reflink>]). Another study, however, showed that gaze cueing is not affected by face masks (Dalmaso et al., [<reflink idref="bib16" id="ref142">16</reflink>]), suggesting that masks do not influence attention. In Experiment 1 here, the same impairment occurred when faces were masked (fitted or loose) as when the bottom half of the image was completely removed. These findings suggest that masks do not impair face matching performance because they attract attention. Instead, our findings suggest that masks impair performance either because they occlude facial features that carry identity information, or because they disrupt holistic processing (as in Freud et al., [<reflink idref="bib25" id="ref143">25</reflink>]; Stajduhar et al., [<reflink idref="bib58" id="ref144">58</reflink>]). We cannot separate these two possible explanations for our results because covering facial features necessarily also interferes with holistic processing. Further research is needed to disentangle these possibilities.</p> <p>Crucial to our results is the finding that both humans and algorithms were poorer at face matching when the images showed people wearing real masks compared to superimposed masks. Comparing two of our previous studies, we found that one study using real face masks (Noyes et al., [<reflink idref="bib45" id="ref145">45</reflink>]) showed a smaller reduction in face matching accuracy than a study using superimposed face masks (Carragher & Hancock, [<reflink idref="bib12" id="ref146">12</reflink>]). We have not replicated this effect here, suggesting that differences between the results of the previous work may be due to different methodologies—Carragher and Hancock ([<reflink idref="bib12" id="ref147">12</reflink>]) used a between subjects design with different participants in each mask condition, whereas participants in Noyes et al. ([<reflink idref="bib45" id="ref148">45</reflink>]) participants all viewed each mask condition. The differences in results may also be due to variations in the baseline matching difficulty of the different identity sets used. This is evidenced when we look again at the original data. In the study using real masks (Noyes et al., [<reflink idref="bib45" id="ref149">45</reflink>]; Experiment 2), unconcealed unfamiliar face matching <emph>d</emph>-prime by controls = 1.10, dropping to 1.03 when one image wore a mask. The equivalent values for the study using superimposed masks (Carragher & Hancock, [<reflink idref="bib12" id="ref150">12</reflink>]) were <emph>d</emph>-prime = 2.74 for unconcealed faces and a substantially greater drop to 1.80 when one image had a superimposed mask. In the current study, we used the same identities in all conditions, overcoming the issue of different baseline difficulties in the tasks (Carragher & Hancock, [<reflink idref="bib12" id="ref151">12</reflink>]; Noyes et al., [<reflink idref="bib45" id="ref152">45</reflink>]).</p> <p>Both super-recognisers and algorithms, in addition to control participants, were impaired at face matching by face coverings, particularly real face masks. This highlights the fact that face masks pose a problem for the very best humans as well as algorithms the likes of which are employed in security settings to perform face matching tasks. A recent study testing forensic face examiners (people who are employed to make face comparisons and whose evidence can be heard in court) showed that even with masked faces the examiners significantly outperformed controls on a face matching task (Noyes et al., [<reflink idref="bib44" id="ref153">44</reflink>]). Taken together these results demonstrate that there is a clear role for very high performing humans and algorithms in security settings, and although face masks reduce matching accuracy, algorithms and specialist humans outperform controls.</p> <p>Our research further highlights the problem that face masks pose for identification, and also emphasises the importance of considering which types of face coverings are used when testing both humans and computers. Since real-world images will involve images of people wearing real face masks, our data suggest it is important to test humans and algorithms with real instead of superimposed masks, as a failure to do so may underestimate the problem posed by face masks.</p> <hd id="AN0175199277-35">Acknowledgements</hd> <p>Not applicable</p> <hd id="AN0175199277-36">Significance statement</hd> <p>Since the outbreak of COVID-19, mask wearing has been required in various settings. It is important from a theoretical and applied perspective to understand the impact of face masks on people's ability to recognise faces, and this has consequences for security/identity verification purposes. To date most research investigating the impact of face masks on face recognition has not used real images of people wearing masks, but has superimposed a mask image on to a preexisting face image. This is true for research using humans as well as computers, and in fact the world standard test of algorithms uses superimposed face masks (https://pages.nist.gov/frvt/html/frvt%5ffacemask.html). Here we ask whether the literature could be underestimating the problem posed by face masks through its use of superimposed (fake masks) as opposed to images of wearing face masks. In face matching tasks using real and superimposed masks, we show that super-recognisers and control participants, as well as algorithms, perform less accurately with real masks than superimposed masks. Therefore, by superimposing masks on to pre-existing stimuli we may be underestimating the problem they pose for face identification.</p> <hd id="AN0175199277-37">Funding</hd> <p>Not applicable.</p> <hd id="AN0175199277-38">Availability of data and materials</hd> <p>The datasets generated and/or analysed during the current study are available in the OFS repository https://osf.io/qgxhs/?view_only=6c6e8368c49d4d4fb634ada0671a7972</p> <hd id="AN0175199277-39">Declarations</hd> <p></p> <hd id="AN0175199277-40">Ethics approval and consent to participate</hd> <p>Experiment 1 was approved by the General University Ethics Panel at the University of Stirling, and all participants gave informed consent. Experiment 2 received ethical approval from the University of Reading (ref: 2021-093-KG), and all participants gave informed consent.</p> <hd id="AN0175199277-41">Consent for publication</hd> <p>The images in Figs. 1 and 3 depict identities who were not included in the experiments, but have given permission for their images to be used.</p> <hd id="AN0175199277-42">Competing interests</hd> <p>The authors declare that they have no competing interests.</p> <hd id="AN0175199277-43">Supplementary Information</hd> <p>Graph: Additional file 1. Supplementary analyses.</p> <hd id="AN0175199277-44">Publisher's Note</hd> <p>Springer Nature remains neutral with regard to jurisdictional claims in published maps and institutional affiliations.</p> <ref id="AN0175199277-45"> <title> References </title> <blist> <bibl id="bib1" idref="ref117" type="bt">1</bibl> <bibtext> Anwar, A, & Raychowdhury, A. (2020). Masked Face Recognition for Secure Authentication. <ulink href="http://arxiv.org/abs/2008.11104">http://arxiv.org/abs/2008.11104</ulink></bibtext> </blist> <blist> <bibl id="bib2" idref="ref89" type="bt">2</bibl> <bibtext> Belanova E, Davis JP, Thompson T. The Part-Whole Effect in super-recognisers and typical-range ability controls. Vision Research. 2021; 187: 75-84. 10.1016/j.visres.2021.06.004. 34225132</bibtext> </blist> <blist> <bibl id="bib3" idref="ref31" type="bt">3</bibl> <bibtext> Bobak AK, Hancock PJ, Bate S. Super-recognisers in action: Evidence from face-matching and face memory tasks. Applied Cognitive Psychology. 2016; 30; 1: 81-91. 10.1002/acp.3170. 30122803</bibtext> </blist> <blist> <bibl id="bib4" idref="ref32" type="bt">4</bibl> <bibtext> Bobak AK, Pampoulov P, Bate S. Detecting superior face recognition skills in a large sample of young British adults. Frontiers in Psychology. 2016; 7: 1378. 10.3389/fpsyg.2016.01378. 27713706. 5031595</bibtext> </blist> <blist> <bibl id="bib5" idref="ref43" type="bt">5</bibl> <bibtext> Boutros, F, Damer, N, Kolf, J. N, Raja, K, Kirchbuchner, F, Ramachandra, R, Kuijper, A, Fang, P, Zhang, C, Wang, F, & Montero, D. (2021). Mfr 2021: Masked face recognition competition. In 2021 IEEE International joint conference on biometrics (IJCB) (pp. 1–10). IEEE.</bibtext> </blist> <blist> <bibl id="bib6" idref="ref1" type="bt">6</bibl> <bibtext> Bruce V. Influences of familiarity on the processing of faces. Perception. 1986; 15; 4: 387-397. 1:STN:280:DyaL2s7ks1ajsQ%3D%3D. 10.1068/p150387. 3822723</bibtext> </blist> <blist> <bibl id="bib7" idref="ref2" type="bt">7</bibl> <bibtext> Bruce V, Henderson Z, Newman C, Burton AM. Matching identities of familiar and unfamiliar faces caught on CCTV images. Journal of Experimental Psychology: Applied. 2001; 7; 3: 207-218. 1:STN:280:DC%2BD3MrmvFKrtg%3D%3D. 11676099</bibtext> </blist> <blist> <bibl id="bib8" idref="ref59" type="bt">8</bibl> <bibtext> Burton AM, Jenkins R, Hancock PJ, White D. Robust representations for face recognition: The power of averages. Cognitive Psychology. 2005; 51; 3: 256-284. 10.1016/j.cogpsych.2005.06.003. 16198327</bibtext> </blist> <blist> <bibl id="bib9" idref="ref93" type="bt">9</bibl> <bibtext> Burton AM, White D, McNeill A. The glasgow face matching test. Behavior Research Methods. 2010; 42; 1: 286-291. 10.3758/BRM.42.1.286. 20160307</bibtext> </blist> <blist> <bibtext> Burton AM, Wilson S, Cowan M, Bruce V. Face recognition in poor-quality video: Evidence from security surveillance. Psychological Science. 1999; 10; 3: 243-248. 10.1111/1467-9280.00144</bibtext> </blist> <blist> <bibtext> Cao, Q, Shen, L, Xie, W, Parkhi, O. M, & Zisserman, A. (2018). Vggface2: A dataset for recognising faces across pose and age. In 2018 13th IEEE international conference on automatic face & gesture recognition (FG 2018) (pp. 67–74). IEEE.</bibtext> </blist> <blist> <bibtext> Carragher DJ, Hancock PJ. Surgical face masks impair human face matching performance for familiar and unfamiliar faces. Cognitive Research: Principles and Implications. 2020; 5; 1: 1-15</bibtext> </blist> <blist> <bibtext> Carragher DJ, Towler A, Mileva VR, White D, Hancock PJ. Masked face identification is improved by diagnostic feature training. Cognitive Research: Principles and Implications. 2022; 7; 1: 1-12</bibtext> </blist> <blist> <bibtext> Clutterbuck R, Johnston RA. Exploring levels of face familiarity by using an indirect face-matching measure. Perception. 2002; 31: 985-994. 10.1068/p3335. 12269591</bibtext> </blist> <blist> <bibtext> Clutterbuck R, Johnston RA. Matching as an index of face familiarity. Visual Cognition. 2004; 11; 7: 857-869. 10.1080/13506280444000021</bibtext> </blist> <blist> <bibtext> Dalmaso M, Zhang X, Galfano G, Castelli L. Face masks do not alter gaze cueing of attention: Evidence from the COVID-19 pandemic. i-Perception. 2021; 12; 6: 20416695211058480. 10.1177/20416695211058480. 34925752. 8673884</bibtext> </blist> <blist> <bibtext> Davis J. P, Bretfelean D, Belanova E, & Thompson T. (2019). Assessing the long-term face memory of highly superior and typical-ability short-term face recognisers. https://doi.org/10.31234/osf.io/var4m</bibtext> </blist> <blist> <bibtext> Davis JP, Valentine T. CCTV on trial: Matching video images with the defendant in the dock. Applied Cognitive Psychology. 2009; 23; 4: 482-505. 10.1002/acp.1490</bibtext> </blist> <blist> <bibtext> Deng, J, Guo, J, Xue, N, & Zafeiriou, S. (2019). ArcFace: Additive angular margin loss for deep face recognition. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (pp. 4690–4699).</bibtext> </blist> <blist> <bibtext> Dhamecha TI, Singh R, Vatsa M, Kumar A. Recognizing disguised faces: Human and machine evaluation. PLoS ONE. 2014; 9; 7. 10.1371/journal.pone.0099212. 25029188. 4100743</bibtext> </blist> <blist> <bibtext> Estudillo AJ, Hills P, Wong HK. The effect of face masks on forensic face matching: An individual differences study. Journal of Applied Research in Memory and Cognition. 2021; 10; 4: 554-563. 10.1037/h0101864</bibtext> </blist> <blist> <bibtext> Fisher G, Cox R. Recognizing human faces. Applied Ergonomics. 1975; 6; 2: 104-109. 1:STN:280:DC%2BD2M%2FkvFejsg%3D%3D. 10.1016/0003-6870(75)90303-8. 15677175</bibtext> </blist> <blist> <bibtext> Fitousi D, Rotschild N, Pnini C, Azizi O. Understanding the impact of face masks on the processing of facial identity, emotion, age, and gender. Frontiers in Psychology. 2021; 12: 4668. 10.3389/fpsyg.2021.743793</bibtext> </blist> <blist> <bibtext> Freud, E, Stajduhar, A, Rosenbaum, R. S, Avidan, G, & Ganel, T. (2021). Recognition of masked faces in the era of the pandemic: No improvement, despite extensive, natural exposure. Preprint PsyArXiv https://psyarxiv.com/x3gzq/</bibtext> </blist> <blist> <bibtext> Freud E, Stajduhar A, Rosenbaum RS, Avidan G, Ganel T. The COVID-19 pandemic masks the way people perceive faces. Scientific Reports. 2020; 10; 1: 1-8. 10.1038/s41598-020-78986-9</bibtext> </blist> <blist> <bibtext> Fysh MC. Individual differences in the detection, matching and memory of faces. Cognitive Research: Principles and Implications. 2018; 3; 1: 1-12</bibtext> </blist> <blist> <bibtext> Fysh MC, Bindemann M. The Kent face matching test. British Journal of Psychology. 2018; 109; 2: 219-231. 10.1111/bjop.12260. 28872661</bibtext> </blist> <blist> <bibtext> Gentry NW, Bindemann M. Examples improve facial identity comparison. Journal of Applied Research in Memory and Cognition. 2019; 8; 3: 376-385. 10.1016/j.jarmac.2019.06.002</bibtext> </blist> <blist> <bibtext> Graham DL, Ritchie KL. Making a spectacle of yourself: The effect of glasses and sunglasses on face perception. Perception. 2019; 48; 6: 461-470. 10.1177/0301006619844680. 31006340</bibtext> </blist> <blist> <bibtext> JASP Team. (2020). JASP (Version 0.14.0)[Computer software].</bibtext> </blist> <blist> <bibtext> Jeevan G, Zacharias GC, Nair MS, Rajan J. An empirical study of the impact of masks on face recognition. Pattern Recognition. 2022; 122. 10.1016/j.patcog.2021.108308</bibtext> </blist> <blist> <bibtext> Jeffreys H. Theory of probability. 19613; Oxford University Press</bibtext> </blist> <blist> <bibtext> Kemelmacher-Shlizerman, I, Seitz, S. M, Miller, D, & Brossard, E. (2016). The megaface benchmark: 1 million faces for recognition at scale. In Proceedings of the IEEE conference on computer vision and pattern recognition (pp. 4873–4882).</bibtext> </blist> <blist> <bibtext> Kemp R, Towell N, Pike G. When seeing should not be believing: Photographs, credit cards and fraud. Applied Cognitive Psychology. 1997; 11; 3: 211-222. 10.1002/(SICI)1099-0720(199706)11:3<211:AID-ACP430>3.0.CO;2-O</bibtext> </blist> <blist> <bibtext> Kramer RSS, Ritchie KL. Disguising superman: How glasses affect unfamiliar face matching. Applied Cognitive Psychology. 2016; 30: 841-845. 10.1002/acp.3261</bibtext> </blist> <blist> <bibtext> Lionnie, R, Apriono, C, & Gunawan, D. (2021, April). Face mask recognition with realistic fabric face mask data set: A combination using surface curvature and glcm. In 2021 IEEE international IOT, electronics and mechatronics conference (IEMTRONICS) (pp. 1–6). IEEE.</bibtext> </blist> <blist> <bibtext> Macmillan NA, Creelman CD. Detection theory: A user's guide. 2004; Psychology Press. 10.4324/9781410611147</bibtext> </blist> <blist> <bibtext> Maurer D, Le Grand R, Mondloch CJ. The many faces of configural processing. Trends in Cognitive Sciences. 2002; 6; 6: 255-260. 10.1016/S1364-6613(02)01903-4. 12039607</bibtext> </blist> <blist> <bibtext> McKelvie SJ. The role of eyes and mouth in the memory of a face. The American Journal of Psychology. 1976; 89; 2: 311-323. 10.2307/1421414</bibtext> </blist> <blist> <bibtext> Megreya AM, Burton AM. Matching faces to photographs: Poor performance in eyewitness memory (without the memory). Journal of Experimental Psychology: Applied. 2008; 14; 4: 364-372. 19102619</bibtext> </blist> <blist> <bibtext> National Institute of Standards and Technology (NIST). FRVT Face Mask Effects. (2022b). Available from: https://pages.nist.gov/frvt/html/frvt_facemask.html</bibtext> </blist> <blist> <bibtext> National Institute of Standards and Technology (NIST). FRVT 1:N Identification. (2022a). Available from: https://pages.nist.gov/frvt/html/frvt1N.html</bibtext> </blist> <blist> <bibtext> Ngan, M, Grother, P, & Hanaoka, K. (2022). Ongoing face recognition vender test (FRVT) Part 6B: Face recognition accuracy with face masks using post-VOVID-19 algorithms. National Institute of Standards and Technology (NIST). Available from https://pages.nist.gov/frvt/reports/facemask/frvt_facemask_report.pdf</bibtext> </blist> <blist> <bibtext> Noyes, E, Moreton, R, Hancock, P. J. B, Ritchie, K. L, Castro Martinez, S, Gray, K. L, & Davis, J. P. (2024). A forensic facial examiner and professional team advantage for masked face identification. https://doi.org/10.31234/osf.io/3s47m</bibtext> </blist> <blist> <bibtext> Noyes E, Davis JP, Petrov N, Gray KL, Ritchie KL. The effect of face masks and sunglasses on identity and expression recognition with super-recognizers and typical observers. Royal Society Open Science. 2021; 8; 3. 10.1098/rsos.201169. 33959312. 8074904</bibtext> </blist> <blist> <bibtext> Noyes E, Hill MQ, O'Toole AJ. Face recognition ability does not predict person identification performance: Using individual data in the interpretation of group results. Cognitive Research: Principles and Implications. 2018; 3; 1: 1-13</bibtext> </blist> <blist> <bibtext> Noyes E, Phillips PJ, O'Toole AJBindemann M, Megreya AM. What is a super-recogniser?. Face processing: Systems, disorders, and cultural differences. 2017; Nova: 173-202</bibtext> </blist> <blist> <bibtext> Phillips PJ, Yates AN, Hu Y, Hahn CA, Noyes E, Jackson K, Cavazos JG, Jeckeln G, Ranjan R, Sankaranarayanan S, Chen JC. Face recognition accuracy of forensic examiners, superrecognizers, and face recognition algorithms. Proceedings of the National Academy of Sciences. 2018; 115; 24: 6171-6176. 1:CAS:528:DC%2BC1cXitlWqsrzK. 10.1073/pnas.1721355115</bibtext> </blist> <blist> <bibtext> Ritchie KL, Flack TR, Maréchal L. Unfamiliar faces might as well be another species: Evidence from a face matching task with human and monkey faces. Visual Cognition. 2023; 30: 1-6</bibtext> </blist> <blist> <bibtext> Ritchie KL, Kramer RSS, Mileva M, Sandford A, Burton AM. Multiple-image arrays in face matching tasks with and without memory. Cognition. 2021; 211. 10.1016/j.cognition.2021.104632. 33621739</bibtext> </blist> <blist> <bibtext> Ritchie KL, Mireku MO, Kramer RSS. Face averages and multiple images in a live matching task. British Journal of Psychology. 2020; 111; 1: 92-102. 10.1111/bjop.12388. 30945267</bibtext> </blist> <blist> <bibtext> Ritchie KL, Smith FG, Jenkins R, Bindemann M, White D, Burton AM. Viewers base estimates of face matching accuracy on their own familiarity: Explaining the photo-ID paradox. Cognition. 2015; 141: 161-169. 10.1016/j.cognition.2015.05.002. 25988915</bibtext> </blist> <blist> <bibtext> Rogers D, Baseler H, Young AW, Jenkins R, Andrews TJ. The roles of shape and texture in the recognition of familiar faces. Vision Research. 2022; 194. 10.1016/j.visres.2022.108013. 35124521</bibtext> </blist> <blist> <bibtext> Russell R, Duchaine B, Nakayama K. Super-recognizers: People with extraordinary face recognition ability. Psychonomic Bulletin & Review. 2009; 16; 2: 252-257. 10.3758/PBR.16.2.252</bibtext> </blist> <blist> <bibtext> Sandford A, Ritchie KL. Unfamiliar face matching, within-person variability, and multiple-image arrays. Visual Cognition. 2021; 29; 3: 143-157. 10.1080/13506285.2021.1883170</bibtext> </blist> <blist> <bibtext> Satchell LP, Davis JP, Julle-Danière E, Tupper N, Marshman P. Recognising faces but not traits: Accurate personality judgment from faces is unrelated to superior face memory. Journal of Research in Personality. 2019; 79: 49-58. 10.1016/j.jrp.2019.02.002</bibtext> </blist> <blist> <bibtext> Schroff, F, Kalenichenko, D, & Philbin, J. (2015). FaceNet: A unified embedding for face recognition and clustering. In Proceedings of the IEEE Conference on computer vision and pattern recognition (CVPR) (pp. 815–823).</bibtext> </blist> <blist> <bibtext> Stajduhar A, Ganel T, Avidan G, Rosenbaum RS, Freud E. Face masks disrupt holistic processing and face perception in school-age children. PsyArXiv. 2021. 10.31234/osf.io/fygjq</bibtext> </blist> <blist> <bibtext> Stanislaw H, Todorov N. Calculation of signal detection theory measures. Behavior Research Methods, Instruments, & Computers. 1999; 31; 1: 137-149. 1:STN:280:DyaK1MvisFymsA%3D%3D. 10.3758/bf03207704</bibtext> </blist> <blist> <bibtext> Taigman, Y, Yang, M, Ranzato, M. A, & Wolf, L. (2014). Deepface: Closing the gap to human-level performance in face verification. In Proceedings of the IEEE conference on computer vision and pattern recognition (pp. 1701–1708).</bibtext> </blist> <blist> <bibtext> Tanaka JW, Farah MJ. Parts and wholes in face recognition. The Quarterly Journal of Experimental Psychology Section A. 1993; 46; 2: 225-245. 1:STN:280:DyaK3szgtlarug%3D%3D. 10.1080/14640749308401045</bibtext> </blist> <blist> <bibtext> Towler A, Kemp RI, White DBindemann M. Can Face identification ability be trained?: Evidence for two routes to expertise. Forensic face matching: research and practice. 2021; Oxford University Press: 89-114. 10.1093/oso/9780198837749.003.0005</bibtext> </blist> <blist> <bibtext> Towler A, White D, Kemp RI. Evaluating the feature comparison strategy for forensic face identification. Journal of Experimental Psychology: Applied. 2017; 23; 1: 47. 10.1037/xap0000108. 28045276</bibtext> </blist> <blist> <bibtext> White D, Kemp RI, Jenkins R, Matheson M, Burton AM. Passport officers' errors in face matching. PLoS ONE. 2014; 9. 1:CAS:528:DC%2BC2cXhs1equrbJ. 10.1371/journal.pone.0103510. 25133682. 4136722</bibtext> </blist> <blist> <bibtext> White D, Phillips PJ, Hahn CA, Hill M, O'Toole AJ. Perceptual expertise in forensic facial image comparison. Proceedings of the Royal Society b: Biological Sciences. 2015; 282; 1814: 20151292. 10.1098/rspb.2015.1292. 4571699</bibtext> </blist> <blist> <bibtext> Żochowska A, Jakuszyk P, Nowicka MM, Nowicka A. Are covered faces eye-catching for us? The impact of masks on attentional processing of self and other faces during the COVID-19 pandemic. Cortex. 2022; 149: 173-187. 10.1016/j.cortex.2022.01.015. 35257944. 8830153</bibtext> </blist> </ref> <aug> <p>By Kay L. Ritchie; Daniel J. Carragher; Josh P. Davis; Katie Read; Ryan E. Jenkins; Eilidh Noyes; Katie L. H. Gray and Peter J. B. Hancock</p> <p>Reported by Author; Author; Author; Author; Author; Author; Author; Author</p> </aug> <nolink nlid="nl1" bibid="bib10" firstref="ref3"></nolink> <nolink nlid="nl2" bibid="bib14" firstref="ref4"></nolink> <nolink nlid="nl3" bibid="bib15" firstref="ref5"></nolink> <nolink nlid="nl4" bibid="bib40" firstref="ref6"></nolink> <nolink nlid="nl5" bibid="bib52" firstref="ref7"></nolink> <nolink nlid="nl6" bibid="bib50" firstref="ref8"></nolink> <nolink nlid="nl7" bibid="bib49" firstref="ref9"></nolink> <nolink nlid="nl8" bibid="bib55" firstref="ref10"></nolink> <nolink nlid="nl9" bibid="bib18" firstref="ref11"></nolink> <nolink nlid="nl10" bibid="bib34" firstref="ref12"></nolink> <nolink nlid="nl11" bibid="bib51" firstref="ref14"></nolink> <nolink nlid="nl12" bibid="bib64" firstref="ref16"></nolink> <nolink nlid="nl13" bibid="bib29" firstref="ref18"></nolink> <nolink nlid="nl14" bibid="bib35" firstref="ref19"></nolink> <nolink nlid="nl15" bibid="bib45" firstref="ref20"></nolink> <nolink nlid="nl16" bibid="bib23" firstref="ref21"></nolink> <nolink nlid="nl17" bibid="bib25" firstref="ref22"></nolink> <nolink nlid="nl18" bibid="bib24" firstref="ref23"></nolink> <nolink nlid="nl19" bibid="bib12" firstref="ref24"></nolink> <nolink nlid="nl20" bibid="bib20" firstref="ref25"></nolink> <nolink nlid="nl21" bibid="bib21" firstref="ref26"></nolink> <nolink nlid="nl22" bibid="bib54" firstref="ref29"></nolink> <nolink nlid="nl23" bibid="bib47" firstref="ref30"></nolink> <nolink nlid="nl24" bibid="bib17" firstref="ref35"></nolink> <nolink nlid="nl25" bibid="bib46" firstref="ref36"></nolink> <nolink nlid="nl26" bibid="bib48" firstref="ref37"></nolink> <nolink nlid="nl27" bibid="bib11" firstref="ref39"></nolink> <nolink nlid="nl28" bibid="bib33" firstref="ref40"></nolink> <nolink nlid="nl29" bibid="bib60" firstref="ref41"></nolink> <nolink nlid="nl30" bibid="bib42" firstref="ref44"></nolink> <nolink nlid="nl31" bibid="bib41" firstref="ref45"></nolink> <nolink nlid="nl32" bibid="bib43" firstref="ref46"></nolink> <nolink nlid="nl33" bibid="bib31" firstref="ref54"></nolink> <nolink nlid="nl34" bibid="bib36" firstref="ref55"></nolink> <nolink nlid="nl35" bibid="bib53" firstref="ref60"></nolink> <nolink nlid="nl36" bibid="bib63" firstref="ref63"></nolink> <nolink nlid="nl37" bibid="bib22" firstref="ref64"></nolink> <nolink nlid="nl38" bibid="bib39" firstref="ref65"></nolink> <nolink nlid="nl39" bibid="bib38" firstref="ref66"></nolink> <nolink nlid="nl40" bibid="bib61" firstref="ref67"></nolink> <nolink nlid="nl41" bibid="bib58" firstref="ref69"></nolink> <nolink nlid="nl42" bibid="bib65" firstref="ref71"></nolink> <nolink nlid="nl43" bibid="bib62" firstref="ref72"></nolink> <nolink nlid="nl44" bibid="bib13" firstref="ref73"></nolink> <nolink nlid="nl45" bibid="bib27" firstref="ref78"></nolink> <nolink nlid="nl46" bibid="bib37" firstref="ref80"></nolink> <nolink nlid="nl47" bibid="bib59" firstref="ref82"></nolink> <nolink nlid="nl48" bibid="bib30" firstref="ref83"></nolink> <nolink nlid="nl49" bibid="bib32" firstref="ref84"></nolink> <nolink nlid="nl50" bibid="bib56" firstref="ref91"></nolink> <nolink nlid="nl51" bibid="bib26" firstref="ref103"></nolink> <nolink nlid="nl52" bibid="bib28" firstref="ref105"></nolink> <nolink nlid="nl53" bibid="bib44" firstref="ref116"></nolink> <nolink nlid="nl54" bibid="bib310" firstref="ref119"></nolink> <nolink nlid="nl55" bibid="bib57" firstref="ref129"></nolink> <nolink nlid="nl56" bibid="bib19" firstref="ref130"></nolink> <nolink nlid="nl57" bibid="bib118" firstref="ref133"></nolink> <nolink nlid="nl58" bibid="bib66" firstref="ref141"></nolink> <nolink nlid="nl59" bibid="bib16" firstref="ref142"></nolink>
Header DbId: eric
DbLabel: ERIC
An: EJ1410458
AccessLevel: 3
PubType: Academic Journal
PubTypeId: academicJournal
PreciseRelevancyScore: 0
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: Face Masks and Fake Masks: The Effect of Real and Superimposed Masks on Face Matching with Super-Recognisers, Typical Observers, and Algorithms
– Name: Language
  Label: Language
  Group: Lang
  Data: English
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22Kay+L%2E+Ritchie%22">Kay L. Ritchie</searchLink> (ORCID <externalLink term="http://orcid.org/0000-0002-1348-760X">0000-0002-1348-760X</externalLink>)<br /><searchLink fieldCode="AR" term="%22Daniel+J%2E+Carragher%22">Daniel J. Carragher</searchLink><br /><searchLink fieldCode="AR" term="%22Josh+P%2E+Davis%22">Josh P. Davis</searchLink><br /><searchLink fieldCode="AR" term="%22Katie+Read%22">Katie Read</searchLink><br /><searchLink fieldCode="AR" term="%22Ryan+E%2E+Jenkins%22">Ryan E. Jenkins</searchLink><br /><searchLink fieldCode="AR" term="%22Eilidh+Noyes%22">Eilidh Noyes</searchLink><br /><searchLink fieldCode="AR" term="%22Katie+L%2E+H%2E+Gray%22">Katie L. H. Gray</searchLink><br /><searchLink fieldCode="AR" term="%22Peter+J%2E+B%2E+Hancock%22">Peter J. B. Hancock</searchLink>
– Name: TitleSource
  Label: Source
  Group: Src
  Data: <searchLink fieldCode="SO" term="%22Cognitive+Research%3A+Principles+and+Implications%22"><i>Cognitive Research: Principles and Implications</i></searchLink>. 2024 9.
– Name: Avail
  Label: Availability
  Group: Avail
  Data: Springer. Available from: Springer Nature. One New York Plaza, Suite 4600, New York, NY 10004. Tel: 800-777-4643; Tel: 212-460-1500; Fax: 212-460-1700; e-mail: customerservice@springernature.com; Web site: https://link.springer.com/
– Name: PeerReviewed
  Label: Peer Reviewed
  Group: SrcInfo
  Data: Y
– Name: Pages
  Label: Page Count
  Group: Src
  Data: 13
– Name: DatePubCY
  Label: Publication Date
  Group: Date
  Data: 2024
– Name: TypeDocument
  Label: Document Type
  Group: TypDoc
  Data: Journal Articles<br />Reports - Research
– Name: Subject
  Label: Descriptors
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Artificial+Intelligence%22">Artificial Intelligence</searchLink><br /><searchLink fieldCode="DE" term="%22Recognition+%28Psychology%29%22">Recognition (Psychology)</searchLink><br /><searchLink fieldCode="DE" term="%22Clothing%22">Clothing</searchLink><br /><searchLink fieldCode="DE" term="%22Health+Behavior%22">Health Behavior</searchLink><br /><searchLink fieldCode="DE" term="%22Observation%22">Observation</searchLink><br /><searchLink fieldCode="DE" term="%22Human+Body%22">Human Body</searchLink><br /><searchLink fieldCode="DE" term="%22Visual+Acuity%22">Visual Acuity</searchLink><br /><searchLink fieldCode="DE" term="%22Visual+Stimuli%22">Visual Stimuli</searchLink><br /><searchLink fieldCode="DE" term="%22COVID-19%22">COVID-19</searchLink><br /><searchLink fieldCode="DE" term="%22Pandemics%22">Pandemics</searchLink>
– Name: DOI
  Label: DOI
  Group: ID
  Data: 10.1186/s41235-024-00532-2
– Name: ISSN
  Label: ISSN
  Group: ISSN
  Data: 2365-7464
– Name: Abstract
  Label: Abstract
  Group: Ab
  Data: Mask wearing has been required in various settings since the outbreak of COVID-19, and research has shown that identity judgements are difficult for faces wearing masks. To date, however, the majority of experiments on face identification with masked faces tested humans and computer algorithms using images with superimposed masks rather than images of people wearing real face coverings. In three experiments we test humans (control participants and super-recognisers) and algorithms with images showing different types of face coverings. In all experiments we tested matching concealed or unconcealed faces to an unconcealed reference image, and we found a consistent decrease in face matching accuracy with masked compared to unconcealed faces. In Experiment 1, typical human observers were most accurate at face matching with unconcealed images, and poorer for three different types of superimposed mask conditions. In Experiment 2, we tested both typical observers and super-recognisers with superimposed and real face masks, and found that performance was poorer for real compared to superimposed masks. The same pattern was observed in Experiment 3 with algorithms. Our results highlight the importance of testing both humans and algorithms with real face masks, as using only superimposed masks may underestimate their detrimental effect on face identification.
– Name: AbstractInfo
  Label: Abstractor
  Group: Ab
  Data: As Provided
– Name: Note
  Label: Notes
  Group: Note
  Data: https://osf.io/qgxhs/?view_only=6c6e8368c49d4d4fb634ada0671a7972
– Name: DateEntry
  Label: Entry Date
  Group: Date
  Data: 2024
– Name: AN
  Label: Accession Number
  Group: ID
  Data: EJ1410458
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=eric&AN=EJ1410458
RecordInfo BibRecord:
  BibEntity:
    Identifiers:
      – Type: doi
        Value: 10.1186/s41235-024-00532-2
    Languages:
      – Text: English
    PhysicalDescription:
      Pagination:
        PageCount: 13
    Subjects:
      – SubjectFull: Artificial Intelligence
        Type: general
      – SubjectFull: Recognition (Psychology)
        Type: general
      – SubjectFull: Clothing
        Type: general
      – SubjectFull: Health Behavior
        Type: general
      – SubjectFull: Observation
        Type: general
      – SubjectFull: Human Body
        Type: general
      – SubjectFull: Visual Acuity
        Type: general
      – SubjectFull: Visual Stimuli
        Type: general
      – SubjectFull: COVID-19
        Type: general
      – SubjectFull: Pandemics
        Type: general
    Titles:
      – TitleFull: Face Masks and Fake Masks: The Effect of Real and Superimposed Masks on Face Matching with Super-Recognisers, Typical Observers, and Algorithms
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: Kay L. Ritchie
      – PersonEntity:
          Name:
            NameFull: Daniel J. Carragher
      – PersonEntity:
          Name:
            NameFull: Josh P. Davis
      – PersonEntity:
          Name:
            NameFull: Katie Read
      – PersonEntity:
          Name:
            NameFull: Ryan E. Jenkins
      – PersonEntity:
          Name:
            NameFull: Eilidh Noyes
      – PersonEntity:
          Name:
            NameFull: Katie L. H. Gray
      – PersonEntity:
          Name:
            NameFull: Peter J. B. Hancock
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 01
              M: 01
              Type: published
              Y: 2024
          Identifiers:
            – Type: issn-electronic
              Value: 2365-7464
          Numbering:
            – Type: volume
              Value: 9
          Titles:
            – TitleFull: Cognitive Research: Principles and Implications
              Type: main
ResultId 1