Conducting Medication Reviews: A Comparative Study Between ChatGPT‐4 and Healthcare Professionals.

Saved in:
Bibliographic Details
Title: Conducting Medication Reviews: A Comparative Study Between ChatGPT‐4 and Healthcare Professionals.
Authors: ten Hoope, Simone M. K., Marongiu, Sabrina, Siegert, Carl E. H., Heerdink, Eibert R., Janssen, Marjo J. A., Karapinar‐Çarkit, Fatma
Source: Journal of the American Geriatrics Society. May2026, Vol. 74 Issue 5, p1378-1385. 8p.
Subjects: Generative artificial intelligence, Academic medical centers, Medication error prevention, Patient care, Polypharmacy, Natural language processing, Retrospective studies, Descriptive statistics, Research methodology, Medical records, Acquisition of data, Statistics, Comparative studies, Hospital care of older people, Confidence intervals, Data analysis software
Geographic Terms: Netherlands
Abstract: Background: The increasing prevalence of patients with hyperpolypharmacy (> 10 medications) has made medication reviews increasingly complex. ChatGPT‐4‐Turbo (ChatGPT), a large language model, has demonstrated potential in healthcare applications and could potentially support medication reviews. Objectives: This study aimed to evaluate the agreement between medication reviews conducted by ChatGPT compared to healthcare professionals (HCPs) in older people. Secondary objectives included: the validity of additional interventions detected by ChatGPT, its ability to structure diagnoses to medication use, and laboratory target values based on patient characteristics. Methods: In this retrospective proof‐of‐concept study, 51 medication reviews previously conducted by a geriatric internist and hospital pharmacist were re‐evaluated using ChatGPT. ChatGPT was trained on polypharmacy guidelines, was then provided with the same primary data as HCPs, and was asked to perform medication reviews. Two pharmacists scored the agreement between ChatGPT and HCPs. The structuring of information and additional interventions suggested by ChatGPT were reviewed within an expert team. Descriptive statistics were used. Outcomes: The primary outcome was the percentage agreement between the interventions suggested by ChatGPT compared to HCPs. Secondary outcomes included the proportion of valid and incorrect interventions suggested by ChatGPT and its ability to structure patient information. Results: HCPs suggested 183 interventions and ChatGPT 202 interventions. ChatGPT achieved a 27.7% agreement with interventions of HCPs. It identified 19 additional valid interventions which HCPs missed (7.6%), but also proposed 84 incorrect interventions (33.7%). While ChatGPT demonstrated strong capability in structuring patient data (86.4% correct diagnoses linked to medication), it struggled with contextualizing appropriate laboratory target values based on patient characteristics (46.7%). Conclusion: ChatGPT had low agreement with HCPs, but found additional interventions that HCPs missed. ChatGPT lacks clinical decision‐making capabilities based on individual patient contexts in older people. ChatGPT may, however, serve as a support tool to structure diagnoses and medication lists. Summary: Key points ○ChatGPT‐4‐Turbo shows 27.7% agreement with interventions suggested by healthcare professionals during medication reviews for older people. It fails to take patient context into consideration.○ChatGPT‐4‐Turbo is able to provide 7.6% new valid interventions that healthcare professionals missed.○ChatGPT‐4‐Turbo can structure information and match diagnosis with medications accordingly (86.4%), but lacks the ability to contextualize appropriate laboratory target values based on patient characteristics (46.7%).Why does this paper matter? ○As the integration of artificial intelligence into healthcare continues to gain momentum, this proof‐of‐concept study highlights the potentials and limitations of large language models such as ChatGPT‐4.○While the model shows promise in organizing clinical information and supporting diagnosis‐medication alignment and finds additional valid interventions missed by healthcare professionals; it remains inadequate for direct use in clinical medication reviews due to its inability to account for patient‐specific contextual factors in older people. [ABSTRACT FROM AUTHOR]
Copyright of Journal of the American Geriatrics Society is the property of Wiley-Blackwell and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.)
Database: Psychology and Behavioral Sciences Collection
FullText Text:
  Availability: 0
Header DbId: pbh
DbLabel: Psychology and Behavioral Sciences Collection
An: 194163431
AccessLevel: 6
PubType: Academic Journal
PubTypeId: academicJournal
PreciseRelevancyScore: 0
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: Conducting Medication Reviews: A Comparative Study Between ChatGPT‐4 and Healthcare Professionals.
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22ten+Hoope%2C+Simone+M%2E+K%2E%22">ten Hoope, Simone M. K.</searchLink><br /><searchLink fieldCode="AR" term="%22Marongiu%2C+Sabrina%22">Marongiu, Sabrina</searchLink><br /><searchLink fieldCode="AR" term="%22Siegert%2C+Carl+E%2E+H%2E%22">Siegert, Carl E. H.</searchLink><br /><searchLink fieldCode="AR" term="%22Heerdink%2C+Eibert+R%2E%22">Heerdink, Eibert R.</searchLink><br /><searchLink fieldCode="AR" term="%22Janssen%2C+Marjo+J%2E+A%2E%22">Janssen, Marjo J. A.</searchLink><br /><searchLink fieldCode="AR" term="%22Karapinar‐Çarkit%2C+Fatma%22">Karapinar‐Çarkit, Fatma</searchLink>
– Name: TitleSource
  Label: Source
  Group: Src
  Data: <searchLink fieldCode="JN" term="%22Journal+of+the+American+Geriatrics+Society%22">Journal of the American Geriatrics Society</searchLink>. May2026, Vol. 74 Issue 5, p1378-1385. 8p.
– Name: Subject
  Label: Subjects
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Generative+artificial+intelligence%22">Generative artificial intelligence</searchLink><br /><searchLink fieldCode="DE" term="%22Academic+medical+centers%22">Academic medical centers</searchLink><br /><searchLink fieldCode="DE" term="%22Medication+error+prevention%22">Medication error prevention</searchLink><br /><searchLink fieldCode="DE" term="%22Patient+care%22">Patient care</searchLink><br /><searchLink fieldCode="DE" term="%22Polypharmacy%22">Polypharmacy</searchLink><br /><searchLink fieldCode="DE" term="%22Natural+language+processing%22">Natural language processing</searchLink><br /><searchLink fieldCode="DE" term="%22Retrospective+studies%22">Retrospective studies</searchLink><br /><searchLink fieldCode="DE" term="%22Descriptive+statistics%22">Descriptive statistics</searchLink><br /><searchLink fieldCode="DE" term="%22Research+methodology%22">Research methodology</searchLink><br /><searchLink fieldCode="DE" term="%22Medical+records%22">Medical records</searchLink><br /><searchLink fieldCode="DE" term="%22Acquisition+of+data%22">Acquisition of data</searchLink><br /><searchLink fieldCode="DE" term="%22Statistics%22">Statistics</searchLink><br /><searchLink fieldCode="DE" term="%22Comparative+studies%22">Comparative studies</searchLink><br /><searchLink fieldCode="DE" term="%22Hospital+care+of+older+people%22">Hospital care of older people</searchLink><br /><searchLink fieldCode="DE" term="%22Confidence+intervals%22">Confidence intervals</searchLink><br /><searchLink fieldCode="DE" term="%22Data+analysis+software%22">Data analysis software</searchLink>
– Name: SubjectGeographic
  Label: Geographic Terms
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Netherlands%22">Netherlands</searchLink>
– Name: Abstract
  Label: Abstract
  Group: Ab
  Data: Background: The increasing prevalence of patients with hyperpolypharmacy (> 10 medications) has made medication reviews increasingly complex. ChatGPT‐4‐Turbo (ChatGPT), a large language model, has demonstrated potential in healthcare applications and could potentially support medication reviews. Objectives: This study aimed to evaluate the agreement between medication reviews conducted by ChatGPT compared to healthcare professionals (HCPs) in older people. Secondary objectives included: the validity of additional interventions detected by ChatGPT, its ability to structure diagnoses to medication use, and laboratory target values based on patient characteristics. Methods: In this retrospective proof‐of‐concept study, 51 medication reviews previously conducted by a geriatric internist and hospital pharmacist were re‐evaluated using ChatGPT. ChatGPT was trained on polypharmacy guidelines, was then provided with the same primary data as HCPs, and was asked to perform medication reviews. Two pharmacists scored the agreement between ChatGPT and HCPs. The structuring of information and additional interventions suggested by ChatGPT were reviewed within an expert team. Descriptive statistics were used. Outcomes: The primary outcome was the percentage agreement between the interventions suggested by ChatGPT compared to HCPs. Secondary outcomes included the proportion of valid and incorrect interventions suggested by ChatGPT and its ability to structure patient information. Results: HCPs suggested 183 interventions and ChatGPT 202 interventions. ChatGPT achieved a 27.7% agreement with interventions of HCPs. It identified 19 additional valid interventions which HCPs missed (7.6%), but also proposed 84 incorrect interventions (33.7%). While ChatGPT demonstrated strong capability in structuring patient data (86.4% correct diagnoses linked to medication), it struggled with contextualizing appropriate laboratory target values based on patient characteristics (46.7%). Conclusion: ChatGPT had low agreement with HCPs, but found additional interventions that HCPs missed. ChatGPT lacks clinical decision‐making capabilities based on individual patient contexts in older people. ChatGPT may, however, serve as a support tool to structure diagnoses and medication lists. Summary: Key points ○ChatGPT‐4‐Turbo shows 27.7% agreement with interventions suggested by healthcare professionals during medication reviews for older people. It fails to take patient context into consideration.○ChatGPT‐4‐Turbo is able to provide 7.6% new valid interventions that healthcare professionals missed.○ChatGPT‐4‐Turbo can structure information and match diagnosis with medications accordingly (86.4%), but lacks the ability to contextualize appropriate laboratory target values based on patient characteristics (46.7%).Why does this paper matter? ○As the integration of artificial intelligence into healthcare continues to gain momentum, this proof‐of‐concept study highlights the potentials and limitations of large language models such as ChatGPT‐4.○While the model shows promise in organizing clinical information and supporting diagnosis‐medication alignment and finds additional valid interventions missed by healthcare professionals; it remains inadequate for direct use in clinical medication reviews due to its inability to account for patient‐specific contextual factors in older people. [ABSTRACT FROM AUTHOR]
– Name: AbstractSuppliedCopyright
  Label:
  Group: Ab
  Data: <i>Copyright of Journal of the American Geriatrics Society is the property of Wiley-Blackwell and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract.</i> (Copyright applies to all Abstracts.)
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=pbh&AN=194163431
RecordInfo BibRecord:
  BibEntity:
    Identifiers:
      – Type: doi
        Value: 10.1111/jgs.70415
    Languages:
      – Code: eng
        Text: English
    PhysicalDescription:
      Pagination:
        PageCount: 8
        StartPage: 1378
    Subjects:
      – SubjectFull: Generative artificial intelligence
        Type: general
      – SubjectFull: Academic medical centers
        Type: general
      – SubjectFull: Medication error prevention
        Type: general
      – SubjectFull: Patient care
        Type: general
      – SubjectFull: Polypharmacy
        Type: general
      – SubjectFull: Natural language processing
        Type: general
      – SubjectFull: Retrospective studies
        Type: general
      – SubjectFull: Descriptive statistics
        Type: general
      – SubjectFull: Research methodology
        Type: general
      – SubjectFull: Medical records
        Type: general
      – SubjectFull: Acquisition of data
        Type: general
      – SubjectFull: Statistics
        Type: general
      – SubjectFull: Comparative studies
        Type: general
      – SubjectFull: Hospital care of older people
        Type: general
      – SubjectFull: Confidence intervals
        Type: general
      – SubjectFull: Data analysis software
        Type: general
      – SubjectFull: Netherlands
        Type: general
    Titles:
      – TitleFull: Conducting Medication Reviews: A Comparative Study Between ChatGPT‐4 and Healthcare Professionals.
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: ten Hoope, Simone M. K.
      – PersonEntity:
          Name:
            NameFull: Marongiu, Sabrina
      – PersonEntity:
          Name:
            NameFull: Siegert, Carl E. H.
      – PersonEntity:
          Name:
            NameFull: Heerdink, Eibert R.
      – PersonEntity:
          Name:
            NameFull: Janssen, Marjo J. A.
      – PersonEntity:
          Name:
            NameFull: Karapinar‐Çarkit, Fatma
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 01
              M: 05
              Text: May2026
              Type: published
              Y: 2026
          Identifiers:
            – Type: issn-print
              Value: 00028614
          Numbering:
            – Type: volume
              Value: 74
            – Type: issue
              Value: 5
          Titles:
            – TitleFull: Journal of the American Geriatrics Society
              Type: main
ResultId 1