Health Care Services / en Can AI help make medical records less biased? New study suggests yes—with caveats /news/2026-07/can-ai-help-make-medical-records-less-biased-new-study-suggests-yes-caveats <span>Can AI help make medical records less biased? New study suggests yes—with caveats </span> <span><span>Heather Carroll</span></span> <span><time datetime="2026-07-10T12:50:49-04:00" title="Friday, July 10, 2026 - 12:50">Fri, 07/10/2026 - 12:50</time> </span> <div class="layout layout--gmu layout--twocol-section layout--twocol-section--70-30"> <div class="layout__region region-first"> <div data-block-plugin-id="field_block:node:news_release:body" class="block block-layout-builder block-field-blocknodenews-releasebody"> <div class="field field--name-body field--type-text-with-summary field--label-visually_hidden"> <div class="field__label visually-hidden">Body</div> <div class="field__item"><p><span class="intro-text">Large language models can identify judgmental language in clinical notes but the settings play a major role in accuracy.&nbsp;</span></p> <figure role="group" class="align-right"> <div> <div class="field field--name-image field--type-image field--label-hidden field__item"> <img loading="lazy" src="/sites/default/files/styles/small_content_image/public/2026-07/teenu_xavier.png?itok=BU560lx-" width="233" height="350"> </div> </div> <figcaption>Teenu Xavier PhD, RN. Photo provided</figcaption> </figure> <p>“Addict,” “non-compliant,” “failed treatment,” and “obese person” are examples of stigmatizing language that can appear in medical records. At 91’s College of Public Health, researchers are exploring whether artificial intelligence (AI) can help identify this kind of language in clinical notes before it impacts patient care.</p> <p>Nurse scientist <a href="https://nursing.gmu.edu/profiles/txavier">Teenu Xavier</a> and colleagues found that large language models (LLMs) show promise in identifying stigmatizing language in clinical documentation, but their performance is highly dependent on their settings. Model size, temperature settings, prompting strategies, and even note type can substantially influence results.</p> <p>One finding was consistent across every model tested: Providing examples of stigmatizing language improved accuracy.</p> <p>“Simply selecting an LLM is not enough when used for clinical documentation,” said Xavier, an assistant professor in the School of Nursing. “Careful attention must be paid to settings and prompting before these tools can be reliably used in health care environments.”</p> <h4><strong>Why does this matter?</strong></h4> <p>The use of stigmatizing language in clinical documentation can reinforce bias and affect a patient’s future care. AI tools may be able to help identify this kind of language, promoting more equitable care and improving patient trust and experience.</p> <p>“Pre-trained models, when optimized for identifying stigmatizing language, could help enable more timely interventions and modifications to the documentation process,” Xavier said. “Our research highlights the need for continued collaboration between health care professionals and AI developers to create tools that improve communication, reduce bias, and improve the overall patient experience.”</p> <h4><strong>What are the detailed study findings?</strong></h4> <ul> <li>The largest LLM (trained on large amounts of data) was the best at predicting “stigmatizing” language (94%), but the worst at predicting “not stigmatizing” correctly (47%).<br>&nbsp;</li> <li>The smallest LLM was the best at predicting “not stigmatizing” correctly (99.7%), but worst at predicting “stigmatizing” correctly (2%).<br>&nbsp;</li> <li>When researchers gave the LLM an example of stigmatizing language, accuracy improved in all models.<br>&nbsp;</li> <li>Emergency provider notes were most accurately (69%) categorized as “stigmatizing” or “non-stigmatizing,” and plan of care notes had the lowest accuracy (56%). Misclassifications most commonly arose in long, clinically dense notes where neutral descriptions of complex illness or adverse events were mistaken by the models for judgmental language.<br>&nbsp;</li> <li>Larger models worked best at lower temperature (how predictable or random the LLMs’s output is when making a classification) and smaller models improved with higher temperature, which means the LLM took more risks in interpretation.</li> </ul> <p><a href="https://academic.oup.com/jamiaopen/article/9/2/ooag037/8572038">“Detecting stigmatizing language with large language models: mind the settings”</a> was published in JAMIA Open in April 2026. Co-authors include Jane M. Carrington from the University of Florida and Joshua Lambert from the University of Cincinnati.</p> <p><em>Thumbnail photo by </em><a class="blue science-text js-contributor-link" href="https://stock.adobe.com/contributor/205162424/issaronow?load_type=author&amp;prev_url=detail"><em>issaronow</em></a><em> from Adobe Stock.</em></p> </div> </div> </div> </div> <div class="layout__region region-second"> <div data-block-plugin-id="inline_block:text" data-inline-block-uuid="c381e12e-0d19-4b4f-ab99-012767634927" class="block block-layout-builder block-inline-blocktext"> <div class="field field--name-body field--type-text-with-summary field--label-hidden field__item"><div style="background-color:#ffeec2;padding:2%;"> <h2>Key Takeaways</h2> <ul> <li>A George Mason study found that&nbsp;large language models&nbsp;show promise in&nbsp;identifying&nbsp;stigmatizing language in clinical documentation, but their performance is highly dependent on their settings, such as model size, temperature settings, prompting strategies, and note type.&nbsp;<br>&nbsp;</li> <li>The ability to detect and correct stigmatizing language early can reduce bias in patient care and lead to improved patient trust and better health outcomes.&nbsp;<br>&nbsp;</li> <li>Continued collaboration between health care workers and AI developers is needed to create tools that enhance communication, reduce bias, and improve the overall patient experience.&nbsp;</li> </ul> </div> </div> </div> <div data-block-plugin-id="inline_block:call_to_action" data-inline-block-uuid="dde0f15e-6633-4359-aff4-d815414b6345"> <div class="cta"> <a class="cta__link" href="/research/AI"> <p class="cta__title">Learn about Artificial Intelligence at George Mason <i class="fas fa-arrow-circle-right"></i> </p> <span class="cta__icon"></span> </a> </div> </div> <div data-block-plugin-id="inline_block:text" data-inline-block-uuid="b1bde1d2-caee-4099-9989-90a515d70643" class="block block-layout-builder block-inline-blocktext"> </div> <div data-block-plugin-id="inline_block:news_list" data-inline-block-uuid="c6116f0c-21ef-4914-a7eb-eb591ea4925a" class="block block-layout-builder block-inline-blocknews-list"> <h2>Related Stories</h2> <div class="views-element-container"><div class="view view-news view-id-news view-display-id-block_1 js-view-dom-id-9e4ea3828aa7c253499d15fe26ecd00800b2f0c65ca65a4d6f16e322430bb1c2"> <div class="view-content"> <div class="news-list-wrapper"> <ul class="news-list"> <li class="news-item"><div class="views-field views-field-title"><span class="field-content"><a href="/news/2026-09/nih-r01-grant-funds-study-physical-activity-and-binge-eating-children" hreflang="en">NIH R01 grant funds study of physical activity and binge eating in children</a></span></div><div class="views-field views-field-field-publish-date"><div class="field-content">September 14, 2026</div></div></li> <li class="news-item"><div class="views-field views-field-title"><span class="field-content"><a href="/news/2026-08/five-strategies-fight-burnout" hreflang="en">Five Strategies to Fight Burnout </a></span></div><div class="views-field views-field-field-publish-date"><div class="field-content">September 14, 2026</div></div></li> <li class="news-item"><div class="views-field views-field-title"><span class="field-content"><a href="/news/2026-08/invisible-front-lines" hreflang="en">The Invisible Front Lines</a></span></div><div class="views-field views-field-field-publish-date"><div class="field-content">September 14, 2026</div></div></li> <li class="news-item"><div class="views-field views-field-title"><span class="field-content"><a href="/news/2026-08/new-pathway-mental-health-professionals" hreflang="en">A New Pathway for Mental Health Professionals </a></span></div><div class="views-field views-field-field-publish-date"><div class="field-content">September 14, 2026</div></div></li> <li class="news-item"><div class="views-field views-field-title"><span class="field-content"><a href="/news/2026-09/gci-catalyst-project-aims-tackle-human-rights-issues-help-ai" hreflang="en">GCI Catalyst Project aims to tackle human rights issues with help of AI </a></span></div><div class="views-field views-field-field-publish-date"><div class="field-content">September 10, 2026</div></div></li> </ul> </div> </div> </div> </div> </div> </div> </div> <div class="layout layout--gmu layout--twocol-section layout--twocol-section--30-70"> <div class="layout__region region-first"> <div data-block-plugin-id="field_block:node:news_release:field_content_topics" class="block block-layout-builder block-field-blocknodenews-releasefield-content-topics"> <h2>Topics</h2> <div class="field field--name-field-content-topics field--type-entity-reference field--label-visually_hidden"> <div class="field__label visually-hidden">Topics</div> <div class="field__items"> <div class="field__item"><a href="/taxonomy/term/21196" hreflang="en">large language models</a></div> <div class="field__item"><a href="/taxonomy/term/5841" hreflang="en">Machine Learning in Health Care</a></div> <div class="field__item"><a href="/taxonomy/term/10561" hreflang="en">Health Care Services</a></div> <div class="field__item"><a href="/taxonomy/term/17226" hreflang="en">College of Public Health</a></div> <div class="field__item"><a href="/taxonomy/term/4656" hreflang="en">Artificial Intelligence</a></div> <div class="field__item"><a href="/taxonomy/term/271" hreflang="en">Research</a></div> <div class="field__item"><a href="/taxonomy/term/18511" hreflang="en">CPH research</a></div> <div class="field__item"><a href="/taxonomy/term/17856" hreflang="en">Nursing Research</a></div> </div> </div> </div> </div> <div> </div> </div> Fri, 10 Jul 2026 16:50:49 +0000 Heather Carroll 345984 at