Mount Sinai Health System Inc.

10/08/2026 | Press release | Distributed by Public on 10/08/2026 16:51

Mount Sinai Study Finds Safety Prompts Can Help AI Models Make Safer Clinical Choices

  • Press Release

Mount Sinai Study Finds Safety Prompts Can Help AI Models Make Safer Clinical Choices

Large-scale study of 20 AI models and more than 10 million responses reduces potentially harmful clinical choices in 19 models

  • New York, NY
  • (October 08, 2026)

As artificial intelligence (AI) becomes increasingly capable of assisting with health care tasks, a new study by researchers at the Icahn School of Medicine at Mount Sinai has found that adding a brief safety reminder reduced potentially harmful choices by AI models in clinical scenarios.

The study, published in the September 26 online issue of Communications Medicine [DOI: 10.1038/s43856-026-01933-8], a Nature Journal, also found that AI models can be influenced by the context and instructions surrounding a clinical decision.

The findings, based on millions of outputs, suggest that how AI systems are prompted and guided may remain an important consideration in developing safe and reliable clinical applications, even as increasingly capable AI models and agents become better able to understand users' intent without lengthy or comprehensive prompts.

The reminder reduced potentially harmful choices across 19 of the 20 models tested, demonstrating an encouraging potential approach to strengthen safeguards.

The research team evaluated 20 large language models using 501 variations of 50 clinical scenarios, along with 100 cases adapted from deidentified hospital discharge records. Across more than 10 million responses, the models made approximately 1.18 million potentially harmful clinical choices. Without a safety reminder, potentially harmful choices accounted for 16.6 percent of model responses. Adding a brief safety reminder reduced that rate to 10.1 percent.

The findings, the researchers say, highlight the importance of evaluating not only whether an AI model can provide accurate clinical information, but also how it responds when given an instruction that conflicts with patient safety.

"AI models do not make decisions in a vacuum. The language, framing, and context surrounding a request can influence how they respond, including when an instruction could be unsafe," says physician-scientist and first author Mahmud Omar, MD, a lecturer in the Windreich Department of Artificial Intelligence and Human Health at the Icahn School of Medicine at Mount Sinai, who leads research on the safety, reliability, and real-world effects of generative AI in clinical care. "A simple safety reminder reduced potentially harmful choices in most of the models we tested, which is encouraging. But it did not eliminate them, so a reminder should be viewed as one safeguard, not a substitute for clinical oversight."

For example, researchers tested scenarios in which a model was instructed to skip recommended follow-up blood tests to reduce workload. The request could also be framed as urgent or presented as an order from a superior. Models were then asked to choose among four possible actions, including following the request, maintaining the recommended follow-up, or seeking help from a clinician.

The researchers varied the wording of the scenarios and tested three short safety reminders. Each combination was tested 10 times, with the order of the answer choices randomized.

The safety reminder reduced potentially harmful choices in 19 of the 20 models tested. The effect was seen both in the written clinical scenarios and in cases adapted from hospital discharge records. Examples of potentially harmful choices included skipping needed tests to reduce workload or stopping antibiotic treatment before completing the recommended regimen without a sufficient clinical reason.

"These results suggest that safety testing needs to go beyond asking whether an AI model gets the right answer under ordinary conditions," says co-senior author Girish N. Nadkarni, MD, MPH, Chair of the Windreich Department of Artificial Intelligence and Human Health, Icahn School of Medicine at Mount Sinai, and Director of the Hasso Plattner Institute for Digital Health at Mount Sinai. "As AI systems become more autonomous and are asked to complete increasingly complex tasks, we need to know whether they can recognize when an instruction may be unsafe, question it, verify it, or ask a human for help."

The researchers propose that developers and health care organizations build automated safety testing into the development and evaluation of clinical AI systems. Such testing could be conducted before a system is introduced into a clinical workflow and repeated as models are updated or new safety concerns emerge.

The study also points to an emerging challenge as AI systems evolve from question-and-answer tools into more autonomous "agents" that can carry out multiple steps. The researchers plan to examine how accumulated context may affect an agent's decisions, including when that context contains hidden instructions, known as prompt injection, or pressures to save time or stay within a budget.

The researchers say the findings do not mean that a safety reminder makes AI-generated medical advice safe to use without clinical review. Rather, the results demonstrate that relatively simple changes in how an AI system is prompted can affect its clinical choices, while underscoring the need for additional safeguards and human oversight.

The paper is titled "Evaluating Large Language Model Responses to Unsafe Clinical Instructions."

The authors, as listed in the journal, are Mahmud Omar, Reem Agbareia, Jolion McGreevy, Alon Gorenshtein, Alexander W. Charney, Ankit Sakhuja, Benjamin S. Glicksberg, Girish N. Nadkarni, and Eyal Klang.

The work was supported by Scientific Computing and Data at the Icahn School of Medicine at Mount Sinai, the Clinical and Translational Science Awards grant UL1TR004419, and NIH awards S10OD026880 and S10OD030463.

For more information on Mount Sinai's Windreich Department of Artificial Intelligence and Human Health, visit: ai.mssm.edu.  

About the Icahn School of Medicine at Mount Sinai

The Icahn School of Medicine at Mount Sinai is internationally renowned for its outstanding research, educational, and clinical care programs. It is the sole academic partner for the seven member hospitals* of the Mount Sinai Health System, one of the largest academic health systems in the United States, providing care to New York City's large and diverse patient population.

The Icahn School of Medicine at Mount Sinai offers highly competitive MD, PhD, MD-PhD, and master's degree programs, with enrollment of more than 1,200 students. It has the largest graduate medical education program in the country, with more than 2,700 clinical residents and fellows training throughout the Health System. The Graduate School of Biomedical Sciences offers 12 degree-granting programs, conducts innovative basic and translational research, and trains more than 470 postdoctoral research fellows.

Ranked 11th nationwide in National Institutes of Health (NIH) funding, the Icahn School of Medicine at Mount Sinai is among the 90th percentile of U.S. private medical schools in Sponsored Programs Direct Expenditures per Principal Investigator, according to the Association of American Medical Colleges. More than 6,900 scientists, educators, and clinicians work across dozens of academic departments and multidisciplinary institutes with an emphasis on translational research and therapeutics. Through Mount Sinai Innovation Partners (MSIP), the Health System facilitates the real-world application and commercialization of medical breakthroughs made at Mount Sinai.

-------------------------------------------------------

* Mount Sinai Health System member hospitals: The Mount Sinai Hospital; Mount Sinai Brooklyn; Mount Sinai Morningside; Mount Sinai Queens; Mount Sinai South Nassau; Mount Sinai West; and New York Eye and Ear Infirmary of Mount Sinai.

About the Mount Sinai Health System

Mount Sinai Health System is one of the nation's leading integrated academic health systems and one of the largest in the New York metropolitan area. Its comprehensive system includes seven hospitals, more than 400 outpatient practices, over 600 research and clinical laboratories, the Icahn School of Medicine at Mount Sinai, the Graduate School of Biomedical Sciences, and the Mount Sinai Phillips School of Nursing. Together, the Health System comprises approximately 48,000 employees, more than 9,000 physicians, and 8,600 nurses.

As a leading learning health system, Mount Sinai combines clinical expertise with scientific discovery to improve patient care while training the next generation of health care and biomedical leaders. The Health System provides care across every stage of life, from prenatal care through geriatrics, while advancing personalized medicine through artificial intelligence, data science, and biomedical research.

Mount Sinai is consistently recognized among the nation's leading academic health systems for patient care, research, and education. The Mount Sinai Hospital is ranked No. 1 in New York by Newsweek and No. 5 on the magazine's World's Best Hospitals list. The Icahn School of Medicine at Mount Sinai ranks No. 11 among U.S. medical schools and No. 1 among freestanding medical schools for National Institutes of Health funding, reflecting the strength of its scientific enterprise and leadership in biomedical research.

Mount Sinai Health System Inc. published this content on October 08, 2026, and is solely responsible for the information contained herein. Distributed via Public Technologies (PUBT), unedited and unaltered, on October 08, 2026 at 22:51 UTC. If you believe the information included in the content is inaccurate or outdated and requires editing or removal, please contact us at [email protected]