Asset Details
MbrlCatalogueTitleDetail
Do you wish to reserve the book?
First, do NOHARM: towards clinically safe large language models
by
Maharaj, Saloni Kumar
, Liang, April S
, Ranji, Sumant
, Chopra, Kanav
, Newman, Kira L
, Weng, Yingjie
, Hong, David I
, Kadiyala, Vinay
, Badhwar, Adi
, Koh, Joel
, McCoy, Liam G
, Hom, Jason
, Koshy, Jacob M
, Ziolkowski, Susan
, Goh, Ethan
, Cosgriff, Christopher V
, Gwiazdon, Matthew
, Bassman Tappuni
, Rodman, Adam
, Agarwal, Anup
, Buchheit, Kathleen M
, Caldwell, Jillian
, Khemani, Sarita
, Rustagi, Arjun
, Dalal, Rahul S
, Chen, Jonathan H
, Milstein, Arnold
, Diep, Robert
, Jesudasen, Sirus
, Shirvani, Daniel
, Chakraborty, Rebanta
, Shah, Nigam H
, Pahalyants, Vartan
, Patil, Advait
, Iberri, David J
, Wei, Nancy
, Ravi, Vishnu
, Jindal, Jenelle
, Shih, Allen
, Kaplan, Tamara B
, Pallais, J Carl
, Schulman, Kevin
, Tran, Jessica
, Wu, David J H
, Wu, David
, Fateme Nateghi Haredasht
, Lee, Ernest Y
, Jain, Priyank
, Marshall, Nicholas
, Galetta, Kristin
, French, Brianna
in
Annotations
/ Benchmarks
/ Large language models
/ Multiagent systems
/ Performance evaluation
/ Physicians
2025
Hey, we have placed the reservation for you!
By the way, why not check out events that you can attend while you pick your title.
You are currently in the queue to collect this book. You will be notified once it is your turn to collect the book.
Oops! Something went wrong.
Looks like we were not able to place the reservation. Kindly try again later.
Are you sure you want to remove the book from the shelf?
First, do NOHARM: towards clinically safe large language models
by
Maharaj, Saloni Kumar
, Liang, April S
, Ranji, Sumant
, Chopra, Kanav
, Newman, Kira L
, Weng, Yingjie
, Hong, David I
, Kadiyala, Vinay
, Badhwar, Adi
, Koh, Joel
, McCoy, Liam G
, Hom, Jason
, Koshy, Jacob M
, Ziolkowski, Susan
, Goh, Ethan
, Cosgriff, Christopher V
, Gwiazdon, Matthew
, Bassman Tappuni
, Rodman, Adam
, Agarwal, Anup
, Buchheit, Kathleen M
, Caldwell, Jillian
, Khemani, Sarita
, Rustagi, Arjun
, Dalal, Rahul S
, Chen, Jonathan H
, Milstein, Arnold
, Diep, Robert
, Jesudasen, Sirus
, Shirvani, Daniel
, Chakraborty, Rebanta
, Shah, Nigam H
, Pahalyants, Vartan
, Patil, Advait
, Iberri, David J
, Wei, Nancy
, Ravi, Vishnu
, Jindal, Jenelle
, Shih, Allen
, Kaplan, Tamara B
, Pallais, J Carl
, Schulman, Kevin
, Tran, Jessica
, Wu, David J H
, Wu, David
, Fateme Nateghi Haredasht
, Lee, Ernest Y
, Jain, Priyank
, Marshall, Nicholas
, Galetta, Kristin
, French, Brianna
in
Annotations
/ Benchmarks
/ Large language models
/ Multiagent systems
/ Performance evaluation
/ Physicians
2025
Oops! Something went wrong.
While trying to remove the title from your shelf something went wrong :( Kindly try again later!
Do you wish to request the book?
First, do NOHARM: towards clinically safe large language models
by
Maharaj, Saloni Kumar
, Liang, April S
, Ranji, Sumant
, Chopra, Kanav
, Newman, Kira L
, Weng, Yingjie
, Hong, David I
, Kadiyala, Vinay
, Badhwar, Adi
, Koh, Joel
, McCoy, Liam G
, Hom, Jason
, Koshy, Jacob M
, Ziolkowski, Susan
, Goh, Ethan
, Cosgriff, Christopher V
, Gwiazdon, Matthew
, Bassman Tappuni
, Rodman, Adam
, Agarwal, Anup
, Buchheit, Kathleen M
, Caldwell, Jillian
, Khemani, Sarita
, Rustagi, Arjun
, Dalal, Rahul S
, Chen, Jonathan H
, Milstein, Arnold
, Diep, Robert
, Jesudasen, Sirus
, Shirvani, Daniel
, Chakraborty, Rebanta
, Shah, Nigam H
, Pahalyants, Vartan
, Patil, Advait
, Iberri, David J
, Wei, Nancy
, Ravi, Vishnu
, Jindal, Jenelle
, Shih, Allen
, Kaplan, Tamara B
, Pallais, J Carl
, Schulman, Kevin
, Tran, Jessica
, Wu, David J H
, Wu, David
, Fateme Nateghi Haredasht
, Lee, Ernest Y
, Jain, Priyank
, Marshall, Nicholas
, Galetta, Kristin
, French, Brianna
in
Annotations
/ Benchmarks
/ Large language models
/ Multiagent systems
/ Performance evaluation
/ Physicians
2025
Please be aware that the book you have requested cannot be checked out. If you would like to checkout this book, you can reserve another copy
We have requested the book for you!
Your request is successful and it will be processed during the Library working hours. Please check the status of your request in My Requests.
Oops! Something went wrong.
Looks like we were not able to place your request. Kindly try again later.
First, do NOHARM: towards clinically safe large language models
Paper
First, do NOHARM: towards clinically safe large language models
2025
Request Book From Autostore
and Choose the Collection Method
Overview
Large language models (LLMs) are routinely used by physicians and patients for medical advice, yet their clinical safety profiles remain poorly characterized. We present NOHARM (Numerous Options Harm Assessment for Risk in Medicine), a benchmark using 100 real primary care-to-specialist consultation cases to measure frequency and severity of harm from LLM-generated medical recommendations. NOHARM covers 10 specialties, with 12,747 expert annotations for 4,249 clinical management options. Across 31 LLMs, potential for severe harm from LLM recommendations occurs in up to 22.2% (95% CI 21.6-22.8%) of cases, with harm of omission accounting for 76.6% (95% CI 76.4-76.8%) of errors. Safety performance is only moderately correlated (r = 0.61-0.64) with existing AI and medical knowledge benchmarks. The best models outperform generalist physicians on safety (mean difference 9.7%, 95% CI 7.0-12.5%), and a diverse multi-agent approach improves safety compared to solo models (mean difference 8.0%, 95% CI 4.0-12.1%). Therefore, despite strong performance on existing evaluations, widely used AI models can produce severely harmful medical advice at nontrivial rates, underscoring clinical safety as a distinct performance dimension necessitating explicit measurement.
Publisher
Cornell University Library, arXiv.org
Subject
MBRLCatalogueRelatedBooks
Related Items
Related Items
We currently cannot retrieve any items related to this title. Kindly check back at a later time.
This website uses cookies to ensure you get the best experience on our website.