Sign in

Danielle Bitterman MD

@daniellebitterman.bsky.social
967 followers 286 following 46 posts

Leading AI for Clinical Trials at AstraZeneca | Physician | Previously Harvard Med & Mass General Brigham

PostsRepliesMedia
Danielle Bitterman MD @daniellebitterman.bsky.social · 04/12/2025
2. A proposal for evaluation of "superhuman" systems in healthcare: ai.nejm.org/doi/full/10....
ai.nejm.org
Humanity’s Next Medical Exam: Preparing to Evaluate Superhuman Systems
The rapid advances in health care AI necessitate a fundamental shift in how we evaluate these systems. Palepu et al. (2025) demonstrate that AI can outperform medical trainees in breast cancer mana...
010
Danielle Bitterman MD @daniellebitterman.bsky.social · 04/12/2025
Check out our related work: 1. Gaps in ability of models to adjust response and differential diagnosis for cancer patients: www.thelancet.com/journals/lan... @shan23chen.bsky.social
thelancet.com
The effect of using a large language model to respond to patient messages
The relentless increase in administrative responsibilities, amplified by electronic health record (EHR) systems, has diverted clinician attention from direct patient care, fuelling burnout.1 In respon...
130
Danielle Bitterman MD @daniellebitterman.bsky.social · 04/12/2025
Our research has found that even when chatbots are given specific patient context, they often drift back toward generic, "average patient" responses. They see the data, but they don't always weigh it like a physician would.
110
Danielle Bitterman MD @daniellebitterman.bsky.social · 04/12/2025
As I shared in the NYT, models often see the data but fail to weigh it like a physician, drifting toward generic "average patient" responses. Context window ≠ Clinical reasoning. www.nytimes.com/2025/12/03/w...
1112
Danielle Bitterman MD @daniellebitterman.bsky.social · 30/11/2025
Check out our editorial on Zazzetti et al (2025)'s paper on synthetic data generation for breast cancer, in JCO CCI! Synthetic data could help with many gaps in clinical AI research, but challenges remain especially (IMO) issues with out-of-domain generalization @shan23chen.bsky.social
031
Danielle Bitterman MD @daniellebitterman.bsky.social · 17/11/2025
Super proud of @shan23chen.bsky.social for his podium presentation on his research into LLM sycophancy in the face of illogical medical queries at #AMIA25! Full paper: www.nature.com/articles/s41... Also cited yesterday in the NYT! www.nytimes.com/2025/11/16/w...
062
Danielle Bitterman MD @daniellebitterman.bsky.social · 18/10/2025
LLMs tend to prioritize helpfulness > reason. We show that safety-aware, compute-efficient fine-tuning helps models reason more critically in healthcare domain, and generalizes to improved safety alignment across other domains. www.nature.com/articles/s41... @shan23chen.bsky.social
nature.com
When helpfulness backfires: LLMs and the risk of false medical information due to sycophantic behavior - npj Digital Medicine
npj Digital Medicine - When helpfulness backfires: LLMs and the risk of false medical information due to sycophantic behavior
085
Reposted by Danielle Bitterman MD
Scott McGrath @smcgrath.phd · 17/10/2025
An overemphasis on helpfulness makes LLMs vulnerable. Research shows models will comply with illogical medical requests, generating false information. This sycophantic tendency can be corrected with specific prompting and fine-tuning. #MedSky #MedAI #MLSky
nature.com
When helpfulness backfires: LLMs and the risk of false medical information due to sycophantic behavior - npj Digital Medicine
npj Digital Medicine - When helpfulness backfires: LLMs and the risk of false medical information due to sycophantic behavior
074
Reposted by Danielle Bitterman MD
Mass General Brigham Innovation @mgbinnovation.bsky.social · 21/08/2025
Mass General physician-scientist @daniellebitterman.bsky.social discusses how AI assists the clinical data pipeline leading to better treatments for patients. Listen to unNatural Selection & register for #WMIF2025 at the link in bio to hear more : www.unnaturalselection.net/podcast/s1e19 #MedTech
unnaturalselection.net
Clinical Reporting: Mass General — unNatural Selection
Signals Over Noise: Cleaning Up Cancer Trial Data
021
Reposted by Danielle Bitterman MD
Jirui Qi @jiruiqi.bsky.social · 20/08/2025
Our paper on multilingual reasoning is accepted to Findings of #EMNLP2025! 🎉 (OA: 3/3/3.5/4) We show SOTA LMs struggle with reasoning in non-English languages; prompt-hack & post-training improve alignment but trade off accuracy. 📄 arxiv.org/abs/2505.22888 See you in Suzhou! #EMNLP
arxiv.org
When Models Reason in Your Language: Controlling Thinking Trace Language Comes at the Cost of Accuracy
Recent Large Reasoning Models (LRMs) with thinking traces have shown strong performance on English reasoning tasks. However, their ability to think in other languages is less studied. This capability ...
073
Danielle Bitterman MD @daniellebitterman.bsky.social · 07/07/2025
Are you driven to use AI to transform patient outcomes in oncology? My lab in the AI in Medicine Program (Mass General Brigham, Harvard Medical School) is seeking Postdoctoral Fellows to pioneer applications of AI—especially LLMs—in cancer care. More here: www.linkedin.com/posts/daniel...
linkedin.com
🚀 Join Us at the Forefront of AI & Cancer Care | Danielle Bitterman
🚀 Join Us at the Forefront of AI & Cancer Care Are you driven to use cutting-edge AI to transform patient outcomes in oncology? My lab within the AI in Medicine Program (Mass General Brigham, Har...
173
Reposted by Danielle Bitterman MD
Jeremy Warner, MD, MS, FAMIA, FASCO @hemoncwarner.bsky.social · 18/06/2025
Reliability of Large Language Model Knowledge Across Brand and Generic Cancer Drug Names | JCO Clinical Cancer Informatics ascopubs.org/doi/abs/10.1... #JCOCCI @daniellebitterman.bsky.social
ascopubs.org
Reliability of Large Language Model Knowledge Across Brand and Generic Cancer Drug Names | JCO Clinical Cancer Informatics
PURPOSETo evaluate the performance and consistency of large language models (LLMs) across brand and generic oncology drug names in various clinical tasks, addressing concerns about potential fluctuati...
011
Danielle Bitterman MD @daniellebitterman.bsky.social · 30/05/2025
Does your LRM reason in your language? Check out new preprint led by ✨ @jiruiqi.bsky.social & @shan23chen.bsky.social. Implications for safety/human oversight & accuracy!
020
Danielle Bitterman MD @daniellebitterman.bsky.social · 22/05/2025
Led by @shan23chen.bsky.social!
010
Danielle Bitterman MD @daniellebitterman.bsky.social · 22/05/2025
Agents are all the rage and we need to track their abilities in the medical domain. Enter MedBrowseComp, the 1st benchmark to assess agents' abilities to reason, navigate the web, and search for verifiable med info! Preprint: arxiv.org/abs/2505.14963 Site: moreirap12.github.io/mbc-browse-a...
131
Reposted by Danielle Bitterman MD
STAT @statnews.com · 14/05/2025
#STATBreakthrough
"I think we have massive opportunity in cancer care to get patients to the right care, the most advanced care earlier, by taking those workforce shortages and using AI to get to solutions."
262
Reposted by Danielle Bitterman MD
STAT @statnews.com · 14/05/2025
#STATBreakthrough
"The other thing I'm scared of, it's a patient's voice is going to be come lost in the conversation of on what type of AI is developed and how we implement it," Danielle
063
Danielle Bitterman MD @daniellebitterman.bsky.social · 14/05/2025
I’m thrilled to be in San Francisco for @statnews.com's Breakthrough West Summit! I’ll be bringing my firsthand perspective as a physician-scientist to speak about how AI is transforming cancer care, alongside leaders in the field. Let's connect if you're here! #STATBreakthroughSummitWest
022
Reposted by Danielle Bitterman MD
STAT @statnews.com · 13/05/2025
AI in Cancer Care Artificial intelligence has the potential to upend oncology, changing everything from diagnosis to treatment options. Get a wide-ranging view of how the use of technology could play out over the next few years. Moderated by @angusrohan.bsky.social #STATBreakthrough
A social card that reads Featured Session: AI in Cancer Care. Then underneath are four headshots and titles. They read: Danielle Bitterman, M.D., Clifford A. Hudis, M.D., Karen Knudsen, Ph.D., and STAT's Angus Chen.
122
Reposted by Danielle Bitterman MD
Tim Miller @tim-miller.bsky.social · 23/04/2025
Exciting news: we are organizing a shared task – 2nd edition of the Chemotherapy Treatment Timelines Extraction from the Clinical Narrative (text mining task) -- collocated with the Clinical NLP Workshop. Do LLMs solve the task? Check out bit.ly/ChemoTimelin...
bit.ly
ChemoTimelines 2025
Treatment regimens are key details in understanding the effects of genetic, epigenetic, and other factors on tumor behavior and responsiveness. As precision oncology progresses, insights into the fine...
031
Reposted by Danielle Bitterman MD
Eric Topol @erictopol.bsky.social · 23/03/2025
A pie graph worth keeping in mind as the NIH budget plummets jamanetwork.com/journals/jam... for 356 new FDA drugs approved
graph of NIH basisfor new drugs
5940141640
Danielle Bitterman MD @daniellebitterman.bsky.social · 18/02/2025
Conference and professional societies: PLEASE make hybrid options available for attendees and presenters at your conferences so that scientists from HHS-funded agencies can attend. These are unmissable opportunities to promote all the great intramural science and scientists from our government.
040
Reposted by Danielle Bitterman MD
Ken Mandl @kenmandl.bsky.social · 06/02/2025
My Perspective in @NEJM_AI. AI could distort clinical decision-making in ways that prioritize profit over patient care. Oversight & regulation must go beyond performance metrics alone to address hidden commercial forces that could shape decision support. ai.nejm.org/doi/full/10....
ai.nejm.org
Unseen Commercial Forces Could Undermine Artificial Intelligence Decision Support
Artificial intelligence (AI) is poised to transform health care, yet without robust safeguards, unseen commercial interests could distort care by prioritizing profit over patient well-being. The ph...
2134
Reposted by Danielle Bitterman MD
Emily Goldberg @dremilygoldberg.bsky.social · 12/02/2025
My opinion as an actual NIH-funded researcher (unlike Vinay) at ucsf: his lies about how NIH dollars are used reflect a complete lack of understanding of how research is performed, a lack of respect for research, and are harmful to the entire biomedical research enterprise #grifter
25710
Reposted by Danielle Bitterman MD
Jeremy Warner, MD, MS, FAMIA, FASCO @hemoncwarner.bsky.social · 09/02/2025
Budgeting for the next year of my grants and they will all need to be rescoped, even before the 15% IDC rate. NCI funding at 83% for new awards and another 10% reduction for renewals (current state). Essentially, we are getting 50% of what we asked for...how is this sustainable? @carlbergstrom.com
1145
Danielle Bitterman MD @daniellebitterman.bsky.social · 05/02/2025
As a cancer doctor I see every day how NIH-funded clinical trials save lives and has made the U.S. a leader in medical innovation. Here's one example: In the 1970s, childhood cancer survival was only 58%. Today it is 85%, largely thanks to NIH/NCI funding of Children's Oncology Group trials.
02114
Reposted by Danielle Bitterman MD
jenna ruddock @ruddock.bsky.social · 03/02/2025
Congressional delegation outside USAID now: “We are here to shed a light on a crime unfolding before our eyes.”
1036347618830
Reposted by Danielle Bitterman MD
Charlotte Clymer @charlotteclymer.bsky.social · 03/02/2025
Senator Andy Kim just went to the USAID building, talked to the security guard there to confirm employees are being barred entry, and then did a press gaggle right there in front to call it out. This is doing something. This is making an effort on messaging. Other Democratic lawmakers: take notes.
14947046615679
Reposted by Danielle Bitterman MD
Alt CDC (they/them) @altcdc.altgov.info · 02/02/2025
Gay? Lesbian? Trans? Intersex? NYC Health has health information for everybody. 🏳️‍🌈🏳️‍⚧️
nyc.gov
LGBTQ+ Health - NYC Health
2706162
Danielle Bitterman MD @daniellebitterman.bsky.social · 29/01/2025
LLM attachment styles: Secure - Claude, Anxious - DeepSeekR1/o1, Avoidant - Gemini Am I wrong?
030
Danielle Bitterman MD @daniellebitterman.bsky.social · 29/01/2025
My lab is funded by the NIH to improve monitoring and management of cancer treatment side effects, so that patients can stay on treatments that are working against their cancer, and have a better quality of life during and after treatment
0376
Danielle Bitterman MD @daniellebitterman.bsky.social · 29/01/2025
In addition to the research funded by the NIH, I am grateful and indebted to the dedicated NIH scientists & staff. Their work advances breakthroughs, scientific careers, and improves & saves lives across the U.S
091
Danielle Bitterman MD @daniellebitterman.bsky.social · 22/01/2025
this is the content I need today
030
Danielle Bitterman MD @daniellebitterman.bsky.social · 15/01/2025
this is amazing - congratulations Sasha!
020
Danielle Bitterman MD @daniellebitterman.bsky.social · 12/01/2025
Also applies broadly to treating trainees/junior faculty kindly and collaboratively.
080
Danielle Bitterman MD @daniellebitterman.bsky.social · 12/01/2025
Love this paper, thanks for this contribution @roxanadaneshjou.bsky.social and @rajpurkar.bsky.social. This is the way to start to understand how LLMs can and should perform for diagnostics - clinical vignettes don't reflect real world medical practice.
082
Danielle Bitterman MD @daniellebitterman.bsky.social · 11/01/2025
Oh I see it was converted. Here is an unofficial version: docs.google.com/document/d/1...
docs.google.com
Supplementary Table 2.docx
The TRIPOD-LLM Statement: A Targeted Guideline For Reporting Large Language Models Use Supplementary Table 2: Fillable TRIPOD-LLM checklist Section Item Checklist Item Research Design LLM Task Page Ti...
110
Danielle Bitterman MD @daniellebitterman.bsky.social · 11/01/2025
There is a word doc already provided in the supplementary materials.
100
Danielle Bitterman MD @daniellebitterman.bsky.social · 11/01/2025
There is an option to download a blank pdf to fill out once you select the study design(s) and task(s). Hope that is helpful!
100
Reposted by Danielle Bitterman MD
TRIPOD Statement @tripodstatement.bsky.social · 08/01/2025
We have a NEW PAPER in @naturemedicine.bsky.social on reporting recommendations for addressing the unique challenges of #largelanguagemodels (LLMs) in biomedical applications www.nature.com/articles/s41... #MLSky #StatsSky #medSky #AISky #artificialintelligence #generativeAI #transparency
1288
Danielle Bitterman MD @daniellebitterman.bsky.social · 08/01/2025
Paper: nature.com/articles/s41... Interactive checklist: tripod-llm.vercel.app
tripod-llm.vercel.app
Tripod + LLM
A checklist for LLM research development.
1116
Danielle Bitterman MD @daniellebitterman.bsky.social · 08/01/2025
TRIPOD-LLM is out! Check out our consensus guidelines for reporting #LLM research in biomedicine. TRIPOD-LLM is intended to be a living guideline to keep up with the rapid advances in LLMs. Kudos to lead author Dr. Jack Gallifant
14213
Reposted by Danielle Bitterman MD
Julia Maués @itsnotpink.bsky.social · 08/01/2025
Physician scientists have my heart full of gratitude. I cannot overstate how important this career is for humanity’s well-being!
062
Reposted by Danielle Bitterman MD
Jeremy Warner, MD, MS, FAMIA, FASCO @hemoncwarner.bsky.social · 20/12/2024
Our latest update of the HemOnc knowledgebase is ready and available at Harvard Dataverse: dataverse.harvard.edu/dataset.xhtm.... Have a look! @peteryang.hemonc.org @ecquis.bsky.social
dataverse.harvard.edu
083
Danielle Bitterman MD @daniellebitterman.bsky.social · 14/12/2024
I have not listened to the full lecture, but the content on the NeurIPS keynote slide is xenophobic and unacceptable. There is growing prejudice against Chinese students and academics - let's let this be a catalyst to begin addressing it more deeply. Listen to the eloquent audience member below.
030
Reposted by Danielle Bitterman MD
Stanford NLP Group @stanfordnlp.bsky.social · 09/12/2024
The extraordinary recent takeover of ML/AI by #NLP is well-known but insufficiently reflected on. Look at the @neuripsconf.bsky.social tutorials in 2024! neurips.cc/virtual/2024... 14 tutorials; 6 have "LLM" in the title; 4 more cover foundation models, with large NLP coverage. That's > 70% 😲
neurips.cc
NeurIPS 2024 TutorialsNeurIPS 2024
16414
Danielle Bitterman MD @daniellebitterman.bsky.social · 09/12/2024
I won't be @ NeurIPS, but Jack Gallifant is representing the lab and some of his other great papers. Stop by to chat w him if you're interested in our work! #NeurIPS2024 CrossCare studies how biased pretraining data -> biased LLMs: neurips.cc/virtual/2024... 12/10 4:30-7:30p W Ballroom A-D #5203
neurips.cc
NeurIPS Poster Cross-Care: Assessing the Healthcare Implications of Pre-training Data on Language Model BiasNeurIPS 2024
010
Danielle Bitterman MD @daniellebitterman.bsky.social · 06/12/2024
020
Danielle Bitterman MD @daniellebitterman.bsky.social · 06/12/2024
I am always worrying about Benzene (my cat)! www.nytimes.com/2024/12/05/w... But please don't stop wearing sunscreen! Sun exposure is a known cancer risk, benzene risks unknown. This article has good tips if you want to minimize benzene exposure. Obligatory Benzene (cat) pic ⬇️
nytimes.com
Is It Time to Worry About Benzene in Personal Care Products?
The carcinogen has been found in sunscreen, deodorants, acne creams and other personal care products. Here’s what to know.
121