Sign in

Danielle Bitterman MD

@daniellebitterman.bsky.social
967 followers 286 following 46 posts

Leading AI for Clinical Trials at AstraZeneca | Physician | Previously Harvard Med & Mass General Brigham

PostsRepliesMedia
Danielle Bitterman MD @daniellebitterman.bsky.social · 04/12/2025
As I shared in the NYT, models often see the data but fail to weigh it like a physician, drifting toward generic "average patient" responses. Context window ≠ Clinical reasoning. www.nytimes.com/2025/12/03/w...
1112
Danielle Bitterman MD @daniellebitterman.bsky.social · 30/11/2025
Check out our editorial on Zazzetti et al (2025)'s paper on synthetic data generation for breast cancer, in JCO CCI! Synthetic data could help with many gaps in clinical AI research, but challenges remain especially (IMO) issues with out-of-domain generalization @shan23chen.bsky.social
031
Danielle Bitterman MD @daniellebitterman.bsky.social · 17/11/2025
Super proud of @shan23chen.bsky.social for his podium presentation on his research into LLM sycophancy in the face of illogical medical queries at #AMIA25! Full paper: www.nature.com/articles/s41... Also cited yesterday in the NYT! www.nytimes.com/2025/11/16/w...
062
Danielle Bitterman MD @daniellebitterman.bsky.social · 18/10/2025
LLMs tend to prioritize helpfulness > reason. We show that safety-aware, compute-efficient fine-tuning helps models reason more critically in healthcare domain, and generalizes to improved safety alignment across other domains. www.nature.com/articles/s41... @shan23chen.bsky.social
nature.com
When helpfulness backfires: LLMs and the risk of false medical information due to sycophantic behavior - npj Digital Medicine
npj Digital Medicine - When helpfulness backfires: LLMs and the risk of false medical information due to sycophantic behavior
085
Reposted by Danielle Bitterman MD
Scott McGrath @smcgrath.phd · 17/10/2025
An overemphasis on helpfulness makes LLMs vulnerable. Research shows models will comply with illogical medical requests, generating false information. This sycophantic tendency can be corrected with specific prompting and fine-tuning. #MedSky #MedAI #MLSky
nature.com
When helpfulness backfires: LLMs and the risk of false medical information due to sycophantic behavior - npj Digital Medicine
npj Digital Medicine - When helpfulness backfires: LLMs and the risk of false medical information due to sycophantic behavior
074
Reposted by Danielle Bitterman MD
Mass General Brigham Innovation @mgbinnovation.bsky.social · 21/08/2025
Mass General physician-scientist @daniellebitterman.bsky.social discusses how AI assists the clinical data pipeline leading to better treatments for patients. Listen to unNatural Selection & register for #WMIF2025 at the link in bio to hear more : www.unnaturalselection.net/podcast/s1e19 #MedTech
unnaturalselection.net
Clinical Reporting: Mass General — unNatural Selection
Signals Over Noise: Cleaning Up Cancer Trial Data
021
Reposted by Danielle Bitterman MD
Jirui Qi @jiruiqi.bsky.social · 20/08/2025
Our paper on multilingual reasoning is accepted to Findings of #EMNLP2025! 🎉 (OA: 3/3/3.5/4) We show SOTA LMs struggle with reasoning in non-English languages; prompt-hack & post-training improve alignment but trade off accuracy. 📄 arxiv.org/abs/2505.22888 See you in Suzhou! #EMNLP
arxiv.org
When Models Reason in Your Language: Controlling Thinking Trace Language Comes at the Cost of Accuracy
Recent Large Reasoning Models (LRMs) with thinking traces have shown strong performance on English reasoning tasks. However, their ability to think in other languages is less studied. This capability ...
073
Danielle Bitterman MD @daniellebitterman.bsky.social · 07/07/2025
Are you driven to use AI to transform patient outcomes in oncology? My lab in the AI in Medicine Program (Mass General Brigham, Harvard Medical School) is seeking Postdoctoral Fellows to pioneer applications of AI—especially LLMs—in cancer care. More here: www.linkedin.com/posts/daniel...
linkedin.com
🚀 Join Us at the Forefront of AI & Cancer Care | Danielle Bitterman
🚀 Join Us at the Forefront of AI & Cancer Care Are you driven to use cutting-edge AI to transform patient outcomes in oncology? My lab within the AI in Medicine Program (Mass General Brigham, Har...
173
Reposted by Danielle Bitterman MD
Jeremy Warner, MD, MS, FAMIA, FASCO @hemoncwarner.bsky.social · 18/06/2025
Reliability of Large Language Model Knowledge Across Brand and Generic Cancer Drug Names | JCO Clinical Cancer Informatics ascopubs.org/doi/abs/10.1... #JCOCCI @daniellebitterman.bsky.social
ascopubs.org
Reliability of Large Language Model Knowledge Across Brand and Generic Cancer Drug Names | JCO Clinical Cancer Informatics
PURPOSETo evaluate the performance and consistency of large language models (LLMs) across brand and generic oncology drug names in various clinical tasks, addressing concerns about potential fluctuati...
011
Danielle Bitterman MD @daniellebitterman.bsky.social · 30/05/2025
Does your LRM reason in your language? Check out new preprint led by ✨ @jiruiqi.bsky.social & @shan23chen.bsky.social. Implications for safety/human oversight & accuracy!
020
Danielle Bitterman MD @daniellebitterman.bsky.social · 22/05/2025
Agents are all the rage and we need to track their abilities in the medical domain. Enter MedBrowseComp, the 1st benchmark to assess agents' abilities to reason, navigate the web, and search for verifiable med info! Preprint: arxiv.org/abs/2505.14963 Site: moreirap12.github.io/mbc-browse-a...
131
Reposted by Danielle Bitterman MD
STAT @statnews.com · 14/05/2025
#STATBreakthrough
"I think we have massive opportunity in cancer care to get patients to the right care, the most advanced care earlier, by taking those workforce shortages and using AI to get to solutions."
262
Reposted by Danielle Bitterman MD
STAT @statnews.com · 14/05/2025
#STATBreakthrough
"The other thing I'm scared of, it's a patient's voice is going to be come lost in the conversation of on what type of AI is developed and how we implement it," Danielle
063
Danielle Bitterman MD @daniellebitterman.bsky.social · 14/05/2025
I’m thrilled to be in San Francisco for @statnews.com's Breakthrough West Summit! I’ll be bringing my firsthand perspective as a physician-scientist to speak about how AI is transforming cancer care, alongside leaders in the field. Let's connect if you're here! #STATBreakthroughSummitWest
022
Reposted by Danielle Bitterman MD
STAT @statnews.com · 13/05/2025
AI in Cancer Care Artificial intelligence has the potential to upend oncology, changing everything from diagnosis to treatment options. Get a wide-ranging view of how the use of technology could play out over the next few years. Moderated by @angusrohan.bsky.social #STATBreakthrough
A social card that reads Featured Session: AI in Cancer Care. Then underneath are four headshots and titles. They read: Danielle Bitterman, M.D., Clifford A. Hudis, M.D., Karen Knudsen, Ph.D., and STAT's Angus Chen.
122
Reposted by Danielle Bitterman MD
Tim Miller @tim-miller.bsky.social · 23/04/2025
Exciting news: we are organizing a shared task – 2nd edition of the Chemotherapy Treatment Timelines Extraction from the Clinical Narrative (text mining task) -- collocated with the Clinical NLP Workshop. Do LLMs solve the task? Check out bit.ly/ChemoTimelin...
bit.ly
ChemoTimelines 2025
Treatment regimens are key details in understanding the effects of genetic, epigenetic, and other factors on tumor behavior and responsiveness. As precision oncology progresses, insights into the fine...
031
Reposted by Danielle Bitterman MD
Eric Topol @erictopol.bsky.social · 23/03/2025
A pie graph worth keeping in mind as the NIH budget plummets jamanetwork.com/journals/jam... for 356 new FDA drugs approved
graph of NIH basisfor new drugs
5940151641
Danielle Bitterman MD @daniellebitterman.bsky.social · 18/02/2025
Conference and professional societies: PLEASE make hybrid options available for attendees and presenters at your conferences so that scientists from HHS-funded agencies can attend. These are unmissable opportunities to promote all the great intramural science and scientists from our government.
040
Reposted by Danielle Bitterman MD
Ken Mandl @kenmandl.bsky.social · 06/02/2025
My Perspective in @NEJM_AI. AI could distort clinical decision-making in ways that prioritize profit over patient care. Oversight & regulation must go beyond performance metrics alone to address hidden commercial forces that could shape decision support. ai.nejm.org/doi/full/10....
ai.nejm.org
Unseen Commercial Forces Could Undermine Artificial Intelligence Decision Support
Artificial intelligence (AI) is poised to transform health care, yet without robust safeguards, unseen commercial interests could distort care by prioritizing profit over patient well-being. The ph...
2134
Reposted by Danielle Bitterman MD
Emily Goldberg @dremilygoldberg.bsky.social · 12/02/2025
My opinion as an actual NIH-funded researcher (unlike Vinay) at ucsf: his lies about how NIH dollars are used reflect a complete lack of understanding of how research is performed, a lack of respect for research, and are harmful to the entire biomedical research enterprise #grifter
25710
Reposted by Danielle Bitterman MD
Jeremy Warner, MD, MS, FAMIA, FASCO @hemoncwarner.bsky.social · 09/02/2025
Budgeting for the next year of my grants and they will all need to be rescoped, even before the 15% IDC rate. NCI funding at 83% for new awards and another 10% reduction for renewals (current state). Essentially, we are getting 50% of what we asked for...how is this sustainable? @carlbergstrom.com
1145
Danielle Bitterman MD @daniellebitterman.bsky.social · 05/02/2025
As a cancer doctor I see every day how NIH-funded clinical trials save lives and has made the U.S. a leader in medical innovation. Here's one example: In the 1970s, childhood cancer survival was only 58%. Today it is 85%, largely thanks to NIH/NCI funding of Children's Oncology Group trials.
02114
Reposted by Danielle Bitterman MD
jenna ruddock @ruddock.bsky.social · 03/02/2025
Congressional delegation outside USAID now: “We are here to shed a light on a crime unfolding before our eyes.”
1036347678831
Reposted by Danielle Bitterman MD
Charlotte Clymer @charlotteclymer.bsky.social · 03/02/2025
Senator Andy Kim just went to the USAID building, talked to the security guard there to confirm employees are being barred entry, and then did a press gaggle right there in front to call it out. This is doing something. This is making an effort on messaging. Other Democratic lawmakers: take notes.
14947047515679
Reposted by Danielle Bitterman MD
Alt CDC (they/them) @altcdc.altgov.info · 02/02/2025
Gay? Lesbian? Trans? Intersex? NYC Health has health information for everybody. 🏳️‍🌈🏳️‍⚧️
nyc.gov
LGBTQ+ Health - NYC Health
2706162
Danielle Bitterman MD @daniellebitterman.bsky.social · 29/01/2025
LLM attachment styles: Secure - Claude, Anxious - DeepSeekR1/o1, Avoidant - Gemini Am I wrong?
030
Danielle Bitterman MD @daniellebitterman.bsky.social · 29/01/2025
In addition to the research funded by the NIH, I am grateful and indebted to the dedicated NIH scientists & staff. Their work advances breakthroughs, scientific careers, and improves & saves lives across the U.S
091
Danielle Bitterman MD @daniellebitterman.bsky.social · 12/01/2025
Also applies broadly to treating trainees/junior faculty kindly and collaboratively.
080
Danielle Bitterman MD @daniellebitterman.bsky.social · 12/01/2025
Love this paper, thanks for this contribution @roxanadaneshjou.bsky.social and @rajpurkar.bsky.social. This is the way to start to understand how LLMs can and should perform for diagnostics - clinical vignettes don't reflect real world medical practice.
082
Reposted by Danielle Bitterman MD
TRIPOD Statement @tripodstatement.bsky.social · 08/01/2025
We have a NEW PAPER in @naturemedicine.bsky.social on reporting recommendations for addressing the unique challenges of #largelanguagemodels (LLMs) in biomedical applications www.nature.com/articles/s41... #MLSky #StatsSky #medSky #AISky #artificialintelligence #generativeAI #transparency
1288
Danielle Bitterman MD @daniellebitterman.bsky.social · 08/01/2025
TRIPOD-LLM is out! Check out our consensus guidelines for reporting #LLM research in biomedicine. TRIPOD-LLM is intended to be a living guideline to keep up with the rapid advances in LLMs. Kudos to lead author Dr. Jack Gallifant
14213
Reposted by Danielle Bitterman MD
Julia Maués @itsnotpink.bsky.social · 08/01/2025
Physician scientists have my heart full of gratitude. I cannot overstate how important this career is for humanity’s well-being!
062
Reposted by Danielle Bitterman MD
Jeremy Warner, MD, MS, FAMIA, FASCO @hemoncwarner.bsky.social · 20/12/2024
Our latest update of the HemOnc knowledgebase is ready and available at Harvard Dataverse: dataverse.harvard.edu/dataset.xhtm.... Have a look! @peteryang.hemonc.org @ecquis.bsky.social
dataverse.harvard.edu
083
Danielle Bitterman MD @daniellebitterman.bsky.social · 14/12/2024
I have not listened to the full lecture, but the content on the NeurIPS keynote slide is xenophobic and unacceptable. There is growing prejudice against Chinese students and academics - let's let this be a catalyst to begin addressing it more deeply. Listen to the eloquent audience member below.
030
Reposted by Danielle Bitterman MD
Stanford NLP Group @stanfordnlp.bsky.social · 09/12/2024
The extraordinary recent takeover of ML/AI by #NLP is well-known but insufficiently reflected on. Look at the @neuripsconf.bsky.social tutorials in 2024! neurips.cc/virtual/2024... 14 tutorials; 6 have "LLM" in the title; 4 more cover foundation models, with large NLP coverage. That's > 70% 😲
neurips.cc
NeurIPS 2024 TutorialsNeurIPS 2024
16414
Danielle Bitterman MD @daniellebitterman.bsky.social · 09/12/2024
I won't be @ NeurIPS, but Jack Gallifant is representing the lab and some of his other great papers. Stop by to chat w him if you're interested in our work! #NeurIPS2024 CrossCare studies how biased pretraining data -> biased LLMs: neurips.cc/virtual/2024... 12/10 4:30-7:30p W Ballroom A-D #5203
neurips.cc
NeurIPS Poster Cross-Care: Assessing the Healthcare Implications of Pre-training Data on Language Model BiasNeurIPS 2024
010
Danielle Bitterman MD @daniellebitterman.bsky.social · 06/12/2024
I am always worrying about Benzene (my cat)! www.nytimes.com/2024/12/05/w... But please don't stop wearing sunscreen! Sun exposure is a known cancer risk, benzene risks unknown. This article has good tips if you want to minimize benzene exposure. Obligatory Benzene (cat) pic ⬇️
nytimes.com
Is It Time to Worry About Benzene in Personal Care Products?
The carcinogen has been found in sunscreen, deodorants, acne creams and other personal care products. Here’s what to know.
121
Reposted by Danielle Bitterman MD
Gary Collins @gscollins.bsky.social · 05/12/2024
It's that time of year to bring out our @bmj.com xmas paper On the 12th Day of Christmas, a Statistician Sent to Me... from 2022 (w/ @richarddriley.bsky.social) with our list of commonly seen statistical faux pas🎄 www.bmj.com/content/379/... #StatsSky #EpiSky
02110
Danielle Bitterman MD @daniellebitterman.bsky.social · 05/12/2024
Check out some of our early experiments on SAE transferability across modalities - moving toward more interpretable vision-language models 🧠
291
Danielle Bitterman MD @daniellebitterman.bsky.social · 02/12/2024
Right there with you... An important point that we all need to reflect on: How are we supporting the next generation of scientists? Young scientists most immediately affected, but there will be ripple effects throughout society
051
Danielle Bitterman MD @daniellebitterman.bsky.social · 27/11/2024
If you are going to NeurIPS, stop by our poster to hear about our Coss-Care project, in the Benchmarks and Dataset Track! 12/11/24 from 4:30pm-7:30pm neurips.cc/virtual/2024... @neuripsconf.bsky.social #NeurIPS
neurips.cc
NeurIPS Poster Cross-Care: Assessing the Healthcare Implications of Pre-training Data on Language Model BiasNeurIPS 2024
050
Danielle Bitterman MD @daniellebitterman.bsky.social · 18/11/2024
Looking forward to discussing risk management of large language models at the FDA Digital Health Advisory Committee Meeting on November 20th. The meeting will be webcast and is open to public comment through January 1, 2025: www.fda.gov/advisory-com...
fda.gov
November 20-21, 2024: Digital Health Advisory Committee Meeting
November 20-21, 2024: Digital Health Advisory Committee Meeting Announcement
0101
Danielle Bitterman MD @daniellebitterman.bsky.social · 16/11/2024
🩺💡The Bitterman lab has spent much of the past year researching #LLMs for healthcare. This post summarizes our inroads into making LLMs safer and reliable for clinicians and patients: huggingface.co/blog/shanche.... We'll be at #EMNLP2024 - come chat if you have similar interests!
huggingface.co
What We Learned About LLM/VLMs in Healthcare AI Evaluation:
A Blog post by Shan Chen on Hugging Face
0165
Danielle Bitterman MD @daniellebitterman.bsky.social · 16/11/2024
🎉 Incredibly proud of @shan23chen.bsky.social for being selected for the 2024 Google PhD Fellowship in Natural Language Processing: blog.google/technology/r... !!! So excited to see how Shan's contributions will continue shaping the future of clinical NLP #HealthAI #NLP 🌟
blog.google
110
Reposted by Danielle Bitterman MD
Shan Chen @shan23chen.bsky.social · 13/11/2024
Here are some reflections on many studies we did this year. Tons of progress has been made, but there are still safety concerns..🧐 Poster 10:30 riverfront at EMNLP2024 🏖️ Happy to chat and connect! 📃 huggingface.co/blog/shanche... 🔊 tinyurl.com/aimpodcast24 @daniellebitterman.bsky.social
huggingface.co
What We Learned About LLM/VLMs in Healthcare AI Evaluation:
A Blog post by Shan Chen on Hugging Face
0122