Raman Dutt @ramandutt4.bsky.social · 22/02/2025This de identification token is consistently present throughout the dataset and contributes nothing towards improving the image quality. This points towards a major flaw in the dataset given MIMIC is one of the most significant medical datasets for T2I generation. 💔 000
Raman Dutt @ramandutt4.bsky.social · 22/02/2025It turns out that this de identification token (“___”) holds the most significant contribution towards memorizing training images. In other words, steps taken to protect patient information are in fact posing a threat to it. 000
Raman Dutt @ramandutt4.bsky.social · 22/02/2025MIMIC dataset contains pairs of images and corresponding text reports. These are raw reports describing the images. In the dataset, the sensitive patient information is hidden or de identified. This is done by replacing it with three underscores (“___”). 200
Raman Dutt @ramandutt4.bsky.social · 22/02/2025Are you working on Text-to-Image generation of Chest X-Rays using the MIMIC dataset? Here is something I found over the last weekend. 🧵 Observations documented in this preprint - arxiv.org/abs/2502.07516arxiv.org 100
Raman Dutt @ramandutt4.bsky.social · 07/12/2024Hello all, What’s the best tool to make nice figures for academic AI papers? 180
Raman Dutt @ramandutt4.bsky.social · 06/12/2024Not an ML-related post but I am just as happy to share this 080
Raman Dutt @ramandutt4.bsky.social · 27/11/2024A new starter pack for Medical AI researchers! go.bsky.app/PddA2uy 051
Raman Dutt @ramandutt4.bsky.social · 26/11/2024Sharing with people who might find this relevant - go.bsky.app/r5eVvT, go.bsky.app/PJKJ8vK, bsky.app/profile/berk... 010
Raman Dutt @ramandutt4.bsky.social · 26/11/2024Hence, we devise MemControl - a framework that searches for the optimal parameters to be fine-tuned to: (1) Improve image generation quality (2) Reduce Memorization! MemControl leads to optimal model capacity that should be used during fine-tuning: Not more, not less! 000
Raman Dutt @ramandutt4.bsky.social · 26/11/2024We also found that fine-tuning different subsets of parameters in a diffusion model can affect generative quality and memorization differently! Each marker in the figure is a diffusion model finetuned on the same data but with different parameter subset. Full FT (green) leads to high memorization! 110
Raman Dutt @ramandutt4.bsky.social · 26/11/2024We provide empirical proof that reducing the model capacity (by fine-tuning fewer parameters) can lead to reduced memorization! Q. How to fine-tune with fewer parameters? 🤔 A. Parameter-Efficient Fine-Tuning (PEFT) ✨ 100
Raman Dutt @ramandutt4.bsky.social · 26/11/2024The conventional way of fine-tuning models (full fine-tuning) can lead to replication of artifacts in X-Rays that can further lead to leakage of patient information, thus endangering patient privacy. Artifact replication is shown in red boxes. 100
Raman Dutt @ramandutt4.bsky.social · 26/11/2024Delighted to share our work "𝐌𝐞𝐦𝐂𝐨𝐧𝐭𝐫𝐨𝐥" now accepted at 𝐖𝐀𝐂𝐕 '𝟐𝟓. We show strong results for medical image generation and also establish an initial benchmark for generative quality and memorization of synthetic chest x-rays! Paper: arxiv.org/abs/2405.19458 Code: github.com/Raman1121/Di... More👇github.comGitHub - Raman1121/Diffusion_Memorization_HPO: A framework to reduce memorization in text-to-image diffusion models using HPOA framework to reduce memorization in text-to-image diffusion models using HPO - Raman1121/Diffusion_Memorization_HPO 160
Raman Dutt @ramandutt4.bsky.social · 26/11/2024MIDL Conference has joined BlueSky! bsky.app/profile/midl...bsky.app 020
Raman Dutt @ramandutt4.bsky.social · 25/11/2024Again, very sorry to hear about what you are going through. Advertising here is a great idea. I personally got some good advice about a condition I was going through. Wish I could be more helpful though. Wishing you the best! 010
Raman Dutt @ramandutt4.bsky.social · 25/11/2024So sorry to hear about this @ian-goodfellow.bsky.social . Do you think any of this might be related to prolonged headphone usage in addition to many other factors? Or if that exacerbates the condition? 120
Raman Dutt @ramandutt4.bsky.social · 25/11/2024For my fellow medical AI researchers, here is a starter pack - go.bsky.app/r5eVvT 152
Raman Dutt @ramandutt4.bsky.social · 25/11/2024I would consider myself slightly cracked haha. Would love to be added! 100
Raman Dutt @ramandutt4.bsky.social · 25/11/2024A starter pack I would highly recommend (not biased at all 😉) 150
Raman Dutt @ramandutt4.bsky.social · 25/11/2024Would love to be added! Currently doing a PhD in Biomedical AI 120
Raman Dutt @ramandutt4.bsky.social · 24/11/2024Omg BlueSky migration success story continues :’) @chriswolfvision.bsky.social 020
Raman Dutt @ramandutt4.bsky.social · 24/11/2024Migration to BlueSky successful! Thank you @mmbronstein.bsky.social 🙏 1151
Raman Dutt @ramandutt4.bsky.social · 24/11/2024Would love to be added! Currently starting the final year of my PhD at the University of Edinburgh! 010
Raman Dutt @ramandutt4.bsky.social · 24/11/2024I would love to be added! Currently at Huawei and Turing Institute working on Autoregressive Multimodal Generation. 210
Reposted by Raman DuttAndrei Bursuc @abursuc.bsky.social · 22/11/2024The return of the Autoregressive Image Model: AIMv2 now going multimodal. Excellent work by @alaaelnouby.bsky.social & team with code and checkpoints already up: arxiv.org/abs/2411.14402 1468
Raman Dutt @ramandutt4.bsky.social · 21/11/2024Here is a starter pack of researchers from the University of Edinburgh working on all areas of AI! go.bsky.app/KRNDkN7 041
Raman Dutt @ramandutt4.bsky.social · 21/11/2024Here is a well-curated list of researchers working on image/ video generation 000
Raman Dutt @ramandutt4.bsky.social · 21/11/2024Would love to be added to the list! Here is my recent work on reducing memorization in diffusion models, inspired by yours :’) Recently accepted at WACV ‘25 Paper - arxiv.org/pdf/2405.19458arxiv.org 100