Sign in

Shoubin Yu

@shoubin.bsky.social
54 followers 447 following 10 posts

Ph.D. Student at UNC CS. Interested in multimodal video understanding&generation. yui010206.github.io

PostsRepliesMedia
Reposted by Shoubin Yu
Vaidehi Patil @vaidehipatil.bsky.social · 07/05/2025
🚨 Introducing our @tmlrorg.bsky.social paper “Unlearning Sensitive Information in Multimodal LLMs: Benchmark and Attack-Defense Evaluation” We present UnLOK-VQA, a benchmark to evaluate unlearning in vision-and-language models, where both images and text may encode sensitive or private information.
1128
Shoubin Yu @shoubin.bsky.social · 22/04/2025
Flying to SG 🇸🇬 to attend #ICLR2025. Check out our 3 papers: ☕️CREMA: Video-language + any modality reasoning 🛡️SAFREE: A training-free concept guard for any visual diffusion models 🧭SRDF: Human-level VL-navigation via self-refined data loop feel free to DM me to grab a coffee&citywalk together 😉
010
Reposted by Shoubin Yu
Archiki Prasad @archiki.bsky.social · 18/04/2025
🚨Real-world retrieval is messy: queries are ambiguous or docs conflict & have incorrect/irrelevant info. How can we jointly address these problems? ➡️RAMDocs: challenging dataset w/ ambiguity, misinformation & noise ➡️MADAM-RAG: multi-agent framework, debates & aggregates evidence across sources 🧵⬇️
3167
Reposted by Shoubin Yu
Elias Stengel-Eskin @esteng.bsky.social · 12/04/2025
🚨Announcing TaCQ 🚨 a new mixed-precision quantization method that identifies critical weights to preserve. We integrate key ideas from circuit discovery, model editing, and input attribution to improve low-bit quant., w/ 96% 16-bit acc. at 3.1 avg bits (~6x compression) 📃 arxiv.org/abs/2504.07389
1157
Shoubin Yu @shoubin.bsky.social · 19/03/2025
Introducing VEGGIE 🥦—a unified, end-to-end, and versatile instructional video generative model. VEGGIE supports 8 skills, from object addition/removal/changing, and stylization to concept grounding/reasoning. It exceeds SoTA and shows 0-shot multimodal instructional & in-context video editing.
154
Reposted by Shoubin Yu
Mohit Bansal @mohitbansal.bsky.social · 27/01/2025
🎉 Congrats to the awesome students, postdocs, & collaborators for this exciting batch of #ICLR2025 and #NAACL2025 accepted papers (FYI some are on the academic/industry job market and a great catch 🙂), on diverse, important topics such as: -- adaptive data generation environments/policies ... 🧵
1189
Reposted by Shoubin Yu
Mohit Bansal @mohitbansal.bsky.social · 23/12/2024
🚨 We have postdoc openings at UNC 🙂 Exciting+diverse NLP/CV/ML topics**, freedom to create research agenda, competitive funding, very strong students, mentorship for grant writing, collabs w/ many faculty+universities+companies, superb quality of life/weather. Please apply + help spread the word 🙏
13715
Shoubin Yu @shoubin.bsky.social · 09/12/2024
I was so lucky to work with Jaemin in my 1st year and learned a lot from him. I can confidently say he's not only a top mind in multimodal AI but also an incredible mentor&collaborator. He is insightful, hands-on, and genuinely knows how to guide and inspire junior students👇👏
122
Reposted by Shoubin Yu
Elias Stengel-Eskin @esteng.bsky.social · 05/12/2024
🚨 I am on the faculty job market this year 🚨 I will be presenting at #NeurIPS2024 and am happy to chat in-person or digitally! I work on developing AI agents that can collaborate and communicate robustly with us and each other. More at: esteng.github.io and in thread below 🧵👇
24714
Reposted by Shoubin Yu
Mohit Bansal @mohitbansal.bsky.social · 03/12/2024
Looking forward to giving this Distinguished Lecture at StonyBrook next week & meeting the several awesome NLP + CV folks there - thanks Niranjan‬ + all for the kind invitation 🙂 PS. Excited to give a new talk on "Planning Agents for Collaborative Reasoning and Multimodal Generation" ➡️➡️ 🧵👇
1238
Reposted by Shoubin Yu
Justin Chih-Yao Chen @cyjustinchen.bsky.social · 02/12/2024
🚨 Reverse Thinking Makes LLMs Stronger Reasoners We can often reason from a problem to a solution and also in reverse to enhance our overall reasoning. RevThink shows that LLMs can also benefit from reverse thinking 👉 13.53% gains + sample efficiency + strong generalization (on 4 OOD datasets)!
11911