Sign in

Sagnik Mukherjee

@sagnikmukherjee.bsky.social
1.1K followers 388 following 30 posts

NLP PhD student @convai_uiuc | Agents, Reasoning, evaluation etc. sagnikmukherjee.github.io scholar.google.com/citations?user=v…

PostsRepliesMedia
Reposted by Sagnik Mukherjee
ConvAI @ UIUC @convai-uiuc.bsky.social · 20/09/2025
Reinforcement Learning Finetunes Small Subnetworks in Large Language Models by @sagnikmukherjee.bsky.social, Lifan Yuan, @dilekh.bsky.social, Hao Peng Read more here: arxiv.org/abs/2505.11711 x.com/saagnikkk/st...
arxiv.org
Reinforcement Learning Finetunes Small Subnetworks in Large Language Models
Reinforcement learning (RL) yields substantial improvements in large language models (LLMs) downstream task performance and alignment with human values. Surprisingly, such large gains result from upda...
021
Sagnik Mukherjee @sagnikmukherjee.bsky.social · 21/05/2025
🚨 Paper Alert: “RL Finetunes Small Subnetworks in Large Language Models” From DeepSeek V3 Base to DeepSeek R1 Zero, a whopping 86% of parameters were NOT updated during RL training 😮😮 And this isn’t a one-off. The pattern holds across RL algorithms and models. 🧵A Deep Dive
1113
Sagnik Mukherjee @sagnikmukherjee.bsky.social · 07/05/2025
🚀Our ICML 2025 paper introduces "Premise-Augmented Reasoning Chains" - a structured approach to induce explicit dependencies in reasoning chains. By revealing the dependencies within chains, we significantly improve how LLM reasoning can be verified. 🧵[1/n]
173
Reposted by Sagnik Mukherjee
Dilek Hakkani-Tur @dilekh.bsky.social · 09/02/2025
AI over-reliance is an important issue for conversational agents. Our work supported mainly by the DARPA FACT program proposes introducing positive friction to encourage users to think critically when making decisions. Great team-work, all! @convai-uiuc.bsky.social @gokhantur.bsky.social
0103
Reposted by Sagnik Mukherjee
Marc Marone @marcmarone.com · 23/11/2024
I noticed a lot of starter packs skewed towards faculty/industry, so I made one of just NLP & ML students: go.bsky.app/vju2ux Students do different research, go on the job market, and recruit other students. Ping me and I'll add you!
10117654
Sagnik Mukherjee @sagnikmukherjee.bsky.social · 21/11/2024
📢📢LLMs are biased towards Western Culture. Well, okay, but what do you mean by "Culture"? In our survey of on cultural bias in LLMs, we reviewed ~90 papers. Interestingly, none of these papers define "culture" explicitly. They use “proxies”. [1/7] [Appeared in EMNLP mains]
171
Reposted by Sagnik Mukherjee
Gokhan Tur @gokhantur.bsky.social · 19/11/2024
Nice overview of the ReSpAct framework for conversational task completion agents @convai-uiuc.bsky.social cobusgreyling.medium.com/building-con...
cobusgreyling.medium.com
Building Conversational AI Agents By Integrating Reasoning, Speaking & Acting With LLMs
AI Agents meet Conversational UI for intuitive & natural conversations.
0152
Reposted by Sagnik Mukherjee
The Data Therapist in the Blue Sky @datatherapist.bsky.social · 18/11/2024
There! I went for it! (Let me know everyone if you want me to add or remove you) go.bsky.app/CUuio7g
58267
Reposted by Sagnik Mukherjee
ConvAI @ UIUC @convai-uiuc.bsky.social · 17/11/2024
Welcome to the official page of ConvAI@UIUC! 🤖 Based in the cornfields of UIUC, and led by Dilek Hakkani-Tur and Gokhan Tur, we do cool research on chatbots, dialogue, embodied agents, and everything in between!
191
Reposted by Sagnik Mukherjee
ConvAI @ UIUC @convai-uiuc.bsky.social · 17/11/2024
We had so much fun at #EMNLP2024 during the poster sessions and in Miami 🎉🎉 Evidence of fun (excursion to the south beach! 🏖️):
0123
Reposted by Sagnik Mukherjee
ACL 2027 @aclmeeting.bsky.social · 19/11/2024
All the ACL chapters are here now: @aaclmeeting.bsky.social @emnlpmeeting.bsky.social @eaclmeeting.bsky.social @naaclmeeting.bsky.social #NLProc
110737