Sign in

Nishant Subramani @ ACL

@nsubramani23.bsky.social
1.4K followers 508 following 23 posts

PhD student @CMU LTI - working on model #interpretability, student researcher @google; prev predoc @ai2; intern @MSFT nishantsubramani.github.io

PostsRepliesMedia
Reposted by Nishant Subramani @ ACL
Matthew Finlayson @mattf.nl · 17/10/2025
We discovered that language models leave a natural "signature" on their API outputs that's extremely hard to fake. Here's how it works 🔍 📄 arxiv.org/abs/2510.14086 1/
arxiv.org
Every Language Model Has a Forgery-Resistant Signature
The ubiquity of closed-weight language models with public-facing APIs has generated interest in forensic methods, both for extracting hidden model details (e.g., parameters) and for identifying...
48723
Nishant Subramani @ ACL @nsubramani23.bsky.social · 06/10/2025
At @colmweb.org all week 🥯🍁! Presenting 3 mechinterp + actionable interp papers at @interplay-workshop.bsky.social 1. BERTology in the Modern World w/ @bearseascape.bsky.social 2. MICE for CATs 3. LLM Microscope w/ Jiarui Liu, Jivitesh Jain, @monadiab77.bsky.social Reach out to chat! #COLM2025
0102
Nishant Subramani @ ACL @nsubramani23.bsky.social · 22/08/2025
Excited to be attending NEMI in Boston today to present 🐁 MICE for CATs: Model-Internal Confidence Estimation for Calibrating Agents with Tools and co-moderate the model steering and control roundtable! Come find me to connect and chat about steering and actionable interp
020
Nishant Subramani @ ACL @nsubramani23.bsky.social · 25/07/2025
At #ACL2025 in Vienna 🇦🇹 till next Saturday! Love to chat about anything #interpretability 🔎, understanding model internals 🔬, and finding yummy vegan food 🥬
050
Nishant Subramani @ ACL @nsubramani23.bsky.social · 14/07/2025
At #ICML2025 🇨🇦 till Sunday! Love to chat about #interpretability, understanding model internals, and finding yummy vegan food in Vancouver 🥬🍜
050
Reposted by Nishant Subramani @ ACL
bearseascape.bsky.social @bearseascape.bsky.social · 04/06/2025
🚨New #interpretability paper with @nsubramani23.bsky.social: 🕵️ Model Internal Sleuthing: Finding Lexical Identity and Inflectional Morphology in Modern Language Models
111
Nishant Subramani @ ACL @nsubramani23.bsky.social · 04/06/2025
🚨 Check out our new #interpretability paper: 🕵🏽 Model Internal Sleuthing led by the amazing @bearseascape.bsky.social who is an undergrad at @scsatcmu.bsky.social @ltiatcmu.bsky.social
041
Nishant Subramani @ ACL @nsubramani23.bsky.social · 02/06/2025
Excited to announce that I started at @googleresearch.bsky.social on the cloud team as a student researcher last month working with Hamid Palangi on actionable #interpretability 🔍 to build better tool using #agents ⚒️🤖
040
Nishant Subramani @ ACL @nsubramani23.bsky.social · 01/05/2025
Presenting this today at the poster session at #NAACL2025! Come chat about interpretability, trustworthiness, and tool-using agents! 🗓️ - Thursday May 1st (today) 📍 - Hall 3 🕑 - 200-330pm
020
Nishant Subramani @ ACL @nsubramani23.bsky.social · 30/04/2025
At #NAACL2025 🌵till Sunday! Love to chat about interpretability, understanding model internals, and finding vegan food 🥬
030
Nishant Subramani @ ACL @nsubramani23.bsky.social · 29/04/2025
🚀 Excited to share a new interp+agents paper: 🐭🐱 MICE for CATs: Model-Internal Confidence Estimation for Calibrating Agents with Tools appearing at #NAACL2025 This was work done @msftresearch.bsky.social last summer with Jason Eisner, Justin Svegliato, Ben Van Durme, Yu Su, and Sam Thomson 1/🧵
1128
Reposted by Nishant Subramani @ ACL
Rumman Chowdhury @ruchowdh.bsky.social · 25/01/2025
Have these people met … society? Read a book? Listened to music? Regurgitating esoteric facts isn’t intelligence. This is more like humanity’s last stand at jeopardy www.nytimes.com/2025/01/23/t...
nytimes.com
A Test So Hard No AI System Can Pass It — Yet
The creators of a new test called “Humanity’s Last Exam” argue we may soon lose the ability to create tests hard enough for A.I. models.
35013
Nishant Subramani @ ACL @nsubramani23.bsky.social · 10/12/2024
👏🏽 Intro 💼 PhD student @ltiatcmu.bsky.social 📜 My research is in model interpretability 🔎, understanding the internals of LLMs to build more controllable and trustworthy systems 🫵🏽 If you are interested in better understanding of language technology or model interpretability, let's connect!
170
Nishant Subramani @ ACL @nsubramani23.bsky.social · 18/11/2024
1) I'm working on using intermediate model generations with LLMs to better calibrate tool using agents ⚒️🤖 than the probabilities themselves! Turns out you can 🥳 2) There's gotta be a nice geometric understanding of what's going on within LLMs when we tune them 🤔
030
Reposted by Nishant Subramani @ ACL
Ana Marasović @anamarasovic.bsky.social · 27/10/2023
Utah is hiring tenure-track/tenured faculty & a priority area is NLP!  Please reach out over email if you have questions about the school and Salt Lake City, happy to share my experience so far.  utah.peopleadmin.com/postings/154...
043