Sign in

Subbarao Kambhampati (కంభంపాటి సుబ్బారావు)

@rao2z.bsky.social
999 followers 16 following 106 posts

AI researcher & teacher at SCAI, ASU. Former President of AAAI & Chair of AAAS Sec T. Here to tweach #AI. YouTube Ch: bit.ly/38twrAV Twitter: rao2z

PostsRepliesMedia
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 04/09/2026
🐜🐜🐜ANThropomorphism..🐜🐜🐜 x.com/rao2z/status...
x.com
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) (@rao2z) on X
In a recent harrowing development, a very normal routine experiment by the frontier company Ant went completely off the rails. 😱 Ant put thousands of their ant agents in a pretty secure ant farm with...
010
Reposted by Subbarao Kambhampati (కంభంపాటి సుబ్బారావు)
Melanie Mitchell @melaniemitchell.bsky.social · 01/08/2026
I recommend this article about AI reasoning, where the author lets us in on his struggles w/ AI cognitive dissonance. Plus some priceless quotes from @rao2z.bsky.social. (My recommendation has *nothing* to do with the fact that I'm quoted in it too 😇) www.quantamagazine.org/is-ai-reason...
quantamagazine.org
Is AI Reasoning Right for the Wrong Reasons? | Quanta Magazine
The idea that artificial intelligence can “reason” is more intuitive than ever. But intuitions can be wrong, and the science is far from settled.
511134
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 14/05/2026
🥳 📢 Our Beyond Semantics paper on LLM "thinking traces" is now officially published at @tmlrorg.bsky.social (with a J2C certification--so you will also see us at #NeurIPS2026..😋) TMLR 👉 openreview.net/forum?id=gDE... Kudos to Karthik Valmeekam, Vardhan Palod, Kaya Stechly and Atharva Gundawar!
030
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 08/05/2026
Thoughts on Mechanistic Interpretability breakthroughs.. www.linkedin.com/posts/subbar...
linkedin.com
#aiaphorisms | Subbarao Kambhampati
Mechanistic Interpretability research be like.. We wanted to see what the bread slice is thinking when it is in the toaster. Is it taking its singing equanimously? thinking religious thoughts? or plo...
010
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 05/05/2026
Here is the recording of a keynote talk I gave at #ICLR2026 UCRL workshop a few days back on the challenges of explainable AI in the era of LLMs. It looks at how LLMs seemingly make some XAI problems go away, while bringing in a new set.. www.youtube.com/watch?v=sM-0...
youtube.com
What does XAI mean in the era of LLMs? (Keynote @ ICLR 2026 Unified Concept Representations wkshp)
YouTube video by Subbarao Kambhampati
010
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 02/05/2026
Thinking Trace Trilogy: Here is our trilogy of refereed papers investigating the interpretability of the thinking traces output by reasoning models: arxiv.org/abs/2504.09762 in #ICML2026; arxiv.org/abs/2505.13792 in #ACL2026 and arxiv.org/abs/2505.13775 in @tmlrorg.bsky.social 2026
100
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 26/04/2026
Speaking in a couple of hours at UCRL workshop #ICLR2026 (Sunday 2:15pm Rio time; Room 209; iclr.cc/virtual/2026... )
020
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 24/04/2026
Looking forward to speak at the #ICLR2026 @iclr-conf.bsky.social workshop on "Unifying Concept Representation Learning" this Sunday (Rm 209; 2:15PM Rio time).. with a title inspired by Bernard Shaw's quote about England and America.. 😛 1/2
130
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 12/04/2026
𝙈𝙞𝙨𝙘𝙤𝙣𝙘𝙚𝙥𝙩𝙞𝙤𝙣𝙨 𝙖𝙗𝙤𝙪𝙩 𝙇𝙇𝙈-𝙈𝙤𝙙𝙪𝙡𝙤 𝙖𝙣𝙙 𝙩𝙝𝙚 𝙍𝙤𝙡𝙚 𝙤𝙛 𝙑𝙚𝙧𝙞𝙛𝙞𝙚𝙧𝙨 𝙞𝙣 𝙍𝙚𝙖𝙨𝙤𝙣𝙞𝙣𝙜 𝙈𝙤𝙙𝙚𝙡𝙨. #SundayHarangue www.linkedin.com/pulse/miscon...
linkedin.com
Misconceptions about LLM-Modulo and the Role of Verifiers in Reasoning Models. #SundayHarangue
I have talked before about LLM-Modulo framework (c.f.
030
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 29/03/2026
Our paper questioning the wide-spread anthropomorphization of LRM intermediate tokens as "reasoning traces" has just been accepted to @tmlrorg.bsky.social (arxiv.org/abs/2505.13775). This work was lead by the dream team of Karthik Valmeekam, Vardhan Palod, Kaya Stechly and Atharva Gundawar. 1/
1212
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 22/03/2026
Verifying LLM problem progress with LLM-Process-Modulo www.linkedin.com/posts/subbar...
linkedin.com
𝐋𝐋𝐌-𝐏𝐫𝐨𝐜𝐞𝐬𝐬-𝐌𝐨𝐝𝐮𝐥𝐨: Our original 𝘓𝘓𝘔-𝘔𝘰𝘥𝘶𝘭𝘰 framework (https://lnkd.in/gyAyKx4E) is a Generate-Test framework, with the LLM generating candidate solutions and a bank of… | Subbarao Kambhampati
𝐋𝐋𝐌-𝐏𝐫𝐨𝐜𝐞𝐬𝐬-𝐌𝐨𝐝𝐮𝐥𝐨: Our original 𝘓𝘓𝘔-𝘔𝘰𝘥𝘶𝘭𝘰 framework (https://lnkd.in/gyAyKx4E) is a Generate-Test framework, with the LLM generating candidate solutions and a bank of verifiers critiquing those sol...
070
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 16/03/2026
World Models: The old, the new and the wishful #SundayHarangue There is a lot of chatter about world models of late--even more than can be explained by Yann's bet. I was going to comment on this clamor in my next class, and thought I will preview it here first..😋 www.linkedin.com/pulse/world-...
linkedin.com
World Models: The old, the new and the wishful #SundayHarangue
There is a lot of chatter about world models of late--even more than can be explained by Yann betting his entire new enterprise on it. I was going to comment on this clamor in my class this week, and ...
020
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 14/03/2026
If you, as a CS Prof, are wondering whether you need PhD students at all now that you have wangled a subscription to Claude Code, your lab probably had a pretty depressing vibe to begin with--and 'em students are likely better off with you hanging out with Claude.. #AIAphorisms
140
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 11/03/2026
Our initial interest in the reasoning capabilities of LLMs arose from our proximal work on Human-AI interaction. It was thus gratifying to give a keynote on the role of LLMs in Human-AI interaction at IEEE CogSIMA conference yesterday Video 👉 youtu.be/yf4RQYlKRJI
youtu.be
Role of LLMs in Human-AI Interaction (Keynote @ Cogsima 2026)
YouTube video by Subbarao Kambhampati
030
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 10/03/2026
We updated our position on anthropomorphization of intermediate tokens in LRMs--with additional results and a call to action.. arxiv.org/abs/2504.09762
0193
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 02/03/2026
𝗣𝗼𝘀𝘁-𝘁𝗿𝗮𝗶𝗻𝗶𝗻𝗴 𝘄𝗶𝘁𝗵 𝗼𝗿 𝘄𝗶𝘁𝗵𝗼𝘂𝘁 𝘀𝗲𝗹𝗳-𝗱𝗶𝘀𝘁𝗶𝗹𝗹𝗮𝘁𝗶𝗼𝗻 𝗶𝘀 𝗮𝗹𝗹 𝗮𝗯𝗼𝘂𝘁 𝗰𝗼𝗺𝗽𝗶𝗹𝗶𝗻𝗴 𝘃𝗲𝗿𝗶𝗳𝗶𝗲𝗿 𝘀𝗶𝗴𝗻𝗮𝗹 𝗶𝗻𝘁𝗼 𝘁𝗵𝗲 𝗯𝗮𝘀𝗲 𝗟𝗟𝗠 𝗴𝗲𝗻𝗲𝗿𝗮𝘁𝗼𝗿 #SundayHarangue 𝗦𝘁𝗮𝗻𝗱𝗮𝗿𝗱 𝗣𝗼𝘀𝘁-𝗧𝗿𝗮𝗶𝗻𝗶𝗻𝗴 ≈ 𝗖𝗼𝗺𝗽𝗶𝗹𝗶𝗻𝗴 𝗩𝗲𝗿𝗶𝗳𝗶𝗲𝗿 𝗦𝗶𝗴𝗻𝗮𝗹 𝗦𝗲𝗹𝗳-𝗱𝗶𝘀𝘁𝗶𝗹𝗹𝗮𝘁𝗶𝗼𝗻 = 𝗖𝗼𝗺𝗽𝗶𝗹𝗶𝗻𝗴 𝗩𝗲𝗿𝗶𝗳𝗶𝗲𝗿 𝗙𝗲𝗲𝗱𝗯𝗮𝗰𝗸 𝘁𝗼𝗼 See 👇 for more.. www.linkedin.com/posts/subbar...
linkedin.com
Planning & Reasoning Abilities of LLMs/LRMs (Lecture 2 @ Melbourne ML Summer School 2026) | Subbarao Kambhampati
𝗣𝗼𝘀𝘁-𝘁𝗿𝗮𝗶𝗻𝗶𝗻𝗴 𝘄𝗶𝘁𝗵 𝗼𝗿 𝘄𝗶𝘁𝗵𝗼𝘂𝘁 𝘀𝗲𝗹𝗳-𝗱𝗶𝘀𝘁𝗶𝗹𝗹𝗮𝘁𝗶𝗼𝗻 𝗶𝘀 𝗮𝗹𝗹 𝗮𝗯𝗼𝘂𝘁 𝗰𝗼𝗺𝗽𝗶𝗹𝗶𝗻𝗴 𝘃𝗲𝗿𝗶𝗳𝗶𝗲𝗿 𝘀𝗶𝗴𝗻𝗮𝗹 𝗶𝗻𝘁𝗼 𝘁𝗵𝗲 𝗯𝗮𝘀𝗲 𝗟𝗟𝗠 𝗴𝗲𝗻𝗲𝗿𝗮𝘁𝗼𝗿 #SundayHarangue 𝗦𝘁𝗮𝗻𝗱𝗮𝗿𝗱 𝗣𝗼𝘀𝘁-𝗧𝗿𝗮𝗶𝗻𝗶𝗻𝗴 ≈ 𝗖𝗼𝗺𝗽𝗶𝗹𝗶𝗻𝗴 𝗩𝗲𝗿𝗶𝗳𝗶𝗲𝗿 𝗦𝗶𝗴𝗻𝗮𝗹 𝗦𝗲𝗹𝗳-𝗱𝗶𝘀𝘁𝗶𝗹𝗹𝗮𝘁𝗶𝗼...
020
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 16/02/2026
Does it really make sense to think of inference efficiency in terms of the number of tokens produced? No. 👇 x.com/i/status/202...
x.com
210
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 13/02/2026
Here are the recordings of two lectures on 𝗣𝗹𝗮𝗻𝗻𝗶𝗻𝗴 & 𝗥𝗲𝗮𝘀𝗼𝗻𝗶𝗻𝗴 𝗖𝗮𝗽𝗮𝗯𝗶𝗹𝗶𝘁𝗶𝗲𝘀 𝗼𝗳 𝗟𝗟𝗠𝘀/𝗟𝗥𝗠𝘀 that I gave this week at Melbourne ML Summer School (lnkd.in/g7rxg9sw). 𝙇𝙚𝙘𝙩𝙪𝙧𝙚 1: youtube.com/watch?v=_PPV... 𝙇𝙚𝙘𝙩𝙪𝙧𝙚 2: youtube.com/watch?v=fKlm...
160
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 28/01/2026
A common theme in our work these past few years has been pushing back on facile anthropomorphizations (and/or efforts that bring questionable/discredited Cognitive Science metaphors) to LLMs.. So I enjoyed giving this talk at @ivado.bsky.social yesterday... www.youtube.com/watch?v=CoyS...
youtube.com
Anthropomorphization Sins in Modern AI (or Perils of Prematurely Applying Lens of Cognition to LLMs)
YouTube video by Subbarao Kambhampati
140
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 06/01/2026
Three of my talks in India last month--at @iitdelhi.bsky.social, @msftresearch.bsky.social India and at IndoML Symposium--were "On the Mythos of LRM Thinking Tokens." Here is a recording of one of them--the talk I gave at MSR India. www.youtube.com/watch?v=fCQX...
youtube.com
On the Mythos of LRM "Thinking Tokens" (Talk @ Microsoft Research, India; 12/16/2025)
YouTube video by Subbarao Kambhampati
000
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 09/12/2025
ICYMI, here is my keynote on the semantics of LRM "thinking traces" at #NeurIPS2025 workshop on Multimodal Algorithmic Reasoning. It's a unified view of the seven papers we presented at the conference workshops. Special thanks to the engaged audience..🙏 www.youtube.com/watch?v=rvby...
youtube.com
Talk on the semantics of "Thinking Traces" (Keynote at NeurIPS2025 MAR Workshop)
YouTube video by Subbarao Kambhampati
010
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 03/11/2025
[On using Continuous Latent Space Vectors in the context windows of Transformers and LLMs] #SundayHarangue 👉 x.com/rao2z/status...
020
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 16/09/2025
My talk at Samsung AI Forum yesterday www.youtube.com/watch?v=L2nA...
youtube.com
LRMs and Agentic AI (Talk at Samsung AI Forum)
YouTube video by Subbarao Kambhampati
020
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 14/09/2025
In the year since LRMs ("reasoning models") hit the scene, we have been trying to understand, analyze and demystify them.. Here are our efforts to date--conveniently all in one place..👇 www.linkedin.com/posts/subbar...
linkedin.com
In the year since LRMs ("reasoning models") hit the scene, we have been trying to understand, analyze and demystify them.. Here are our efforts to date--conveniently all in one… | Subbarao K...
In the year since LRMs ("reasoning models") hit the scene, we have been trying to understand, analyze and demystify them.. Here are our efforts to date--conveniently all in one place.. (𝗙𝗶𝗿𝘀𝘁..) 𝗘𝘃𝗮𝗹...
051
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 10/09/2025
𝐏𝐞𝐫𝐟𝐨𝐫𝐦𝐚𝐭𝐢𝐯𝐞 𝐓𝐡𝐢𝐧𝐤𝐢𝐧𝐠? The anthropomorphization of LRM intermediate tokens as thinking begat a cottage industry to "get efficiency by shortening thinking." We ask: 𝗜𝘀 𝗖𝗼𝗧 𝗹𝗲𝗻𝗴𝘁𝗵 𝗿𝗲𝗮𝗹𝗹𝘆 𝗮 𝗿𝗲𝗳𝗹𝗲𝗰𝘁𝗶𝗼𝗻 𝗼𝗳 𝗽𝗿𝗼𝗯𝗹𝗲𝗺 𝗵𝗮𝗿𝗱𝗻𝗲𝘀𝘀 𝗼𝗿 𝗶𝘀 𝗶𝘁 𝗺𝗼𝗿𝗲 𝗽𝗲𝗿𝗳𝗼𝗿𝗺𝗮𝘁𝗶𝘃𝗲? 👉 www.linkedin.com/posts/subbar...
060
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 31/08/2025
Rejecting papers in #AI Conferences because of "resource constraints" is shooting ourselves in the foot as a community; use Findings.. #SundayHarangue 👇 x.com/rao2z/status...
x.com
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) on X: "Rejecting papers in #AI Conferences because of "resource constraints" is shooting ourselves in the foot as a community; use Findings.. #SundayHarangue By now, we have all know that top AI conferences are oversubscribed (in terms of paper submissions), and have heard that that" / X
Rejecting papers in #AI Conferences because of "resource constraints" is shooting ourselves in the foot as a community; use Findings.. #SundayHarangue By now, we have all know that top AI conferences are oversubscribed (in terms of paper submissions), and have heard that that
210
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 22/07/2025
Proofs are not reasoning traces & I/O Format Language shouldn't be much of an issue for LLMs + other things #SundayHarangue (Special IMO edition). 🧵 👇 x.com/rao2z/status...
x.com
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) on X: "Proofs are not reasoning traces & I/O Format Language shouldn't be much of an issue for LLMs #SundayHarangue (Special IMO edition). 1/ My feed these last couple of days of IMO discussions has been full of comments that seem to conflate LRM intermediate tokens (aka reasoning" / X
Proofs are not reasoning traces & I/O Format Language shouldn't be much of an issue for LLMs #SundayHarangue (Special IMO edition). 1/ My feed these last couple of days of IMO discussions has been full of comments that seem to conflate LRM intermediate tokens (aka reasoning
041
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 19/07/2025
Both LLMs and LRMs are upper bounded by humanity's knowledge closure. True scientific discoveries are, by definition, outside of that closure. Ergo, LLMs/LRMs are great force multipliers to us; but don't support "Nobel this weekend" hype.. 👉 www.linkedin.com/posts/subbar...
linkedin.com
Neither LLMs nor LRMs have the ability to go beyond the humanity's knowledge closure--which is needed for true discoveries. | Subbarao Kambhampati
Neither LLMs nor LRMs have the ability to go beyond the humanity's knowledge closure--which is needed for true discoveries. Both are beholden to the collected knowledge of the humanity (whether de...
092
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 13/07/2025
Computational Complexity is the wrong measure for LRMs (as it was for LLMs)--think distributional distance instead #SundayHarangue (yes, we're back!) 👉 x.com/rao2z/status...
020
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 23/06/2025
A̶̶̶I̶̶̶ ̶ ̶ ̶ ̶(̶A̶r̶t̶i̶f̶i̶c̶i̶a̶l̶ ̶I̶n̶t̶e̶l̶l̶i̶g̶e̶n̶c̶e̶)̶ ̶̶̶A̶̶̶G̶̶̶I̶̶̶ ̶(̶A̶r̶t̶i̶f̶i̶c̶i̶a̶l̶ ̶G̶e̶n̶e̶r̶a̶l̶ ̶I̶n̶t̶e̶l̶l̶i̶g̶e̶n̶c̶e̶)̶ ̶̶̶A̶̶̶S̶̶̶I̶̶̶ ̶(̶A̶r̶t̶i̶f̶i̶c̶i̶a̶l̶ ̶S̶u̶p̶e̶r̶ ̶I̶n̶t̶e̶l̶l̶i̶g̶e̶n̶c̶e̶) ASDI (Artificial Super Duper Intelligence) Don't get stuck with yesterday's hypeonyms! Dare to get to the next level! #AIAphorisms
031
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 19/06/2025
For anyone interested, here are the videos of the three ~50min each lectures on the reasoning/planning capabilities of LLMs/LRMs that I gave at #ACDL2025 in Riva Del Sole resort last week. 1/ www.youtube.com/playlist?lis...
youtube.com
ACDL Summer School Lectures on Planning/Reasoning Abilities of LLMs/LRMs - YouTube
132
Reposted by Subbarao Kambhampati (కంభంపాటి సుబ్బారావు)
Melanie Mitchell @melaniemitchell.bsky.social · 09/06/2025
...it basically confirmed what is already well-established: LLMs (& LRMs & "LLM agents") have trouble w/ problems that require many steps of reasoning/planning. See, e.g., lots of recent papers by Subbarao Kambhampati's group at ASU. (2/2)
2525
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 16/06/2025
An AGI-wannabe reasoning model whining that it couldn't handle a problem because its context window isn't big enough is like a superman-wannabe little kid protesting that he couldn't add those numbers because he doesn't have enough fingers and toes.. #AIAphorisms
030
Reposted by Subbarao Kambhampati (కంభంపాటి సుబ్బారావు)
Dr Abeba Birhane @abeba.blacksky.app · 01/06/2025
"our counter-intuitive results demonstrate ways in which common interpretations of Large Reasoning Models may be anthropomorphizations or simplifications" arxiv.org/abs/2505.13775
arxiv.org
Beyond Semantics: The Unreasonable Effectiveness of Reasonless Intermediate Tokens
Recent impressive results from large reasoning models have been interpreted as a triumph of Chain of Thought (CoT), and especially of the process of training on CoTs sampled from base LLMs in order to...
25411
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 30/05/2025
The transformer expressiveness results are often a bit of a red herring as there tends to be a huge gap between what can be expressed in transformers, and what can be learned with gradient descent. Mind the Gap, a new paper with Lucas Saldyt dives deeper into this issue 👇👇 x.com/SaldytLucas/...
x.com
Lucas Saldyt on X: "Neural networks can express more than they learn, creating expressivity-trainability gaps. Our paper, “Mind The Gap,” shows neural networks best learn parallel algorithms, and analyzes gaps in faithfulness and effectiveness. @rao2z https://t.co/8YjxPkXFu0" / X
Neural networks can express more than they learn, creating expressivity-trainability gaps. Our paper, “Mind The Gap,” shows neural networks best learn parallel algorithms, and analyzes gaps in faithfulness and effectiveness. @rao2z https://t.co/8YjxPkXFu0
020
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 28/05/2025
Anthropomorphization of intermediate tokens as reasoning/thinking traces isn't quite a harmless fad, and may be pushing LRM research into questionable directions.. So we decided to put together a more complete argument. Paper 👉 arxiv.org/pdf/2504.09762 (Twitter thread: x.com/rao2z/status...)
0101
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 25/05/2025
This RLiNo? paper (arxiv.org/abs/2505.13697) lead by Soumya Samineni and Durgesh_kalwar dives into the MDP model used in the RL post-training methods inspired by DeepSeek R1, and asks if some of the idiosyncrasies of RL aren't just consequences of the simplistic structural assumptions made
140
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 21/05/2025
Do Intermediate Tokens Produced by LRMs (need to) have any semantics? Our new study 👇 Thread 👉 x.com/rao2z/status...
280
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 13/05/2025
Delighted to share that Siddhant Bhambri & Mudit Verma's critical evaluation and refutation of the reasoning claims of ReACT has been accepted to #TMLR (Transactions on Machine Learning) 👉https://openreview.net/forum?id=aFAMPSmNHR
141
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 12/05/2025
Solving Single Agent Fully Observable Deterministic (SAFODP) Problems with Dec-POMDP approaches #SundayHarangue #allegory x.com/rao2z/status...
x.com
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) on X: "Solving Single Agent Fully Observable Deterministic (SAFODP) Problems with Dec-POMDP approaches #SundayHarangue #allegory Imagine you modeled your decision problem into a Dec-POMDP problem ('cuz that's as expressive a decision model as you can get! )--but with some https://t.co/vDWYTHbnQA" / X
Solving Single Agent Fully Observable Deterministic (SAFODP) Problems with Dec-POMDP approaches #SundayHarangue #allegory Imagine you modeled your decision problem into a Dec-POMDP problem ('cuz that's as expressive a decision model as you can get! )--but with some https://t.co/vDWYTHbnQA
020
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 09/05/2025
IMHO, the whole idea of connecting "length of intermediate tokens" produced by LRMs to inference time compute is a mind-boggling demonstration of circular reasoning--that comes from the assumptions about MDP model and reward model.. 👇 x.com/rao2z/status...
x.com
130
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 05/05/2025
It ain't "The Bitter Lesson" if you are in the loop curating the training data for your LLM, y'all.. Pick your lesson, will ya? #SundayHarangue (h/t @kstechly.bsky.social)
043
Reposted by Subbarao Kambhampati (కంభంపాటి సుబ్బారావు)
Adrian Chan @gravity7.bsky.social · 19/04/2025
Don't use summarizers for the papers by @rao2z.bsky.social because the reasoning traces therein are, unlike the LRMs & LLMs under investigation, substantively meaningful, semantically well-ordered, and stylistically compelling and engaging! #AI #LLMs #CoT arxiv.org/abs/2504.09762
arxiv.org
(How) Do reasoning models reason?
We will provide a broad unifying perspective on the recent breed of Large Reasoning Models (LRMs) such as OpenAI o1 and DeepSeek R1, including their promise, sources of power, misconceptions and limit...
182
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 16/04/2025
Here is a recording of my talk at @msftresearch.bsky.social last week titled "(How) Do LLMs Reason/Plan?" (Also gave a version of it at as a distinguished lecture at Oracle today..) www.youtube.com/watch?v=0u2h...
youtube.com
(How) Do LLMs Reason/Plan? (Talk given at Microsoft Research; 4/11/25)
YouTube video by Subbarao Kambhampati
051
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 15/04/2025
A preprint available at arxiv.org/abs/2504.09762
030
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 13/04/2025
Our invited commentary for the Annals of New York Academy of Sciences titled "(How) Do reasoning models reason?" is now online 👉 nyaspubs.onlinelibrary.wiley.com/doi/epdf/10.... It is a written version of my recent talks (and #SundayHarangues) on the recent developments in LRMs..
122
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 09/04/2025
Woo hoo.. Our first #TMLR paper!🤗 On the planning and scheduling abilities of LRMs o1 & R1 (w/ Karthik, Kaya, Atharva) 👉 openreview.net/forum?id=FkK... Even a jaded researcher like me has to admit that Transactions on Machine Learning Research is a veritable oasis among #AI publication venues! 🙏
071
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 30/03/2025
AI Hype: The phenomenon where experts without expertise hype up imminent arrival of expertise without experts. #AIAphorisms
0104
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 16/03/2025
Test-time-scaling, Post-training and Distillation are just compiling the verifier signal into the LLM at different phases #SundayHarangue See 👉 x.com/rao2z/status... Or 👉 www.linkedin.com/posts/subbar...
060
Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) @rao2z.bsky.social · 10/03/2025
Intermediate tokens being dubbed as "Reasoning Traces" is the new anthropomorphization fashion.. See this video #SundayHarangue 👉https://youtube.com/watch?v=CQ5JS3v61Ns&list=PLNONVE5W8PCRbf3WmbcqgXPToJuA2NUfP&t=3787s that wonders whether LRMs should instead be called LMMs--Large Mumbling Models..
291