Sign in

Keshav Ramji

@keshavramji.bsky.social
961 followers 227 following 13 posts

working toward continually self-improving AI, reasoning and alignment @ IBM Research AI | Prev: Penn CS + Wharton

PostsRepliesMedia
Reposted by Keshav Ramji
Ramon Astudillo @ramon-astudillo.bsky.social · 18/08/2026
> the token usage is 2.3x GPT Luna Max and almost 2x Kimi K3! Imagine this is not benchmaxxed and test-time compute trade-off can be stretched to this level ... and text can definitely be compressed 🙂 cc @keshavramji.bsky.social arxiv.org/abs/2604.22709
arxiv.org
Thinking Without Words: Efficient Latent Reasoning with Abstract Chain-of-Thought
While long, explicit chains-of-thought (CoT) have proven effective on complex reasoning tasks, they are costly to generate during inference. Non-verbal reasoning methods have emerged with shorter gene...
091
Keshav Ramji @keshavramji.bsky.social · 29/04/2026
Thanks for sharing our work! Please check it out -- we have a Twitter (X) thread as well, bsky thread coming soon!
140
Reposted by Keshav Ramji
Sung Kim @sungkim.bsky.social · 28/04/2026
Does a LLM really need to think in English or Chinese? How about it thinks using a short sequence of reserved "abstract" tokens through reinforcement learning? They find out that it is as performant as verbalized CoT at a fraction of the cost, achieving major gains in inference-time efficiency.
1221325
Keshav Ramji @keshavramji.bsky.social · 08/10/2025
If these topics excite you, reach out about joining us next summer!
000
Keshav Ramji @keshavramji.bsky.social · 23/05/2025
Excited to share our new paper on language model self-improvement! Paper: arxiv.org/abs/2505.16927 We introduce Self-Taught Principle Learning (STaPLe), a new approach for LMs to generate their own constitutions, by learning the principles that are most effective to self-correct their responses.
141
Keshav Ramji @keshavramji.bsky.social · 24/04/2025
I'm at #ICLR2025 🇸🇬 and will be presenting Conformal Language Model Reasoning with Coherent Factuality (arXiv to come soon) this afternoon (4/24, poster session 2)! This work is with my amazing collaborators Max Rubin-Toles, Maya Gambhir, @aaroth.bsky.social, and @surbhigoel.bsky.social!
141
Reposted by Keshav Ramji
Conference on Language Modeling @colmweb.org · 17/12/2024
Announcement #1: our call for papers is up! 🎉 colmweb.org/cfp.html And excited to announce the COLM 2025 program chairs @yoavartzi.com @eunsol.bsky.social @ranjaykrishna.bsky.social and @adtraghunathan.bsky.social
06624
Keshav Ramji @keshavramji.bsky.social · 12/12/2024
I'll be at NeurIPS tomorrow and Saturday 🇨🇦! DM me if you're working on alignment, reasoning, data-centric methods, or uncertainty quantification and would like to chat!
090