Sign in

Benno Krojer

@bennokrojer.bsky.social
2.7K followers 1K following 1.8K posts

AI PhDing at Mila/McGill. Happily residing in Montreal 🥯❄️ Academic stuff: language grounding, vision+language, interp, rigorous & creative evals, cogsci Other: many sports, urban explorations, puzzles/quizzes bennokrojer.com

PostsRepliesMedia
Benno Krojer @bennokrojer.bsky.social · 05/06/2026
If I had to summarize this paper in one philosophical question: Why do we call some things by the same name and others by different names? Related to old philosophy of logic/language is Quine's work (1960): en.wikipedia.org/wiki/Inscrut...
en.wikipedia.org
Inscrutability of reference - Wikipedia
040
Benno Krojer @bennokrojer.bsky.social · 05/06/2026
My first last-author paper is out! If you saw this dog below and someone showed you the second image, would you consider them the same word/concept? (more examples in Ada's thread) We study if VLMs agree with humans on this and revisit old questions around shape vs. texture bias in vision
1124
Reposted by Benno Krojer
Elinor @elinorpd.bsky.social · 21/04/2026
I'll be presenting OvertonBench at #ICLR2026 in Rio later this week! 📍Sat, Apr 25, 10:30am in Pavilion 4 (#4109) Please DM me if you'd like to chat about pluralistic / value alignment, societal impacts, epistemology, fairness, evals, etc
081
Benno Krojer @bennokrojer.bsky.social · 21/04/2026
I'll be at ICLR! Will present our LatentLens paper at the Re-Align workshop and happy to chat (ideally at the beach 🏖️)
0100
Benno Krojer @bennokrojer.bsky.social · 31/03/2026
What are your favorite papers that can serve as excellent examples how to write great scientific paper, present results, great figures, make it engaging and easy to follow? Doesn't necessarily have to be the most cited or impactful ones
260
Reposted by Benno Krojer
Elinor @elinorpd.bsky.social · 10/03/2026
Inspired by @bennokrojer.bsky.social, we included a Behind the Scenes section 🎬 The goal is to make science more transparent 🔍, share lessons learned 🧠, and provide a more realistic lens on the research journey 👣 8/ bsky.app/profile/benn...
161
Reposted by Benno Krojer
Gaurav Kamath @grvkamath.bsky.social · 04/03/2026
🚨New Paper!🚨 How do reasoning LLMs handle inferences that have no deterministic answer? We find that they diverge from humans in some significant ways, and fail to reflect human uncertainty… 🧵(1/10)
35820
Benno Krojer @bennokrojer.bsky.social · 27/02/2026
People often say (myself too): Interpretability on AI is so much easier than neuroscience! We can inspect everything and even retrain (vs carefully poke a little into the brain)! One big advantage in neuroscience I often forget: We're quite literally *inside* the thing we're studying
471
Benno Krojer @bennokrojer.bsky.social · 23/02/2026
You can now "pip install latentlens" 🔨 It comes with: * pre-computed embeddings for several popular LLMs and VLMs * a txt file with sentences describing WordNet concepts, which we recommend as a standard corpus to get embeddings from * ... Try it out and let us know what we can improve!
272
Benno Krojer @bennokrojer.bsky.social · 23/02/2026
Finally getting into this classic Let's see if by the end I'll have a clearer idea what type of science some fields of AI are, like interpretability What are our paradigms?
270
Benno Krojer @bennokrojer.bsky.social · 20/02/2026
Google decided to show this as my first sentence from my website (and not any of the sentences actually at the top of the website)
010
Reposted by Benno Krojer
Vaibhav @vaibhavadlakha.bsky.social · 11/02/2026
What does it mean for visual tokens to be "interpretable" to LLM? And how to we measure it? These, and many more pressing questions are addressed! Introducing LatentLens -- a new, more faithful tool for interpretability! Honoured to have collaborated with @bennokrojer.bsky.social on this!
051
Benno Krojer @bennokrojer.bsky.social · 11/02/2026
For every one of my papers, I try to include a "Behind the Scenes" section I think this paper in particular has a lot going on behind the scenes; from lessons learned to personal reflections let me share some
130
Benno Krojer @bennokrojer.bsky.social · 11/02/2026
This will be my last paper of the phd, can't believe it's been almost 5 years! It is the work i am most proud of and believe has the most potential. Feels right to wrap it up with this one
130
Benno Krojer @bennokrojer.bsky.social · 11/02/2026
🚨New paper Are visual tokens going into an LLM interpretable 🤔 Existing methods (e.g. logit lens) and assumptions would lead you to think “not much”... We propose LatentLens and show that most visual tokens are interpretable across *all* layers 💡 Details 🧵
1337
Reposted by Benno Krojer
Declan Campbell @thisisadax.bsky.social · 05/02/2026
The visual world is composed of objects, and those objects are composed of features. But do VLMs exploit this compositional structure when processing multi-object scenes? In our 🆒🆕 #ICLR2026 paper, we find they do – via emergent symbolic mechanisms for visual binding. 🧵👇
18326
Reposted by Benno Krojer
Desmond Elliott @delliott.bsky.social · 04/02/2026
📢 I am hiring a highly-motivated Ph.D student at the University of Copenhagen to work on tokenization-free NLP. Read our previous work in this topic: aclanthology.org/2025.emnlp-m... aclanthology.org/2023.emnlp-m... openreview.net/forum?id=FkS... Apply by March 8: employment.ku.dk/phd/?show=1563
A photograph of sunny Copenhagen in the summer!
0209
Benno Krojer @bennokrojer.bsky.social · 04/02/2026
nooo claude's --verbose is gone/broken, how will i now catch wrong assumptions before it goes on for 10 minutes?
010
Benno Krojer @bennokrojer.bsky.social · 03/02/2026
A paper that should get more attention, for those interested in building truly multimodal models (aka not just plugging stuff post hoc into LLMs): arxiv.org/abs/2412.06646 TLDR: Counterintuitively, native multimodal modes seem *less* unified internally (interp)
arxiv.org
The Narrow Gate: Localized Image-Text Communication in Native Multimodal Models
Recent advances in multimodal training have significantly improved the integration of image understanding and generation within a unified model. This study investigates how vision-language models (VLM...
030
Reposted by Benno Krojer
Chanda Prescod-Weinstein 🌌 @chanda.blacksky.app · 01/02/2026
Epstein’s economic power among academics was made possible by a capitalist system that makes higher education dependent on the charity economy rather than a public good supported by taxing the rich
42113962918
Reposted by Benno Krojer
Elinor @elinorpd.bsky.social · 23/01/2026
🎉 Excited to share our new paper which was accepted to #AAAI2026! As LLMs become increasingly used as sources of factual knowledge, we ask: Do they perform equitably across users of different backgrounds? 🧵⬇️ 1/6
121
Benno Krojer @bennokrojer.bsky.social · 17/01/2026
Figuring out the right rules+practices for my CLAUDE.md file and generally how to prompt Claude Code, has taught me more about software/ML engineering paradigms/concepts/best-practices than anything I've learned in uni
260
Benno Krojer @bennokrojer.bsky.social · 14/01/2026
Listenining to Michelle Obama's audiobook "Becoming" (loving it btw) in 2026 is wild, an almost comical contrast to now... What hopeful times it was back then, as I'm now at the chapters (2006-2008) describing their 2008 run for president
010
Benno Krojer @bennokrojer.bsky.social · 12/01/2026
Wild output from Gemini i got two weeks ago i wonder if the appearance on tv is meant to represent the sin of pride
030
Benno Krojer @bennokrojer.bsky.social · 11/01/2026
With the latest coding agents there's almost no excuse anymore to publish a paper without any cool demo or interactive data explorer For my current paper it made it so much easier to quickly grasp different interp tools and their effects throughout the whole project
1111
Benno Krojer @bennokrojer.bsky.social · 14/12/2025
New blog post 📜 "Better late than never: Getting into interpretability in 2025" It's been a great year pivoting into interp and i wanted to reflect on it bennokrojer.com/interp.html
bennokrojer.com
Better late than never: Getting into interpretability in 2025
3160
Benno Krojer @bennokrojer.bsky.social · 22/11/2025
Really enjoyed the discussions in the UT Austin NLP group!
030
Benno Krojer @bennokrojer.bsky.social · 15/10/2025
Couldn't have wished for a better place to do my PhD, come apply!
070
Benno Krojer @bennokrojer.bsky.social · 07/10/2025
I'll be at COLM! Excited to chat about about anything vision+language, interpretability, cogsci/psych, embedding spaces, visual reasoning, video/world models
050
Benno Krojer @bennokrojer.bsky.social · 22/09/2025
Devoured this book in 18 hours, usually not a big fan of audio books! It covered lots from crowdworker rights, the ideologies (doomers, EA, ...) and the silicon valley startup world to the many big egos and company-internal battles Great work by @karenhao.bsky.social
050
Benno Krojer @bennokrojer.bsky.social · 12/09/2025
Lmao
020
Reposted by Benno Krojer
Elinor @elinorpd.bsky.social · 21/08/2025
Congratulations @bennokrojer.bsky.social on passing your PhD proposal exam! A great presentation and exciting work!
051
Benno Krojer @bennokrojer.bsky.social · 13/08/2025
very happy to see the trend of a Behind the Scenes section catching on! transparent & honest science 👌 love the detailed montreal spots mentioned consider including such a section in your next appendix! (paper by @a-krishnan.bsky.social arxiv.org/pdf/2504.050...)
181
Benno Krojer @bennokrojer.bsky.social · 29/07/2025
Super cool work on quantifying with NLP how language evolves through generations In linguistics, the "apparent time hypothesis" famously discusses this but never empirically tests it
031
Reposted by Benno Krojer
Verna Dankers @vernadankers.bsky.social · 01/07/2025
I miss Edinburgh and its wonderful people already!! Thanks to @tallinzen.bsky.social and @edoardo-ponti.bsky.social for inspiring discussions during the viva! I'm now exchanging Arthur's Seat for Mont Royal to join @sivareddyg.bsky.social's wonderful lab @mila-quebec.bsky.social 🤩
3151
Benno Krojer @bennokrojer.bsky.social · 25/06/2025
Started a new podcast with @tomvergara.bsky.social ! Behind the Research of AI: We look behind the scenes, beyond the polished papers 🧐🧪 If this sounds fun, check out our first "official" episode with the awesome Gauthier Gidel from @mila-quebec.bsky.social : open.spotify.com/episode/7oTc...
open.spotify.com
02 | Gauthier Gidel: Bridging Theory and Deep Learning, Vibes at Mila, and the Effects of AI on Art
Behind the Research of AI · Episode
1176
Benno Krojer @bennokrojer.bsky.social · 20/06/2025
Turns out condensing your research into 3min is very hard but also teaches you a lot Finally the video from Mila's speed science competition is on YouTube! From a soup of raw pixels to abstract meaning t.co/RDpu1kR7jM
090
Benno Krojer @bennokrojer.bsky.social · 13/06/2025
Excited to share the results of my recent internship! We ask 🤔 What subtle shortcuts are VideoLLMs taking on spatio-temporal questions? And how can we instead curate shortcut-robust examples at a large-scale? We release: MVPBench Details 👇🔬
1165
Benno Krojer @bennokrojer.bsky.social · 30/05/2025
Top 99% in Boston 💪 Love these interactive maps
130
Benno Krojer @bennokrojer.bsky.social · 29/05/2025
When you forget your leftover rice in the fridge for 3 weeks
040
Benno Krojer @bennokrojer.bsky.social · 28/05/2025
Maybe it's just that I'm now paying more attention to the good parts again but since this post bluesky seems more fun again. Still not many paper discussions going on but saw some fun general posts
150
Benno Krojer @bennokrojer.bsky.social · 27/05/2025
It was tough just logging back into my retired twitter account and to see a timeline that is so much fuller with interesting research discourse... And yes I've tried to customizing my feeds and whatnot but no feed can fix a lack of posts
030
Benno Krojer @bennokrojer.bsky.social · 22/05/2025
I finish my work day with the conclusion that code assistants are maybe net negative for my short-term progress and most likely negative for my long-term progress and learning Also sycophancy is annoying af
030
Benno Krojer @bennokrojer.bsky.social · 19/05/2025
Attend my AI 2025 bootcamp
2112
Reposted by Benno Krojer
Ted Underwood @tedunderwood.com · 12/05/2025
This pattern is going to repeat in one domain after another, and gradually force us to admit that 60% of every job is networking, knowing who to trust, and doing poorly justified risk/benefit assessment.
2709
Benno Krojer @bennokrojer.bsky.social · 09/05/2025
with great responsibility comes a great amount of reimbursements to file
040
Benno Krojer @bennokrojer.bsky.social · 06/05/2025
are there any analysis/interp papers on image editing? papers that are more insights than performance
000
Benno Krojer @bennokrojer.bsky.social · 05/05/2025
Day 13: (the original mega-thread has become too long and nested so reposting now as a new strategy) Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models A few notes below 👇 I took less digital notes this time as i was sitting outside in the sun reading 🌞
161
Reposted by Benno Krojer
Casey Newton @caseynewton.bsky.social · 01/05/2025
This is one of the most-shared posts on Bluesky in the past day and it's just completely false. You might think ChatGPT is a *bad* search engine, or prefer another search engine. But it has had integrated web search since last year.
Skeet from Ann Leckie reading: "Say it after me: Chat GPT is not a search engine. It does not scan the web for information, it just generates statistically likely sentences. You cannot use it a search engine, or as a substitute for searching.

Now. Please never use an LLM for information searches ever again."
861878197
Benno Krojer @bennokrojer.bsky.social · 01/05/2025
A must-read for anyone in NLP right now
161