Sign in

Ivan Kartáč

@ivankartac.bsky.social
206 followers 456 following 93 posts

Researcher in NLP & computational linguistics. PhD student @ Charles University, Prague. Working on evaluation, explainability, and reasoning. ivankartac.github.io

PostsRepliesMedia
Reposted by Ivan Kartáč
Saad Mahamood @saad.me.uk · 22h
I am pleased to announce that our @nejlt.bsky.social letter looking at the secular trends of LLMs and Natural Language Generation research has been published today: nejlt.ep.liu.se/article/view...
nejlt.ep.liu.se
The Future of Natural Language Generation in the Age of Large Language Models | Northern European Journal of Language Technology
121
Reposted by Ivan Kartáč
Vilém Zouhar @zouhar.bsky.social · 04/09/2026
Machine translation is not solved and it will take a while for it to be done arxiv.org/abs/2609.04173
arxiv.org
Last Translation Benchmark
For scientific progress, we need benchmarks that test the limits of state-of-the-art models, and evaluation methods that inform us about failure cases. As models get stronger, standard benchmarks for ...
34612
Ivan Kartáč @ivankartac.bsky.social · 03/09/2026
Actually I forgot to mention that this was a swarm of taggers
020
Ivan Kartáč @ivankartac.bsky.social · 03/09/2026
When writing a position paper, it must feel really good to finally be allowed to put all those exclamation marks in an academic text
010
Reposted by Ivan Kartáč
Ben Recht @beenwrekt.bsky.social · 30/08/2026
The overreaction to this Hugging Face thing is driving me nuts. Come flay me for being wrong but... 1/x
14436126
Reposted by Ivan Kartáč
Victoria Bosch @initself.bsky.social · 28/08/2026
Are brains and artificial neural networks converging onto universal representations? There is a seductive idea making the rounds in NeuroAI / machine learning: train systems well enough, and they all converge on the same representation of reality (i.e. a unique world model). We have thoughts™ 1/n
cell.com
The Umwelt Representation Hypothesis: rethinking Universality
Recent studies reveal striking representational alignment between artificial neural networks (ANNs) and biological brains, leading to proposals that all sufficiently capable systems converge on univer...
818176
Reposted by Ivan Kartáč
Tom McCoy @rtommccoy.bsky.social · 20/08/2026
Since many are starting grad school soon, let me re-share my One Big Tip™️ for research! Research involves many skills - collaborating, writing, presenting, etc. But many of these skills can be unified under a single overarching ability: theory of mind Blog post link in reply
Illustration of the blog post's main argument, summarized as: "Theory of Mind as a Central Skill for Researchers: Research involves many skills.If each skill is viewed separately, each one takes a long time to learn. These skills can instead be connected via theory of mind – the ability to reason about the mental states of others. This allows you to transfer your abilities across areas, making it easier to gain new skills."
28019
Ivan Kartáč @ivankartac.bsky.social · 25/08/2026
Fun fact: EMNLP proceedings had only 14 papers thirty years ago aclanthology.org/events/emnlp...
aclanthology.org
Conference on Empirical Methods in Natural Language Processing (1996) - ACL Anthology
292
Reposted by Ivan Kartáč
Julian Togelius @togelius.bsky.social · 22/08/2026
A personal essay about how I’ve been feeling and thinking about this new technology that I’m contributing to and what it might to do to us all. togelius.blogspot.com/2026/08/losi...
togelius.blogspot.com
Losing my religion
In spring 2025 I had a crisis of faith. I thought about what the technology I'm helping to create might do to our future, and got scared. M...
99522
Reposted by Ivan Kartáč
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 19/08/2026
Still think being able to understand the algorithm you're interacting with is very underrated as important www.eugenevinitsky.com/posts/audita...
eugenevinitsky.com
Quit social media that won't tell you their algorithm
Opaque feeds shape what you believe and you can never check whether they are working for you. We should insist on auditable algorithms, and support the platforms willing to build them.
58112
Reposted by Ivan Kartáč
Ben Recht @beenwrekt.bsky.social · 17/08/2026
Microconferences: A modest proposal for alternative systems to generate, evaluate, and share knowledge.
argmin.net
Microconferences
A proposal to create alternative systems for generating, evaluating, and sharing knowledge
2508
Reposted by Ivan Kartáč
Gautam Kamath @gautamkamath.com · 18/08/2026
I saw someone post a dozen+ AI slop papers purporting to solve niche open problems. This is antisocial behaviour and worse than if the problems stayed open. 0 people understand the solution, and there is no readable writeup. Also, incentive for either is removed.
2363
Ivan Kartáč @ivankartac.bsky.social · 17/08/2026
Reading generic LLM-written text is so painful. Not sure if it's just how the most recent models use language, but regardless of the ideas presented, I find it really hard to focus on semantics when I see the text is generated/rewritten by an LLM.
120
Ivan Kartáč @ivankartac.bsky.social · 15/08/2026
It’s nice that Prover9 now has a new version after so many years (prover9.org), but they should have kept the logo
Original Prover9 logo
010
Reposted by Ivan Kartáč
Ivan Kartáč @ivankartac.bsky.social · 29/07/2026
We should make the topic of climate change more sci-fi, so that Silicon Valley EAs and rationalists can finally pivot to it.
021
Ivan Kartáč @ivankartac.bsky.social · 13/08/2026
It’s great to see this approach to teaching about AI evaluation. Isolated capabilities are definitely interesting, but we should be doing more evaluations in real-world settings and focus on whether systems actually help users with their tasks.
020
Reposted by Ivan Kartáč
Maria Antoniak @mariaa.bsky.social · 12/08/2026
Check it out! Lea is live and open to the public 🥳 Lea is an experiment to build better social media by researchers and for researchers, where we own our own data and can build our own feeds, events, and communities. This is still very much a proof-of-concept; we welcome ideas + collaborators!
2024285
Ivan Kartáč @ivankartac.bsky.social · 12/08/2026
Hyped to try this out!
030
Reposted by Ivan Kartáč
Paper Skygest Team @paper-feed.bsky.social · 19/08/2025
**Please repost** If you're enjoying Paper Skygest -- our personalized feed of academic content on Bluesky -- we'd appreciate you reposting this! We’ve found that the most effective way for us to reach new users and communities is through users sharing it with their network
2114142
Reposted by Ivan Kartáč
Timothee Mickus @linguistickus.bsky.social · 10/08/2026
I was asked to write down my thoughts about ACL 2026 by several different people and why I wasn't thrilled by the experience, so here's a stupidly long ass rant timotheemickus.github.io/hustle%20and...
152
Ivan Kartáč @ivankartac.bsky.social · 06/08/2026
My rule-based part-of-speech tagger hacked two companies last night!
031
Reposted by Ivan Kartáč
Sireesh Gururaja @siree.sh · 04/08/2026
It has been *wild* to me to see the way that people in my department have fully gone back to business-as-usual posting on Twitter for their papers. There were a couple of brief blips where people tried bluesky and LinkedIn, but that's nearly all gone now 😔
161
Reposted by Ivan Kartáč
By Cory Doctorow (GPG 0xBF3D9110957E5F4C) @doctorow.pluralistic.net · 01/08/2026
Neoclassical econ assumes rationality. The corollary of, "If you're so smart, why aren't you rich?" is "you're rich, so you must be very smart!" Thus, people assume that if powerful, well-compensated CEOs insist that "AI is changing everything," well then, *AI must be changing everything*. 1/
A sepia-toned 1950s era boardroom in which men in suits sit around a circular table. The chair of the meeting has been replaced with a man in a straitjacket, making a funny face. The heads of the remaining board-members have been replaced with robots from 1930s pulp magazines. The image has been hand-tinted.
19774303
Ivan Kartáč @ivankartac.bsky.social · 01/08/2026
Funny how they now call >100B parameter models “small”. And apparently 30B today is “nano”.
050
Reposted by Ivan Kartáč
Melanie Mitchell @melaniemitchell.bsky.social · 01/08/2026
I recommend this article about AI reasoning, where the author lets us in on his struggles w/ AI cognitive dissonance. Plus some priceless quotes from @rao2z.bsky.social. (My recommendation has *nothing* to do with the fact that I'm quoted in it too 😇) www.quantamagazine.org/is-ai-reason...
quantamagazine.org
Is AI Reasoning Right for the Wrong Reasons? | Quanta Magazine
The idea that artificial intelligence can “reason” is more intuitive than ever. But intuitions can be wrong, and the science is far from settled.
511134
Reposted by Ivan Kartáč
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 31/07/2026
Fantastic paper demonstrating how LLM editing is gradually distorting our writing and the way we think to write: arxiv.org/abs/2603.18161
arxiv.org
How LLMs Distort Our Written Language
Large language models (LLMs) are used by over a billion people globally, most often to assist with writing. In this work, we demonstrate that LLMs not only alter the voice and tone of human writing, b...
317544
Reposted by Ivan Kartáč
jeffery --dangerously-skip-permissions @jefferyharrell.bsky.social · 29/07/2026
This needs to be the post of the day, I swear.
0171
Ivan Kartáč @ivankartac.bsky.social · 29/07/2026
We should make the topic of climate change more sci-fi, so that Silicon Valley EAs and rationalists can finally pivot to it.
021
Reposted by Ivan Kartáč
Maria Antoniak @mariaa.bsky.social · 26/07/2026
This reverse chronological #NLP feed is still running, and I've made a companion feed that is ranked by engagement and recency: bsky.app/profile/did:... Feedback and ideas are welcome!
0122
Reposted by Ivan Kartáč
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 19/07/2026
The AI community failed to exit twitter. What implications can we draw about what people believe from this?
rl-blogging.leaflet.pub
What it means that the AI community can't quit twitter
Some quickly jotted thoughts working through the implications of the AI community remaining on twitter
2222231
Reposted by Ivan Kartáč
Alex Turner @turntrout.bsky.social · 15/07/2026
I resigned from Google DeepMind bc it broke its founding promise by selling AI to the military without restrictions against killer robots or mass spying. For months, I worked to stop this but watched powerful ethicists and institutions choose silence. Here's what happened. 🧵
3456118
Reposted by Ivan Kartáč
depths of wikipedia @depthsofwikipedia.bsky.social · 29/05/2026
89155594913
Reposted by Ivan Kartáč
Nils Feldhus @nfel.bsky.social · 13/07/2026
📢 Call for Papers: YNLG 2026 The Young Researchers in Natural Language Generation workshop is a 2d in-person event part of INLG 2026 @inlg.bsky.social in Utrecht 🇳🇱, with poster sessions, keynote talks, roundtable discussions, and a one-day hackathon. Due: August 10, 2026 ynlg-workshop.github.io
0109
Ivan Kartáč @ivankartac.bsky.social · 11/07/2026
My assumption has been that in academia, humanities are much less inclined to use GenAI for their work. To what extent is this true? Or are people in humanities just less willing to acknowledge the use?
101
Ivan Kartáč @ivankartac.bsky.social · 10/07/2026
I often see arxiv pre-prints in reference lists of many papers even for sources that already have published versions. This is a really nice tool (by @zdenekkasner.cz) which will automatically replace arxiv versions or fix incomplete references through DBLP API: github.com/kasnerz/reffix
081
Ivan Kartáč @ivankartac.bsky.social · 09/07/2026
This is a great intro to experimental methods: experimentology.io I think NLP has a lot to learn from psychology in this respect, especially as evaluation becomes more and more important these days.
1133
Reposted by Ivan Kartáč
Marzena Karpinska @markar.bsky.social · 08/07/2026
I think I will be posting this after each #ARR cycle: Please 🥺🙏 let's prohibit AI review writing. Otherwise, we will get lazy reviewers claiming they wrote bullet points and used AI only to put that in prose.
1194
Reposted by Ivan Kartáč
Institute of Formal and Applied Linguistics @ufal.mff.cuni.cz · 05/07/2026
#ACL2026 continues with the main conference (main + findings + demos + industry + SRW), and @ufal.mff.cuni.cz folks will present 8️⃣ papers. Stop by and check with our colleagues. All times in PDT.
882
Reposted by Ivan Kartáč
Tomer Ullman @tomerullman.bsky.social · 02/07/2026
Now out (for realz) in Cognition: "People Make Graded Judgments About The Inconceivable" (by Hu, Sosa, & me) Free preprint: www.tomerullman.org/papers/grade... Journal link: bit.ly/gradedInconCog @jennhu.bsky.social @cognitionjournal.bsky.social
17117
Reposted by Ivan Kartáč
Maria Antoniak @mariaa.bsky.social · 02/07/2026
Not attending ACL in person? Follow along via my new NLP feed! It should match all mentions of #ACL2026, #EMNLP2026, #COLM2026, #NLProc, #NLP, and more.
2319
Ivan Kartáč @ivankartac.bsky.social · 01/07/2026
Heading to San Diego for #ACL2026, where I’ll be presenting two papers (see 🧵). Stop by to chat about evaluating reasoning embedded in task-oriented dialogue, or how to use small LLMs in modular neuro-symbolic approaches to syllogistic reasoning!
180
Reposted by Ivan Kartáč
Daniel Scalena @danielsc4.it · 01/07/2026
I'd never have guessed models commit to their final answer this early, often within the first 20% of reasoning, across math/logic tasks and model families. The rest is mostly hedging that doesn't change their mind. And turns out they encode this internally, we can decode it! 🧵👇
093
Ivan Kartáč @ivankartac.bsky.social · 15/06/2026
It’s only a question of time until Anthropic makes headlines telling us they found Claude doing Zen meditation with Extended Not-thinking.
030
Reposted by Ivan Kartáč
kaijie-mo.bsky.social @kaijie-mo.bsky.social · 11/06/2026
“Dimicillin” isn’t real. We made it up. Yet many LLMs still call it an antibiotic. Across 9 models and 653 drugs, we find that drug-name affixes alone can drive pharmacological reasoning. Models often rely on morphology over facts. We trace this shortcut from behavior to mechanism. 🧵
1176
Ivan Kartáč @ivankartac.bsky.social · 11/06/2026
Do you sometimes have to explain to engineers that the main role of science is not to produce software? Once in a while I see people comment on some paper along the lines of “but it’s not efficient” or “I can’t use this in production” as if this was what research is about.
151
Reposted by Ivan Kartáč
Ehud Reiter @ehudreiter.bsky.social · 08/06/2026
New blog: I am worried by NLP research culture NLG and NLP are mostly much better in 2026 than when I got my PhD in 1990. Unfortunately research culture has gotten *worse” in this period, which really worries me as I retire. ehudreiter.com/2026/06/08/n...
ehudreiter.com
I am worried by NLP research culture
In most ways NLG and NLP are much better in 2026 than when I got my PhD in 1990. Unfortunately research culture has gotten *worse” in this period, which really worries me as I retire. We have…
1144
Reposted by Ivan Kartáč
Andrew Lampinen @lampinen.bsky.social · 26/05/2026
We've updated the preprint of our Naturalistic Computational Cognitive Science paper (arxiv.org/abs/2502.20349) — we've tried to clarify and streamline the arguments, and added some new examples: 1/5
arxiv.org
Naturalistic Computational Cognitive Science: Towards generalizable models and theories that capture the full range of natural behavior
How can cognitive science build generalizable theories that span the full scope of natural situations and behaviors? We argue that progress in Artificial Intelligence (AI) offers timely opportunities ...
13415
Reposted by Ivan Kartáč
Marzena Karpinska @markar.bsky.social · 23/05/2026
this is how massive illusion of 'creativity' gets crushed... please read it to understand why models may appear to produce coherent text but are in fact Frankenstein factory ...
052
Reposted by Ivan Kartáč
Martin Haspelmath @haspelmath.bsky.social · 23/05/2026
Maybe one of the biggest obstacles for progress in science comes from entrenched stereotypes? In linguistics, we have, for example, (1) the word stereotype, (2) the grammar/dictionary stereotype, (3) the building-block stereotype, and (4) the speaker directionality stereotype dlc.hypotheses.org/4343
dlc.hypotheses.org
Four stereotypes that have guided morphosyntactic thinking
Thinking about language structures is made difficult not only by their incredible complexity, but also by entrenched ways of thinking about grammatical and lexical patterns. Linguists do not investiga...
1143
Reposted by Ivan Kartáč
Manoel Horta Ribeiro @manoelhortaribeiro.bsky.social · 19/05/2026
In a new blog post, I argue that the anti-ai movement ought to distinguish between claims about the technology and the "project of AI," as defined by Vetsi et al. in their new paper. 🔗: doomscrollingbabel.manoel.xyz/p/the-anti-a...
3416