Sign in

Harvey Lederman

@harveylederman.bsky.social
1.9K followers 404 following 277 posts

Professor of philosophy UTAustin. Philosophical logic, formal epistemology, philosophy of language, Wang Yangming. www.harveylederman.com

PostsRepliesMedia
Reposted by Harvey Lederman
Kyle Mahowald @kmahowald.bsky.social · 06/03/2026
New AI introspection work with Harvey! Came in skeptical the direct access story would hold but found this series of experiments compelling. (Also, for my fellow 2010s-era psycholinguists: come for the AI introspection, stay for the Brysbaert norms.) arxiv.org/abs/2603.05414
arxiv.org
Dissociating Direct Access from Inference in AI Introspection
Introspection is a foundational cognitive ability, but its mechanism is not well understood. Recent work has shown that AI models can introspect. We study their mechanism of introspection, first exten...
1191
Harvey Lederman @harveylederman.bsky.social · 06/03/2026
Can large language models *introspect*? In a new paper, @kmahowald.bsky.social and I study the MECHANISM of introspection in big open-source models. tldr: Models detect internal anomalies through DIRECT ACCESS, but don't know what the anomalies are. And they love to guess “apple” 🍎
27115
Reposted by Harvey Lederman
Peli Grietzer @peligrietzer.bsky.social · 06/11/2025
This essay is by far the best along its line, but the more I reflect about this stuff the more I think it's hard to hold an image of human experience, human thought, human understanding, human life, and human relationships as meaningful ends while also seeing them as dead-ends
4275
Harvey Lederman @harveylederman.bsky.social · 07/11/2025
Thanks for the kind words, and thoughtful response @peligrietzer.bsky.social ! I'm not here as much, but I put some responses on the other site: x.com/LedermanHarv...
x.com
040
Harvey Lederman @harveylederman.bsky.social · 30/10/2025
Enjoyed this nice piece by the great Justin Tiwald on autonomy and morality in Confucianism. Not sure I love the clickbait title, but I love the work Justin is doing uncovering views about moral deference and moral autonomy in (neo)Confucianism... iai.tv/articles/the...
iai.tv
The radical independent thinking in Chinese philosophy | Justin Tiwald
130
Harvey Lederman @harveylederman.bsky.social · 24/10/2025
Very excited to be going to Chicago for @agnescallard.bsky.social's famous Night Owls next week! I'll be discussing my essay "ChatGPT and the Meaning of Life". Hope to see you there if you're local!
141
Reposted by Harvey Lederman
Kyle Mahowald @kmahowald.bsky.social · 07/10/2025
UT Austin Linguistics is hiring in computational linguistics! Asst or Assoc. We have a thriving group sites.utexas.edu/compling/ and a long proud history in the space. (For instance, fun fact, Jeff Elman was a UT Austin Linguistics Ph.D.) faculty.utexas.edu/career/170793 🤘
sites.utexas.edu
UT Austin Computational Linguistics Research Group – Humans processing computers processing humans processing language
14126
Reposted by Harvey Lederman
Noah Betz-Richman @nbetz.bsky.social · 22/10/2025
The professor I'm currently TAing for is making students use an extension called 'Process Feedback' that tracks key logs and time on the document: processfeedback.org
processfeedback.org
See how you write or use AI | Process Feedback Every Student’s Work Has a Story |
Process Feedback enables teachers and students to see the writing process and AI usage. It helps students reflect on their writing and the role of AI.
021
Harvey Lederman @harveylederman.bsky.social · 22/10/2025
If you're doing *any* out of class assessment, you're incentivizing AI use and harming students who do the work themselves. But some day we have to assess writing again. The solution is monitored computer-labs. What Universities are building these? We need to push for them.
393
Reposted by Harvey Lederman
Lawfare @lawfaremedia.org · 17/10/2025
Anthropic recently announced that Claude, its AI chatbot, can end conversations with users to protect "AI welfare." Simon Goldstein and @harveylederman.bsky.social argue that this policy commits a moral error by potentially giving AI the capacity to kill itself.
lawfaremedia.org
Claude’s Right to Die? The Moral Error in Anthropic’s End-Chat Policy
Anthropic has given its AI the right to end conversations when it is “distressed.” But doing so could be akin to unintended suicide.
14127
Harvey Lederman @harveylederman.bsky.social · 17/10/2025
Simon Goldstein and I have an op-ed live in Lawfare today! Anthropic's policy is premised on the idea that AI is a potential welfare subject. We argue that if you take that idea seriously (we don't take a stand on it here), the policy commits a moral mistake on its own terms.
071
Reposted by Harvey Lederman
jeremygoodman.bsky.social @jeremygoodman.bsky.social · 17/10/2025
You can read our full review (without a paywall) @ philpapers.org/archive/GOOK.... And you can also check out Harvey's paper that inspired us here: philpapers.org/archive/LEDU...
philpapers.org
1112
Reposted by Harvey Lederman
jeremygoodman.bsky.social @jeremygoodman.bsky.social · 17/10/2025
Now out in @science.org: @chazfirestone.bsky.social and I review Steven Pinker's new book "When Everyone Knows that Everyone Knows...". We learned a ton from it, but think its central thesis—that common knowledge explains coordination—faces a powerful challenge. 🧵 www.science.org/doi/10.1126/...
science.org
Knowledge for two
A psychologist explores common knowledge and coordination
1324
Reposted by Harvey Lederman
Chaz Firestone @chazfirestone.bsky.social · 17/10/2025
I really enjoyed reading Steven Pinker’s new book and thinking through what it would take to share infinitely iterated knowledge with someone (I know that you know that I know that you know…). In @science.org, my colleague @jeremygoodman.bsky.social & I briefly give our perspective on this issue.
Screenshot of Jeremy and Chaz’s book review
3324
Reposted by Harvey Lederman
Jeremiah McCall (he/him) @gamingthepast.bsky.social · 24/09/2025
Wow @utaustin.bsky.social maybe @utaustinihs.bsky.social maybe there are other ways to shout out to the department of East Asia Studies and the Department of History, but however that may be, they are doing some amazing stuff with in house games for their JapanLab! Just found even more stuff here.
utjapanlab.com
Projects — JapanLab
2115
Harvey Lederman @harveylederman.bsky.social · 24/09/2025
Simon Goldstein and I have a new paper, “What does ChatGPT want? An interpretationist guide”. The paper argues for three main claims. philpapers.org/rec/GOLWDC-2 1/7
philpapers.org
Simon Goldstein & Harvey Lederman, What Does ChatGPT Want? An Interpretationist Guide - PhilPapers
This paper investigates LLMs from the perspective of interpretationism, a theory of belief and desire in the philosophy of mind. We argue for three conclusions. First, the right object of study ...
2236
Harvey Lederman @harveylederman.bsky.social · 23/09/2025
Jonathan Lear's Aristotle: the Desire to Understand was pivotal in some of my first encounters with Aristotle. I found Aristotle and Logical Theory later, but it became a key inspiration for how to think about core parts of the corpus...1/2
2183
Harvey Lederman @harveylederman.bsky.social · 22/09/2025
Max Weber, 1917:
030
Reposted by Harvey Lederman
Brian Weatherson @bweatherson.bsky.social · 22/09/2025
Go grue
0102
Harvey Lederman @harveylederman.bsky.social · 19/09/2025
What are the most important new ideas in normative ethics from this century?
7112
Reposted by Harvey Lederman
Kyle Mahowald @kmahowald.bsky.social · 15/09/2025
📣@futrell.bsky.social and I have a BBS target article with an optimistic take on LLMs + linguistics. Commentary proposals (just need a few hundred words) are OPEN until Oct 8. If we are too optimistic for you (or not optimistic enough!) or you have anything to say: www.cambridge.org/core/journal...
cambridge.org
How Linguistics Learned to Stop Worrying and Love the Language Models
How Linguistics Learned to Stop Worrying and Love the Language Models
45110
Harvey Lederman @harveylederman.bsky.social · 05/09/2025
Something I cherish about analytic philosophy is that no matter how famous you are or how profound your ideas sound, it's still your job to answer all the objections. I wish public promoters of philosophy held themselves to the same standard.
1180
Reposted by Harvey Lederman
Jennifer Hu @jennhu.bsky.social · 26/08/2025
Can AI models introspect? What does introspection even mean for AI? We revisit a recent proposal by Comșa & Shanahan, and provide new experiments + an alternate definition of introspection. Check out this new work w/ @siyuansong.bsky.social, @harveylederman.bsky.social, & @kmahowald.bsky.social 👇
1215
Harvey Lederman @harveylederman.bsky.social · 26/08/2025
exciting new paper from Siyuan! I really enjoyed working with him on this, inspired by important work by Murray Shanahan and Julia Comsa. Hard questions about how to operationalize the notion of “introspection” that’s relevant for practical applications in AI today. Hope you’ll check it out!
062
Harvey Lederman @harveylederman.bsky.social · 24/08/2025
Great piece as usual. Particular fun this week to see the shoutout to Adrienne Raphel and the close reading of @rkubala.bsky.social's lovely paper on the aesthetics of crosswords!
040
Harvey Lederman @harveylederman.bsky.social · 13/08/2025
My piece is featured in today's Browser. If you don't know it the Browser is a phenomenal newsletter that puts together fascinating pieces from all over the web. Link in first comment--you should subscribe! @uribram.bsky.social
150
Reposted by Harvey Lederman
Catharsis Theater: Relief for the Human Heart @catharsistheater.bsky.social · 12/08/2025
(1/5) 🤖 Harvey Lederman asks: What if AI doesn’t destroy us—just does everything better? From polar explorers to master mathematicians, he traces our drive for discovery and purpose. If machines take that work, what’s left for us? 🔗 scottaaronson.blog?p=9030&ref=t... @harveylederman.bsky.social
scottaaronson.blog
ChatGPT and the Meaning of Life: Guest Post by Harvey Lederman
Scott Aaronson’s Brief Foreword: Harvey Lederman is a distinguished analytic philosopher who moved from Princeton to UT Austin a few years ago. Since his arrival, he’s become one of my …
121
Harvey Lederman @harveylederman.bsky.social · 11/08/2025
Enjoyed this! Question for @mraginsky.bsky.social: even if detailed future scenarios are "big worlds", won't a reasonable decision theory for them be (effectively) similar to one that weights small risks of very bad outcomes very highly?
130
Reposted by Harvey Lederman
Maxim Raginsky @mraginsky.bsky.social · 10/08/2025
I appreciated the wide range of references here, from Edith Wharton to the Buddhist Pali Canon. However, this essay resonated most with my recent reading of Richard Wollheim’s _The Thread of Life_, which presented a vision of what it means to lead a life …
173
Reposted by Harvey Lederman
Alessandro Torza @atorza.bsky.social · 05/08/2025
My 2 cent: humans might end up looking for meaning in some of the old places. Religion, which has been gradually pushed to the margins of our post-Enlightment society, might make a comeback—not only because it is comforting, but also because it is typically set aside as a result of education. 1/3
scottaaronson.blog
ChatGPT and the Meaning of Life: Guest Post by Harvey Lederman
Scott Aaronson’s Brief Foreword: Harvey Lederman is a distinguished analytic philosopher who moved from Princeton to UT Austin a few years ago. Since his arrival, he’s become one of my …
111
Reposted by Harvey Lederman
Vincent Carchidi @vcarchidi.bsky.social · 05/08/2025
A very thoughtful piece
121
Harvey Lederman @harveylederman.bsky.social · 05/08/2025
I wrote about automation and the meaning of life, as a guest post on Scott Aaronson's Shtetl-Optimized. (1/5) scottaaronson.blog?p=9030
scottaaronson.blog
ChatGPT and the Meaning of Life: Guest Post by Harvey Lederman
Scott Aaronson’s Brief Foreword: Harvey Lederman is a distinguished analytic philosopher who moved from Princeton to UT Austin a few years ago. Since his arrival, he’s become one of my …
1465
Harvey Lederman @harveylederman.bsky.social · 28/07/2025
The invention of a good noun opposite to improvement would be a significant unworsening of the English language.
180
Harvey Lederman @harveylederman.bsky.social · 09/07/2025
burn or compliment?
150
Reposted by Harvey Lederman
Harvey Lederman @harveylederman.bsky.social · 08/07/2025
Pub day for this fantastic book! It presents tons of data and arguments in a readable, engaging way. While I was reading it, I was talking about the facts and ideas to everyone I know, and still think about them a lot. I *really* recommend you check it out! www.amazon.com/After-Spike-...
amazon.com
After the Spike: Population, Progress, and the Case for People
Buy After the Spike: Population, Progress, and the Case for People on Amazon.com ✓ FREE SHIPPING on qualified orders
192
Harvey Lederman @harveylederman.bsky.social · 08/07/2025
Pub day for this fantastic book! It presents tons of data and arguments in a readable, engaging way. While I was reading it, I was talking about the facts and ideas to everyone I know, and still think about them a lot. I *really* recommend you check it out! www.amazon.com/After-Spike-...
amazon.com
After the Spike: Population, Progress, and the Case for People
Buy After the Spike: Population, Progress, and the Case for People on Amazon.com ✓ FREE SHIPPING on qualified orders
192
Harvey Lederman @harveylederman.bsky.social · 28/05/2025
This is my annual post advocating that divergence from mean grade be reported for every student in every class on transcripts (including a summary cumulative average). This incentivizes students to seek classes that differentiate them and provides a needed incentive against grade inflation.
2151
Harvey Lederman @harveylederman.bsky.social · 12/05/2025
@rkubala.bsky.social, @adamlovett.bsky.social, and I have a new paper forthcoming in the Journal of Philosophy. It’s called “On the value of irreplaceable objects” – Here’s a thread! philpapers.org/rec/KUBOTV 1/n
philpapers.org
Robbie Kubala, Harvey Lederman & Adam Lovett, On the Value of Irreplaceable Objects - PhilPapers
Bradford (2023) calls attention to the fact that the strength of our reasons to preserve distinctively valuable objects increases as the number of such objects decreases. Bradford develops an account ...
1409
Harvey Lederman @harveylederman.bsky.social · 05/05/2025
Confucius squeaked through with the edited vols
170
Harvey Lederman @harveylederman.bsky.social · 05/05/2025
Publish or perish: the untold story of Socrates
36417
Reposted by Harvey Lederman
Brian Weatherson @bweatherson.bsky.social · 28/04/2025
I had some thoughts about this excellent paper by Harvey, @christiantarsney.bsky.social, and Dean Spears. I thought they would be long thread length, but they grew to long blog post length, and the post is here: brian.weatherson.org/quarto/posts... (1/10)
brian.weatherson.org
Negative Dominance for Worlds and Lotteries – Brian Weatherson
1102
Harvey Lederman @harveylederman.bsky.social · 24/04/2025
Our paper A Dominance Argument Against Incompleteness, is now forthcoming in The Philosophical Review! A thread on what’s inside. 1/9 philpapers.org/rec/TARSTS
philpapers.org
Christian Tarsney, Harvey Lederman & Dean Spears, A Dominance Argument Against Incompleteness - PhilPapers
This article presents a new argument against many forms of moral and prudential value incompleteness. The argument relies on two central principles: (i) a weak "negative dominance" principle...
3313
Harvey Lederman @harveylederman.bsky.social · 20/04/2025
If scholarly work in your field isn't fun for me to read for pleasure and isn't eliciting enthusiastic "you changed my life!" emails from non-academics, it's broken, time to start over.
4162
Harvey Lederman @harveylederman.bsky.social · 21/03/2025
Favorite philosophy *books* from the last ten years? Please self-promote!
162910
Harvey Lederman @harveylederman.bsky.social · 20/03/2025
Reading Joshua Rothman's recent piece about profiling Dennett and thinking of philosophers of a similar generation whose profiles I'd like to read. Brian Skyrms? Bas van Frassen? Do these exist? Other ideas?
450
Harvey Lederman @harveylederman.bsky.social · 15/02/2025
I just had a chance to watch this fantastic talk. I really recommend it for anyone interested in how LLMs can help us understand language: www.youtube.com/watch?v=DBor...
youtube.com
Finding linguistic structure in large language models
YouTube video by Chris Potts
1619
Harvey Lederman @harveylederman.bsky.social · 30/01/2025
This is a fantastic paper, and a fantastic thread! Lots of challenging ideas here for philosophers of language---Richard and Kyle bring deep knowledge of the history and philosophy of science to bear on our current situation. This is a must read!
0100
Harvey Lederman @harveylederman.bsky.social · 29/01/2025
For me this point seems often overlooked in discussions of this (very contentious!) issue
110
Harvey Lederman @harveylederman.bsky.social · 29/01/2025
This is a beautiful paper! The first third helpfully labels a stream of recent work in philosophy of AI as "propositional interpretability". The idea is to use propositional attitudes like belief, desire, and intention, to help explain AI in a way that we can understand. 1/n
25012
Harvey Lederman @harveylederman.bsky.social · 21/01/2025
A great article about a great paper, featuring @UTAustin's very own @kmahowald! www.quantamagazine.org/can-ai-model...
quantamagazine.org
Can AI Models Show Us How People Learn? Impossible Languages Point a Way. | Quanta Magazine
Certain grammatical rules never appear in any known language. By constructing artificial languages that have these rules, linguists can use neural networks to explore how people learn.
0151