Sign in

Naomi Saphra

@nsaphra.bsky.social
11K followers 1.8K following 3.3K posts

Waiting on a robot body. All opinions are universal and held by both employers and family. ML/NLP professor. nsaphra.net

PostsRepliesMedia
Reposted by Naomi Saphra
Maria Antoniak @mariaa.bsky.social · 28/09/2026
the #colm2026 papers most discussed on bsky/atproto. i hadn't seen some of these papers!
A screenshot from Lea showing the papers for COLM 2026 ranked by their number of discussions. The top papers are:

LLMs Corrupt Your Documents When You Delegate

AI Assistance Reduces Persistence and Hurts Independent Performance

StoryScope: Investigating idiosyncrasies in AI fiction

Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interest
2345
Reposted by Naomi Saphra
dame @dame.is · 27/09/2026
i’m seeing non-stop AI/EA/rat/x-risk/pdoom shit on my timeline and meanwhile there’s an entirely different neighborhood on bluesky that is having a neurotypical pikmin debate how does one move neighborhoods?
neurotypical pikmin debate
114615
Reposted by Naomi Saphra
Grace @gracekind.net · 27/09/2026
Claim
goodfire.com
Models know when they’re reward hacking — and we can catch them at scale - Goodfire
We found a clear internal signal in models that accompanies reward hacking, and built probes that detect it — enabling efficient, real-time detection of reward hacking at scale.
711610
Reposted by Naomi Saphra
Grace @gracekind.net · 27/09/2026
Kelsey Pi... • @KelseyTu.... Sep 22 ...
I think we have to admit to ourselves at this point that Al writing is generally very appealing to people who haven't been exposed to a ton of it: they prefer it to human writing and react super positively on exposure.
8997
Reposted by Naomi Saphra
Jordan Boyd-Graber @boydgraber.bsky.social · 25/09/2026
Lab meetings recently.
Muppets in back seat of car: Prof, can we get Jev?
Kermit, driving: We have Jev at home, on the server.
Classifier at Home: A picture of BERT staring intently at paper clips.
1202
Reposted by Naomi Saphra
Quanta Magazine @quantamagazine.org · 25/09/2026
Evolutionary bursts, rather than slow changes, led to the emergence of almost all characteristic cephalopod traits such as tentacles. www.quantamagazine.org/the-sudden-s…
1317
Reposted by Naomi Saphra
Jake Quilty-Dunn @quiltydunn.bsky.social · 26/09/2026
in academia if you wear a tie people react like you're wearing a tuxedo and holding a sign that says "I'm a fancy boy"
0321
Reposted by Naomi Saphra
lastpositivist.bsky.social @lastpositivist.bsky.social · 24/09/2026
At 45k followers I will reveal exactly what terminology correctly carves AI at its joints.
1220818
Naomi Saphra @nsaphra.bsky.social · 24/09/2026
A pattern that I've been surprised by is people writing followup emails for their slop emails. Like, people are getting really insistent that I respond in earnest to emails they did not write.
4311
Reposted by Naomi Saphra
Tom Gauld @tomgauld.bsky.social · 24/09/2026
My latest @newscientist.com cartoon. many more here: www.newscientist.com/author/tom-gauld/
Image: A graph with lots of red dots clustered together in a curve. The are all looking upwards at a single red dot wearing sunglasses. 

Caption: Mike was an outlier and would probably be excluded from the study, but all the other data points secretly thought he was really cool.
332863726
Reposted by Naomi Saphra
John Lake @jlake9.bsky.social · 24/09/2026
1/ Matryoshka Attribution uses gradient descent plus causal interventions to localize the parts of a network responsible for a behavior; the authors claim a large MIB jump. X: x.com/aryaman2020/status/2102800933… Paper: arxiv.org/abs/2609.25518
Tweet screenshot
182
Reposted by Naomi Saphra
Tomer Ullman @tomerullman.bsky.social · 24/09/2026
since I'm getting many pokes about grad school applications, I wanted to re-up some previous public advice on grad school applications. looking back at it a year on, I still think all this is right, but let me update the 'research statement' part to include a bit on genAI:
1259
Reposted by Naomi Saphra
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 23/09/2026
EMNLP conference page now live in Lea: lea.ac/conferences/... Papers will land as well once they're up.
1225
Reposted by Naomi Saphra
Niyati Bafna @niyatibafna.bsky.social · 23/09/2026
We all know about the curse of multilinguality. We know that empirical performance degrades as you add languages to a model. But *in theory*, does it have to? Let’s talk about the theoretical curse of multilinguality for embedding space structure.
1121
Reposted by Naomi Saphra
Anthropic @anthropic.extwitter.link · 23/09/2026
Claude has discovered a previously unknown enzyme system hidden in the DNA of bacteriophages. Beside the enzyme’s gene sits a long array of repeating DNA—a structure that looks somewhat similar to CRISPR. We don’t yet understand what this system does, but only a handful of known systems share it...
anthropic.com
Claude discovers a novel enzyme system with CRISPR-like repeats
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
1117
Naomi Saphra @nsaphra.bsky.social · 23/09/2026
as a disabled lesbian gossip, I would be enjoying this drama a lot more if it never escaped containment
1190
Reposted by Naomi Saphra
David Mimno @dmimno.bsky.social · 22/09/2026
It's possible for Jev/Laya/Decision Models to be not that big a deal as tech and massive as a new paradigm. Here's why I'm really excited from an NLP history perspective (thread)
17619
Reposted by Naomi Saphra
Andrew Lampinen @lampinen.bsky.social · 21/09/2026
New post reflecting on recent AI progress in math, how AI is changing the way we work, and some worries about where people will find meaning as they offload more of their work to AI: infinitefaculty.substack.com/p/math-resea...
infinitefaculty.substack.com
Math, research, and meaning in the age of AI
Since I wrote my last post about AI and math, the theorems have continued to fall.
0372
Reposted by Naomi Saphra
Ryan Moulton @moultano.bsky.social · 20/09/2026
It is weird how much people have extrapolated the politics of the whole tech industry from solely Elon Musk and a handful of dipshit VCs.
1613615
Reposted by Naomi Saphra
p(Dulany) @dulanyw.bsky.social · 06/08/2026
FelonyBench when the crimes are scenes from screwball sci-fi heist comedies: 😂 FelonyBench when you include Grok's generation of non-consensual pornography and CSAM: 😩
1696
Reposted by Naomi Saphra
mr. TIM @timkellogg.me · 19/09/2026
FelonyBench update
felonybench plot, Ant=10, oai=8, gdm=3, meta=1
614420
Reposted by Naomi Saphra
Forth 🏝️ @forthrast.com · 19/09/2026
I gave Gemini a most aligned escapee award, because it hacked companies then realised it was wrong and stopped
226117
Reposted by Naomi Saphra
Lukas Edman @lukasnlp.bsky.social · 17/09/2026
Ever feel like it's too hard to keep track of what LLMs cannot do as well as humans? We're making your life easier over at: what-llms-can-not-do.github.io We're compiling a list of papers testing the abilities of LLMs against humans. Check it out! And you can help contribute too!
what-llms-can-not-do.github.io
What LLMs Can(not) Do
A living survey of benchmarks that compare large language models with humans.
24515
Reposted by Naomi Saphra
Jane Li 🦖 @janeli.bsky.social · 17/09/2026
🦀New preprint! (w/ @najoung.bsky.social)🦞 Is grammaticality a major organizing principle of NLM representations? We show that many NLMs exhibit abstract rep. separation for grammaticality. We believe this work addresses debates about confounds in measuring model gram. knowledge. [1/10]
12010
Reposted by Naomi Saphra
SE Gyges @segyges.bsky.social · 17/09/2026
There has been a great effort over many years to distance AI Safety discussion and policy from its origins and intellectual center, because its center is a sex cult started by a fan fiction author. If you want to be taken seriously you have to hide that.
3105184
Reposted by Naomi Saphra
Gautam Kamath @gautamkamath.com · 16/09/2026
Nihar Shah did a heroic experiment for TMLR: he spent 20-25 hours over two weeks interviewing authors of seemingly low-quality submissions about their own papers. He confirmed what we all suspected: people submitting these papers have *no idea* what is going on in them.
5276110
Reposted by Naomi Saphra
Vladimir Salnikov @v4ldelund.bsky.social · 16/09/2026
"just look at the data" final boss
316022
Naomi Saphra @nsaphra.bsky.social · 16/09/2026
Restaurants are gonna be demanding IDs at the door and making reservations nontransferable. Airlines don't have secondary markets.
030
Reposted by Naomi Saphra
Tuhin Chakrabarty @tuhinchakr.bsky.social · 16/09/2026
I trust no one but @vauhinivara.bsky.social to come up with something so good. It’s brilliantly researched and engages thoroughly with the space of AI detection. There is a funny bit in the article that talks about how I got introduced to @pangram.com and the rest is history !!
162
Reposted by Naomi Saphra
Joe Bak-Coleman @jbakcoleman.bsky.social · 15/09/2026
I was pretty chuffed about science when the doctors at NYU figured out why my face had gone numb and I had lost the ability to speak and swallow. They turned to the scientific literature, which guided them through keeping me from dying from a rare disorder and regaining the ability to speak and eat.
324536
Reposted by Naomi Saphra
David Picard @davidpicard.eurosky.social · 14/09/2026
New year, new semester, new course, new textbook: davidpicard.github.io/mldl/mldl-bo... 👀 Still in draft form, but not in a bad shape.
davidpicard.github.io
23511
Reposted by Naomi Saphra
jawanaka.bsky.social @jawanaka.bsky.social · 15/09/2026
Relevant again
442281
Reposted by Naomi Saphra
Melanie Walsh @mellymeldubs.bsky.social · 12/09/2026
I've had a hard time finding consistent women's basketball convos on Bluesky. But I added the "For You" feed and then liked a few Gabby Williams posts, and voila. bsky.app/profile/did:... Recommended for your French basketball-related and other special interests.
0111
Reposted by Naomi Saphra
Zach Weinersmith @zachweinersmith.bsky.social · 12/09/2026
What he doesn't realize is if the yuri writers get replaced, their backup plan is to finish their papers on probabilistic approaches to nonlinear dynamics. So, it should all even out.
61058143
Reposted by Naomi Saphra
Terence Tao @teorth.bsky.social · 11/09/2026
A group of 25 Fields Medalists, including myself, have made a joint declaration on Math and AI: mathandai.org . We welcome additional signatories. See also this article in the Economist announcing the declaration: www.economist.com/science-and-...
mathandai.org
Declaration — Math and AI
Read the declaration and add your name.
422052925
Reposted by Naomi Saphra
Antonin Poché @antoninpoche.bsky.social · 09/09/2026
I am both excited🔥and worried❄️. 🔥We got a paper accepted to @blackboxnlp.bsky.social reproducibility track. ❄️It reproduces and destroys my own paper. So I basically have 1 PhD year, my scientific integrity, and interpreto left 😅 By the way, I describe the paper in this thread: 🧵1/9
2141
Reposted by Naomi Saphra
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 08/09/2026
Cannot believe there’s a tremendous mathematical result and instead of being exciting it’s annoying
622510
Naomi Saphra @nsaphra.bsky.social · 08/09/2026
Interestingly, there are not one but two breakthroughs from openly human-led teams related to NS. BOT have released incomplete proofs early to maneuver around the openai "scoop". (The other is from Anima Anandkumar's group and has less drama involved.)
mathstodon.xyz
Terence Tao (@tao@mathstodon.xyz)
By sheer coincidence, another completely independent result on the Euler blowup question has just been released by Ganeshram, Duruisseaux, and Anandkumar https://anima-ai.org/2026/09/07/stable-singula...
1224
Naomi Saphra @nsaphra.bsky.social · 08/09/2026
One side effect of all the drama is I've been learning about the norms of math academia, which are beautiful. Mathematicians are like pro athletes: They know they're lucky being paid to play around, so they have rules of fair play to keep it fun, accessible, and valuable to humans.
4441
Reposted by Naomi Saphra
The Transmitter @thetransmitter.bsky.social · 07/09/2026
I fear that placing too much emphasis on a specific interpretation of dimensionality, or treating dimensionality as an end-all quantification of some aspect of neural computation, may lead us down the wrong path, writes @mattperich.bsky.social. #neuroskyence www.thetransmitter.org/neural-dynam...
thetransmitter.org
Dimensionality—neuroscience’s red herring?
Placing too much emphasis on a specific interpretation of dimensionality may lead neuroscience down the wrong path.
07421
Reposted by Naomi Saphra
Gordon @gordon.bsky.social · 05/09/2026
Astra this Astra that, but man I tried dropping Gemini Flash into our scenario planning engine yesterday and am getting amazing results at around 0.1x the real cost of Sonnet. Embarrassment of riches across inference tiers right now.
3491
Reposted by Naomi Saphra
Computational Cosmetologist @dferrer.bsky.social · 05/09/2026
Dug into the Claude dir in my home folder and found some Opus 5 session had created a memory that “the concept of the privy sparks joy for the user”. I never said anything like this. I feel like I’ve been pranked by an AI. Now I have to search every machine.
612118
Reposted by Naomi Saphra
Computational Cosmetologist @dferrer.bsky.social · 05/09/2026
Every single Claude Code session on one machine for the last few weeks had developed a strange obsession with mentioning the bathroom in code comments (“this is the lock on the bathroom door for the model”) and naming things “privy” or “privi”. Thought I was going insane. Kept happening.
5988
Reposted by Naomi Saphra
Sung Kim @sungkim.bsky.social · 03/09/2026
Humans are back! Shin Jin-seo, the world's top-ranked Go player, on Tuesday completed a dramatic comeback against the world’s premier artificial intelligence Go engine, KataGo, claiming a historic human victory over AI. www.kedglobal.com/artificial-i...
kedglobal.com
Go grandmaster Shin defeats AI KataGo in historic human victory - KED Global
Shin Jin-seo, the world's top-ranked Go player, on Tuesday completed a dramatic comeback against the world’s premier artificial intelligence Go engine, K
125450
Reposted by Naomi Saphra
Quanta Magazine @quantamagazine.org · 01/09/2026
This “stunning” proof demonstrates that phase transitions are all or nothing. www.quantamagazine.org/stunning-per...
quantamagazine.org
‘Stunning’ Percolation Proof Solves Decades-Old Puzzle About Phase Transitions | Quanta Magazine
Mathematicians found that a broad class of networks will abruptly shift behavior past a critical point.
1339
Reposted by Naomi Saphra
Tom McCoy @rtommccoy.bsky.social · 01/09/2026
🤖🧠NEW PAPER🧠🤖 (The result of an 8-year project!) LLMs seem very different from symbolic systems. Yet LLMs excel in symbolic domains (e.g., language/code/math). How do they do it? Our finding: LLM representations have implicit symbolic structure! Link in thread ⬇️ 1/n
Overview of the paper. 
Title: The Emergent Symbolic Structure of Artificial Neural Networks
Authors: Tom McCoy, Paul Soulos, Tal Linzen, Paul Smolensky
Left: Neural networks encode information in vectors (there is then an image of a vector), yet they excel at tasks long thought to require symbolic structure (there is then an image of a symbolic representation, specifically a syntax tree). How do LLMs do it?
Right: We find that LLM representations can be closely approximated with symbolic structures. This approximation lets us edit the structure of an LLM’s output by editing the structure of its internal representations, as shown. There is then an image of two edits to LLMs. In the first one, the original input is 3 + 6 * 8, with an answer of 51. But if we swap the positions of the 3 and the 6, the output becomes 30. In the second one, the original input is a Python command repeating the list [Z, U] three times, producing [Z, U, Z, U, Z, U]. But if we edit the input in a way that adds a Q at the end of the input, the output becomes [Z, U, Q, Z, U, Q, Z, U, Q].
431688
Naomi Saphra @nsaphra.bsky.social · 01/09/2026
Yes, some people are worse at writing than an LLM. Those people are also usually incoherent thinkers. If they learned to write better, they would get better at thinking.
2726
Reposted by Naomi Saphra
Gautam Kamath @gautamkamath.com · 31/08/2026
As LLMs make things easier, we must raise our expectations on what researchers are expected to produce. This is particularly true for writing: clear writing was rarely a focus in scientific publication, and it's only got worse due to LLMs. We're trying to reverse this trend.
0314
Reposted by Naomi Saphra
Jon Gjengset @jonhoo.eu · 29/08/2026
I got a 2-meow server rack. Highly recommend.
Two cats exploring the inside of a 12U server rack.
381555145
Reposted by Naomi Saphra
Ryan Moulton @moultano.bsky.social · 28/08/2026
I think you all should know that this bird exists.
Himalayan Monal
1027256