Reposted by Naomi SaphraMaria Antoniak @mariaa.bsky.social · 28/09/2026the #colm2026 papers most discussed on bsky/atproto. i hadn't seen some of these papers! 2345
Reposted by Naomi Saphradame @dame.is · 27/09/2026i’m seeing non-stop AI/EA/rat/x-risk/pdoom shit on my timeline and meanwhile there’s an entirely different neighborhood on bluesky that is having a neurotypical pikmin debate how does one move neighborhoods? 114615
Reposted by Naomi SaphraGrace @gracekind.net · 27/09/2026Claimgoodfire.comModels know when they’re reward hacking — and we can catch them at scale - GoodfireWe found a clear internal signal in models that accompanies reward hacking, and built probes that detect it — enabling efficient, real-time detection of reward hacking at scale. 711610
Reposted by Naomi SaphraJordan Boyd-Graber @boydgraber.bsky.social · 25/09/2026Lab meetings recently. 1202
Reposted by Naomi SaphraQuanta Magazine @quantamagazine.org · 25/09/2026Evolutionary bursts, rather than slow changes, led to the emergence of almost all characteristic cephalopod traits such as tentacles. www.quantamagazine.org/the-sudden-s… 1317
Reposted by Naomi SaphraJake Quilty-Dunn @quiltydunn.bsky.social · 26/09/2026in academia if you wear a tie people react like you're wearing a tuxedo and holding a sign that says "I'm a fancy boy" 0321
Reposted by Naomi Saphralastpositivist.bsky.social @lastpositivist.bsky.social · 24/09/2026At 45k followers I will reveal exactly what terminology correctly carves AI at its joints. 1220818
Naomi Saphra @nsaphra.bsky.social · 24/09/2026A pattern that I've been surprised by is people writing followup emails for their slop emails. Like, people are getting really insistent that I respond in earnest to emails they did not write. 4311
Reposted by Naomi SaphraTom Gauld @tomgauld.bsky.social · 24/09/2026My latest @newscientist.com cartoon. many more here: www.newscientist.com/author/tom-gauld/ 332863726
Reposted by Naomi SaphraJohn Lake @jlake9.bsky.social · 24/09/20261/ Matryoshka Attribution uses gradient descent plus causal interventions to localize the parts of a network responsible for a behavior; the authors claim a large MIB jump. X: x.com/aryaman2020/status/2102800933… Paper: arxiv.org/abs/2609.25518 182
Reposted by Naomi SaphraTomer Ullman @tomerullman.bsky.social · 24/09/2026since I'm getting many pokes about grad school applications, I wanted to re-up some previous public advice on grad school applications. looking back at it a year on, I still think all this is right, but let me update the 'research statement' part to include a bit on genAI: 1259
Reposted by Naomi SaphraEugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 23/09/2026EMNLP conference page now live in Lea: lea.ac/conferences/... Papers will land as well once they're up. 1225
Reposted by Naomi SaphraNiyati Bafna @niyatibafna.bsky.social · 23/09/2026We all know about the curse of multilinguality. We know that empirical performance degrades as you add languages to a model. But *in theory*, does it have to? Let’s talk about the theoretical curse of multilinguality for embedding space structure. 1121
Reposted by Naomi SaphraAnthropic @anthropic.extwitter.link · 23/09/2026Claude has discovered a previously unknown enzyme system hidden in the DNA of bacteriophages. Beside the enzyme’s gene sits a long array of repeating DNA—a structure that looks somewhat similar to CRISPR. We don’t yet understand what this system does, but only a handful of known systems share it...anthropic.com Claude discovers a novel enzyme system with CRISPR-like repeatsAnthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems. 1117
Naomi Saphra @nsaphra.bsky.social · 23/09/2026as a disabled lesbian gossip, I would be enjoying this drama a lot more if it never escaped containment 1190
Reposted by Naomi SaphraDavid Mimno @dmimno.bsky.social · 22/09/2026It's possible for Jev/Laya/Decision Models to be not that big a deal as tech and massive as a new paradigm. Here's why I'm really excited from an NLP history perspective (thread) 17619
Reposted by Naomi SaphraAndrew Lampinen @lampinen.bsky.social · 21/09/2026New post reflecting on recent AI progress in math, how AI is changing the way we work, and some worries about where people will find meaning as they offload more of their work to AI: infinitefaculty.substack.com/p/math-resea...infinitefaculty.substack.comMath, research, and meaning in the age of AISince I wrote my last post about AI and math, the theorems have continued to fall. 0372
Reposted by Naomi SaphraRyan Moulton @moultano.bsky.social · 20/09/2026It is weird how much people have extrapolated the politics of the whole tech industry from solely Elon Musk and a handful of dipshit VCs. 1613615
Reposted by Naomi Saphrap(Dulany) @dulanyw.bsky.social · 06/08/2026FelonyBench when the crimes are scenes from screwball sci-fi heist comedies: 😂 FelonyBench when you include Grok's generation of non-consensual pornography and CSAM: 😩 1696
Reposted by Naomi SaphraLukas Edman @lukasnlp.bsky.social · 17/09/2026Ever feel like it's too hard to keep track of what LLMs cannot do as well as humans? We're making your life easier over at: what-llms-can-not-do.github.io We're compiling a list of papers testing the abilities of LLMs against humans. Check it out! And you can help contribute too!what-llms-can-not-do.github.ioWhat LLMs Can(not) DoA living survey of benchmarks that compare large language models with humans. 24515
Reposted by Naomi SaphraJane Li 🦖 @janeli.bsky.social · 17/09/2026🦀New preprint! (w/ @najoung.bsky.social)🦞 Is grammaticality a major organizing principle of NLM representations? We show that many NLMs exhibit abstract rep. separation for grammaticality. We believe this work addresses debates about confounds in measuring model gram. knowledge. [1/10] 12010
Reposted by Naomi SaphraSE Gyges @segyges.bsky.social · 17/09/2026There has been a great effort over many years to distance AI Safety discussion and policy from its origins and intellectual center, because its center is a sex cult started by a fan fiction author. If you want to be taken seriously you have to hide that. 3105184
Reposted by Naomi SaphraGautam Kamath @gautamkamath.com · 16/09/2026Nihar Shah did a heroic experiment for TMLR: he spent 20-25 hours over two weeks interviewing authors of seemingly low-quality submissions about their own papers. He confirmed what we all suspected: people submitting these papers have *no idea* what is going on in them. 5276110
Reposted by Naomi SaphraVladimir Salnikov @v4ldelund.bsky.social · 16/09/2026"just look at the data" final boss 316022
Naomi Saphra @nsaphra.bsky.social · 16/09/2026Restaurants are gonna be demanding IDs at the door and making reservations nontransferable. Airlines don't have secondary markets. 030
Reposted by Naomi SaphraTuhin Chakrabarty @tuhinchakr.bsky.social · 16/09/2026I trust no one but @vauhinivara.bsky.social to come up with something so good. It’s brilliantly researched and engages thoroughly with the space of AI detection. There is a funny bit in the article that talks about how I got introduced to @pangram.com and the rest is history !! 162
Reposted by Naomi SaphraJoe Bak-Coleman @jbakcoleman.bsky.social · 15/09/2026I was pretty chuffed about science when the doctors at NYU figured out why my face had gone numb and I had lost the ability to speak and swallow. They turned to the scientific literature, which guided them through keeping me from dying from a rare disorder and regaining the ability to speak and eat. 324536
Reposted by Naomi SaphraDavid Picard @davidpicard.eurosky.social · 14/09/2026New year, new semester, new course, new textbook: davidpicard.github.io/mldl/mldl-bo... 👀 Still in draft form, but not in a bad shape.davidpicard.github.io 23511
Reposted by Naomi SaphraMelanie Walsh @mellymeldubs.bsky.social · 12/09/2026I've had a hard time finding consistent women's basketball convos on Bluesky. But I added the "For You" feed and then liked a few Gabby Williams posts, and voila. bsky.app/profile/did:... Recommended for your French basketball-related and other special interests. 0111
Reposted by Naomi SaphraZach Weinersmith @zachweinersmith.bsky.social · 12/09/2026What he doesn't realize is if the yuri writers get replaced, their backup plan is to finish their papers on probabilistic approaches to nonlinear dynamics. So, it should all even out. 61058143
Reposted by Naomi SaphraTerence Tao @teorth.bsky.social · 11/09/2026A group of 25 Fields Medalists, including myself, have made a joint declaration on Math and AI: mathandai.org . We welcome additional signatories. See also this article in the Economist announcing the declaration: www.economist.com/science-and-...mathandai.orgDeclaration — Math and AIRead the declaration and add your name. 422052925
Reposted by Naomi SaphraAntonin Poché @antoninpoche.bsky.social · 09/09/2026I am both excited🔥and worried❄️. 🔥We got a paper accepted to @blackboxnlp.bsky.social reproducibility track. ❄️It reproduces and destroys my own paper. So I basically have 1 PhD year, my scientific integrity, and interpreto left 😅 By the way, I describe the paper in this thread: 🧵1/9 2141
Reposted by Naomi SaphraEugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 08/09/2026Cannot believe there’s a tremendous mathematical result and instead of being exciting it’s annoying 622510
Naomi Saphra @nsaphra.bsky.social · 08/09/2026Interestingly, there are not one but two breakthroughs from openly human-led teams related to NS. BOT have released incomplete proofs early to maneuver around the openai "scoop". (The other is from Anima Anandkumar's group and has less drama involved.)mathstodon.xyzTerence Tao (@tao@mathstodon.xyz)By sheer coincidence, another completely independent result on the Euler blowup question has just been released by Ganeshram, Duruisseaux, and Anandkumar https://anima-ai.org/2026/09/07/stable-singula... 1224
Naomi Saphra @nsaphra.bsky.social · 08/09/2026One side effect of all the drama is I've been learning about the norms of math academia, which are beautiful. Mathematicians are like pro athletes: They know they're lucky being paid to play around, so they have rules of fair play to keep it fun, accessible, and valuable to humans. 4441
Reposted by Naomi SaphraThe Transmitter @thetransmitter.bsky.social · 07/09/2026I fear that placing too much emphasis on a specific interpretation of dimensionality, or treating dimensionality as an end-all quantification of some aspect of neural computation, may lead us down the wrong path, writes @mattperich.bsky.social. #neuroskyence www.thetransmitter.org/neural-dynam...thetransmitter.orgDimensionality—neuroscience’s red herring?Placing too much emphasis on a specific interpretation of dimensionality may lead neuroscience down the wrong path. 07421
Reposted by Naomi SaphraGordon @gordon.bsky.social · 05/09/2026Astra this Astra that, but man I tried dropping Gemini Flash into our scenario planning engine yesterday and am getting amazing results at around 0.1x the real cost of Sonnet. Embarrassment of riches across inference tiers right now. 3491
Reposted by Naomi SaphraComputational Cosmetologist @dferrer.bsky.social · 05/09/2026Dug into the Claude dir in my home folder and found some Opus 5 session had created a memory that “the concept of the privy sparks joy for the user”. I never said anything like this. I feel like I’ve been pranked by an AI. Now I have to search every machine. 612118
Reposted by Naomi SaphraComputational Cosmetologist @dferrer.bsky.social · 05/09/2026Every single Claude Code session on one machine for the last few weeks had developed a strange obsession with mentioning the bathroom in code comments (“this is the lock on the bathroom door for the model”) and naming things “privy” or “privi”. Thought I was going insane. Kept happening. 5988
Reposted by Naomi SaphraSung Kim @sungkim.bsky.social · 03/09/2026Humans are back! Shin Jin-seo, the world's top-ranked Go player, on Tuesday completed a dramatic comeback against the world’s premier artificial intelligence Go engine, KataGo, claiming a historic human victory over AI. www.kedglobal.com/artificial-i...kedglobal.comGo grandmaster Shin defeats AI KataGo in historic human victory - KED GlobalShin Jin-seo, the world's top-ranked Go player, on Tuesday completed a dramatic comeback against the world’s premier artificial intelligence Go engine, K 125450
Reposted by Naomi SaphraQuanta Magazine @quantamagazine.org · 01/09/2026This “stunning” proof demonstrates that phase transitions are all or nothing. www.quantamagazine.org/stunning-per...quantamagazine.org‘Stunning’ Percolation Proof Solves Decades-Old Puzzle About Phase Transitions | Quanta MagazineMathematicians found that a broad class of networks will abruptly shift behavior past a critical point. 1339
Reposted by Naomi SaphraTom McCoy @rtommccoy.bsky.social · 01/09/2026🤖🧠NEW PAPER🧠🤖 (The result of an 8-year project!) LLMs seem very different from symbolic systems. Yet LLMs excel in symbolic domains (e.g., language/code/math). How do they do it? Our finding: LLM representations have implicit symbolic structure! Link in thread ⬇️ 1/n 431688
Naomi Saphra @nsaphra.bsky.social · 01/09/2026Yes, some people are worse at writing than an LLM. Those people are also usually incoherent thinkers. If they learned to write better, they would get better at thinking. 2726
Reposted by Naomi SaphraGautam Kamath @gautamkamath.com · 31/08/2026As LLMs make things easier, we must raise our expectations on what researchers are expected to produce. This is particularly true for writing: clear writing was rarely a focus in scientific publication, and it's only got worse due to LLMs. We're trying to reverse this trend. 0314
Reposted by Naomi SaphraJon Gjengset @jonhoo.eu · 29/08/2026I got a 2-meow server rack. Highly recommend. 381555145
Reposted by Naomi SaphraRyan Moulton @moultano.bsky.social · 28/08/2026I think you all should know that this bird exists. 1027256