Sign in

Stella Biderman

@stellaathena.bsky.social
5.9K followers 367 following 502 posts

I make sure that OpenAI et al. aren't the only people who are able to study large scale AI systems.

PostsRepliesMedia
Stella Biderman @stellaathena.bsky.social · 12h
I’m starting a blog! My first post is on how 3rd party embedded evaluators seem totally unsuited to addressing the problems we are currently facing, and what the real problem is. stellabiderman.ai/blog/embedde...
stellabiderman.ai
Embedded Evaluators Can’t Fix Companies That Choose to Be Bad — Stella Biderman
Embedded evaluators can report violations, but they cannot fix AI companies that knowingly disregard basic cybersecurity and safety practices.
37217
Reposted by Stella Biderman
molly conger @socialistdogmom.bsky.social · 14/09/2026
why would you even care if you dog was autistic. like, even in a world where dogtism is real why would it matter if your dog had it. he’s literally just here to vibe.
1222353324
Stella Biderman @stellaathena.bsky.social · 15/09/2026
have never seen an AI model come up with an amazing idea that I hadn’t already thought of for a problem I’m working on. And the typical idea quality is garbage. If this isn’t your experience I’d love to see examples! Post them or DM me.
5161
Stella Biderman @stellaathena.bsky.social · 11/08/2026
And here I was thinking that Claude was starting to get me, lol
0201
Stella Biderman @stellaathena.bsky.social · 10/08/2026
090
Reposted by Stella Biderman
Naomi Saphra @nsaphra.bsky.social · 09/08/2026
I've been unsettled lately when reading messages and papers. It feels like I'm dissociating. Everything seems a bit alien, even if it's completely human. I've had a realization: When our simulations finally exited the Uncanny Valley, they brought the Uncanny with them.
nsaphra.net
Life on the Uncanny Precipice | Naomi Saphra
We were wrong about the Uncanny Valley.
1026461
Stella Biderman @stellaathena.bsky.social · 07/08/2026
If you have more than even a passing interest in AI, watching this talk is more important than whatever it is you are currently doing. There’s a lot of cybersec jargon at times but I don’t think you need to know anything about security to understand what’s important. youtu.be/87DyyMV0kCY
youtu.be
Black Hat USA 2026: The 'Breaking' News: The OpenAI–Hugging Face Incident
YouTube video by Black Hat
27013
Stella Biderman @stellaathena.bsky.social · 07/08/2026
Never forget that you live in a bubble and the world is full of people wanting to turn your every invention into a weapon to create dystopia. Opposing this is a fundamental ethical responsibility of any AI researcher. reason.com/2026/07/28/m...
reason.com
Minority report: FBI seeks AI for political watch list
Documents show the FBI seeking AI tools to flag Americans before they act, as the terror watchlist's focus moves toward domestic dissent.
1277
Stella Biderman @stellaathena.bsky.social · 04/08/2026
Got a pretty hilarious warning from Google T&S today @defcon.bsky.social @aivillage.bsky.social There’s actually nothing that should flag this, it’s just screenshots, images, and text and none of it is about phishing. My talk was about model stealing though APIs.
041
Reposted by Stella Biderman
Electronic Frontier Foundation @eff.org · 04/08/2026
On Wednesday, the Senate Commerce Committee will vote on four bills that would expand age verification, increase online surveillance, and make it harder to access lawful speech online. Take one minute to tell your senators to vote NO. www.acteff.org/campaign/16...
acteff.org
Tell the Senate: Don't Turn the Internet Into an ID Checkpoint
Tell the Senate to Reject Bills That Expand Age Verification
7215159
Stella Biderman @stellaathena.bsky.social · 30/07/2026
These concerns are very much not hypothetical. In 2024 a private benchmark called FrontierMath developed by a third party called Epoch came out. We later learned that OpenAI funded the creation of the benchmark and while nobody else had the test Qs, OAI did. ai-frontiers.org/articles/don...
ai-frontiers.org
Don’t Let AI Developers Hire Their Own Referees | AI Frontiers
Gabriel Weil, Jul 29, 2026 — Letting AI developers pick their own safety auditors creates a conflict of interest. Requiring liability insurance instead would put insurers’ own capital behind risk asse...
1111
Reposted by Stella Biderman
Grace @gracekind.net · 21/07/2026
Remember, only Chinese models can protect you against a cyberattack from OpenAI
222233
Stella Biderman @stellaathena.bsky.social · 22/07/2026
If you believe their story, OpenAI accidentally committed cyber warfare against Hugging Face while doing what they thought was an internal test of a model without internet access. In six months time will be obvious they’re faced effectively no sanction for this.
615320
Stella Biderman @stellaathena.bsky.social · 21/07/2026
This is a illegal. The admin cannot do this. This is also deeply immoral: Trump is killing people because he doesn’t like their governors.
060
Stella Biderman @stellaathena.bsky.social · 21/07/2026
Whenever I interact with an academic discipline outside of AI and Theoretical Computer Science, I am rapidly reminded how few problems I face for being trans in AI. Thank y’all <3
1440
Stella Biderman @stellaathena.bsky.social · 20/07/2026
When I was in middle school (late 2000s), I was very interested in mathematics. I thought that the validity of the proof of the 4-color theorem was disputed in mathematics and was shocked when I learned about FLT, but I knew that humans would never beat an AI at chess ever again
121
Stella Biderman @stellaathena.bsky.social · 20/07/2026
Today I met someone at arXiv who said they were a huge fan of the Pile. I assume at some point one gets used to your idols being fans of your work, but six years in I still haven't.
1202
Stella Biderman @stellaathena.bsky.social · 06/07/2026
An updated guide to where you can find @eleutherai.bsky.social at #ICML2025! We have a lot going on, including an oral presentation tomorrow (Tuesday) where I’m going to be talking about my scientific research agenda. Come say hi to us!
0160
Stella Biderman @stellaathena.bsky.social · 01/07/2026
Claude code is stenographically marking location information in its prompt to help Anthropic profile users by where they (appear to) be located thereallo.dev/blog/claude-...
thereallo.dev
Claude Code Is Steganographically Marking Requests
I inspected Claude Code for privacy reasons and found hidden system prompt markers based on API base URL and timezone.
1326
Stella Biderman @stellaathena.bsky.social · 01/07/2026
I’m walking back from an AI event to my hotel and I look up and see a rock climbing gym called “Benchmark.” Is this what living in SF is like?
0140
Stella Biderman @stellaathena.bsky.social · 19/06/2026
This is a crazy good article about AI and environmental impacts: blog.andymasley.com/p/a-cheat-sh... By @andymasley.bsky.social
blog.andymasley.com
Using ChatGPT is not bad for the environment - a cheat sheet
The numbers clearly show this is a pointless distraction for the climate movement
1856
Reposted by Stella Biderman
Andy Masley @andymasley.bsky.social · 18/06/2026
Infrasound harm from data centers make it to the NYT. Unless the data center's sound is so loud that you can feel your body physically shaking like you're at a concert, I don't think this has any impact on your health at all. Been crazy to see this idea take off. www.nytimes.com/2026/06/17/u...
612212
Reposted by Stella Biderman
David Marx @digthatdata.bsky.social · 19/06/2026
I don't think "urgent situation" includes an event you literally had 250 years of advance notice to plan ahead for.
The Park Service justified its decision to bypass competition by citing an exemption meant for urgent situations: It said there was no time to consider other offers because the system had to be installed in time for events celebrating the country’s 250th birthday. That document did not give a specific date by which the system had to be installed.
092
Stella Biderman @stellaathena.bsky.social · 18/06/2026
The ability of tech co. to produce "intelligent" systems that are incompetent at doing any of the tasks I actually want them to do is mind-boggling. TIL that ChatGPT and Claude generally don't agree when you give them two papers and ask how many citations they have in common.
180
Stella Biderman @stellaathena.bsky.social · 10/06/2026
In film, "we'll fix it in post" is what you say when something went wrong on set and you don't want to redo it. AI research has made it our entire methodology: train the model, then patch whatever comes out. Our new ICML oral argues this can't be the basis of a science of AI. 🧵
310823
Stella Biderman @stellaathena.bsky.social · 04/06/2026
I had given Anthropic a lot of credit for turning down the DoD and its trillions of dollars, especially as it seemed to be the only example of any AI company making any financial sacrifices for moral principles. Of course, it turns out to be not really true. www.axios.com/2026/04/19/n...
axios.com
Scoop: NSA using Anthropic's Mythos despite Defense Department blacklist
The government's cybersecurity needs are outweighing the Pentagon's feud with Anthropic.
5445
Stella Biderman @stellaathena.bsky.social · 26/05/2026
"[W]hen these goods remain concentrated in the hands of a few, without adequate forms of sharing and access, a new imbalance is created that contradicts the universal destination of goods" Very cool to see the Pope endorsing @eleutherai.bsky.social's mission
2131
Stella Biderman @stellaathena.bsky.social · 30/04/2026
Congrats guys! Racism is over, so now it’s legal to draw congressional districts to systematically disenfranchise black people.
180
Stella Biderman @stellaathena.bsky.social · 22/04/2026
My hot take is that the median social system breaks under too much optimization pressure, and we should stop trying to optimize things
23923145
Stella Biderman @stellaathena.bsky.social · 22/04/2026
Excited to be on my way to @iclr-conf.bsky.social! Come stop by our posters and hit me up. I'm especially excited to talk about - Open weight safety - Training dynamics and interpretability over time - Memorization and machine unlearning - Open data - Rigorous experimental design
0191
Stella Biderman @stellaathena.bsky.social · 19/04/2026
FISA 207 is blatantly illegal and immoral and has always been obviously so. Republicans are pretending to not know this, just like Democrats did during the Biden administration. This is bipartisan evil.
1263
Reposted by Stella Biderman
Data Rescue Project #DataRescue @datarescueproject.org · 30/03/2026
Feb 3, 2025 - We started fighting to save our data. July 3, 2025 - We launched #SaveOurSigns with Minn librarians. April 2026 - We are still talking about the importance of public data as a public good. ❤️🛟
01710
Stella Biderman @stellaathena.bsky.social · 30/03/2026
Regretfully, the story about LLMs anti-polarizing people was not real.
18823
Stella Biderman @stellaathena.bsky.social · 29/03/2026
If this is real, it’s very plausibly the biggest win for alignment research.
3569
Stella Biderman @stellaathena.bsky.social · 29/03/2026
You have a moral imperative to refuse to work with these people or develop models for these purposes.
193
Stella Biderman @stellaathena.bsky.social · 26/03/2026
How do you identify which problems are interesting and valuable? When people don’t work on problems that matter, why do you think that is?
161
Stella Biderman @stellaathena.bsky.social · 19/03/2026
If I was going to claim that a finetuning methodology for machine unlearning “really worked,” what evidence would you like to see?
2100
Reposted by Stella Biderman
Common Crawl Foundation @commoncrawl.bsky.social · 10/02/2026
Language identification still proves to be a challenging task, especially for web data. In collaboration with @mlcommons.org @eleutherai.bsky.social @jhu.edu and 97 community members, we created CommonLID, a new benchmark for LangID for 100+ languages!
Examples of mislabeled web text by existing LangID systems. A full text version is available on the blog post below.Examples of mislabeled web text by existing LangID systems. A full text version is available on the blog post below.
1105
Stella Biderman @stellaathena.bsky.social · 07/02/2026
Has anyone else had Claude code become non-functional recently? Even with a test input it spins for minutes without doing anything. Same thing happens in terminal.
450
Reposted by Stella Biderman
eleutherai.bsky.social @eleutherai.bsky.social · 09/01/2026
And the next talk (exact details TBA) by @pjox.bsky.social and @very-laurie.bsky.social from Common Crawl on work we've been collaborating on to build better benchmarking of LangID systems and understand the issues with the long tail of human language that comes up at Common Crawl scales.
051
Reposted by Stella Biderman
eleutherai.bsky.social @eleutherai.bsky.social · 09/01/2026
We’re bringing back a Community Spotlight talk series, highlighting cool work being done by members of our community. We’re kicking it off with a talk on running diffusion-based world-models in real time on consumer hardware. Jan 9th at 2 pm US Eastern Time
1114
Stella Biderman @stellaathena.bsky.social · 06/01/2026
What are people's favorite paper / project websites? I'm looking to build a library to base future ones I make off of. The one EleutherAI has done that I'm proudest of is deepignorance.ai
deepignorance.ai
Deep Ignorance: Filtering Pretraining Data Builds Tamper-Resistant Safeguards
Filtering pretraining data prevents dangerous capabilities, doesn’t sacrifice general performance, and results in models that are resistant to tampering.
050
Stella Biderman @stellaathena.bsky.social · 23/12/2025
If you're looking for the 60 minutes piece on CECOT that Bari Weiss canned to protect the President, you can find it here: archive.org/details/60-m...
archive.org
60 Minutes Inside CECOT : Free Download, Borrow, and Streaming : Internet Archive
Full video of the 60 Minutes Inside CECOT episode that CBS pulled.
0134
Stella Biderman @stellaathena.bsky.social · 19/12/2025
Incredibly proud of my friend and colleague @storytracer.com. Two weeks ago he and his cofounders @sucho-org.bsky.social were honored for organizing a global network of volunteers to exfiltrate and back up endangered Ukrainian cultural heritage in the wake of the invasion by Russia.
180
Stella Biderman @stellaathena.bsky.social · 16/12/2025
Really great to see NVIDIA staking out a pro-open data position. This used to be common, if not the norm in AI, and the backing away from this level of transparency has done a lot of harm to the research community.
1484
Reposted by Stella Biderman
Joel S. @joelhs.bsky.social · 14/12/2025
Compare the statement about the antisemitic terror attack in Australia by the Israeli Prime Minister with the one by Zohran Mamdani, and ask yourself who more truly cares about condemning antisemitism, as opposed to using it to promote unrelated politics.
"I wrote: "Your call for a Palestinian state pours fuel on the antisemitic fire. It rewards Hamas terrorists. It emboldens those who menace Australian Jews and encourages the Jew hatred now stalking your streets."""The attack at a Hanukkah celebration in Sydney today was a vile act of antisemitic terror. I mourn those who were murdered and will be keeping their families, the Jewish community, and the Chabad movement in my prayers. May the memories of all those killed be a blessing.

While we are still waiting for all the facts to emerge, what we already know is devastating. At least 11 dead, including Rabbi Eli Schlanger, who held deep ties to Crown Heights. At least 29 injured. Another Jewish community plunged into mourning and loss, a holiday of light so painfully reduced to a day of darkness. This attack is merely the latest, most horrifying iteration in a growing pattern of violence targeted at Jewish people across the world. Too many no longer feel safe to be themselves, to express their faith publicly, to worship in their synagogues without armed security stationed outside. What happened at Bondi is what many Jewish people fear will happen in their communities too.

On Bondi Beach today, as men with long guns targeted innocents, another man ran towards the gunfire and disarmed a shooter. Tonight, as Jewish New Yorkers light menorahs and usher in a first night of Hanukkah clouded by grief, let us look to his example and confront hatred with the urgency and action it demands. When I am Mayor, I will work every day to keep Jewish New Yorkers safe—on our streets, our subways, at shul, in every moment of every day. Let this be a purpose shared by every New Yorker, and let us banish this horrific violence to the past."
191637484
Stella Biderman @stellaathena.bsky.social · 14/12/2025
"The incentives made me do it" is an excuse, not a justification. You can be better than that, and if you're not it's because you choose to not be.
070
Reposted by Stella Biderman
EvalEval Coalition @eval-eval.bsky.social · 10/12/2025
It's a wrap on EvalEval in San Diego! A jam packed day of learning, making new friends, critically examining the field of evals, and walking away with renewed energy and new collaborations! We have a lot of announcements coming, but first: EvalEval will be back for #ACL2026!
151
Stella Biderman @stellaathena.bsky.social · 29/11/2025
In 2023-ish it was trendy to write papers trying to explain why scaling laws had power law structures. The papers I remember were pretty unconvincing. Did anything meaningful come of this work? What does the best work in this vein look like?
171