Sign in

Yacine Jernite

@yjernite.bsky.social
514 followers 216 following 45 posts

Head of ML & Society at Hugging Face 🤗

PostsRepliesMedia
Reposted by Yacine Jernite
Dr Sasha Luccioni @sashamtl.bsky.social · 06/08/2025
Think AI's rising energy demands aren't your problem? Think again! ⚡ In our commentary, @yjernite.bsky.social and I explain how AI & data center expansion is making energy bills rise in the US, what the mechanisms driving this are, and why it matters. www.techpolicy.press/how-your-uti...
techpolicy.press
How Your Utility Bills Are Subsidizing Power-Hungry AI | TechPolicy.Press
The next few years will be pivotal for determining the future of AI and its impact on energy grids worldwide, write Sasha Luccioni and Yacine Jernite.
22714
Reposted by Yacine Jernite
Tech Policy Press @techpolicypress.bsky.social · 06/08/2025
As tech firms keep adding the largest and most compute-intensive AI models into more and more aspects of our digital lives, they are increasingly dependent on a growing share of existing energy and natural resources, leading to rising costs for everyone else, write Sasha Luccioni and Yacine Jernite.
techpolicy.press
How Your Utility Bills Are Subsidizing Power-Hungry AI | TechPolicy.Press
The next few years will be pivotal for determining the future of AI and its impact on energy grids worldwide, write Sasha Luccioni and Yacine Jernite.
13719
Reposted by Yacine Jernite
Giada Pistilli @giadapistilli.com · 29/07/2025
From Replika to everyday chatbots, people form emotional bonds with AI. But what happens when an AI tells you "I understand how you feel" and you actually believe it? With @frimelle.bsky.social and @yjernite.bsky.social, we dug into something: how AI systems handle our emotional lives.
huggingface.co
AI Companionship: Why We Need to Evaluate How AI Systems Handle Emotional Bonds
A Blog post by Giada Pistilli on Hugging Face
143
Reposted by Yacine Jernite
Dr Sasha Luccioni @sashamtl.bsky.social · 17/07/2025
Friends! Does anyone know of any model distillation with public logs (W&B or other)? I'm trying to figure out the energy tradeoffs between model training and distillation..
075
Reposted by Yacine Jernite
Avijit Ghosh @evijit.io · 15/07/2025
New blog post alert! 🚨"What is the Hugging Face Community Building?", with @yjernite.bsky.social and Irene Soliaman The AI narrative focuses on big players, but the real story is happening in the open source AI ecosystem across 1.8M models, 450K datasets, and 560K apps, on @hf.co.
1143
Reposted by Yacine Jernite
Dr Sasha Luccioni @sashamtl.bsky.social · 19/06/2025
One of the biggest frustrations I have is the lack of transparency around AI's energy use and environmental impacts. I know the numbers are out there... but somehow we're not seeing them 🫠 Thank you @wired.com for covering this topic in such depth and detail ! www.wired.com/story/ai-car...
wired.com
How Much Energy Does AI Use? The People Who Know Aren’t Saying
A growing body of research attempts to put a number on energy use and AI—even as the companies behind the most popular models keep their carbon emissions a secret.
15928
Reposted by Yacine Jernite
Daniel van Strien @danielvanstrien.bsky.social · 17/06/2025
“AI Scraping Bots Are Breaking Open Libraries, Archives, and Museums” – interesting piece via @404media.co Not a perfect fix, but making ML-ready datasets from collections can help. If you want help getting your data on @hf.co, I'd be happy to help.
Screenshot of the header of the article with text:

AI Scraping Bots Are Breaking Open Libraries, Archives, and Museums
0144
Reposted by Yacine Jernite
Daniel van Strien @danielvanstrien.bsky.social · 16/06/2025
Institutional Books: Massive Historical Text Corpus - 983K books, 242B tokens, 386M pages - 19th-20th century texts in 254 languages - Refined OCR with quality scores & metadata - Noncommercial early-access release huggingface.co/datasets/ins...
huggingface.co
institutional/institutional-books-1.0 · Datasets at Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
03815
Yacine Jernite @yjernite.bsky.social · 11/06/2025
Great blog post on *Digital Sovereignty and OS AI* led by the fantastic @frimelle.bsky.social! Digital sovereignty for AI needs to properly account for: 📚 data 🧑‍🔬 technology 💽 infrastructure ⚖️ regulation Open/transparent AI contributes to all, read for some concrete examples! hf.co/blog/frimell...
huggingface.co
Open Source AI: A Cornerstone of Digital Sovereignty
A Blog post by Lucie-Aimée Kaffee on Hugging Face
021
Reposted by Yacine Jernite
Lucie-Aimée Kaffee @frimelle.bsky.social · 11/06/2025
❗️New policy blogpost! The EU is speaking a lot about sovereignty. A cornerstone of digital sovereignty is and has to be open source. As AI becomes more central, the ability to govern, adapt, and understand these systems is no longer optional.
151
Reposted by Yacine Jernite
Alex Hanna @alexhanna.bsky.social · 06/06/2025
Off-mark from @sanders.senate.gov. We need progressive legislators to not buy into the clown show from AI CEOs. Labor replacement is not real, labor displacement is. We need regulation to protect workers and anticipate the kind of worker speedups that employers buying into the hype will cause.
A tweet from Bernie Sanders. It reads:
The CEO of Anthropic (a powerful AI company) predicts that AI could wipe out HALF of entry-level white collar jobs in the next 5 years. 

We must demand that increased worker productivity from AI benefits working people, not just wealthy stockholders on Wall St. AI IS A BIG DEAL.
414127
Reposted by Yacine Jernite
Dr Sasha Luccioni @sashamtl.bsky.social · 28/05/2025
How can we make informed choices based on performance AND energy when using AI in real-life tasks like question answering? By evaluating them and picking the models that optimize both factors! Check out my new blog post on the subject: huggingface.co/blog/sasha/e...
huggingface.co
Bigger isn't always better: how to choose the most efficient model for context-specific tasks 🌱🧑🏼‍💻
A Blog post by Sasha Luccioni on Hugging Face
193
Yacine Jernite @yjernite.bsky.social · 07/05/2025
I'm consistently impressed by @giadapistilli.com's extensive insights into AI technology 🤗 Her latest blog on design factors of AI "companions" shows that those go way beyond model performance and give some nice hands-on tool to do your own analysis - must read! huggingface.co/blog/giadap/...
huggingface.co
AI Personas: The Impact of Design Choices
A Blog post by Giada Pistilli on Hugging Face
020
Reposted by Yacine Jernite
Giada Pistilli @giadapistilli.com · 07/05/2025
Ever notice how some AI assistants feel like tools while others feel like companions? Turns out, it's not always about fancy tech upgrades, because sometimes it's just clever design. huggingface.co/blog/giadap/...
huggingface.co
AI Personas: The Impact of Design Choices
A Blog post by Giada Pistilli on Hugging Face
1105
Reposted by Yacine Jernite
Dr Sasha Luccioni @sashamtl.bsky.social · 29/04/2025
We just integrated the new Qwen3-8B into Chat UI Energy and asked it to do a simple multiplication problem. What's the energy cost? → Without reasoning: wrong (😅), but low energy use → With reasoning: correct (!!)… but using 42x more energy!
2245
Reposted by Yacine Jernite
Giada Pistilli @giadapistilli.com · 17/04/2025
🤗 New from us! Just published a blog post exploring how we're rethinking consent in the AI ecosystem. Here's what we're seeing in the @hf.co Hub that differs from traditional closed systems...
huggingface.co
Consent by Design: Approaches to User Data in Open AI Ecosystems
A Blog post by Giada Pistilli on Hugging Face
162
Yacine Jernite @yjernite.bsky.social · 16/04/2025
Today in Privacy & AI Tooling - introducing a nifty new tool to examine where data goes in open-source apps on @hf.co 🤗 HF Spaces have tons (100Ks!) of cool demos leveraging or examining AI systems - and because most of them are OSS we can see exactly how they handle user data 📚🔍 1/4 🧵
Interface of Space Privacy Analyzer app, describing how it reviews Hugging Face Spaces for data privacy concerns, with a pre-loaded example for the Hugging Face demo app for SmolVLM2A TLDR report generated by the Spaces Privacy app outlining the different types of data used and where they go when using the app
174
Reposted by Yacine Jernite
Avijit Ghosh @evijit.io · 15/04/2025
Thrilled to share that our paper: "It's not a representation of me": Examining Accent Bias and Digital Exclusion in Synthetic AI Voice Services - has been accepted at @facct.bsky.social 2025! - with @shiramichel.bsky.social , Sufi Kaur, Sarah Gilespie, Jeffrey Gleason and Dr. Christo Wilson.
2157
Yacine Jernite @yjernite.bsky.social · 10/04/2025
New blog post led by @evijit.io on wrangling public data for AI - and helping public orgs have more control over how AI systems serve their mission by shaping how their data's used📚 Have a read especially if your org's being asked to do more AI (common theme these days 🤗) hf.co/blog/evijit/...
huggingface.co
Empowering Public Organizations: Preparing Your Data for the AI Era
A Blog post by Avijit Ghosh on Hugging Face
020
Reposted by Yacine Jernite
Avijit Ghosh @evijit.io · 10/04/2025
🚨 New Article: Empowering Public Organizations: Preparing Your Data for the AI Era, with @yjernite.bsky.social Let’s discuss how public organizations can unlock the full potential of their data in the age of AI.
131
Yacine Jernite @yjernite.bsky.social · 09/04/2025
This is *extremely* cool I'm increasingly excited about using the OLMo based apps for daily use - I find the playground genuinely better than the commercial apps whenever I need some originality, and the transparency/privacy guarantees are just so much stronger
041
Reposted by Yacine Jernite
Alex Hanna @alexhanna.bsky.social · 07/04/2025
We wrote about this in thecon.ai: with enterprise deals, employees will be initially highly encouraged to use synthetic text generation machines. But at some point, it'll become required to justify the cost.
thecon.ai
THE AI CON
How to Fight Big Tech's Hype and Create the Future We Want
2278
Reposted by Yacine Jernite
Center for Democracy & Technology @cdt.org · 03/04/2025
AB 566 has officially passed out of the committee in the CA Assembly! 🎉 This bill ensures that web browsers & mobile OS vendors provide a simple, automated way for users to send opt-out signals—closing a loophole in CA privacy law. A huge step toward meaningful data rights!
1103
Yacine Jernite @yjernite.bsky.social · 03/04/2025
One of the reasons it's so hard to easily debunk the most outlandish claims on "loss of control" risks is the lack of self-consistency - e.g. same labs showing model outputs are NOT "human thinking" and still doubling down on calling the phenomenon "deception" 🙃 www.anthropic.com/research/rea...
anthropic.com
Reasoning models don't always say what they think
Research from Anthropic on the faithfulness of AI models' Chain-of-Thought
010
Reposted by Yacine Jernite
Center for Democracy & Technology @cdt.org · 01/04/2025
📢 Today, the California Assembly will hold a hearing on AB 566 — a bill requiring web browser & mobile operating system vendors to incl. a setting that enables users to send automated signals indicating that they wish to opt-out of sales of their personal data. cdt.org/insights/mea...
cdt.org
Meaningful Opt-Out Rights Require Companies to Do Their Part. State Governments Might Have to Make Them.
Update: CDT submitted a letter on March 7, 2025 to the California State Assembly’s Committee on Consumer Protection in support of AB 566, which requires vendors of web browsers and mobile operating systems to include a setting that enables users to send an automated signal to businesses with which they interact through their browser or […]
1167
Reposted by Yacine Jernite
Data & Society @datasociety.bsky.social · 27/03/2025
We’re thrilled to announce nine new Data & Society affiliates! Welcome, Sareeta Amrute, Omer Bilgin, Minsu Longiaru, John Edgar Lopez, Sanjay Pinto, Lana Swartz, Zoë West, @davidthewid.bsky.social, and Sara Ziff! datasociety.net/announcement...
Boxes showing the headshots of the new affiliates in black and white.
0133
Yacine Jernite @yjernite.bsky.social · 26/03/2025
Super helpful update on the latest DeepSeek-v3 from my amazing @hf.co colleagues working on Open-R1 with a neat section about security TLDR: yes, it's safe to download - and if you're using it to code/agent just follow the same basic security principles as for any other model hf.co/blog/open-r1...
huggingface.co
Open R1: Update #4
A Blog post by Open R1 on Hugging Face
081
Yacine Jernite @yjernite.bsky.social · 26/03/2025
I cannot overstate how impressed I am by my colleague @giadapistilli.com's writing on digital consent in the age of AI She just published a fantastic blog post on the topic, which is by far the best piece I've seen in a while outlining the breadth of the issue, must-read 🙌 hf.co/blog/giadap/...
huggingface.co
I Clicked “I Agree”, But What Am I Really Consenting To?
A Blog post by Giada Pistilli on Hugging Face
011
Reposted by Yacine Jernite
Giada Pistilli @giadapistilli.com · 26/03/2025
We've all become experts at clicking "I agree" without a second thought. In my latest blog post available on @hf.co, I explore why these traditional consent models are increasingly problematic in the age of generative AI.
huggingface.co
I Clicked “I Agree”, But What Am I Really Consenting To?
A Blog post by Giada Pistilli on Hugging Face
1114
Yacine Jernite @yjernite.bsky.social · 20/03/2025
New @hf.co policy blog on our AI Action Plan response just out 🤗 TLDR: open research/development and smaller/efficient adaptable models that are developed and evaluated in context just make the most sense; whether you're looking for performance, reliability, or security! hf.co/blog/ai-acti...
huggingface.co
AI Policy @🤗: Response to the White House AI Action Plan RFI
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
092
Yacine Jernite @yjernite.bsky.social · 18/03/2025
Getting really tired of "scientific" reports on the governance of "frontier"/state-of-the-art AI models that somehow seem to forget the fundamentals of science when assessing claims related to "loss of control" or "deceptive" models 🙃 Reminder from our recent CoP comment 👇 hf.co/blog/frimell...
"[1] The work cited in the FAQ to support those concerns is primarily developed by commercial entities whose success depends on the perceived performance of the models, shows strong anthropomorphization bias in its framing, and should be understood to be of marginal value at best given the lack of transparency on the systems tested, especially in terms of training data. "
000
Reposted by Yacine Jernite
Shayne Longpre @shaynelongpre.bsky.social · 13/03/2025
What are 3 concrete steps that can improve AI safety in 2025? 🤖⚠️ Our new paper, “In House Evaluation is Not Enough” has 3 calls-to-actions to empower evaluators: 1️⃣ Standardized AI flaw reports 2️⃣ AI flaw disclosure programs + safe harbors. 3️⃣ A coordination center for transferable AI flaws. 1/🧵
1118
Reposted by Yacine Jernite
Lucie-Aimée Kaffee @frimelle.bsky.social · 13/03/2025
🚨 Our comments on the 3rd draft of the EU Code of Practice for GPAI developers are out! 🚨 Some promising updates—but also concerning choices that could harm transparency & open AI development. A quick breakdown of the blogpost with @yjernite.bsky.social 🧵👇
111
Yacine Jernite @yjernite.bsky.social · 13/03/2025
So as I was saying yesterday ;) allenai.org/blog/olmo2-32B Huge congrats to the @ai2.bsky.social team. This is a fantastic achievement, and a strong reminder not to discount meaningfully open models when talking about the state of the art in AI!!!
allenai.org
OLMo 2 32B: First fully open model to outperform GPT 3.5 and GPT 4o mini | Ai2
Introducing OLMo 2 32B, the most capable and largest model in the OLMo 2 family.
050
Yacine Jernite @yjernite.bsky.social · 12/03/2025
Amazing work from HF colleagues that's super relevant to AI policy - open development matches "frontier" performance 🤗 For policymakers, it means we need regulation that works for both open and closed model - not closed-first that disproportionately harms open! Good news is... 1/2
A graph comparing the performance of different AI systems on a benchmark derived from the Informatics Olympiads, showing an open Hugging Face model coming third ahead of DeepSeek-R1 and Claude-3.7 Sonnet
2138
Reposted by Yacine Jernite
Dr Sasha Luccioni @sashamtl.bsky.social · 21/02/2025
Thank you to @theglobeandmail.com for including me in your 2025 "Changemakers" list, among such an inspiring list of leaders. I am proud to be Canadian and to bring my passion for AI (and for saving the planet 🌎) to our country 🇨🇦 www.theglobeandmail.com/business/rob...
0206
Reposted by Yacine Jernite
Alexander Doria @dorialexander.bsky.social · 18/02/2025
I'm very happy to announce a strategic partnership between @wikimediafoundation.org enterprise and Pleias for open, ethical and trustworthy AI innovation. enterprise.wikimedia.com/blog/pleias-...
15219
Reposted by Yacine Jernite
Luca Soldaini 🎀 @soldaini.net · 18/02/2025
This was such a fun project to work on! We release efficient classifiers 🌐 to partition large corpora, and use them to improve sampling for LLM pretraining great work lead by @awettig.bsky.social 👇
0182
Reposted by Yacine Jernite
Lucie-Aimée Kaffee @frimelle.bsky.social · 12/02/2025
How should AI tools be designed to support rather than replace workers? At the Reshaping Work conference, I led a roundtable exploring AI’s impact on labor. We published a blogpost on our key takeaways on responsible AI and the future of work w/ Franco Bastida 🔗 www.rsm.nl/discovery/20... 🧵👇
rsm.nl
Start-Up Approaches to Responsible AI: Worker-Centric InnovationRotterdam school of Management, Erasmus University logoRotterdam school of Management, Erasmus University compact logo
Explore how start-ups are reshaping AI development through transparency, worker inclusivity, and ethical approaches that prioritise human augmentation over replacement.
153
Reposted by Yacine Jernite
Luca Soldaini 🎀 @soldaini.net · 11/02/2025
They made me do video 😬 but for a good reason! We are launching an iOS app–it runs OLMoE locally 📱 We're gonna see more on-device AI in 2025, and wanted to offer a simple way to prototype with it App: apps.apple.com/us/app/ai2-o... Code: github.com/allenai/OLMo... Blog: allenai.org/blog/olmoe-app
54913
Reposted by Yacine Jernite
ROOST @roost.tools · 11/02/2025
"Small-scale developers...within the broader development chain often lack the resources to create new safety tools from scratch, face disproportionate challenges in adopting off-the-shelf solutions...and are sometimes even denied access to those tools altogether." huggingface.co/blog/yjernit...
huggingface.co
ROOST: Safety Tooling needs Open Tech🐓🤗
A Blog post by Yacine Jernite on Hugging Face
0102
Reposted by Yacine Jernite
Kathy Baxter @baxterkb.bsky.social · 11/02/2025
Boris Gamazaychikov, @salesforce.com Head of #AI #Sustainability announced the AI Energy Score we launched at the AI Action Summit in Paris. 🌍 This offers a standardized way to measure & compare the energy efficiency of AI models. 🫶 www.linkedin.com/posts/bgamaz...
linkedin.com
1174
Reposted by Yacine Jernite
Giada Pistilli @giadapistilli.com · 10/02/2025
Dans un segment pour l'émission "En société" diffusé hier sur France 5, nous avons retracé l'histoire du domaine de l'intelligence artificielle. Au-delà de la chronologie, une question fondamentale a émergé : si l'IA peut surpasser l'humain dans certaines tâches, quelle est notre valeur ajoutée ?
131
Yacine Jernite @yjernite.bsky.social · 10/02/2025
Congrats to the entire @roost-tools.bsky.social team for their successful launch! It's been fantastic to see this project take shape, open tools are very much needed if we're to develop technology that is safer for all - glad to be a partner with @hf.co 🤗 huggingface.co/blog/yjernit...
huggingface.co
ROOST: Safety Tooling needs Open Tech🐓🤗
A Blog post by Yacine Jernite on Hugging Face
184
Reposted by Yacine Jernite
Center for Democracy & Technology @cdt.org · 07/02/2025
Today, CDT submitted testimony to the Connecticut Joint Committee on Government Administration and Elections, urging modifications to H.B.6846 to protect users’ free expression rights. cdt.org/insights/cdt...
cdt.org
CDT Submits Testimony on Connecticut Bill Creating Criminal Penalties for Election Deepfakes
Today, CDT submitted testimony to the Connecticut Joint Committee on Government Administration and Elections, urging modifications to H.B.6846 to protect users’ free expression rights. H.B.6846 would create misdemeanor and felony criminal penalties for any speaker that knowingly distributes certain unlabeled AI-generated images within 90 days of an general or primary election with the intent of […]
163
Yacine Jernite @yjernite.bsky.social · 07/02/2025
I'll be in Paris from Sunday to Wednesday for AI Summit events! Reach out to catch up, talk about data, openness and transparency in AI, and building agency for all AI stakeholders in the current context 🤗
020
Yacine Jernite @yjernite.bsky.social · 04/02/2025
Cool investigation of smol models on the BBQ benchmark, including the fully-open SmolLMv2 and the smallest DeepSeek-R1 distillation, brought to you by @evijit.io - Great to see that these much smaller models have accuracies comparable to Claude v1 and v2 ;) huggingface.co/blog/evijit/...
huggingface.co
Smol but Mighty: Can Small Models Reason well? 🤔
A Blog post by Avijit Ghosh on Hugging Face
020
Reposted by Yacine Jernite
Shayne Longpre @shaynelongpre.bsky.social · 03/02/2025
Our updated Responsible Foundation Model Development Cheatsheet (250+ tools & resources) is now officially accepted to @tmlrorg.bsky.social (TMLR) 2025! It covers: - data sourcing, - documentation, - environmental impact, - risk eval - model release & licensing - ++
272
Yacine Jernite @yjernite.bsky.social · 31/01/2025
If indeed current commercial systems cost "only a few 10M$" to train (thx Anthropic) we need to talk about how this money will be used to (under)pay for further data concentration through unaccountable licensing/contracting & subsidized compute to gather user data - not "innovation" for all
cnbc.com
OpenAI in talks to raise funding that would value AI startup at up to $340 billion
SoftBank would contribute as much as $25 billion to OpenAI's funding round and become the largest investor.
000
Yacine Jernite @yjernite.bsky.social · 23/01/2025
I see OpenAI is back to the "we did work, it was good work" style of communication with the Operator "system card" 🙃 I think specifying that the system is tested in "two dozen" unspecified languages without reporting any language-specific results takes the cake there openai.com/index/operat...
openai.com
Operator System Card
Drawing from OpenAI’s established safety frameworks, this document highlights our multi-layered approach, including model and product mitigations we’ve implemented to protect against prompt engineerin...
010