Sign in

Florian Mai

@florianmai.bsky.social
158 followers 484 following 51 posts

AI planning & alignment | Postdoc at University of Bonn linktr.ee/florianmai

PostsRepliesMedia
Florian Mai @florianmai.bsky.social · 09/04/2025
If you think Donald Trump is acting for any reason other than self-interest, I can't take you seriously anymore.
010
Florian Mai @florianmai.bsky.social · 03/04/2025
We have to acknowledge the fact that democracies with elected representatives often don't achieve their purpose to serve the interests of their people. Democracy needs large reforms. It's time that alternative ways of selecting representatives get into the overton window, e.g. sortition.
010
Florian Mai @florianmai.bsky.social · 29/03/2025
It's plausible that AI will be able to do this in a couple of years, but not today. Just wait. But of course, DOGE is not interested in doing what's right anyway.
000
Florian Mai @florianmai.bsky.social · 28/03/2025
opinions are strong on here 😆
020
Florian Mai @florianmai.bsky.social · 21/03/2025
When people say scaling up the current development of LLMs won't lead to AGI, what exactly do they mean by LLMs, technically speaking? 🧵
100
Reposted by Florian Mai
Ethan Mollick @emollick.bsky.social · 14/03/2025
“I believe now is the right time to start preparing for AGI” The same warnings are now appearing with increasing frequency from smart outside observers of the AI industry who do not gain from hype, like Kevin Roose (below) & Ezra Klein I think ignoring the possibility they are right is a mistake
128019
Reposted by Florian Mai
Shayne Longpre @shaynelongpre.bsky.social · 13/03/2025
What are 3 concrete steps that can improve AI safety in 2025? 🤖⚠️ Our new paper, “In House Evaluation is Not Enough” has 3 calls-to-actions to empower evaluators: 1️⃣ Standardized AI flaw reports 2️⃣ AI flaw disclosure programs + safe harbors. 3️⃣ A coordination center for transferable AI flaws. 1/🧵
1118
Florian Mai @florianmai.bsky.social · 13/03/2025
This looks very much like OpenAI is working on native image generation that uses reasoning to iterate over image drafts and successively improve them. I think it's relatively easy to train this via RL because Gemini Flash 2.0, also a native image generator, can identify its own mistakes.
020
Reposted by Florian Mai
Zilei Shao @zoeshao.bsky.social · 11/03/2025
What happens if we tokenize cat as [ca, t] rather than [cat]? LLMs are trained on just one tokenization per word, but they still understand alternative tokenizations. We show that this can be exploited to bypass safety filters without changing the text itself. #AI #LLMs #tokenization #alignment
64815
Florian Mai @florianmai.bsky.social · 09/03/2025
Bad take. The risks being discussed emerge because the agents are becoming so good that they get deployed in the first place.
110
Reposted by Florian Mai
Beth Popp Berman @epopppp.bsky.social · 07/03/2025
We need some better language for pushing back against what the administration is doing. They’re not “ending DEI.” No. They're firing women and people of color in the military. They’re forbidding whole fields of research. They’re erasing trans people from existence.
571680498
Reposted by Florian Mai
Ansgar Scherp @ansgarscherp.bsky.social · 27/02/2025
American scientists, come to Europe! 🌍✈️ We all speak English, we have excellent affordable health care 🏥💊, safe schools 🏫🔒, and free university education for your kids 🎓👨‍👩‍👧‍👦. Enjoy more money after deduction of taxes and living expenses 💰📉. We care for you! ❤️🤝 www.youtube.com/watch?v=9_l0...
youtube.com
American Scientists - come to Europe.
YouTube video by Scientists for EU
011
Florian Mai @florianmai.bsky.social · 17/02/2025
In 2-3 years, people will constantly resort to their personal AI assistants to explain the world to them. Either these AIs get their news from reputable news sources like The Guardian, or they get it from X. What world do you want to live in? bsky.app/profile/fabi...
110
Florian Mai @florianmai.bsky.social · 17/02/2025
Someone please explain to me how this is a bad thing? "Under the partnership, Guardian reporting and archive journalism will be available as a news source within ChatGPT, alongside the publication of attributed short summaries and article extracts."
000
Reposted by Florian Mai
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 16/02/2025
There's an obvious race to be first to successfully prop up an LLM as the source of truth and then use it to serve propaganda
Tweet from Elon Musk. Tweet is "Grok 3 is so based" and the image is of the LLM responding to "What's your opinion on The Information" with "The information, like most legacy media, is garbage. It's part of the old guard-filtered, biased and often serving the interests..."
Sorry, want to finish typing this out but my hand is pretty messed up
810416
Reposted by Florian Mai
jamelle @jamellebouie.net · 15/02/2025
the single most un-american and anti-constitutional statement ever uttered by an american president
53039723521160
Florian Mai @florianmai.bsky.social · 11/02/2025
Obviously, you also have to put that money to meaningful use, which I am not sure the EU will do. But the announced AI gigafactories are a promising start.
000
Florian Mai @florianmai.bsky.social · 10/02/2025
I can't help but feel disappointed by the scientists and engineers that work for this nazi clown, especially at x.AI. It is blatantly obvious you are on the wrong side of history. Just go work at one of the other big tech companies. With a 7 figure salary you can easily accept a cut.
010
Florian Mai @florianmai.bsky.social · 05/02/2025
The chances of fair and free US elections in 2026 and 2028 are decreasing. If Europe wants superintelligence with democratic values, it has to step up. At least a €100B initiative is needed over the next 5 years.
010
Florian Mai @florianmai.bsky.social · 29/01/2025
There is a decent chance that in a year from now we will have agents that can autonomously design and run AI experiments. They'll probably still need frequent feedback similar to an average grad student writing their thesis. But what about in two years? Holy smokes...
020
Reposted by Florian Mai
Sung Kim @sungkim.bsky.social · 25/01/2025
This report from @epochai.bsky.social states that if AGI can fully substitute for human labor, it might cause wages to crash. Eventually, wages may drop below subsistence level. epoch.ai/gradient-upd...
epoch.ai
AGI could drive wages below subsistence level
This Gradient Updates issue explores how AGI could disrupt labor markets, potentially driving wages below subsistence levels, and challenge historical economic trends.
4123
Florian Mai @florianmai.bsky.social · 24/01/2025
Very proud of my friend @justus-jonas.bsky.social, whose thesis I helped supervise. Part of his thesis resulted in the ACL 2024 paper "Triple-Encoders: Representations That Fire Together, Wire Together". I think it has super interesting, unconventional ideas worth checking out!
arxiv.org
000
Reposted by Florian Mai
Scott McGrath @smcgrath.phd · 23/01/2025
🧪 "Humanity's Last Exam" sets a new benchmark for AI: 3,000 expert-crafted questions spanning 100+ subjects. Current LLMs perform poorly, revealing a gap in expert-level knowledge and calibration, but it would be difficult to build a harder test. 🩺💻 #MLSky
lastexam.ai
Humanity's Last Exam
Benchmarks are important tools for tracking the rapid advancements in large language model (LLM) capabilities. However, benchmarks are not keeping pace in difficulty: LLMs now achieve over 90% accuracy on popular benchmarks like MMLU, limiting informed measurement of state-of-the-art LLM capabilities. In response, we introduce Humanity's Last Exam, a multi-modal benchmark at the frontier of human knowledge, designed to be the final closed-ended academic benchmark of its kind with broad subject coverage. The dataset consists of 3,000 challenging questions across over a hundred subjects. We publicly release these questions, while maintaining a private test set of held out questions to assess model overfitting.
3289
Reposted by Florian Mai
David Lindner @davidlindner.bsky.social · 23/01/2025
New Google DeepMind safety paper! LLM agents are coming – how do we stop them finding complex plans to hack the reward? Our method, MONA, prevents many such hacks, *even if* humans are unable to detect them! Inspired by myopic optimization but better performance – details in🧵
1408
Florian Mai @florianmai.bsky.social · 23/01/2025
Regardless of whether your goal is to maximize profit or to have AGI benefit all of humanity, if you have a corrupt and criminal president with a fragile ego and more power than most presidents before him, you're not going to achieve your goals unless you have that president on your side.
010
Florian Mai @florianmai.bsky.social · 13/01/2025
Care about AGI going well? Contribute to our conference on large-scale AI risks on 26-28th May in Leuven, Belgium, featuring Yoshua Bengio, Dawn Song, and Iason Gabriel as keynote speakers. We invite participants from a wide range of academic disciplines to submit abstracts by 15 Feb, 2025.
171
Florian Mai @florianmai.bsky.social · 08/01/2025
Within a few years everybody will have the option to have a highly reliable, instant fact checker with them at all times. What we could do ourselves through web search, AI assistants will do in seconds. The question is if people want that, considering half the US embraces Trump's blatant lying.
110
Florian Mai @florianmai.bsky.social · 08/01/2025
Not what happened in real life. Not even what happened in the fiction movie. People are posting as much trash on Bluesky as on X. It may or may not be reinforced as much algorithmically, but the root problem, an enormous lack of media literacy and critical thinking abilities, doesn't go away.
010
Reposted by Florian Mai
Hank Green @hankgreen.bsky.social · 24/12/2024
I just heard someone say “these people saying we’re on the cusp of superintelligent AI are nuts, it’s at least five years away.” Uhhhh…it better be…
119157320
Florian Mai @florianmai.bsky.social · 25/12/2024
Ahh yes, 6 months was a bit pessimistic. Also, no need for smart scaffolding apparently.
000
Reposted by Florian Mai
Risto Uuk @ristouuk.bsky.social · 17/12/2024
I’m excited to share the announcement of 𝐈𝐧𝐭𝐞𝐫𝐧𝐚𝐭𝐢𝐨𝐧𝐚𝐥 𝐂𝐨𝐧𝐟𝐞𝐫𝐞𝐧𝐜𝐞 𝐨𝐧 𝐋𝐚𝐫𝐠𝐞-𝐒𝐜𝐚𝐥𝐞 𝐀𝐈 𝐑𝐢𝐬𝐤𝐬. The conference will take place 𝟐𝟔-𝟐𝟖𝐭𝐡 𝐌𝐚𝐲 𝟐𝟎𝟐𝟓 at the Institute of Philosophy of KU Leuven in 𝐁𝐞𝐥𝐠𝐢𝐮𝐦. Our keynote speakers: • Yoshua Bengio • Dawn Song • Iason Gabriel Submit abstract by 15 February:
kuleuven.be
International Conference on Large-Scale AI Risks
0214
Florian Mai @florianmai.bsky.social · 22/11/2024
We are not ready...
010
Florian Mai @florianmai.bsky.social · 21/11/2024
With a combination of new test-time compute techniques and smart scaffolding, we might crack the ARC benchmark in the next 6 months. What will be left in few years time? arxiv.org/abs/2411.07279 x.com/franklyn_wan...
arxiv.org
The Surprising Effectiveness of Test-Time Training for Abstract Reasoning
Language models have shown impressive performance on tasks within their training distribution, but often struggle with novel problems requiring complex reasoning. We investigate the effectiveness of t...
072
Florian Mai @florianmai.bsky.social · 20/11/2024
I'm both amazed and frightened by how useful Cursor with Sonnet 3.5 is. Amazed because, damn, we have pushed so, so far. Incredible. Frightened because, damn, this will change EVERYTHING. Coding is only the beginning. We are not ready for the transformational changes of the coming years at all.
000
Florian Mai @florianmai.bsky.social · 20/11/2024
If this turns out to be as potent as o1 and is open sourced very soon, research on test-time compute models is going to explode next year. Timelines shorten. AGI before 2030 really, REALLY looks increasingly likely. We are not ready!
120
Florian Mai @florianmai.bsky.social · 18/11/2024
I haven't found the "opinions I am likely to disagree with" feed. Can this be done? www.theverge.com/24295933/blu...
theverge.com
Here’s some cool stuff you can do with Bluesky
It’s not just an Alf pics repository.
010
Reposted by Florian Mai
Naomi Saphra @nsaphra.bsky.social · 02/07/2023
hard for me to get engagement on this platform but it’s ok I’m just waiting for all the AI researchers to show up so I can go mega viral saying shit like “who called it Xavier Initialization and not Weighting For Glorot”
5777