Ethan Mollick @emollick.bsky.social · 17hAlso a loose affiliation of millionaires and billionaires. 2100
Ethan Mollick @emollick.bsky.social · 18hJust thinking that Paul Simon wrote that we were living in the "days of miracle and wonder" because of the availability of long-distance calls and slow motion cameras. 81117
Ethan Mollick @emollick.bsky.social · 18hIts reasonable for mathematicians to point out that more proofs are not always the same thing as advancing mathematics. But it suggests a need for new goals for what math is trying to do. I think the same thing will happen everywhere. 100x more PowerPoint or code is not always progress - what is? 6656
Ethan Mollick @emollick.bsky.social · 22hRandomized trials with GPT-4o: "AI access raises test scores, and a smaller gain persists a week later. Gains remain for students who use AI as a tutor (“augmentation”) and fade for students who have AI write for them (“automation”)" Across all experiments, positive effects arxiv.org/abs/2607.08849 36412
Ethan Mollick @emollick.bsky.social · 09/10/2026Paywalled copy here: www.thelancet.com/journals/lan...thelancet.comConversational diagnostic artificial intelligence in ambulatory primary care: a prospective feasibility studyAlthough further research is needed, this study shows the initial feasibility of conversational AI in a real-world setting—assessed via conversation safety and quality, as well as user acceptance—and ... 090
Ethan Mollick @emollick.bsky.social · 09/10/2026From The Lancet: in an urgent care setting, the advice of the obsolete Gemini 2.5 Pro & Gemini 2.5 Flash (without access to patient medical records) were rated of similar quality to doctors by other physicians. There was no safety issues spotted. Models have gotten significantly better since. 310516
Ethan Mollick @emollick.bsky.social · 08/10/2026This document is going to be an assigned reading in college classes that cover this moment in time, there's a lot happening in a few paragraphs... www.ahmath.org/statements 2535874
Ethan Mollick @emollick.bsky.social · 08/10/2026The few attempts to measure similar things (METR Long Tasks, GDPval) all became saturated earlier this year as AIs began to do days of work at a high quality level. 2140
Ethan Mollick @emollick.bsky.social · 08/10/2026As a researcher who did early some work on the productivity impacts of AI chatbots using RCTs, I’d note a lack of similar studies since the dawn of true agents last fall Partially that is newness & partially research design challenges, but I suspect we are missing some large & important effects 6451
Ethan Mollick @emollick.bsky.social · 08/10/2026Anyhow, I think we are now done with the last round of skeptical responses to AI after Navier-Stokes. Yes, AI is now capable of producing novel proofs to hard problems that have eluded us The next line will be that this is only constrained to math because of RL training. I suspect not, we will see. 9996
Ethan Mollick @emollick.bsky.social · 07/10/2026Some early first-hand accounts of the experience of encountering a narrow superhuman intelligence as mathematicians grapple with the hundreds of big AI proofs released by OpenAI. Problems solved in inhuman ways that make us wonder what it means to actually know things... scottaaronson.blog?p=10169 49915
Ethan Mollick @emollick.bsky.social · 07/10/2026When people talk about AI easily replacing workers, I think about this ethnography of copier repair technicians in the 1990s. Reading it, you see how work is complicated and improvised by workers and informal and badly documented. (But eventually technology did largely eliminate this job) 4674
Ethan Mollick @emollick.bsky.social · 07/10/2026(And I am seeing rapid increases in the ability of AI to do novel work in my field of economic sociology, with nearly autonomous research getting very close to top journal level) 5543
Ethan Mollick @emollick.bsky.social · 07/10/2026Context, OpenAI cracked or made important progress in hundreds of critical problems today : openai.com/index/sharin...openai.comSharing AI progress in mathematicsOpenAI publishes new results on open problems in mathematics from an internal frontier model and shares Lean proof formalizations and research details on GitHub. 2565
Ethan Mollick @emollick.bsky.social · 07/10/2026After a brief but institution-eroding slop science era, it increasingly looks like we are going to have 2 revolutions from AI: 1) Everything ever published will be re-read and re-judged in ways that human scientists never anticipated 2) Novel discoveries will start to come faster than we can absorb 916311
Ethan Mollick @emollick.bsky.social · 06/10/2026I think people are muddling through to some extent. But also some of the early warnings about things like mass psychosis seem to not be appearing in the data. There are lots of reasons why we may be underestimating negative impacts but also many observers expected a lot of obvious stuff by now. 280
Ethan Mollick @emollick.bsky.social · 06/10/2026And I also think the positive use cases (like the evidence shows current AIs provide good health information & financial advice & lead to startups etc.) are also undermeasured and underdiscussed. We do need to worry about many dangers & many risks, but its worth building on successes as well. 1341
Ethan Mollick @emollick.bsky.social · 06/10/2026I don't think that is the primary principle at work explaining the lack of incidents. 030
Ethan Mollick @emollick.bsky.social · 06/10/2026I think most people who have been closely watching AI for the past few years are likely surprised by how relatively rare terrible AI incidents have been, given a billion users overall and hundreds of millions of corporate users. Doesn’t mean that will continue, but still its surprising. 1317811
Ethan Mollick @emollick.bsky.social · 06/10/2026Yes, it makes it much easier to use computer control if it is running locally. 220
Ethan Mollick @emollick.bsky.social · 06/10/2026Here is a good explanation from the Claude team about why they are making the switch. 3311
Ethan Mollick @emollick.bsky.social · 06/10/2026Big swing happening from this summer's trend to have AI running on virtual machines on your own computer (Claude and ChatGPT apps) to moving them back into you having a dedicated VM the cloud (Claude Cowork just announced they were doing this, also Grok Bot, Dot, and others). Cloud is mostly better. 6644
Ethan Mollick @emollick.bsky.social · 05/10/2026AI policy right now, especially from the Labs, makes me think of this poem. If you believe superintelligence is near, you don't need to make hard decisions about how to build AI to make the world better. Just wait for ASI & it will decide for us, for better or worse. But if that doesn't happen... 4615
Ethan Mollick @emollick.bsky.social · 04/10/2026I am quite impressed with how pretty the results are. Here it is on github: github.com/emollick/brut 1140
Ethan Mollick @emollick.bsky.social · 04/10/2026"create the ultimate brutalist city builder. you should have lots of controls over buildings in ways that are delightful and intuitive to people..." Here is Fable 5.1, a little over a year since GPT-5 below & a single shot (compared with many rounds with GPT-5.6) Quite good: brut-city.netlify.app 4657
Ethan Mollick @emollick.bsky.social · 03/10/2026Eh. The idea is the whole sovereign thing is a bit of a house of cards. 220
Ethan Mollick @emollick.bsky.social · 03/10/2026🎵 Wave of distillation 🎵(sung to the tune of the Pixies) US labs say Chinese labs distill their models. And Europe's new “sovereign model” is fine-tuned on data generated by… GLM and Qwen. 4619
Ethan Mollick @emollick.bsky.social · 03/10/2026I think my upcoming book, Co-Existence, might be the first to include a blurb written specifically for AI readers, in this case from Tyler Cowen (whom I thought AIs would respect) The book website (with elaborate pre-order bonus) also has a page for AIs. Its out October 20: co-existence.ai 2382
Ethan Mollick @emollick.bsky.social · 03/10/2026People asked for me to upload the Bitter Lesson video from this piece to YouTube, so here you are: youtu.be/OAwat51S_Sk 2599
Ethan Mollick @emollick.bsky.social · 02/10/2026“We find that on medium-length, well-defined accounting tasks, frontier AI models are now faster and more accurate than junior accountants, even the best one in our study.” Eighteen months ago they scored well below human accountants Good discussion here: www.mercor.com/blog/human-b... 513212
Ethan Mollick @emollick.bsky.social · 01/10/2026I wrote about the thing I underestimated most about progress in AI: its ability to self-organize to accomplish tasks. Also, what that means for agents like Muse and Dots, along with a music video explaining why we keep relearning The Bitter Lesson. open.substack.com/pub/oneusefu...open.substack.comThe Dot and the SwarmBenefitting from the Bitter Lesson 59017
Ethan Mollick @emollick.bsky.social · 30/09/2026They can use the channels built for humans, but which have far too much friction for any human to use. Your Clawlike is happy to navigate a phone tree or stay on hold or have an awkward conversation with an agent where they get rejected... and then try again. 3745
Ethan Mollick @emollick.bsky.social · 30/09/2026As a heads-up: every firm's customer service agents are about to be overwhelmed with Dots & Muses negotiating for better deals using voice/chat channels made for humans. Stories of people delegating this sort of work to their agents & saving money are popping up and are going to only go more viral. 1017919
Ethan Mollick @emollick.bsky.social · 29/09/2026Leaving aside Anthropic's incentives for publishing this research, there is no doubt that open weights models will soon create the same security threats that closed source models have been demonstrating, except without guardrails. We are close, so plan accordingly. www.anthropic.com/research/glm... 714627
Ethan Mollick @emollick.bsky.social · 29/09/2026It also has a natural analog on teams as a persistent entity with a job that you can delegate to using existing team communication tools like Slack. I expect more Clawlikes soon, though I don’t think they are the final form of this type of AI. 0150
Ethan Mollick @emollick.bsky.social · 29/09/2026I only used the newly announced ChatGPT Dots briefly before launch, so can’t offer a detailed review, but I found it to be a good entry in the rapidly-expanding Clawlike category with Muse & Grokbot. Using a capable model with access to your data as an assistant & second opinion is remarkably useful 3470
Ethan Mollick @emollick.bsky.social · 29/09/2026Hey, Claude: "I asked an earlier Claude to Remove the Squid from All Quiet on the Western Front. Now you can make movies and such. I need you to show how far you have come" & I pasted in the post below. That was it. There is some clever stuff here, models have a real sense of humor at this point 4905
Ethan Mollick @emollick.bsky.social · 28/09/2026Since people liked this, and the weak point was the text to speech, I asked Opus to update with an ElevenLabs voice. This is the same video, but with much better narration. 0182
Ethan Mollick @emollick.bsky.social · 28/09/2026Sources and footnotes here: built-for-something-else.netlify.appbuilt-for-something-else.netlify.appBuilt for something else: sources and notesEvery claim in the video “Built for something else,” in the order you hear it, with the sources behind it. 1312
Ethan Mollick @emollick.bsky.social · 28/09/2026Yeah, I only used bad TTS, Elevenlabs would have been better 070
Ethan Mollick @emollick.bsky.social · 28/09/2026Sorry, I sent that quickly, did not mean to provoke by being flippant. Mythos was Fable without guardrails (I have been told this by Anthropic as well), I do not believe that open weights models have caught up to that, and the cyber evaluators (AISI, etc) don't seem to either. But we will see! 140
Ethan Mollick @emollick.bsky.social · 28/09/2026I was reflecting on how many random contingencies went into creating the current accelerating AI moment and had Opus 5.5 put together a video, inspired by old science show Connections, to try and point out how a series of initially unrelated ideas led to the moment we are in now. 1012524
Ethan Mollick @emollick.bsky.social · 27/09/2026Qualitatively, there is now a larger gap between open & closed models than there has been in awhile. Fable/Astra class models are agentic in a way pre-Fable models are not, for better or worse. None of the open models have crossed that line, yet. When they do, it will be a jump. 7754
Ethan Mollick @emollick.bsky.social · 27/09/2026Also please feel free to fight in the comments about "pure LLMs" versus multimodal LLMs versus multimodal LLMs with tool use or whatever. I totally get all the caveats, but you also get what I mean. 41090
Ethan Mollick @emollick.bsky.social · 27/09/2026It is impossible to imagine a general intelligence that is artificial, though. What would you even call that? 3630
Ethan Mollick @emollick.bsky.social · 27/09/2026It is strange how much LLMs turned out to be able to solve such a wide range of hard problems that would not, instinctively, seem to be problems that a model of human language would be able to solve This is from a Stanford project that let Astra drive a robot in a kitchen tml.stanford.edu/homebody/ 2339745