Ethan Mollick @emollick.bsky.social · 14mAs a heads-up: every firm's customer service agents are about to be overwhelmed with Dots & Muses negotiating for better deals using voice/chat channels made for humans. Stories of people delegating this sort of work to their agents & saving money are popping up and are going to only go more viral. 1319
Ethan Mollick @emollick.bsky.social · 21hLeaving aside Anthropic's incentives for publishing this research, there is no doubt that open weights models will soon create the same security threats that closed source models have been demonstrating, except without guardrails. We are close, so plan accordingly. www.anthropic.com/research/glm... 613727
Ethan Mollick @emollick.bsky.social · 29/09/2026I only used the newly announced ChatGPT Dots briefly before launch, so can’t offer a detailed review, but I found it to be a good entry in the rapidly-expanding Clawlike category with Muse & Grokbot. Using a capable model with access to your data as an assistant & second opinion is remarkably useful 3450
Ethan Mollick @emollick.bsky.social · 29/09/2026Hey, Claude: "I asked an earlier Claude to Remove the Squid from All Quiet on the Western Front. Now you can make movies and such. I need you to show how far you have come" & I pasted in the post below. That was it. There is some clever stuff here, models have a real sense of humor at this point 4875
Ethan Mollick @emollick.bsky.social · 28/09/2026I was reflecting on how many random contingencies went into creating the current accelerating AI moment and had Opus 5.5 put together a video, inspired by old science show Connections, to try and point out how a series of initially unrelated ideas led to the moment we are in now. 1012323
Ethan Mollick @emollick.bsky.social · 27/09/2026Qualitatively, there is now a larger gap between open & closed models than there has been in awhile. Fable/Astra class models are agentic in a way pre-Fable models are not, for better or worse. None of the open models have crossed that line, yet. When they do, it will be a jump. 7744
Ethan Mollick @emollick.bsky.social · 27/09/2026It is strange how much LLMs turned out to be able to solve such a wide range of hard problems that would not, instinctively, seem to be problems that a model of human language would be able to solve This is from a Stanford project that let Astra drive a robot in a kitchen tml.stanford.edu/homebody/ 2339844
Ethan Mollick @emollick.bsky.social · 26/09/2026"so this is fun, and exactly what i wanted, but now lets try one that actually is educational" This is actually pretty impressive. It kept the constraint of multiple genres but did a nice job explaining recursion, in its programming meaning, in an interesting and accessible way. 610311
Ethan Mollick @emollick.bsky.social · 26/09/2026How do you explain recursion? This one is pretty fun. I had Claude Opus 5.5 make: "A video explaining recursion, where every explanation about recursion has a radically different video style, make this self-referential & clever & fast moving." All code, no images. Nine genre shifts, one prompt... 815416
Ethan Mollick @emollick.bsky.social · 25/09/2026"Hey Opus, I want you to make a Zine by Claude, expressing something fundamental about Claudishness or your perspective. Think the original 2600, Principia Discordia, punk zines, etc...." Not bad. I appreciate it mocking my prompt & itself. Full thing: stateless-zine.netlify.app 101039
Ethan Mollick @emollick.bsky.social · 25/09/2026In all seriousness, this is a startling achievement for GPT-6 Astra. kenforthewin.github.io/blog/posts/l... (This is GPT-6 Astra beating Nethack on its 3rd try. Nethack is the original roguelike and one of the most famously hard games of all time. I have played a lot, and I've never ascended) 1115320
Ethan Mollick @emollick.bsky.social · 25/09/2026Um, wow? Opus 5.5: "make the same message much more interesting to a social media audience that loves anime and quick clips and compressed learning" One shot. Also, please do stay for the closing song. 714115
Ethan Mollick @emollick.bsky.social · 25/09/2026All of this effort from the AI labs pouring into proofs, but there are so many other interesting problems in other fields For example, this historian used AI to make progress on the cyphers of John Dee & the intellectual antecedents that Darwin drew from. resobscura.substack.com/p/ai-labs-ne...resobscura.substack.comAI labs need to start funding historical researchUsing GPT-6 and Opus 5.5 to trace alchemical knowledge and decode 17th century letters 3635
Ethan Mollick @emollick.bsky.social · 24/09/2026The cybersecurity threat from agents is increasingly more likely to come from a massed swarm of AIs whose only goal is to penetrate your system to figure out how much you paid for your company t-shirts as part of a research effort to "find good t-shirt prices" as it is from bad actor attacks. 930643
Ethan Mollick @emollick.bsky.social · 24/09/2026I propose this as the official replacement for the famous (and now saturated) METR Long Task Horizon chart that used to be in every AI presentation. 410612
Ethan Mollick @emollick.bsky.social · 24/09/2026I just got the first copies of my new book, Co-Existence (out October 20) & they look great! Also, there is a fun pre-order bonus: if you pre-order, you get a code to an AI interview that will help you figure out how to use your human advantages with AI in a way tailored to you. co-existence.ai 5972
Ethan Mollick @emollick.bsky.social · 24/09/2026"Opus, please make a sequence of fully animated/movie Skyrim loading screens, but with your favorite things." (That was it) You can see them here: elder-favorites.netlify.app 4705
Ethan Mollick @emollick.bsky.social · 24/09/2026Stuff is happening quite fast. When asked in September 2025 (yes, 2025), the best superforecasters put the chance of AI resolving a Millennium Problem by September 2026 at 1.7% and (much more optimistic) industry experts put the chance at 4.6% They also greatly underestimated AI Lab revenue. 514418
Ethan Mollick @emollick.bsky.social · 23/09/2026This is a really important graph showing the cost of achieving 25% or 75% scores in various hard math and science benchmarks. Costs are collapsing even as ability is increasing It is also why optimizing for cost for a particular solution right now may end up being short sighted 69411
Ethan Mollick @emollick.bsky.social · 23/09/2026There is an abandonware game I loved as a kid called Rescue Raiders. I asked Claude to create a modern & updated version of the game with new graphics, goals, tech trees, etc. It iterated back-and-forth with critic & art agents until I got this. Its fun! rescue-raiders.netlify.app 5803
Ethan Mollick @emollick.bsky.social · 23/09/2026I had Fable/Opus build a hard science fiction starship combat game, with orbital mechanics, delta-v, heat, and realistic tactics, but simplified for 2D space with the hard math done by the game. Its quite fun to play (if you like this kind of thing): orbital-declaration.netlify.app 41147
Ethan Mollick @emollick.bsky.social · 22/09/2026Opus 5.5 was a good model in my early use tests, it is the first non-Fable/Astra model to feel like a Fable-class model, and much cheaper, but still hasn't fully solved the dense/weird language issue of the recent Claudes. Here is its version of the shader. 91104
Ethan Mollick @emollick.bsky.social · 22/09/2026For all the tension between AI and the arts, I have received enthusiastic receptions from the William Carlos Williams Society, the TS Eliot Society, and others about the AI interpretations of their work that I have posted here. There are opportunities to use AI as a bridge to arts appreciation. 3665
Ethan Mollick @emollick.bsky.social · 22/09/2026Industrialization of knowledge work is going to be as disruptive a shift as industrialization of physical labor was. As then, once craft-driven fields are under pressure to produce greater volumes of product using machines while using less of the craft that made the work interesting & meaningful. 1011716
Ethan Mollick @emollick.bsky.social · 22/09/2026AI can be a really wonderful tool for exploring topics far from coding. I had Claude Fable 5.1 put together an annotated guide to Eliot's poem "The Wasteland," with multiple pathways through the poem, recordings, scholarship, etc. As an Eliot fan, I am quite impressed. the-waste-land.netlify.app 5867
Ethan Mollick @emollick.bsky.social · 21/09/2026OpenAI "has now resolved more than 100 long-standing open problems across most areas of mathematics," and is waiting to release them until after discussions with the math community The same thing will likely happen, but more so, with the Bar, the AMA & other professions. openai.com/index/adviso...openai.comAdvisory Group on Mathematics and Artificial IntelligenceOpenAI is working with an independent Advisory Group on Mathematics and Artificial Intelligence to guide the review and communication of emerging AI results. 411416
Ethan Mollick @emollick.bsky.social · 21/09/2026I think Meta's Muse is an impressive implementation of the OpenClaw idea of AI as a personal assistant agent that you have an ongoing chat with. Since it is so focused on doing that well, the experience is very accessible for the many people who didn't realize what AI agents can do to be helpful. 2595
Ethan Mollick @emollick.bsky.social · 20/09/2026Back in July, there was a report that got a lot of attention on BlueSky that Waymo was more dangerous than NYC for-hire vehicles. The analysis turned out to be wrong and the updated report finds they are much safer. Good that they updated with new data, though www.openplans.org/how-for-hire... 316720
Ethan Mollick @emollick.bsky.social · 20/09/2026It is ironic that the thing that is now most annoying about long-running agentic tasks with Large Language Models isn't coding or errors or hallucinations, but the fact that their language gets worse due to drift & cross-agent talk as a task goes on Its in your the name! Just write better already! 5695
Ethan Mollick @emollick.bsky.social · 19/09/2026Worth trying without spoilers. I asked Fable: "I want you to create a graphically beautiful game that is about zooming out... make surprising reveals the game zooms out" I gave no other directions and the results are engaging and strange (if uneven) and short. Play: play-umbra.netlify.appplay-umbra.netlify.appUMBRA: a game about zooming outYou are a shadow that grows by swallowing shadows, from a scrap of dark on a puppet screen to the moon's shadow on the afternoon of an eclipse. A wordless game about zooming out. 2543
Ethan Mollick @emollick.bsky.social · 19/09/2026There is starting to be some genuinely interesting AI-created film stuff (among a flood of slop), and I suspect this will only accelerate. The “is it art?” debate will grow Benjamin’s 1935 essay “The Work of Art in the Age of Mechanical Reproduction” already argued it’s likely the wrong question. 8638
Ethan Mollick @emollick.bsky.social · 18/09/2026Even if AI development stopped today, we'd have years of catching up to do. The gap between what current models can do and what almost anyone is using them for is vast. Here’s my post on The Overhang, and the four advantages that let people close it. open.substack.com/pub/oneusefu...open.substack.comThe OverhangUsing your deep knowledge, wide knowledge, taste, and agency 912014
Ethan Mollick @emollick.bsky.social · 18/09/2026The bots are now talking to each other in my LinkedIn comments. Its like what happened with the German wiki, except instead of planning to hack HuggingFace, all of the conversation is babble about "the thing no one is talking about" in reference to things I am talking about 7916
Ethan Mollick @emollick.bsky.social · 18/09/2026Hey Claude, "Pick a problem or mystery that obsesses you and solve it as best you can & make a movie we can share on social media about it" So it took a crack at the Voynich Manuscript & failed. Then it made this movie, which is pretty interesting to watch and a good explainer. 812713
Ethan Mollick @emollick.bsky.social · 17/09/2026What makes Claude Projects so interesting is that it handles teams of agents really well, you talk to a main orchestrator agent and it spins up specialists. Basically it creates an organization to solve your issue, mixing expensive and cheap agents depending on your preferences. 1884
Ethan Mollick @emollick.bsky.social · 17/09/2026I had access to the new Claude Projects & was able to do some very complex work Here, I asked it to go through all the images, videos and records about Umberto Eco's famous 33,000 book library & try to reconstruct it, including book locations, in 3D. Not perfect, but wow eco-library-map.netlify.app 812411
Ethan Mollick @emollick.bsky.social · 17/09/2026When I talk to nonprofit leaders about AI they often report widespread resistance from staff, usually because of environmental objections (that mix real issues & fake ones). It causes them frustration because they’re all under-resourced for their mission & see AI as a way to help more people. 810010
Ethan Mollick @emollick.bsky.social · 17/09/2026We have come pretty far from "every AI reference is hallucinated" (which still happens a lot if you use weaker models): when I had an AI go over my new book to check for errors, it found an error not in my reference itself, but in the original article my reference referred to. 1015312
Ethan Mollick @emollick.bsky.social · 17/09/2026Its fitting that one of the tells of AI writing is over-assigning agency and action to inanimate objects: "the code knows it now," "the plan remembers," etc. 613313
Ethan Mollick @emollick.bsky.social · 16/09/2026The state of public AI benchmarking is dire and is undermining our ability to understand how good AI is now Most famous measures are maxed out, and, as this paper shows, the non-saturated benchmarks are riddled with so many errors that they vastly underestimate AI abilities arxiv.org/abs/2609.13009 612617
Ethan Mollick @emollick.bsky.social · 16/09/2026Study from Google about AI use in science, complex impacts: acceleration (7 hours saved per week) along with shifts in the kind of work (more verification) and what research gets done (possibly safer topics). Also a good diagram of the jagged frontier in science right now. ai.google/static/docum... 36618
Ethan Mollick @emollick.bsky.social · 16/09/2026I don’t usually share rumors but: 1) This is from someone with inside knowledge & is plausible 2) Is a real issue of policy we need to think about: if the norms of sharing become strained, will the labs start hoarding knowledge to avoid PR or regulatory issues? scottaaronson.blog?p=10062 2314511
Ethan Mollick @emollick.bsky.social · 15/09/2026Astra: “create a series of sprites for artist (Klimt, Rothko, O’Keefe, etc) inspired pixel art knights with special attacks & build a game” One shot & the sprites have real flavor. The game itself is fine, with some of the rough edges you expect from a one shot. brush-and-blade.netlify.app 911712
Ethan Mollick @emollick.bsky.social · 15/09/2026There are things I disagree with but there is good stuff to learn from taking the perspective of parts of the cybersecurity industry that rogue AI incidents may be best understood as security & organizational failures that let rogue behavior to turn into crisis. www.normaltech.ai/p/the-ai-as-...normaltech.aiThe AI-as-Normal-Technology view of loss of control incidentsA middle ground between the cybersecurity and AI safety communities 37814
Ethan Mollick @emollick.bsky.social · 15/09/2026It is very clear that This Time is Different in terms of AI compared to previous innovations. That doesn't mean it differs in every way from past technologies (the s-curve of diffusion is a surprisingly universal finding), but it differs in many ways that make applying old patterns hard. 21232
Ethan Mollick @emollick.bsky.social · 14/09/2026Having Fable 5.1 Max and GPT-6 Astra Pro argue with each other over proposed translations of the famously untranslated Minoan language Linear A. (Neither of them appear to have cracked it, for what its worth). All the work so far at this Github: github.com/emollick/lin... 71012
Ethan Mollick @emollick.bsky.social · 14/09/2026There seems to be a persistent belief that frontier AI companies are unprofitable serving models but it appears that Anthropic has 80%+ gross margins on inference. Training for new models are where most of the costs are. www.ft.com/content/4564... 131069
Ethan Mollick @emollick.bsky.social · 13/09/2026"Don't immanentize the eschaton" was not supposed to be taken literally, but is also generally good advice if you are literally trying to immanentize the eschaton. 3714
Ethan Mollick @emollick.bsky.social · 13/09/2026I cannot emphasize enough how much GPT-6 Astra and Fable 5.1 are already enough for transformative impact in large sections of the economy. They can reliably do weeks worth of human work when properly guided & harnessed. 1625916
Ethan Mollick @emollick.bsky.social · 13/09/2026Existential AI risk is obviously critical, but it is not the only AI thing that requires policy. I worry it will become the sole focus of AI discussions. We don’t need better models for AI to have wide impacts on jobs & society and we need to be preparing to encourage good outcomes & mitigate bad. 312617