Jeff Clune @jeffclune.com · 16/05/2025Why do we tolerate loud motorcycles? I can't walk around with a crazy loud speaker, yet we allow people to have (and companies to sell) insanely loud machines that torture everyone else. They can be made quiet or silent...we should start heavily taxing and fining noise polluting machines. 12477
Jeff Clune @jeffclune.com · 29/04/2025I greatly enjoyed “The Spectrum of AI Risks” panel at the Singapore Conference on AI. Thanks @teganmaharaj.bsky.social for great moderating, Max Tegmark for the invitation, and the organizers and other panelists for a great event! PS. Do I really have sad resting panel face? 😐 281
Jeff Clune @jeffclune.com · 27/04/2025Tom @rockt.ai did a great job in his #ICLR2025 keynote on open-endedness of explaining the ideas we are all so passionate about. A huge thanks for the kind words and for featuring our work, including Jenny Zhang's OMNI. cc @kennethstanley.bsky.social @joelbot3000.bsky.social 0100
Jeff Clune @jeffclune.com · 25/04/2025Very excited for this keynote by @_rockt! Awesome to see open-endedness go from a niche (😉) area to a keynote at #ICLR ! 🌱🌿🌳🌲🍀🌍✨ 📈 🧬🧪 cc @joelbot3000.bsky.social @kennethstanley.bsky.social 0101
Jeff Clune @jeffclune.com · 12/02/2025Human surveys confirm a majority of the tasks are clear & valid. Moreover, the automated scoring aligns closely with human judgment for all but the hardest tasks. [7/9] 120
Jeff Clune @jeffclune.com · 12/02/2025Different scientist models uncover wildly creative, out-of-the-box probes. For example, Claude Sonnet 3.5 found GPT-4o can successfully design alien communication protocols!! 🌌🤖[6/9] 110
Jeff Clune @jeffclune.com · 12/02/2025We also explored different scientist/subject pairings. Results with Llama3-8B as the subject instead of GPT-4o reveal unique failure modes & emerging skills, offering interesting insights into each model’s capabilities. [5/9] 110
Jeff Clune @jeffclune.com · 12/02/2025ACD mimics community exploration: endlessly generating tasks (in code with automated scoring) probing for new capabilities or weaknesses—covering topics from string games to complex puzzles. In a GPT-4o self-eval, ACD uncovered thousands of capabilities (visualized here)! [3/9] 110
Jeff Clune @jeffclune.com · 12/02/2025ACD automatically creates a concise "Capability Report" of discovered capabilities and failure modes, enabling quick inspection and easier dissemination of results or flagging issues pre-deployment. [2/9] 110
Jeff Clune @jeffclune.com · 12/02/2025Introducing Automated Capability Discovery! ACD automatically identifies surprising new capabilities and failure modes in foundation models, via "self-exploration" (models exploring their own abilities). Led by @cong-ml.bsky.social & @shengranhu.bsky.social 🔬🤖🧠🔎 [1/9] 1193
Jeff Clune @jeffclune.com · 22/01/2025Research Idea: Predict abstract properties of text in addition to next-word prediction as helpful auxiliary losses. Has this been tried? (1/n) 180
Jeff Clune @jeffclune.com · 13/01/2025A copy of Tim's @rockt.ai has arrived. As I say on the back cover, I read it cover to cover and recommend it to those looking to ramp up on modern AI. Congrats and thanks Tim! 2180
Jeff Clune @jeffclune.com · 08/01/2025It's an honor that The AI Scientist is #1 on this list! www.linkedin.com/feed/update/... Congrats @chris-lu.bsky.social @cong-ml.bsky.social @RobertTLange @hardmaru.bsky.social @jfoerst.bsky.social 0233
Jeff Clune @jeffclune.com · 07/01/2025I strongly believe the development of superintelligence is inevitable. I do not believe humanity has the ability to not invent it, given its economic, scientific, and military value. Many people say things like “we should not build AGI” without realizing that outcome is virtually impossible. 1/ 3171
Jeff Clune @jeffclune.com · 17/12/2024I guess AI and human scientists stand on the shoulders of the same giants. 0100
Jeff Clune @jeffclune.com · 17/12/2024NO WAY! This is nearly EXACTLY one of the paper ideas The AI Scientist came up with! It was my favorite idea it generated, & made us very impressed with its creativity & good taste in an ML paper proposal. Cool to see humans agree! The humans def executed better though. x.com/BrantonDeMos... 1213
Jeff Clune @jeffclune.com · 16/12/2024Dream finish to #NeurIPS2024: a dinner with Ted Chiang (one of my favorite writers), the excellent @alisongopnik.bsky.social, many amazing open-ended researchers (including @cedcolas) & other excellent people. Thanks to the tremendous IMOL organizers! Until next time! 🦎🧬🦖🦣 0220
Jeff Clune @jeffclune.com · 16/12/2024Lots of interest in ADAS! Thanks everyone, and congrats Shengran Hu and @cong-ml.bsky.social! 🚀🚀🚀 0103
Jeff Clune @jeffclune.com · 16/12/2024Jenny did a great job presenting OMNI-EPIC during the IMOL Oral. Congrats Jenny Zhang! arxiv.org/abs/2405.15568 071
Jeff Clune @jeffclune.com · 16/12/2024A huge congratulations to Shengran Hu and @cong-ml.bsky.social on ADAS winning an Outstanding Paper Award at the #NeurIPS2024 OWA workshop!! And nice jumps everyone!! 🦘 🦘 🦘 1121
Jeff Clune @jeffclune.com · 15/12/2024Is Developing AGI a socially responsible goal? I enjoyed the thoughtful, passionate SoLaR panel on this critically important topic w/ @yoshuabengio.bsky.social, @mmitchell.bsky.social, & @jfoerst.bsky.social. It's clear everyone is both worried & cares deeply about making this go as well as possible 2151
Jeff Clune @jeffclune.com · 14/12/2024Great to catch up with @Yoshua_Bengio, Aaron Courville, and @hugo_larochelle at #NeurIPS2024, with lots of talk about AI Safety. Thanks @CIFAR_News for supporting our research! 0140
Jeff Clune @jeffclune.com · 10/12/2024Our work Automated Design of Agentic Systems (w/ Shengran Hu & @cong-ml.bsky.social) will have ✨two orals✨ @ #NeurIPS2024 workshops (LanGame Sat 10:20, OWA Sun 4:50). Please come visit us😃 We would also love to chat about open-endedness, LLM agents, etc. Come by if you want to meet! 0122
Jeff Clune @jeffclune.com · 25/11/2024ML / AI Twitter is reforming here, which is awesome! But is this an expansion or switch? I wonder if people will keep posting here too, or if the entire community will jump ship. Blue Bird --> aXshes --> Butterfly rising? 1211
Jeff Clune @jeffclune.com · 25/11/2024ML / AI Twitter is reforming here, which is awesome. But is this an expansion or switch? I wonder if people will keep posting here too, or if the entire community will jump ship. Blue Bird --> aXshes --> Butterfly rising? 030
Jeff Clune @jeffclune.com · 22/11/2024It's so interesting watching a new community re-form. It feels like seeing people greet and welcome each other "on the other side of the veil" or in the heavens from stories/myths. Any coincidence heaven was/is imagined to exist up in the clear Blue Sky? 181
Jeff Clune @jeffclune.com · 20/11/2024Apparently some newspaper accidentally swapped the captions for far side and Dennis the Menace. The result is awesome. 181