Reposted by Max LittleGeorgia Tomova @georgiatomova.bsky.social · 13/05/2026Why we should rethink causal mediation, and what to do instead? Come to hear the answer from Vanessa Didelez at the next CIIG seminar! The seminar will be hybrid. If you are in London, come join us in person at UCL! Otherwise, you can join on Zoom as usual. Registration links in comment below. 44615
Reposted by Max LittlePeter Tennant @pwgtennant.bsky.social · 12/05/2026Does engaging with the arts help to slow the aging process? According to this new paper from Fancourt et al the answer is yes! Alas, the paper is actually unsuitable for such claims. Let’s examine by discussing how you'd establish whether arts engagement actually caused future health! 1/18 🧵academic.oup.comDoes leisure activity matter for epigenetic ageing? Analyses of arts engagement and physical activity in the UK Household Longitudinal StudyAbstractBackground and Objectives. Over the past decade, ageing clocks have become widely adopted as important tools for understanding biological ageing an 813160
Reposted by Max LittlePeter Tennant @pwgtennant.bsky.social · 25/02/2026This recent RCT of an "AI stethoscope" claims the technology "shows promise" for diagnosing cardiovascular conditions. It does not. It is a textbook example of the risks of conducting unprincipled 'per protocol analyses'. Once again, peer review at a major medical journal has failed. 🧵 1/ 8436186
Reposted by Max LittleDarren Dahly @statsepi.bsky.social · 02/09/2025Depends on how you spin it I guess (screenshots from 2 different articles). People working at universities are pushed so incredibly hard to ensure that every study is a breakthrough that they just...lie. All the time. Probably without even realizing it. www.standard.co.uk/news/tech/im... 4308
Reposted by Max LittleRodney Brooks @rodneyabrooks.bsky.social · 01/01/2025Every Jan 1 I post a scorecard on predictions I made, with dates, on Jan 1, 2018 on cars (self-driving), robots, AI, & ML, and on human spaceflight. Besides telling which turned out right and which wrong in the last year I also talk a lot of smack about these topics. rodneybrooks.com/predictions-...rodneybrooks.comPredictions Scorecard, 2025 January 01 – Rodney Brooks 711045
Reposted by Max LittlePasquale Minervini @neuralnoise.com · 21/12/2024LLMs can be also seen as big bags of heuristics: arxiv.org/abs/2410.21272arxiv.orgArithmetic Without Algorithms: Language Models Solve Math With a Bag of HeuristicsDo large language models (LLMs) solve reasoning tasks by learning robust generalizable algorithms, or do they memorize training data? To investigate this question, we use arithmetic reasoning as a rep... 082
Reposted by Max LittleThe Viking (Gunnar Blohm) @gunnarblohm.bsky.social · 21/12/2024www.science.org/content/arti...science.orgU.S. science funding agencies roll out policies on free access to journal articlesNIH and DOE are first to act, with implementation by all set to begin by end of 2025 1128
Reposted by Max LittleGary Marcus @garymarcus.bsky.social · 21/12/2024𝗢𝟯 𝘄𝗮𝘀 𝘁𝗿𝗮𝗶𝗻𝗲𝗱 𝗼𝗻 𝟳𝟱% 𝗼𝗳 𝘁𝗵𝗲 𝗽𝘂𝗯𝗹𝗶𝗰 𝘀𝗲𝘁 𝗳𝗼𝗿 𝗔𝗥𝗖-𝗔𝗚𝗜. OpenAI did not disclose this in the video. Sam said they didn’t target the test. Never trust a staged demo. Never trust a product you haven’t tried. Never trust OpenAI. 2438655
Reposted by Max LittleGary Marcus @garymarcus.bsky.social · 21/12/2024o3, AGI, and the art of the demo. Long read on what OpenAI didn’t tell you yesterday. garymarcus.substack.com/p/o3-agi-the...garymarcus.substack.como3, AGI, the art of the demo, and what you can expect in 2025OpenAI’s new model was revealed yesterday; its most fervent believers think AGI has already arrived. Here’s what you should pay attention to in the coming year. 96312
Max Little @maxal.bsky.social · 21/12/2024Likewise, a simple adversarial strategy beats "superhuman" Go-playing algorithms: goattack.far.ai It's wise to remember that there is no scientific consensus on what "intelligence", actually is.goattack.far.aiAdversarial policies in Go 051
Max Little @maxal.bsky.social · 21/12/2024Just for those who don't know: the vast majority of open problems in maths, are not numerical in nature. 010
Reposted by Max LittleTimothy Gowers @wtgowers.bsky.social · 21/12/2024The questions have numerical answers, so it is easy to check whether it gets them right. 111
Reposted by Max LittleGary Marcus @garymarcus.bsky.social · 21/12/2024How many times do we have to see this same movie, where an AI beats some benchmark and influencers gleefully shout “It’s So Over” without even trying out the AI and then on careful inspection the AI turns out to not be robust or reliable? Thousands? (It’s already been hundreds.) 7749
Reposted by Max LittleTimothy Gowers @wtgowers.bsky.social · 20/12/2024It seems that OpenAI's latest model, o3, can solve 25% of problems on a database called FrontierMath, created by EpochAI, where previous LLMs could only solve 2%. On Twitter I am quoted as saying, "Getting even one question right would be well beyond what we can do now, let alone saturating them." 8868
Reposted by Max LittleStephanie M. Lee @stephaniemlee.bsky.social · 18/12/2024It's widely agreed that scholars are supposed to say when they use ChatGPT. Yet phrases like "I am an AI language model"—with no disclosure—are popping up in papers. I wrote about how journals seemingly aren't enforcing their AI policies, according to a new study: www.chronicle.com/article/scho...chronicle.comScholars Are Supposed to Say When They Use AI. Do They?Journals have policies about disclosing ChatGPT writing, but enforcing them is another matter, according to a new study. 15221
Reposted by Max LittleRodney Brooks @rodneyabrooks.bsky.social · 18/12/2024This seems like a pretty balanced commentary. They certainly get this right: "connection between capability improvements & AI’s social or economic impacts is extremely weak. The bottlenecks for impact are the pace of product development and the rate of adoption" www.aisnakeoil.com/p/is-ai-prog...aisnakeoil.comIs AI progress slowing down?Making sense of recent technology trends and claims 1184
Max Little @maxal.bsky.social · 17/12/2024Good reporting here, but sadly, these tragedies were predictable. Those of us who actually work on machine learning know that deep-learning based computer vision simply isn't reliable enough for safety-critical applications such as self-driving cars. @garymarcus.bsky.social @filippie509.bsky.social 091
Reposted by Max LittleGary Marcus @garymarcus.bsky.social · 14/12/2024The late Suchir Balaji’s blog post on AI, copyright and fair use, reposted in his memory. suchir.net/fair_use.htmlsuchir.netWhen does generative AI qualify for fair use? 412436
Reposted by Max LittleMaths Facts @mathsfacts.bsky.social · 14/12/2024The bootstrap can be used to generate a new random sample from an existing random sample. Its validity can be guaranteed by the Glivenko-Cantelli theorem, which demonstrates how the empirical cumulative distribution (CDF, top panel), converges on the CDF of the sample (bottom panel). 001
Reposted by Max LittleMaths Facts @mathsfacts.bsky.social · 14/12/2024For an increasing function 𝑓:ℝ→ℝ, max(𝑓(𝑎),𝑓(𝑏))=𝑓(max(𝑎,𝑏)). An important special case is 𝑓(𝑥)=𝑥+𝑐, for which we obtain max(𝑎+𝑐,𝑏+𝑐)=𝑐+max(𝑎,𝑏). 001
Reposted by Max LittleFilip Piekniewski @filippie509.bsky.social · 14/12/2024Since 2016 Waymo raised ~$25B, so they burn ~$3B/y or little over 8mln/day. With ~700 cars, assuming they operate each car every day, it costs them over 11k dollars to operate each of their cars per day. $11k PER DAY per CAR. If you don't find this ridiculous IDK what else to say. 283
Reposted by Max LittleGary Marcus @garymarcus.bsky.social · 14/12/2024Suchir Balaji was a good young man. I spoke to him six weeks ago. He had left OpenAI and wanted to make the world a better place. This is tragic. 816146
Reposted by Max Littlecwcyau.bsky.social @cwcyau.bsky.social · 13/12/2024Health Data Research UK PhD meet! Work from Ant Lee and Jianqiao Mao (latter with @maxal.bsky.social) 021
Reposted by Max LittleFilip Piekniewski @filippie509.bsky.social · 10/12/2024As first predicted some 10 years ago that is how "self driving cars" will end - as glorified driver assistance features. The graveyard of autonomous vehicle efforts is pretty crowded already with pretty much only Waymo remaining, until life support from Google mothership ends. 192
Reposted by Max LittleGary Marcus @garymarcus.bsky.social · 09/12/2024What if all the hype just didn’t turn out to be true? Evidence of productivity gains is mixed - yet hypey takes continue to dominate in the media. 7429
Reposted by Max LittleJose Afonso Furtado @jafurtado.bsky.social · 08/12/2024Don’t Ride This Bike! Generative AI’s persistent trouble with compositionality and parts, by Gary Marcus @garymarcus.bsky.social and Ernest Davis / Marcus on AI - Substack garymarcus.substack.com/p/dont-ride-...garymarcus.substack.comDon’t Ride This Bike! Generative AI’s persistent trouble with compositionality and partsWhen the text-to-image AI generation system DALL-E2 was released in April 2022, the two of us, together with Scott Aaronson, ran some informal experiments to probe its abilities. 0113
Max Little @maxal.bsky.social · 02/12/2024Not quite: AI got people excited about interpolation, it seems. Numerical analysts suddenly feel seen. 030
Max Little @maxal.bsky.social · 27/11/2024Fully-funded PhD position available. If you are interested in machine learning for signal processing of biosignals, please do get in touch. 010
Reposted by Max LittleDorothy Bishop @deevybee.bsky.social · 25/11/2024New blogpost: deevybee.blogspot.com/2024/11/why-...deevybee.blogspot.comWhy I have resigned from the Royal SocietyThe Royal Society is a venerable institution founded in 1660, whose original members included such eminent men as Christopher Wren, Robert H... 1632444998
Max Little @maxal.bsky.social · 23/11/2024The current incarnation of AI companies earnestly engaged in the Sisyphean task of labeling "all the things, in every circumstance". 000
Max Little @maxal.bsky.social · 10/11/2024Nice, terse quote: "There is no principled solution to hallucinations in systems that traffic only in the statistics of language without explicit representation of facts and explicit tools to reason over those facts." 000
Reposted by Max LittleGary Marcus @garymarcus.bsky.social · 09/11/2024Game over, I won. GPT is hitting a period of dimishing returns, just like I said it would. Analysis and breaking news here: open.substack.com/pub/garymarc...open.substack.comCONFIRMED: LLMs have indeed reached a point of diminishing returnsScience, sociology, and the likely financial collapse of the Generative AI bubble 76311
Max Little @maxal.bsky.social · 24/09/2024Can we be confident that rigour-enhancing practices causally increase the replicability rate of scientific studies in social science? A high-profile claim of that specific cause-effect relationship was found to have insufficient rigour, now retracted. statmodeling.stat.columbia.edu/2024/09/24/w...statmodeling.stat.columbia.edu Well, today we find our heroes flying along smoothly… | Statistical Modeling, Causal Inference, and Social Science 020
Reposted by Max LittleJessica Hullman @jessicahullman.bsky.social · 24/09/2024In other science reform implosion news, the contested article including OSF & Data Colada authors on how preregistration & other rigor-enhanching practices led to high observed replicability (the one that couldn't produce its own preregistration) was just retracted www.nature.com/articles/s41...nature.comRETRACTED ARTICLE: High replicability of newly discovered social-behavioural findings is achievable - Nature Human BehaviourFour labs discovered and replicated 16 novel findings with practices such as preregistration, large sample sizes and replication fidelity. Their findings suggest that with best practices, high replica... 613259
Reposted by Max Littlesamuel mehr @mehr.nz · 14/09/2024this is pretty amazing: @lucinauddin.bsky.social is taking on the giant, hugely profitable publishers of academic journals, on grounds of antitrust If you'd like to join the case as a plaintiff you can sign up at www.lieffcabraser.com/antitrust/ac... summary of the case: 1116899
Reposted by Max LittleDavid Ho @davidho.bsky.social · 14/09/2024In case you don’t believe the same equations describe motion in the atmosphere and the ocean. 321074331
Reposted by Max LittleRodney Brooks @rodneyabrooks.bsky.social · 13/09/2024"By creating the illusion of complete autonomy, companies can fuel interest in their technology and raise the billions of dollars they need to build a viable robot taxi service." Waymo & Cruise recently admitted they use remote assist. Zoox is more up front about it. www.nytimes.com/2024/09/11/i...nytimes.comWhen Self-Driving Cars Don’t Actually Drive ThemselvesAn immersive article shows readers what a New York Times reporter has tracked for nearly a decade: Robot taxis still need human help. 1113
Reposted by Max LittleBlake Richards @tyrellturing.bsky.social · 12/09/20241/ Here's a critical problem that the #neuroai field is going to have to contend with: Increasingly, it looks like neural networks converge on the same representational structures - regardless of their specific losses and architectures - as long as they're big and trained on real world data. 🧠📈 🧪 512635
Reposted by Max LittleGary Marcus @garymarcus.bsky.social · 05/09/2024“AI worse than humans in every way at summarizing information … the technology might actually make more work for people, not less” 2184