Coraline !!!!!! @tech5l.bsky.social · 11mThis is a much scarier safety issue than effective altruist nonsense about super intelligence or RSI... 000
Coraline !!!!!! @tech5l.bsky.social · 14mThis is structurally inherent to the Transformer architecture and cannot ever be changed. The prescribing of a mind or "intelligence" to these things is very risky because people will start to trust it in undue matters. 100
Coraline !!!!!! @tech5l.bsky.social · 15mRight. They have a weak understanding over residual stream layers, but this pertains to trying to figure out the connections of things (attention) and how these relate to statistical "knowledge" (the perceptron) in order to weight for next token (or multi token prediction). There's no mind. 110
Reposted by Coraline !!!!!!Jak @jaksays.bsky.social · 5hLLMs make the typing bit easy But the typing bit is the simple part of development - the hard parts are requirements gathering, problem analysis and verification, and as per this study those are parts LLMs actually make harder Those are also the only parts that add business value 2263
Coraline !!!!!! @tech5l.bsky.social · 42mBut it's findings are extremely obvious to any SWE working in enterprise today lol. If anything it's much worse now. 000
Reposted by Coraline !!!!!!Sean Barrett @nothings.bsky.social · 3h2024: "This study doesn't include the newest models so its conclusions aren't valid." 2025: "This study doesn't include the newest models so its conclusions aren't valid." 2026: "This study doesn't include the newest models so its conclusions aren't valid." 1161
Coraline !!!!!! @tech5l.bsky.social · 46mNo answer I guess lol. Another vibe coder with zero computer science training... 010
Reposted by Coraline !!!!!!Butlerian Jihadist of House Dornick @shimminykricket.blacksky.app · 8hBasically AI allows software engineers to produce shitty code *faster* and in higher volumes which increases the amount time spent on code reviews and revisions nullifying the productivity gained from producing the code in the first place. 161267223
Coraline !!!!!! @tech5l.bsky.social · 1hHow am I being a prick. I'm a ML engineer and natural language transformers is my area of expertise. I'm trying to engage in the discussion earnestly 010
Coraline !!!!!! @tech5l.bsky.social · 1hThis benchmark doesn't measure hallucinations. It measures "correctness" 130
Coraline !!!!!! @tech5l.bsky.social · 1hNot really a coherent question. Hallucination rate is softly correlated with model size but it's kind of tricky to measure to what extent or why. Some models achieve low hallucinations rate by using a refusal mechanism (Eg minimax) 100
Coraline !!!!!! @tech5l.bsky.social · 1hDo you know what the difference is between an algorithmic process and a stochastic process? Lol 120
Coraline !!!!!! @tech5l.bsky.social · 1hOn here opus 5.5 has an hallucinations rate of 59%. What about that comes across as improvement to you? Do you think it's reliable if something is going to give you a wrong answer 59% of the time? 160
Reposted by Coraline !!!!!!Butlerian Jihadist of House Dornick @shimminykricket.blacksky.app · 8hHarvard just dropped a study AI productivity in software engineering and it found the average increase in productivity from using Claude code is.....zero and for exactly the reasons you'd think fion.ac/jellyfish.pdf 3933561297
Reposted by Coraline !!!!!!Emily M. Bender @emilymbender.bsky.social · 9hSoftware engineering could and should be done with the same care as any other type of engineering. This is a post in honor of Margaret Hamilton (and also a subtweet of all vibe-coders everywhere). 8662206
Coraline !!!!!! @tech5l.bsky.social · 9hEven less so when everything is done by the chatbot and no one is checking anything 020
Coraline !!!!!! @tech5l.bsky.social · 9hThis remains me of Mythos when we were being told it was so strong and powerful at hacking yet in reality it was absolutely nothing like it was sold to us. It's like that but in mathematics. 120
Coraline !!!!!! @tech5l.bsky.social · 9hThe problems discussed in the paper expand extremely when you're just fully automating the entire loop, deciding it's lean valid and throwing it on the pile of "solved problems". This is exactly what OpenAI are doing. 120
Coraline !!!!!! @tech5l.bsky.social · 11hOh my apologies. It's such an important time to be critical right now against this anti science onslaught from OpenAI! 140
Coraline !!!!!! @tech5l.bsky.social · 11hOn the scepticism side this take down of AI and lean is really important arxiv.org/abs/2610.08144arxiv.orgNavier-Stokes lost in translation: Why Lean verification of AI autoformalisation does not guarantee correct natural language proofsAutoformalisation is increasingly used to verify mathematical texts, including those generated by AI, as in OpenAI's announced proof of blow-up of solutions to the Navier-Stokes equations. In this pro... 1179
Reposted by Coraline !!!!!!Antonio E. Porreca 🐳 @aeporreca.org · 12hThe Association for Human Mathematics (AHM) statement on OpenAI’s October 6 release of mathematical documents just dropped! www.ahmath.org/statements 12300151
Coraline !!!!!! @tech5l.bsky.social · 12harxiv.org/abs/2610.08144 this is rather timelyarxiv.orgNavier-Stokes lost in translation: Why Lean verification of AI autoformalisation does not guarantee correct natural language proofsAutoformalisation is increasingly used to verify mathematical texts, including those generated by AI, as in OpenAI's announced proof of blow-up of solutions to the Navier-Stokes equations. In this pro... 000
Reposted by Coraline !!!!!!Kybosh @kybosh.bsky.social · 06/10/2026I like seeing the dev of the SNES core on the Pocket actually telling one of these dudes off. Holy shit. 4789542167
Coraline !!!!!! @tech5l.bsky.social · 07/10/2026I hate them and the whole slop crusade they're doing in math which ultimately isn't contributing to human understanding in math or helping develop new directions. It's just clankers clanking and OpenAI desperately trying to save themselves from financial oblivion. Not to mention the theft 010
Coraline !!!!!! @tech5l.bsky.social · 07/10/2026This is accurate but consider: pathetic attempt to prove your company is worth trillions as its financial future is extremely dubious 041
Coraline !!!!!! @tech5l.bsky.social · 07/10/2026Lol I mean they're definitely not checking the value of the proofs in any depth. They're fully reliant on Lean 030
Coraline !!!!!! @tech5l.bsky.social · 07/10/2026Before all this math slop l was actually a lean enjoyer but we've got so much relatively or completely useless AI Math results now because "lean valid" that I'm really not sure anymore that lean valid carries any substantial weight as to the ontological value of a proof 070
Reposted by Coraline !!!!!!Timnit Gebru @timnitgebru.blacksky.app · 06/10/2026In how many ways can this message be repeated? @emilymbender.bsky.social has been saying this for how long now? www.bloomberg.com/news/videos/...bloomberg.comWatch Bender: AI’s ‘Existential’ Risk Is ‘Fake’ - BloombergUniversity of Washington professor Emily Bender argues the AI safety debate is too focused on hypothetical existential threats while overlooking harms already happening today. She points to environmen... 622679
Coraline !!!!!! @tech5l.bsky.social · 06/10/2026It really exposes the extremely deluded bubble they live in. Mfs actually believe LLMs are emerging deity like structures. They're so delusional 010
Coraline !!!!!! @tech5l.bsky.social · 06/10/2026It's a little more recent but you could also argue very strongly that since the 2000s silicon valley culture around AI has effectively been Sci fi eugenics 000
Coraline !!!!!! @tech5l.bsky.social · 06/10/2026As a ML person who's extremely sick of the silicon valley cultism, the horrible state of ML research and the ridiculous bubble around AI, thank you so much for writing this. I feel heard 010
Reposted by Coraline !!!!!!Mike Cook @mtrc.bsky.social · 22/09/2026New Post: Why I Love AI Next month it'll be twenty years since I went to university, learned to program, and studied AI. I wrote about why I loved it, why I still love it, and why, unfortunately, it no longer exists. www.possibilityspace.org/blog/posts/i... 32558165
Reposted by Coraline !!!!!!David Gerard @davidgerard.circumstances.run.ap.brid.gy · 04/10/2026 30127313967
Coraline !!!!!! @tech5l.bsky.social · 04/10/2026Concerning that this needs to be stated at all. ELIZA effect has cooked 99% of discourse around NLP Transformers 000
Reposted by Coraline !!!!!!Timnit Gebru @timnitgebru.blacksky.app · 03/10/2026Even more depressing than the knowledge that congress is now overrun by effective altruist and overall TESCREAL cult member staffers, is the number of billionaires in these cults that are going to be created by the OpenAI and Anthropic IPOs. 619448
Coraline !!!!!! @tech5l.bsky.social · 04/10/2026I am sceptical Yud knows any of these things to any real depth lel 050
Coraline !!!!!! @tech5l.bsky.social · 04/10/2026This is much more reasonable to call brute forcing than "intelligence". The question of if LLMs "understand" is tricky. I would argue they only understand in that they have pairwise QKV clusters which can weight their probability distributions but they don't have an ontological sense of mind. 000
Coraline !!!!!! @tech5l.bsky.social · 04/10/2026Well also you can brute force over lean and in theory use a soft form of Evolutionary algorithms. This is probably what they did for NS aside from also stealing the problem space lead from mathematicians. You can run millions of instances with lean and kill processes that eventually hit a wall 100
Reposted by Coraline !!!!!!Chanda Prescod-Weinstein 🌌 @chanda.blacksky.app · 02/10/2026thebulletin.org/2026/09/how-...thebulletin.orgHow the bad science of AI doomerism is good for big businessThere is no clear path showing that today’s AI systems will lead to superintelligence, much less a basis for calculating the probability of doom. 17324