Shiminsky @shiminsky.bsky.social · 28/09/2026Catching up on AI links post trip, I have this nagging sense that I'm in an Adam Curtis documentary: 'By the mid 2020s, frontier AI labs intensified their false sense of control over their creations'. 010
Reposted by ShiminskyTom McCoy @rtommccoy.bsky.social · 01/09/2026🤖🧠NEW PAPER🧠🤖 (The result of an 8-year project!) LLMs seem very different from symbolic systems. Yet LLMs excel in symbolic domains (e.g., language/code/math). How do they do it? Our finding: LLM representations have implicit symbolic structure! Link in thread ⬇️ 1/n 431788
Shiminsky @shiminsky.bsky.social · 30/06/2026Public Domain Review, where have you been my entire life? Thank you for sharing @davatron5000.bsky.social! 040
Shiminsky @shiminsky.bsky.social · 23/06/2026It was fun applying creativity techniques with AI generated front-end designs: shimin.io/journal/inha... Also got a chance to present this at AI Tinkerers Seattle earlier this month, give the skill a try if you are sick of projects looking like AI Slop.shimin.ioInhabited Design, a Skill for the AI Slop Site Problem — Shimin ZhangPurple gradients are the new em dash. Building inhabited-design — a Claude Code skill that samples a different designer to inhabit on every run — and the detours through attractors, personas, self-ref... 100
Shiminsky @shiminsky.bsky.social · 22/06/2026Insightful post from @terriblesoftware.org, and something I've been thinking about: there's an inherent asymmetry with AI generated content where cost of reading > cost of review. We should start measuring performance by fewest LOC generated. 010
Shiminsky @shiminsky.bsky.social · 11/06/2026Here's my prompt that just got flagged: "what do you know about the theory of creativity? do research if you have to, give me a full list of theories and hypothesis" and "what are some of the most impactful recent studies?" It's like they actually don't want folks to use the thing!! 000
Shiminsky @shiminsky.bsky.social · 10/06/2026Can't tell if we are all cooked, or this is the start of something beautiful. 220
Shiminsky @shiminsky.bsky.social · 16/05/2026Nibbling on the same thought today, future code education will be on a new level of abstraction and we haven’t figured out what that looks like yet. Very excited to sign up for @danabra.mov’s new course when it’s out! 080
Shiminsky @shiminsky.bsky.social · 13/05/2026Setting up a team of AI philosophers to debate whether Magikarp is the best Pokemon, average Tuesday night fun stuff. 100
Reposted by ShiminskyAi2 @ai2.bsky.social · 08/05/2026EMO’s expert clusters look very different from a traditional MoE: they organize around semantic domains like health, news, politics, & film/music. Traditional MoEs often cluster around surface patterns like prepositions and articles, making selective expert use tougher. 1694
Reposted by ShiminskyJesse Vincent @s.ly · 01/05/2026A short post about my favorite adversarial review prompt: blog.fsck.com/2026/05/01/a...blog.fsck.comMy favorite adversarial review promptI'm Jesse. I make stuff. Software, hardware. Very occasionally, trouble. 281
Shiminsky @shiminsky.bsky.social · 04/05/2026It’s about time we stop infantilizing farmers, they are just like you and me, acting according to their self interests. (I’m just a gardener but married into a farming family) 19214
Shiminsky @shiminsky.bsky.social · 24/04/2026Are you tired of keeping track of 4 different variants SWE Benchmarks? I am. So I did something kinda dumb -- but at least it was fun -- asking 11 frontier models to blind grade each other's open ended prediction about AI's future. Post at: shimin.io/journal/what... Findings below (given n=1)shimin.ioWhat I learned asking 11 AI models to grade each other's AI predictions — Shimin ZhangAn experiment on model personalities, a delusion index, and the open-weight dark horse contender I didn't see coming. 110
Reposted by ShiminskyWilliam B. Fuckley @opinionhaver.bsky.social · 22/04/2026here's a mildly provocative take: every valid LLM worry about societal effects is reducible to some approximation of 'it makes what were once costly things trivially easy to do' 3030033
Shiminsky @shiminsky.bsky.social · 21/04/2026The Latest Pelican Bicycle Benchmark result from @simonwillison.net for Opus 4.7 was so shocking that I had to do some follow up experiments. It turns out Opus 4.7 is ....just kinda lazy?? It uses almost no reasoning tokens compare to Qwen, and 40x less than Opus 4.6 shimin.io/journal/clau...shimin.ioClaude 4.7 isn't dumb, it's just lazy — Shimin ZhangSome follow up experiments with Claude 4.7 based on Simon Willison's Pelican Benchmark Shocker. 2473
Shiminsky @shiminsky.bsky.social · 21/04/2026I'd been missing my Pi Agent harness since the Anthropic subscription crack down. Tonight I finally got it back with a local Qwen 3.6 setup locally and it feel great to be reunited with my favorite harness! Thank you @mariozechner.at @mitsuhiko.at for your work on Pi! 130
Shiminsky @shiminsky.bsky.social · 08/04/2026@addyosmani.bsky.social's multi agent experience agrees with my own, going from 2 to 5 parallel agents is a quick way to transform from a judge to a ticket usher.addyosmani.comYour parallel Agent limitRunning multiple agents in parallel is not just a question of throughput. It is a new kind of cognitive labor that requires managing multiple mental models, ... 021
Shiminsky @shiminsky.bsky.social · 08/04/2026Another banger from @anildash.com : Actually, people love to work hard anildash.com/2026/04/06/p...anildash.comActually, people love to work hard - Anil DashA blog about making culture. Since 1999. 0249
Shiminsky @shiminsky.bsky.social · 09/03/2026Reminder when the only fatigue we felt was about too many new JS frameworks? Those were the days. 010
Shiminsky @shiminsky.bsky.social · 09/03/2026Great post from @antirez.bsky.social tracing through the history of the open source movement and positions 'clean room clones' as an extension and not revolution. 010
Shiminsky @shiminsky.bsky.social · 08/03/2026AI sycophancy will reap many promising careers in next next few years. 000
Shiminsky @shiminsky.bsky.social · 21/02/2026Folks speak of the dangers of AI psychosis, what about the dangers of human psychosis from repeated attempts at convincing LLMs that the Earth is flat? 020
Reposted by ShiminskySK Winnicki, PhD 🏳️⚧️ @skwinnicki.bsky.social · 12/02/2026Found an additional graphic that gets even more of these quotes together. I've kept "I hate myself, I hate clover, and I hate bees" pinned above my desk since I first started studying evolutionary biology as an undergraduate. So relatable to get extremely frustrated with your study system. 7549661784
Reposted by ShiminskySimon Willison @simonwillison.net · 12/02/2026Genuinely very impressed by the SVG of a pelican riding a bicycle I just got out of Google's new Gemini 3 Deep Think model simonwillison.net/2026/Feb/12/...simonwillison.netGemini 3 Deep ThinkNew from Google. They say it's "built to push the frontier of intelligence and solve modern challenges across science, research, and engineering". It drew me a really good SVG of … 1925521