Josh @joshsaintjacque.com · 23/09/2026I disagree a little bit because I think Astra is worth using, but this is fundamentally correct. 012
Josh @joshsaintjacque.com · 22/09/2026I was wondering what Sol's position would be with Astra being so capable and Luna being so cheap, and it looks like they're carving out a niche where it's really cheap to work with Sol-level intelligence. 030
Josh @joshsaintjacque.com · 22/09/2026DeepSWE. Luna/max is functionally equivalent to Opus/medium but costs ~25x less. Wild. 131
Josh @joshsaintjacque.com · 22/09/2026When it rains, it pours... OpenAI released Sol and Luna 6, and cut the pricing in half for both. Luna was already a fantastic value. 140
Josh @joshsaintjacque.com · 22/09/2026Super happy to see how much better it is at producing text that doesn't sound like a robot. 000
Josh @joshsaintjacque.com · 22/09/2026Dominates Fable at real world knowledge task at all effort levels. I'll be interested to see if the benchmarks translate to reality when orchestrating long/loosely defined tasks. 100
Josh @joshsaintjacque.com · 22/09/2026Opus/Medium seems like the sweet spot for agentic coding, but it's interesting they didn't plot it against Fable for business workflows. Seems like it struggles more there, though it still performs well. 100
Josh @joshsaintjacque.com · 22/09/2026Opus 5.5 is out. Supposedly performers better than Fable (!) at least in certain areas. And it's cheaper. 100
Josh @joshsaintjacque.com · 10/09/2026My current orchestration flow looks roughly like this. Anyone doing anything different? 000
Josh @joshsaintjacque.com · 04/09/2026Someone needs to teach Anthropic that this is how you do PR. 000
Josh @joshsaintjacque.com · 02/09/2026I'm almost irrationally happy that GitHub's CLI finally supports image and video uploads. Agents can now attach proof of work to PRs without resorting to hacks. 000
Josh @joshsaintjacque.com · 16/05/2026Had so much fun making this. The View Transition API is so powerful. #css #js #rails 4935
Josh @joshsaintjacque.com · 20/09/2025It's neat that Tailwind has child selectors and all but I wonder if at this point you should just use a stylesheet. 010
Josh @joshsaintjacque.com · 16/09/2025I love how extreme these opinions are. Like, sometimes an OS update is just an OS update. 120
Josh @joshsaintjacque.com · 29/01/2025Wiring up an image gen AI model to a bluesky feed might be a good use case 020
Josh @joshsaintjacque.com · 13/01/2025Here's the same prompt using `isDog`. Still looks like it got it, no? For reference, this is a clean thread via 4o. 100
Josh @joshsaintjacque.com · 13/01/2025That’s a good point. People aren’t used to consulting a computer the way they would a person (understanding it can be wrong) and verifying accordingly. As an aside, in my test it took o1 (the reasoning model) to correctly answer this question. 010
Josh @joshsaintjacque.com · 09/01/2025The situation at Meta is actually awful. They prohibit "allegations of stupidity, intellectual capacity, and mental illness"... unless it's targeted at someone because of their gender or sexual orientation. Call my daughter stupid? Not ok. Call her stupid because she's a girl? Perfectly fine. 000
Josh @joshsaintjacque.com · 28/12/2024This is the level of discourse on AI right now. Block function working overtime. 120
Josh @joshsaintjacque.com · 28/12/2024Meanwhile current gen models produce routinely produce complex code. Anyone who has worked with Claude/GPT-4o/o1 experiences this first hand every day. The take that their useless because their training set isn't flawless is wild. 030
Josh @joshsaintjacque.com · 04/12/2024I would want to learn more about the pluses/minuses of this approach. I'm sure there are some good reasons for it, but in my head I'm comparing to the simplest "Rails-y" approach I can think of. 120