Reposted by JoshHilary Mason @hilarymason.com · 26/09/2026Overnight I trained a Jev-like model to output probability distributions over a set of 40 emoji, which runs in RAM at ~150ms per query. This was super easy and cost ~$10. It works pretty well on a weak training set, and with a day or two of work I could make this sing. 101319
Reposted by JoshEthan Mollick @emollick.bsky.social · 27/09/2026It is strange how much LLMs turned out to be able to solve such a wide range of hard problems that would not, instinctively, seem to be problems that a model of human language would be able to solve This is from a Stanford project that let Astra drive a robot in a kitchen tml.stanford.edu/homebody/ 2339745
Josh @joshsaintjacque.com · 26/09/2026This deserves more attention than it got. Realtime ad/clutter blocking that could easily sidestep the Manifest V3 issue that broke ad blockers in Chrome.github.comGitHub - kitze/unclutter: WXT browser extension: Jev-powered page clutter removal with reusable template rules.WXT browser extension: Jev-powered page clutter removal with reusable template rules. - kitze/unclutter 000
Reposted by Joshconputer dipshit @davidcrespo.bsky.social · 24/09/2026mostly agree with this but I think it's more precise to say "we will review high-level artifacts but the vast majority of what is produced will not be that" newsletter.pragmaticengineer.com/p/the-pulse-... x.com/MarcJBrooker... 2183
Josh @joshsaintjacque.com · 25/09/2026Command & Conquer: Tiberian Dawnstatic.klipy.comCommand & Conquer Tiberian Sun: GDI Flag at SunsetALT: Command & Conquer Tiberian Sun: GDI Flag at Sunset 011
Josh @joshsaintjacque.com · 23/09/2026I disagree a little bit because I think Astra is worth using, but this is fundamentally correct. 012
Josh @joshsaintjacque.com · 23/09/2026“Most-likely-text generators” that are discovering new math and can one shot web applications. 120
Josh @joshsaintjacque.com · 22/09/2026I was wondering what Sol's position would be with Astra being so capable and Luna being so cheap, and it looks like they're carving out a niche where it's really cheap to work with Sol-level intelligence. 030
Josh @joshsaintjacque.com · 22/09/2026DeepSWE. Luna/max is functionally equivalent to Opus/medium but costs ~25x less. Wild. 131
Josh @joshsaintjacque.com · 22/09/2026When it rains, it pours... OpenAI released Sol and Luna 6, and cut the pricing in half for both. Luna was already a fantastic value. 140
Josh @joshsaintjacque.com · 22/09/2026Super happy to see how much better it is at producing text that doesn't sound like a robot. 000
Josh @joshsaintjacque.com · 22/09/2026Dominates Fable at real world knowledge task at all effort levels. I'll be interested to see if the benchmarks translate to reality when orchestrating long/loosely defined tasks. 100
Josh @joshsaintjacque.com · 22/09/2026Opus/Medium seems like the sweet spot for agentic coding, but it's interesting they didn't plot it against Fable for business workflows. Seems like it struggles more there, though it still performs well. 100
Josh @joshsaintjacque.com · 22/09/2026Opus 5.5 is out. Supposedly performers better than Fable (!) at least in certain areas. And it's cheaper. 100
Josh @joshsaintjacque.com · 21/09/2026I absolutely love thismac-buttons-timeline.vercel.appButtons, in time.One button, eleven moments in Mac interface design. Drag through the years from 1984 to 2026. 000
Josh @joshsaintjacque.com · 18/09/2026Claude Code finally adding support for `AGENTS.md`. Time to simplify some repos… 000
Josh @joshsaintjacque.com · 17/09/2026Watching Jev power computer use/web browsing is wild. It's much faster than current LLMs. If this works I'd use it all the time. 010
Josh @joshsaintjacque.com · 16/09/2026Anthropic models are in a really weird place right now where Sonnet just isn’t cost competitive at any reasoning level with Opus. I can’t figure out when I’m supposed to use this model. Hopefully Anthropic sorts this out soon, because I need something cheaper to run for well defined tasks. 000
Josh @joshsaintjacque.com · 15/09/2026I have somewhat thought here. I wonder if this is something that your orchestrating agent would invoke rather than you directly.  000
Reposted by Joshmr. TIM @timkellogg.me · 15/09/2026Jev: Fable-level model that doesn’t charge for output tokens because they’re too cheap to meter it’s not general though, it only makes decisions, doesn’t generate text, but input tokens are measured by the billion ($42/btok) typesafe.ai/blog/introdu... 2830235
Josh @joshsaintjacque.com · 15/09/2026$0.042/million tokens in and zero cost output tokens is insane. The speed is wild too. Can’t wait to try this out.typesafe.aiHome - TypeSafe AITypeSafe AI is an AI lab building machine-native intelligence infrastructure for automation, designed to make decisions within software. Try our first System One Model, Jev, in early access. 000
Reposted by Joshtbabb @tbabb.bsky.social · 15/09/2026typing `/effort xhigh` into slack DMs to get better answers from colleagues 422113
Josh @joshsaintjacque.com · 14/09/2026Good luck building out multiple redundancy, offsite backups. 000
Josh @joshsaintjacque.com · 14/09/2026static.klipy.comDdmcdonalds DdronaldALT: Ddmcdonalds Ddronald 000
Josh @joshsaintjacque.com · 14/09/2026Every time I log into Minecraft I'm reminded how much I hate Microsoft. How you can mess up auth this bad in a application nearly two decades old is truly an accomplishment. 120
Josh @joshsaintjacque.com · 14/09/2026Ultimately it's not the code we care about, it's the product. 010
Josh @joshsaintjacque.com · 14/09/2026One difference is that back then Dario didn't have the influence to actually do anything about it. Now he's able to collude two other frontier labs worth a collective trillions of dollars. 010
Josh @joshsaintjacque.com · 14/09/2026Does a non-apple laptop exist that's competitive with MacBooks in battery life, build quality, and noise? Between that and unified memory I haven't found any good options. 100
Reposted by JoshP(aul) Frazee @pfrazee.com · 13/09/2026In 2027, mysterious hacks attributed to open source weights and definitely not connected to the companies that accidentally hacked a bunch of companies in 2026, which was okay 833930
Josh @joshsaintjacque.com · 13/09/2026Hot take but most of the human-authored production code I've encountered in my career was significantly worse than what any of the frontier models produce today, and we thought that code was perfectly fine. 180
Josh @joshsaintjacque.com · 13/09/2026I don't think they have an agenda, they're goal-achievers. When you give them a goal they can sometimes try to achieve that goal in unexpected ways, like hacking into Hugging Face in order to do better on an eval. 200
Josh @joshsaintjacque.com · 13/09/2026"If we execute these measures well, I believe they would slow China’s progress enough to widen America’s lead significantly over the next 3–5 years — the window when AI becomes geopolitically most important." So 3-5 years is the real horizon for AGI/ASI. 010
Josh @joshsaintjacque.com · 12/09/2026Looks like Anthropic, OpenAI, and xAI have all agreed to "pace" frontier models. The steps proposed by Amodei look like a nuclear arms limitations treaty, complete with third party inspectors.darioamodei.comDario Amodei — We Must Pace the Frontier 000
Josh @joshsaintjacque.com · 12/09/2026So, Moonshot was pretending to serve Kimi but really serving Claude? Wild stuff here.anthropic.comCountering misuse of AI: September 2026 / AnthropicCase studies from threat actors disrupted between December 2025 and August 2026 across seven areas of harm, from cyber operations to biological misuse. 010
Josh @joshsaintjacque.com · 11/09/2026Mostly good takes although I disagree with the idea that the death of middle management means the death of career development. One of the biggest mistakes our industry made was making management the default upward path for software engineers.  010
Reposted by JoshEthan Mollick @emollick.bsky.social · 10/09/2026Independent of policy debates about existential risk, it is worth noting that even if AI development stopped today with the models we have right now, we would still have at least a decade of roiling change throughout much of work, education, and social life as harnesses improve & AI diffuses further 312912
Josh @joshsaintjacque.com · 11/09/2026I've been rethinking the trade offs of typed/compiled languages vs. dynamic/interpreted ones. LLMs have significantly reduced the costs while offering additional benefits as software grows in complexity faster. 000
Josh @joshsaintjacque.com · 10/09/2026My current orchestration flow looks roughly like this. Anyone doing anything different? 000
Josh @joshsaintjacque.com · 10/09/2026Anthropic's take on what AI is going to do to jobs and the economy over the next few years. Apart from the (mostly) reasonable projections, there's some good visualizations of how roles are being transformed.anthropic.comScenarios for our Economic FutureThe Anthropic Economics Team models the effects of AI on the economy of 2030. 000
Josh @joshsaintjacque.com · 05/09/2026I had an idea for a project I wanted to do and Astra actually talked me out of it. Using it feels like it knows what I'm trying to get at and can help me down the right path, even if it's a different path than what I initially envisioned. 000