Sign in

Aaron Sterling

@aaronsterling.bsky.social
742 followers 2K following 2K posts

CEO, Thistleseeds. Personal account. Current primary project: tech for substance use disorder programs.

PostsRepliesMedia
Reposted by Aaron Sterling
Axe Ghost. On Steam! @axeghostgame.bsky.social · 06/06/2026
this finding matches my experience: the valuable thing that knowledgeable human devs can bring to a project, that the agents aren't yet good at, is ontology creation.
063
Reposted by Aaron Sterling
Avik Dey @avikdey.bsky.social · 06/06/2026
Think of this as a pattern for any LLM integrated system that claims to provide guarantees. Agents propose, domain verifiers validate, approved proposals are committed and every step and decision is logged. Yes, most domain specific verifiers can be hard. So is building most deterministic system.
051
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
Someone just wrote me to ask if I could be their Arxiv endorser, meaning someone who vouches for the author uploading non-slop. I declined, because I am not clear on the recent submission rules. It appears publishing a preprint yesterday put me on a list of verified authors.
020
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
Eh. The technology for the microwave oven came from unsuccessful military research to send force beams at a distance.
000
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
Why are you coming at me so hard when you appear to have no experience with the difference between a prototype and mass-scale production? The last mile is long. The inability of technologists to grapple with that is why there is little business productivity boost from agents right now.
120
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
Mass-produced proprioception sufficient to play soccer in every neighborhood is not out of reach, but it is unlikely to happen while your children are alive.
510
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
That's only an issue if you gauge jobs by what's available in this historical snapshot. Go back 30 years, or more. Soccer clubs, park rangers, continuing education classes like pottery or paleontology. Return to a society with third spaces, and the number of jobs needed explodes.
150
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
I find a good first step is, "List 5 peer-reviewed publications that are central to the focus of this project." Then add them. The LLM responses sometimes change dramatically from what was available just through memory and search. Might be more dramatic if you added content not currently online.
020
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
My current view is that humans have to provide the meaning of a system to an LLM, and, once the LLM has that, it can often proceed faster and more accurately than a human would. It's close to, "We provide inspiration, LLMs provide perspiration."
110
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
I'd love to hear more about that someday!
010
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
Thanks! Curiously, I posted something about LLMs and ontologies earlier today. bsky.app/profile/aaro...
010
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
For checkability-over-correctness, the meta-agent can't be an LLM agent. It reviews a deterministic checklist of system invariants. The LLM calls handle ambiguity and the agents formulate concrete proposals; the meta-agent ensures proposals don't break anything.
130
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
@danabra.mov you might get a kick out of this!
010
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
Shamelessly pinging followers who I believe would enjoy reading this: @kirancodes.me @mariaa.bsky.social @alexcbecker.net @tedunderwood.com @scythiamarrow.bsky.social @eugenevinitsky.bsky.social @timkellogg.me @shallit.bsky.social @village11.bsky.social @zabong69.bsky.social
250
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
@julietshen.bsky.social could you please send this on to the Risk Agent people (or give me contact info and I will)? Their construction is similar to mine, and I cite their work.
010
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
New from me. arxiv.org/abs/2606.04903
It is worth pausing for a moment to review ontology creation by humans and by LLMs, because there is empirical data that might look contradictory at first glance, but, in fact, paints a unifying
picture. LLMs are not as good as humans at ontology creation (sometimes called “ontology learning”), as shown in [4, 15, 6]. However, at least according to the OntoURL benchmarks,
LLMs are better than humans at reasoning over an ontology that already exists[29].
Despite the previous results, the quality of LLM-generated ontologies can be higher than
the quality of ontologies created by novice human engineers[19]. This is not a contradiction,
because an LLM’s ability to 1-shot ontology creation is directly related to how completely
humans already ontologized the space through documentation. The need for humans to pre-
ontologize the space can be seen in [21], which presented LLMs with well-structured gibberish,
and the LLMs were unable to ontologize the gibberish, showing an inability to reason over
semantic relations between concepts.
One goal of Ontology-First Agent Design is to focus human expert input where it is most
needed: creation of the ontology (at the start), and refinements to the ontology to improve the
system’s functionality (a feedback loop at the end). The LLM does the work in the middle,
where it is most effective.
6425
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
I don't understand the goal of your post. Would you rather people not try?
200
Aaron Sterling @aaronsterling.bsky.social · 04/06/2026
Similar example. I think we're a couple years away from having personable interfaces so people uncomfortable around tech will be able en masse to spin up tools for themselves. bsky.app/profile/gord...
000
Reposted by Aaron Sterling
Larry Hunter @proflhunter.bsky.social · 04/06/2026
Great thread! Empirical evidence about LLM use in scientific articles.
041
Reposted by Aaron Sterling
Carl T. Bergstrom @carlbergstrom.com · 04/06/2026
9. Here's a surprise: controlling for the other covariates (correct me if I have that wrong, Kyle), we see the *most* LLM use in the high impact journals, not low impact journals.
LLM use by JIF.
58915
Aaron Sterling @aaronsterling.bsky.social · 03/06/2026
Zitron's primary flaw, in my eyes, is the Western Chauvinism of his analysis. Every single US and EU AI company could collapse tomorrow, and AI would continue to grow in quality and reach.
010
Aaron Sterling @aaronsterling.bsky.social · 03/06/2026
Are you opposed to universities providing free condoms? From harm reduction criteria alone, students are safer now.
000
Aaron Sterling @aaronsterling.bsky.social · 03/06/2026
It depends what you mean by bubble. Anthropic, Google and Microsoft will be fine, as will most Chinese companies. Everything else, who knows. If you mean that genAI will go away, that's not happening. It's getting safer and more skilled every quarter. Can run models locally on your laptop.
210
Aaron Sterling @aaronsterling.bsky.social · 03/06/2026
I might be missing something, but that reads like a win, especially compared to some other academic agreements. Heightened privacy, no requirement to use it. I don't see the downside from the announcement text.
100
Aaron Sterling @aaronsterling.bsky.social · 03/06/2026
Powerful people have been going on TV saying things like, "You, yes you personally, AI will take your job in 12 months."
000
Reposted by Aaron Sterling
daniel:// stenberg:// @bagder.mastodon.social.ap.brid.gy · 02/06/2026
While the curl project does not ban the use of AI tools - recognizing they can enhance development - AIs are merely tools. Humans must always drive the process, taking full responsibility for presenting, reviewing, and understanding every change.
0187
Aaron Sterling @aaronsterling.bsky.social · 02/06/2026
It's strong, and "atproto" is objectively weak because it collides with "@proto". Imagine you're giving dictation, or getting your mom to write it down. Clarity.
010
Aaron Sterling @aaronsterling.bsky.social · 01/06/2026
Years ago, he invited me to submit because of a guest blog post I wrote. He uses blog posts to keep on top of novelty in different fields.
030
Aaron Sterling @aaronsterling.bsky.social · 01/06/2026
That comment section is lousy.
010
Aaron Sterling @aaronsterling.bsky.social · 01/06/2026
Nice find.
100
Reposted by Aaron Sterling
Phillip Carter @phillipcarter.dev · 31/05/2026
Having been a target of a social media-driven OSS pile on, the only thing you can do is continue without taking their written slop into consideration and locking the thread. In my case it was proceeding with a Code of Conduct and daring the assholes to fork. They gave up really quickly
0375
Aaron Sterling @aaronsterling.bsky.social · 31/05/2026
So I think there is something to investigate, measure and discover. I think in the long term it might be empirically determinable whether consciousness is an illusion, for example. I expect investigations into how LLMs think will give ideas for better experiments into how people think.
120
Aaron Sterling @aaronsterling.bsky.social · 31/05/2026
That faith in human separation from machines has the same "Claudean shape" (threw that in just to annoy you) as religious belief. I do think that is true for a lot of people. But there's a difference is that religious belief varies geographically, while goals of self-control are near universal.
110
Aaron Sterling @aaronsterling.bsky.social · 31/05/2026
That's an interesting failure type. I bet it parsed "many" and ""many"" as different words, probably one the word and the other a descriptor.
030
Aaron Sterling @aaronsterling.bsky.social · 31/05/2026
So people think the pop version of Dennett is enough. Or maybe it's all they know and they think they are in the same situation as before. Not realizing there's a difference between experiences we don't know how to measure yet, and unmeasurable claims like divine visitation.
110
Aaron Sterling @aaronsterling.bsky.social · 31/05/2026
It's not just the reuse of the God of the Gaps argument that feels religious to me, either. I've seen at least one interpretation of the Papal Encyclical that is just prosperity theology, but instead of being paid in money, your faith buys you consciousness.
110
Aaron Sterling @aaronsterling.bsky.social · 31/05/2026
Maybe because the current situation feels a lot like New Atheism to me. All of the "only humans have consciousness" arguments I've seen here (probably not representative) are "consciousness of the gaps." They are utterly vulnerable to LLMs improving, even if LLMs never achieve consciousness.
110
Aaron Sterling @aaronsterling.bsky.social · 30/05/2026
In production today, are volume backups stored on the same volume they are backing up? And did you publish a retrospective of the severe production incident?
100
Aaron Sterling @aaronsterling.bsky.social · 30/05/2026
Would you be willing to address this? bsky.app/profile/aaro...
110
Aaron Sterling @aaronsterling.bsky.social · 29/05/2026
@lu.is is an expert in the area, and he is active here. (A good follow btw.)
020
Aaron Sterling @aaronsterling.bsky.social · 28/05/2026
I doubt he's wrong. What party is a candidate more likely to be in if their slogan is, "Vote for me and I'll fuck these motherfuckers up."
000
Aaron Sterling @aaronsterling.bsky.social · 28/05/2026
I'm honestly surprised npm has not delisted them. The supply chain attack is coming from inside the house.
060
Aaron Sterling @aaronsterling.bsky.social · 28/05/2026
Final tedious steps like exposition, redirection and refinement of research programs, using the new techniques to disprove an expected-to-be-true statement about Sum-Product over the reals?
020
Reposted by Aaron Sterling
Flo 🔶 @faz.ms · 28/05/2026
The bitter math lesson: you can essentially solve all of math by just doing more matrix multiplication
2683
Reposted by Aaron Sterling
⚡️🌙 @dystopiabreaker.xyz · 27/05/2026
one of my favorite cheeky papers, from 1999, is this one where the authors argue that sometimes it is better to simply wait and do nothing, because your astrophysical simulations will complete faster if you simply wait for the next epoch of compute arxiv.org/pdf/astro-ph...
We show that, in the context of Moore's Law, overall productivity can be increased for large enough computations by 'slacking' or waiting for some period of time before purchasing a computer and beginning the calculation.
511613
Aaron Sterling @aaronsterling.bsky.social · 27/05/2026
"Follow your dreams, if they're hiring." -- Chris Rock
072
Aaron Sterling @aaronsterling.bsky.social · 27/05/2026
en.wikipedia.org/wiki/BLUF_(c...
en.wikipedia.org
BLUF (communication) - Wikipedia
000
Aaron Sterling @aaronsterling.bsky.social · 27/05/2026
It's real IMO. Most anti-AI will be performative in five years, once use of LLMs is more accurate and doesn't rely on chat interfaces. The genuine problems with LLMs, like algorithmic bias around country or race, will remain as serious problems that probably still won't receive enough attention.
070
Aaron Sterling @aaronsterling.bsky.social · 27/05/2026
There was a huge dropoff in the quality of my feed before and after DOGE, maybe as simple as academics no longer able to afford to research, or being afraid of making any public comments about anything. I agree it's gotten quieter again, but maybe because the sun is out and people are outside.
110
Aaron Sterling @aaronsterling.bsky.social · 27/05/2026
"--Claude" would have been funnier.
010