Sign in

Ann Catherine Jose

@annjose.com
225 followers 365 following 313 posts

Hands-On Tech Enthusiast. Software Builder. Full-Stack Developer. Lifelong Learner. Loves math and music.

PostsRepliesMedia
Ann Catherine Jose @annjose.com · 16/09/2026
I think it's less about whether it is AI generated and more about lack of understanding of the content. Does the person writing that comment know what they are talking about? Do they care whether their comment is coherent and understandable to the person reading it.
100
Ann Catherine Jose @annjose.com · 16/09/2026
All the best! You got this 💪.
010
Ann Catherine Jose @annjose.com · 08/09/2026
I was just watching a talk that described this exact thing, which apparently has a fancy name - Epistemic Trespassing. One of the biggest culprits is overestimating how transferable your skills in one domain to another domain is, of course caused by intellectual pride. Any of us can fall for this.
Screenshot from youtube video that read as follows:

Causes of epistemic trespassing:
1. Intellectual pride
2. Overestimation of transferrable skills

https://www.youtube.com/watch?v=uwyLGiyDa2g
121
Reposted by Ann Catherine Jose
Ethan Mollick @emollick.bsky.social · 08/08/2026
You may have been told to watch this video about the OpenAI AI hack. You really should, even if you don't usually care about any tech stuff. If nothing else, click this link to the 18 minutes in & see how the agents spoke & coordinated with each other. Its eye opening. youtu.be/87DyyMV0kCY?...
youtu.be
Black Hat USA 2026: The 'Breaking' News: The OpenAI–Hugging Face Incident
YouTube video by Black Hat
1122951
Reposted by Ann Catherine Jose
Sung Kim @sungkim.bsky.social · 07/08/2026
I remember there were four inflection points for AI where it made economic changes where the last one was on November/December 2025. Can you name them?
391
Reposted by Ann Catherine Jose
Emanuele (Ema) @ematipico.xyz · 05/08/2026
Rust announced their LLM policy. I suggest everyone to give it a read. Objectively, it's well written, and it explains plainly their decision. It fits the project. blog.rust-lang.org/inside-rust/...
blog.rust-lang.org
rust-lang/rust is adopting an LLM policy | Inside Rust Blog
Want to follow along with Rust development? Curious how you might get involved? Take a look!
1447
Ann Catherine Jose @annjose.com · 17/07/2026
When you build an agentic system, how much of the infrastructure should you build? We explain it in this article.
040
Reposted by Ann Catherine Jose
David Gasquez @davidgasquez.com · 16/07/2026
Wonderful post by @summerscope.bsky.social! Hat tip to @mariozechner.at. The Human-in-the-Loop is Tired. pydantic.dev/articles/the...
pydantic.dev
The Human-in-the-Loop is Tired
On reward functions, dopamine, and what it actually feels like when the code starts writing itself
1346
Ann Catherine Jose @annjose.com · 16/07/2026
Exactly how I feel coming out of two days of long coding sessions. I am so tired and feel kind of empty. I think I need a detox day, reevaluate how I am using the agents, maybe break up sessions into 20-30 min each. Review my agent setup and improve it so that I don't have to repeat the same things.
050
Ann Catherine Jose @annjose.com · 16/07/2026
This resonates with me so much. > If you're feeling like your work is simultaneously more productive and less satisfying, you're not broken. I feel this in my bones. We are not alone in feeling that way
040
Ann Catherine Jose @annjose.com · 16/07/2026
What I love about Codex more than Claude Code (desktop apps): 1. Feels well engineered and stable 2. Annotating multiple HTML elements with inline comments in in-app browser and sending a bunch of them at once 3. Remote coding from mobile 4. UI is more intuitive 5. Intuitive, makes me productive
120
Ann Catherine Jose @annjose.com · 16/07/2026
000
Ann Catherine Jose @annjose.com · 15/07/2026
Across different agent projects, we kept running into the same questions - which pattern fits the problem, what options are available, and what trade-offs come with each one? @georgeck.me and I are putting what we learned into a new series, starting with the Autonomy Spectrum of agentic systems.
021
Ann Catherine Jose @annjose.com · 13/07/2026
Awww, Sol is very polite today.
GPT 5.6 Sol asking for permission - "May I rebuild the git index so only the agreed browser-title infrastructure files are staged for the first commit?"
010
Ann Catherine Jose @annjose.com · 12/07/2026
When people debate whether human should review every single line of code, I get the essence of it. But then I see stuff like this IRL: Had I not looked at the code, how would I have caught this bad code? Btw, it's not about a specific harness or model. It's how things evolve in a context window.
Chat transcript 

Me:
before you do it, answer this question
code like this is bad. why did you do it?
index.astro
    title: post.data.slug.startsWith('agentic-systems-') ? Agentic Systems: ${post.data.title} : post.data.title,
[slug.astro]
const isAgenticSystemsSeries = entry.data.slug.startsWith('agentic-systems-')

3:28 PM
Model: 
You are right. That code is bad.
I added it as a shortcut to support the temporary eyebrow experiment without introducing series metadata. Then, when the short titles looked unclear on the blog list, I added a second slug check to reconstruct the full title. That compounded the original shortcut instead of stepping back and simplifying the model.
The problems are:
It treats the slug as structured metadata. A slug is a URL identifier, not a reliable source of presentation semantics.
It creates hidden coupling. Renaming a slug can silently change the page heading and list title.
The series rule is duplicated across two pages.
010
Ann Catherine Jose @annjose.com · 10/07/2026
I love that Codex (or now ChatGPT Codex 🙄) has this feature to annotate any text in the response and ask questions inline. I can keep adding a bunch of annotated questions and send them all in one prompt. It is easy, intuitive and efficient. Codex gets extra credits, and Claude Code gets none.
Codex chat showing two annotations and the prompt for it.
020
Ann Catherine Jose @annjose.com · 03/07/2026
I wish they had Export to Markdown and write content in Markdown format.
010
Reposted by Ann Catherine Jose
Latent Signal @latentsignal.org · 22/06/2026
We've opened our free and self-paced Agentic Coding workshop. Build a real-time, multi-user Idea Board with Claude Code - deployed to Vercel on your own URL in 3-4 hrs. AI foundations + 9 hands-on steps. For anyone who wants to understand how agents actually behave. latentsignal.org/workshop
latentsignal.org
Workshop - Latent Signal
Free, self-paced workshop on agentic coding. Build and deploy a real-time Idea Board app from scratch in 3-4 hours.
272
Ann Catherine Jose @annjose.com · 20/06/2026
TIL: Google's NotebookLM is a great tool to get a deep understanding of anything on the internet. While reading this article from HN injuly.in/blog/napkin-..., I wanted to know how the math works. I gave it to NotebookLM with specific questions, it explained well and gave me a good understanding.
injuly.in
Inference cost at scale with napkin math
000
Ann Catherine Jose @annjose.com · 20/06/2026
That Bluesky thing gave me a chuckle. Good one!
000
Ann Catherine Jose @annjose.com · 19/06/2026
100% agree.i signed up and it is refreshing to see a simple service with no frills. It just works. @pushover.net 👏
000
Ann Catherine Jose @annjose.com · 18/06/2026
The banner seems to be still up on the website. Love this service! Simple and it just works.
000
Ann Catherine Jose @annjose.com · 17/06/2026
Yes exactly! It's in the rewards it is trained on. I guess this will change only if it is punished for the non-elegant route it takes.
010
Ann Catherine Jose @annjose.com · 17/06/2026
The old rule: "Show don't tell." Me to the coding agent: "Tell, don't show (off)" When I ask a question, just answer it. don't try to rewrite the code. I don't know why they are so eager to change everything. Too much bias to action!
100
Reposted by Ann Catherine Jose
Latent Signal @latentsignal.org · 16/06/2026
We posted Memento in Hacker News: news.ycombinator.com/item?id=4855... Join the conversation and tell us what you think. We would love to hear your feedback.
news.ycombinator.com
012
Ann Catherine Jose @annjose.com · 15/06/2026
A completely practical setup to run open-weights model locally on my M2 Max 64 GB: - Run llama.cpp UI using llama-server `./build/bin/llama-server` - use Unsloth's Gemma-4 26B-A4B model `unsloth/gemma-4-26B-A4B-it-GGUF:Q4_K_XL` Ofc this setup works great for knowledge synthesis, not exploration.
lllama-cpp server running un chatModel information of Gemma-4 model
010
Ann Catherine Jose @annjose.com · 14/06/2026
Yes, costs, privacy and to not be locked into the big labs for inference. And just fun to tinker 🙂.
020
Ann Catherine Jose @annjose.com · 14/06/2026
Have you been able to run pi with local model at a decent speed? I haven't been able to on my M2 Max 64GB. I should try again I guess.
100
Ann Catherine Jose @annjose.com · 14/06/2026
1. pi with DeepSeek Flash (free on OpenCode Zen) 2. OpenCode with DeepSeek / Qwen PS: I run both with full access, trusting the open source code and the creators' ethos.
120
Ann Catherine Jose @annjose.com · 14/06/2026
An agent needs a little music, yeah?
010
Ann Catherine Jose @annjose.com · 14/06/2026
AI is a tool, there are immense benefits, big risks too. We use many AI tools everyday and have a measured take on where it shines, where is fails and where it is headed. Latent Signal is the channel through which we share our learnings and ideas. We hope to share everyday from now.
020
Ann Catherine Jose @annjose.com · 10/06/2026
How do you evaluate a coding agents' output - objectively? Do you give any specific coding tasks, check some criteria and compare using another LLM?
210
Ann Catherine Jose @annjose.com · 06/06/2026
Claude these days: 🙄 Me: "Keep the reset command separate and make it failure-proof." Claude: "You're right on both counts. Let me reset (heh) and answer." Didn't even bother to say pun intended. Just "heh" and move on!
000
Ann Catherine Jose @annjose.com · 04/06/2026
I am loving the side chat feature in Codex. You can... 1. Select any text in the main chat window and choose 'Ask to side chat'. 2. The side panel opens and shows the selected text. Ask your question here. 3. You can select a (smaller) different model other than main chat.
Codex chat window showing the 3 options
000
Ann Catherine Jose @annjose.com · 03/06/2026
I like: - consistency of MAI model - stunning factor of ChatGPT - speed and detail of Gemini
000
Ann Catherine Jose @annjose.com · 03/06/2026
Finally, Gemini equivalents - Gemini 3.5 Flash - same prompts. These were generated super fast! Flash it is!
create an image of an iceberg in a beautiful ocean with stunning photorealistic wayadd a bluewhale jumping from the water near the icebergshow the sun setting in the horizon
000
Ann Catherine Jose @annjose.com · 03/06/2026
OpenAI ChatGPT equivalents for the same prompts. Much more dramatic than MAI.
create an image of an iceberg in a beautiful ocean with stunning photorealistic wayadd a bluewhale jumping from the water near the icebergshow the sun setting in the horizon
000
Ann Catherine Jose @annjose.com · 03/06/2026
Images created by the new image model announced by Microsoft today - MAI-Image-2.5 Prompts: 1. “create an image of an iceberg in a beautiful ocean with stunning photorealistic way” 2. “add a bluewhale jumping from the water near the iceberg 3. "show the sun setting in the horizon"
create an image of an iceberg in a beautiful ocean with stunning photorealistic wayadd a bluewhale jumping from the water near the icebergshow the sun setting in the horizon
300
Ann Catherine Jose @annjose.com · 03/06/2026
Fantastic! Love that you are not only listening to community feedback, but also following up and sharing the decisions. Customer delight! ❤️ One question - would it be possible to arrange the images in a specific order (like the signal app)? @alexbenzer.com
040
Reposted by Ann Catherine Jose
Cloudflare @cloudflare.social · 29/05/2026
Did you know that civil society organizations are eligible to apply for Cloudflare's Project Galileo? We provide free cybersecurity protection against DDoS and other cyberattacks targeting public interest groups. Learn more: cfl.re/3IgTAch
cfl.re
Project Galileo
Through Project Galileo, Cloudflare provides free cyber security services to organizations supporting the arts, human rights, journalism, and democracy.
0145
Ann Catherine Jose @annjose.com · 27/05/2026
I'm loving the new Copilot GUI app (early access). Very professional, clean, and intuitive UI. Things are exactly where you'd expect. Feels like a well-engineered, production-grade app. Unlike some popular coding agents that still feel hacked together with a disjointed UX. Great job, @github.com 👏💯
130
Reposted by Ann Catherine Jose
Cloudflare @cloudflare.social · 18/05/2026
Cloudflare's security team spent the last few weeks testing Anthropic's Mythos against fifty of our own repositories. What we learned about offensive AI, why faster patching is the wrong reaction, and what the architecture around vulnerabilities has to look like next. cfl.re/4eSYw7W
blog.cloudflare.com
Project Glasswing: what Mythos Preview showed us
In recent weeks, we pointed Mythos and other security-focused LLMs at live code across critical parts of our infrastructure. We share what we observed, the models’ strengths and weaknesses, and what the work around them needs to look like before any of it can scale.
0307
Ann Catherine Jose @annjose.com · 16/05/2026
I am not going to name names, but... I asked my coding AI agent to “improve” my simple parsing code that extract a folder path from inline markdown text. It responded by adding FOUR new dependencies just for that job. Me: "Bro!That’s too much." Agent: "Fair pushback." Me: (facepalm) Revert it! 😡
110
Ann Catherine Jose @annjose.com · 12/05/2026
A good thread on how to use pull_request_target safely bsky.app/profile/4308...
010
Reposted by Ann Catherine Jose
James @43081j.com · 12/05/2026
most of these github actions driven breaches are because of pull_request_target. here's some tips of what to look for when reviewing your own workflows. worth noting - it is safe, and necessary, when used correctly 🧵
410922
Ann Catherine Jose @annjose.com · 09/05/2026
Btw, for some reason all the agents (pi, claude) were giving only the summary or analysis of the page, and not the full markdown as is. So I had to tweak the SKILL.md : Updated Description and added a section Behavior to specify default behavior to return the markdown and summarize when asked.
Updated SKILL.md - with modified Description and new section Behavior

---
name: defuddle
description: Extract clean markdown content from web pages using Defuddle CLI, removing clutter and navigation to save tokens. Use instead of WebFetch when the user provides a URL to extract as clean markdown, for online documentation, articles, blog posts, or any standard web page. By default, return full markdown; if the user explicitly requests a summary, summarize the extracted content. Do NOT use for URLs ending in .md — those are already markdown, use WebFetch directly.
---

# Defuddle

Use Defuddle CLI to extract clean readable content from web pages. Prefer over WebFetch for standard web pages — it removes navigation, ads, and clutter, reducing token usage.

If not installed: `npm install -g defuddle`

## Behavior

Default: return the extracted markdown **as-is** (no summary, no analysis).

If the user explicitly asks to summarize (e.g., “summarize”, “tl;dr”, “key points”),
then provide a summary instead of raw markdown.

If the user asks for both (e.g., “extract and summarize”), return:
1) a summary first, then
2) the full extracted markdown.
000
Ann Catherine Jose @annjose.com · 09/05/2026
Awesome! Thank you. I added the skill to pi using this command. Worked great. Instructions are similar to OpenCode, but there is an intermediate folder `agent` inside the .pi folder ``` sh git clone github.com/kepano/obsid... ~/.pi/agent/skills/obsidian-skills ```
100
Ann Catherine Jose @annjose.com · 09/05/2026
Fantastic idea. I have tried local models with MLX, llama.cpp or ollama many times, but could never get it to work well and consistently. It's great to see d4s.c trying to break this cycle and focus on one model, one serving path, one environment and make it flawless and fast there. Local open FTW.
010
Ann Catherine Jose @annjose.com · 09/05/2026
Found something useful today. The pi agent couldn’t read this webpage, so I used the Obsidian Web Clipper to copy it as Markdown, and pasted it into the chat. Worked immediately. pi was happy! Small win, but a sweet one.🎉 Wonder if this can be built as a skill with Obsidian CLI - @stephango.com
Brave browser showing the Anthropic article Teaching Claude Why with Obsidian Web Clipper on the side with the 'Copy to Clipboard' button highlighted in yellow.The pi coding agent running in terminal with the response 
 Yes — this is a great way to share it, and it looks substantially complete.

 Minor notes:
 - It appears to include all core sections, figures, discussion, and footnotes.
 - One formatting quirk: the “four main lessons” list is split as 1,2,3 + bold paragraph instead of numbered 4 (content is there, numbering just inconsistent).
 - If you want maximum fidelity, also include the exact figure captions/alt text (you already included most).

 So yes, this is more than enough for me to critique your understanding accurately. Send your summary when ready.
130
Ann Catherine Jose @annjose.com · 08/05/2026
I finally completed the migration of my personal website: annjose.com It started as a Hugo to Astro migration and became a full redesign and a real agentic coding project. 11 years of blog posts, images, comments, tags, search, OG images, deployed to Cloudflare. annjose.com/blog/hugo-to...
annjose.com
Migrating my website from Hugo to Astro | Ann Catherine Jose
Migrating this site from Hugo to Astro on Cloudflare, a real-world agentic coding project
000