Sign in

Birgitta B

@birgitta410.bsky.social
679 followers 95 following 45 posts

Software Dev at Thoughtworks, Berlin and elsewhere | birgitta.info

PostsRepliesMedia
Birgitta B @birgitta410.bsky.social · 11/08/2026
Does telling a coding agent to do TDD by itself make a difference? Or is it one of the rare examples where what's good for the human is irrelevant or bad for an agent? My thoughts and hypotheses, including observations from running a few batches with & without TDD: martinfowler.com/articles/exp...
martinfowler.com
TDD inside the agent loop - theater or actual value?
Notes from my Thoughtworks colleagues on AI-assisted software delivery
5167
Birgitta B @birgitta410.bsky.social · 03/08/2026
I actually agree that it should all be standalone terms - I wish we could scratch "spec-driven development" (and harness engineering for that matter, and lots of other terms that emerged in the heat of the AI speed) and start with fresh words!
000
Birgitta B @birgitta410.bsky.social · 27/07/2026
I joined Priyanka Raghavan on Software Engineering Radio for a conversation about harness engineering for coding agents, and the evolving trust model for AI-generated code se-radio.net/2026/07/se-r...
se-radio.net
SE Radio 730: Birgitta Boeckeler on Harness Engineering for AI Agents – Software Engineering Radio
240
Reposted by Birgitta B
Ben Winters @benwinters.bsky.social · 16/07/2026
This is a great and normal idea that will end well
316849
Birgitta B @birgitta410.bsky.social · 13/07/2026
I added an appendix to my article on "Maintainability sensors for coding agents", about the little sidecar application I built to run all the computational sensors I had next to the agent, and give both human and agent the ability to monitor their state martinfowler.com/articles/sen...
Preview of 3 figures from the article: A high level overview of the sensors sidecar orchestrating the sensors and parsing their outputs; a screenshot of the human view of sensors, listed in a table with their status and trend since last snapshot; and the start of the agent's view of the sensors, in text form.
062
Birgitta B @birgitta410.bsky.social · 09/07/2026
I gave "local models" another go over the past few weeks, to find out how viable it is to use a model for agentic coding on a typical high end developer machine (M3/M5 with 48/64GB RAM in my case), and how ready this setup feels for "plug and play" broader adoption martinfowler.com/articles/exp...
Diagram showing the factors described in the linked post and how they affect the speed of response and quality of outcomes of a model
011
Reposted by Birgitta B
Martin Fowler @martinfowler.com · 07/07/2026
NEW POST @birgitta410.bsky.social recently spent some time trying out running local LLMs for some programming tasks. In this memo she outlines the factors that influence how viable they are for the job. martinfowler.com/articles/exp...
martinfowler.com
Viability of local models for coding
Notes from my Thoughtworks colleagues on AI-assisted software delivery
1244
Birgitta B @birgitta410.bsky.social · 05/07/2026
„Learn AI so you can complain about AI better“
040
Reposted by Birgitta B
Charity Majors @charity.wtf · 24/06/2026
as promised, here is my post on AI and ethics. I think unilateral disarmament in the face of powerful new tools is neither wise nor effective. I think those of us in tech have a particular responsibility to engage, participate and build better solutions. charitydotwtf.substack.com/p/make-ai-bo...
charitydotwtf.substack.com
Make AI Boring Again
On the ethics of engagement, the problem with purity politics, and a world worth fighting for
97325
Birgitta B @birgitta410.bsky.social · 27/05/2026
Third update to my article on "Maintainability sensors for coding agents": The test suite as a regression sensor. Assuming the tests are testing the right things (!), how complete are they? Are coverage metrics enough? (Spoiler alert: They are not) martinfowler.com/articles/sen...
martinfowler.com
Maintainability sensors for coding agents
A practical walkthrough of computational sensors on the path to production, with a deep dive on ESLint and static analysis as feedback for coding agents.
172
Birgitta B @birgitta410.bsky.social · 20/05/2026
I started an article on providing sensors about the code's maintainability to coding agents. Part 1 and 2 out so far, about static code analysis, coupling, and modularity review martinfowler.com/articles/sen...
martinfowler.com
042
Reposted by Birgitta B
Emily M. Bender @emilymbender.bsky.social · 19/05/2026
Wow some terrible reporting about Google's latest horrible ideas about how to distort information access in the name of "convenience" (or something): techcrunch.com/2026/05/19/g... A short thread 🧵>>
techcrunch.com
Google Search as you know it is over | TechCrunch
Google is transforming Search from a list of links into an AI-powered experience filled with conversational answers, autonomous agents, and interactive interfaces — a shift that could further reduce t...
17376159
Birgitta B @birgitta410.bsky.social · 27/04/2026
I see lots of AI coding tooling being built by very experienced developers, seemingly assuming themselves as the users. Shouldn't we be user researching the heck out of less experienced developers and build with them in mind? How would they learn, design, navigate a codebase, review?
140
Birgitta B @birgitta410.bsky.social · 27/04/2026
I recorded a conversation with Chris Ford about learnings from using computational sensors with a coding agent on an application, in particular static code analysis and test quality www.youtube.com/watch?v=uLWO...
youtube.com
Harness engineering beyond skills: Using sensors to keep your coding agent in check
YouTube video by Thoughtworks
141
Reposted by Birgitta B
Betty C. Jung @bettycjung.bsky.social · 19/04/2026
Sam Altman’s Creepy Eyeball-Scanning Company Gets in Bed With Zoom and Tinder Will your boss require an eyeball scan the next time you need to jump on Zoom? gizmodo.com/sam-altmans-...
gizmodo.com
Sam Altman's Creepy Eyeball-Scanning Company Gets in Bed With Zoom and Tinder
Will your boss require an eyeball scan the next time you need to jump on Zoom?
23322
Birgitta B @birgitta410.bsky.social · 02/04/2026
New article where I offer definitions & a mental model how to think about harness engineering as coding agent users. Building blocks at our disposal, dimensions and goals to consider; emerging practices, open questions; and of course, what role do humans play martinfowler.com/articles/har...
martinfowler.com
Harness engineering for coding agent users
A mental model for building trust in coding agents through feedforward guides, feedback sensors, and iterative harness engineering.
1136
Birgitta B @birgitta410.bsky.social · 28/03/2026
„But if you care about the internet […] these two verdicts should scare the hell out of you. Because the legal theories […] will be weaponized against everyone. […] Meta can afford that. Google can afford that. You know who can’t? Basically everyone else who runs a platform where users post things“
010
Birgitta B @birgitta410.bsky.social · 17/02/2026
I enjoyed reading an OpenAI team's report on building a "harness" for a no-manual-coding-allowed codebase. Refreshing to read concrete ideas about where the rigor might go, rather than just hoping “better models” will magically solve maintainability martinfowler.com/articles/exp...
martinfowler.com
Harness Engineering
Notes from my Thoughtworks colleagues on AI-assisted software delivery
242
Birgitta B @birgitta410.bsky.social · 06/02/2026
If you’re confused by the growing number of context configuration options for coding agents (commands, skills, rules, …), then my write up might be helpful to you martinfowler.com/articles/exp...
martinfowler.com
Context Engineering for Coding Agents
Notes from my Thoughtworks colleagues on AI-assisted software delivery
251
Reposted by Birgitta B
Stephen Turner @stephenturner.us · 28/12/2025
I wrote a short essay on what I'm calling the AI Attribution Error. doi.org/10.59350/c3g...
811331
Birgitta B @birgitta410.bsky.social · 17/10/2025
„…the tech industry should stop focusing so heavily on these one-size-fits-all tools, and instead concentrate on narrow, specialized A.I. tools engineered for particular problems. Because, frankly, they’re often more effective.“ www.nytimes.com/2025/10/16/o...
nytimes.com
Opinion | Silicon Valley Is Investing in the Wrong A.I.
0172
Birgitta B @birgitta410.bsky.social · 15/10/2025
I tried to make sense of "spec-driven development" by looking at 3 tools: Amazon's Kiro, GitHub's spec-kit, and the Tessl Framework martinfowler.com/articles/exp...
martinfowler.com
Understanding Spec-Driven-Development: Kiro, spec-kit, and Tessl
Notes from my Thoughtworks colleagues on AI-assisted software delivery
1279
Reposted by Birgitta B
Meredith Whittaker @meredithmeredith.bsky.social · 06/10/2025
📣 Germany's close to reversing its opposition to mass surveillance & private message scanning, & backing the Chat Control bill. This could end private comms-& Signal-in the EU. Time's short and they're counting on obscurity: please let German politicians know how horrifying their reversal would be.
3022381606
Birgitta B @birgitta410.bsky.social · 25/09/2025
One of the challenges with service templates is that once a team instantiated a service with a template, it’s tedious to feed template updates back to those services. I wonder if anchoring AI agents to a template or reference application could help make that easier? martinfowler.com/articles/exp...
Overview diagram showing a coding agent connected to a reference application via an MCP server. The agent can find latest changes, get the latest commit from the reference application.
1101
Birgitta B @birgitta410.bsky.social · 23/09/2025
To vibe or not to vibe? I wrote a new memo about the constant little risk assessments I make during AI-assisted coding, thinking about probability and impact if AI gets it wrong, and if I will be able to detect that. martinfowler.com/articles/exp...
An illustration showing the two extreme cases of the 3 dimensions: Low probability + low impact + high detectability is the perfect case for vibe coding; High probability + high impact + low detectability is the case that needs the most human scrutiny
030
Reposted by Birgitta B
Gergely Orosz @gergely.pragmaticengineer.com · 04/09/2025
Wild to me that a CEO sets goals about outcomes that have nothing to do with the business (are customers more satisfied? Is the product more reliable? Etc.) Setting the goal of what % of code should be AI-generated is as useful as setting the goal of how many lines of code devs should write per day
1427530
Birgitta B @birgitta410.bsky.social · 05/08/2025
We recently ran an experiment to explore how far GenAI can currently be pushed toward autonomously developing high-quality, up-to-date software without human intervention, and gather observations about where it breaks down. martinfowler.com/articles/pus...
martinfowler.com
How far can we push AI autonomy in code generation?
An experiment to test the limits of autonomous code generation by LLMs
12710
Reposted by Birgitta B
Schöpflin Stiftung @schoepflinstiftung.bsky.social · 29/07/2025
Gemeinsam mit Publix und @algorithmwatch.org laden wir ein zum Gespräch mit @meredithmeredith.bsky.social, Präsidentin des Messengers Signal, über die Frage, wie wir Technologie wieder stärker an menschlichen Bedürfnissen ausrichten können – vor allem beim Schutz unserer Privatsphäre. ➡️ t.ly/lIzA3
Das Individuum in der Maschine: Meredith Whittaker über die Rückgewinnung der Privatsphäre im Zeitalter der KI
0529
Birgitta B @birgitta410.bsky.social · 21/07/2025
I totally get it, I’m also tired of much of the public discourse. But here is one more argument for experienced devs like the author: If we want to guide & teach the „juniors“, we have a responsibility to know first hand what works and what doesn’t, because they are using LLMs if we like it or not.
030
Birgitta B @birgitta410.bsky.social · 09/07/2025
I've seen a surge of discussions recently about large AI-generated change sets that are impossible to review by humans, paired with speculation if we still need to care about the code in the future. I expect to continue to care, especially if I'm on call for it martinfowler.com/articles/exp...
martinfowler.com
I still care about the code
Notes from my Thoughtworks colleagues on AI-assisted software delivery
043
Birgitta B @birgitta410.bsky.social · 24/06/2025
I wrote a guest post on The Pragmatic Engineer newsletter.pragmaticengineer.com/p/two-years-... rounding up 2 years of using AI coding assistants - how they evolved; ways of working; impact I see on speed, quality and team flow; and some thoughts on the future
newsletter.pragmaticengineer.com
Learnings from two years of using AI tools for software engineering
How to think about today’s AI tools, approaches that work well, and concerns about using them for development. Guest post by Birgitta Böckeler, Distinguished Engineer at Thoughtworks
1165
Reposted by Birgitta B
Randall Munroe @xkcd.com · 23/06/2025
Tukey xkcd.com/3104/
Comic. [block quote] “Far better an approximate answer to the *right* question, which is often vague, than an *exact* answer to the wrong question, which can always be made precise.” -John W. Tukey, The Future of Data Analysis (1962) [caption] Happy Approximate Birthday to John Tukey, author of my favorite statistics quote, who was born 110.000 years ago sometime this week.
192892427
Reposted by Birgitta B
Pete Hodgson @thepete.net · 05/06/2025
I blurghed some thoughts blog.thepete.net/blog/2025/06... Starting to realize what makes me nervous about Devin et al. It becomes a new incantation of the perennial problem of letting a stakeholder get work done by shoulder-tapping their favorite young engineer
blog.thepete.net
Your CEO Should Not Be Slacking Your Coding Agent
Autonomous coding agents may seem magical, but they generate drag on your team. Allowing stakeholders to directly ask an AI to do work leads to the same old disruptions as directly requesting work fro...
0114
Birgitta B @birgitta410.bsky.social · 04/06/2025
In the past few weeks, a new surge of "background" coding agents came out. I wrote down an example of using OpenAI's Codex, this will hopefully help you understand better what they do under the hood, and which agent category they fall into. martinfowler.com/articles/exp...
martinfowler.com
3131
Birgitta B @birgitta410.bsky.social · 02/06/2025
When I code with an AI agent, I revert to the last comfortable checkpoint as soon as I feel like losing control. I wonder what that will be like for the delivery managers & POs of the future, what will THEY do when things around them change so fast and erratically that they feel like losing control?
000
Birgitta B @birgitta410.bsky.social · 01/06/2025
Good scope management is the linchpin of agile software delivery. Generative AI is like a scope chaos monkey, generating more code, more story details, more requirements than necessary
060
Reposted by Birgitta B
Martin Fowler @martinfowler.com · 25/03/2025
NEW POST To work effectively with agentic coding assistants, Birgitta Böckeler found she needs to intervene, correct and steer all the time. She describes examples of these interventions indicating the skills we need to correct the tools' missteps martinfowler.com/articles/exp...
915234
Birgitta B @birgitta410.bsky.social · 20/02/2025
Nice write-up by @harper.lol on his AI-assisted coding workflow. I personally prefer in-IDE tools, but the concepts are reusable. ❤️ this:"I really want someone to solve this problem in a way that makes coding with an LLM a multiplayer game. Not a solo hacker experience." harper.blog/2025/02/16/m...
harper.blog
My LLM codegen workflow atm
A detailed walkthrough of my current workflow for using LLms to build software, from brainstorming through planning and execution.
120
Birgitta B @birgitta410.bsky.social · 18/02/2025
I have thoughts and open questions about the role reasoning models might play or not play in coding assistance. A lot of stake is put into how reasoning models are a step change in coding assistance, but I don't see it - yet? martinfowler.com/articles/exp...
martinfowler.com
Exploring Generative AI
Notes from my Thoughtworks colleagues on AI-assisted software delivery
193
Birgitta B @birgitta410.bsky.social · 19/11/2024
New GenAI memo: I wrote down my thoughts and observations about the multi-file editing features that are currently being released in lots of coding assistants: martinfowler.com/articles/exp...
martinfowler.com
Exploring Generative AI
Notes from my Thoughtworks colleagues on AI-assisted software delivery
182
Birgitta B @birgitta410.bsky.social · 26/08/2024
New "GenAI memo": In this one I explore the potential of AI assistance for tech stack migrations. I describe building an agent that changes the testing framework used in a test. As a side effect you can also gain a better understanding of how AI agents work: martinfowler.com/articles/exp...
martinfowler.com
Exploring Generative AI
Notes from my Thoughtworks colleagues on AI-assisted software delivery
000
Birgitta B @birgitta410.bsky.social · 26/08/2024
My talk at GOTO Amsterdam is now live on YouTube: "AI Assistance Beyond Code: What Do We Need to Make it Work?" youtu.be/8jwiABwGC6c?...
youtu.be
AI Assistance Beyond Code: What Do We Need to Make it Work? • Birgitta Böckeler • GOTO 2024
YouTube video by GOTO Conferences
011
Birgitta B @birgitta410.bsky.social · 15/08/2024
In my newest "GenAI memo", I explore how today's AI tools can assist with onboarding to existing, potentially messy codebases. I do this by trying to understand and solve an issue in a real life codebase: martinfowler.com/articles/exp...
martinfowler.com
Exploring Generative AI
Notes from my Thoughtworks colleagues on AI-assisted software delivery
050