disreGUARD @disreguard.com · 10/02/2026There is no silver bullet for prompt injection, but `sig` addresses one of the most tricky challenges: LLMs are dealing with a wall of text. Distinctions between trusted instructions and untrusted text are flimsy at best. By involving the model, we can add texture to make trust boundaries clearer. 000
disreGUARD @disreguard.com · 10/02/2026`sig` is a simple tool. The concept works like this: - sign genuine system and user instructions - require mutations to system prompts to be signed - give model a tool to verify instructions are genuine - gate critical tool usage on the model calling verify() 100
disreGUARD @disreguard.com · 09/02/2026Can we give an agent a tool to check which instructions it should actually follow? And if we do, can we make sure it uses it? Yes, and yes. disreguard.com/blog/posts/s...disreguard.comsig: instruction signing for prompt injection defenseWe can create a clear trust boundary by signing instructions and giving models a tool to participate in making secure choices 131
Reposted by disreGUARDAdam Avenir @adamavenir.com · 09/02/2026For the last year, I've been heads down working on building infrastructure to better defend against prompt injection. @disreguard.com is a security research lab focused on tools and methods for ergonomic approaches for the rigorous defense-in-depth needed for agents. disreguard.com/blog/posts/p...disreguard.comInjection is inevitable. Disaster is optional.Prompt injection is an infrastructure problem. We can't prevent it, but we can massively reduce the risk and impact. 002