RLMs & Program Agents
I'm experimenting with this idea, Program Agents, inside of DeepSeek Harness (DSH)
They're like RLMs, except without the LLM. Agents are writing such large blocks of code, what if they just never exited?
github.com/tkellogg/dsh...