The LLM can't tell the difference between your messages and theirs, however many times you say "No mistakes"
Perhaps it is time to implement the lessons learned from Perl's taint checking[1], but this time for AI agent harnesses instead[2].
[1] https://en.wikipedia.org/wiki/Taint_checking
[2] https://arxiv.org/html/2607.03423v1
Perhaps it is time to implement the lessons learned from Perl's taint checking[1], but this time for AI agent harnesses instead[2].
[1] https://en.wikipedia.org/wiki/Taint_checking
[2] https://arxiv.org/html/2607.03423v1