EvoSafeHarness: Evolving Model- and Domain-Specific Harnesses for Securing Agents
This paper develops a framework to create customizable safety harnesses for AI agents, which can adapt to different models and domains to prevent harmful behavior. Practitioners may care about this research to improve the safety of AI systems in real-world applications.