safety
fact
neutral
RedEvoAgent outperforms fixed and agentic baselines on multiple benchmarks, target models, and execution harnesses while improving tool efficiency and transferring across attacker models and target execution harnesses
Experiments on multiple benchmarks, target models, and target execution harnesses show that RedEvoAgent outperforms fixed and agentic baselines, improves tool efficiency, and transfers across attacker models and target execution harnesses.
Artificial Intelligence30 Aug 2026